ContaVect
Revision as of 20:40, 10 June 2022 by Israel.herrera (talk | contribs)
Description
Contavect is a python2.7 object oriented script, developed to quantify and characterize DNA contaminants from gene therapy vector production after NGS sequencing. This automated pipeline can however be used for wider purpose requiring to identify map NGS datasets consisting of a mix of DNA sequences on multiple references. It combine several features such as reference homologies masking, fastq filtering/adapter trimming, short read alignments, SAM file splitting and generating human readable output.
Environment Modules
Run module spider contavect
to find out what environment modules are available for this application.
System Variables
- HPC_CONTAVECT_DIR - installation directory
Additional Information
The sample config file can be copied from the $HPC_CONTAVECT_CONF
$ cp $HPC_CONTAVECT_CONF .