ContaVect
Revision as of 16:12, 23 November 2016 by Maxprok (talk | contribs) (Created page with "Category:SoftwareCategory:BiologyCategory:Phylogenetics {|<!--CONFIGURATION: REQUIRED--> |{{#vardefine:app|contavect}} |{{#vardefine:url|https://github.com/a-slide...")
Description
Contavect is a python2.7 object oriented script, developed to quantify and characterize DNA contaminants from gene therapy vector production after NGS sequencing. This automated pipeline can however be used for wider purpose requiring to identify map NGS datasets consisting of a mix of DNA sequences on multiple references. It combine several features such as reference homologies masking, fastq filtering/adapter trimming, short read alignments, SAM file splitting and generating human readable output.
Required Modules
System Variables
- HPC_{{#uppercase:contavect}}_DIR - installation directory