bioRxiv · 10.1101/477869
CoCo: RNA-seq Read Assignment Correction for Nested Genes and Multimapped Reads
Abstract
MotivationNext generation sequencing techniques revolutionized the study of RNA expression by permitting whole transcriptome analysis. However, sequencing reads generated from nested and multi-copy genes are often either misassigned or discarded, which greatly reduces both quantification accuracy and gene coverage.\n\nResultsHere we present CoCo, a read assignment pipeline that takes into account the multitude of overlapping and repetitive genes in the transcriptome of higher eukaryotes. CoCo uses a modified annotation file that highlights nested genes and proportionally distributes multimapped reads between repeated sequences. CoCo salvages over 15% of discarded aligned RNA-seq reads and significantly changes the abundance estimates for both coding and non-coding RNA as validated by PCR and bed-graph comparisons.\n\nAvailabilityThe CoCo software is an open source package written in Python and available from http://gitlabscottgroup.med.usherbrooke.ca/scott-group/coco.\n\nContactmichelle.scott@usherbrooke.ca
Source connections
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Deschamps-Francoeur, G., Boivin, V., Abou Elela, S., Scott, M. S.. 2018-11-29. CoCo: RNA-seq Read Assignment Correction for Nested Genes and Multimapped Reads. https://doi.org/10.1101/477869
Cite the original work for its findings. Save a collection to share your selection of sources.