bioRxiv · 10.1101/2020.12.15.422742
Ranked Choice Voting for Representative Transcripts with TRaCE
Abstract
SummaryGenome sequencing projects annotate protein-coding gene models with multiple transcripts, aiming to represent all of the available transcript evidence. However, downstream analyses often operate on only one representative transcript per gene locus, sometimes known as the canonical transcript. To choose canonical transcripts, TRaCE (Transcript Ranking and Canonical Election) holds an election in which a set of RNA-seq samples rank transcripts by annotation edit distance. These sample-specific votes are tallied along with other criteria such as protein length and InterPro domain coverage. The winner is selected as the canonical transcript, but the election proceeds through multiple rounds of voting to order all the transcripts by relevance. Based on the set of expression data provided, TRaCE can identify the most common isoforms from a broad expression atlas or prioritize alternative transcripts expressed in specific contexts. Availability and ImplementationTranscript ranking code can be found on GitHub at {{https://github.com/warelab/TRaCE}} Contactolson@cshl.edu, ware@cshl.edu Supplementary informationAdditional data are available in the github repository.
Source connections
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Olson, A. J., Ware, D.. 2020-12-16. Ranked Choice Voting for Representative Transcripts with TRaCE. https://doi.org/10.1101/2020.12.15.422742
Cite the original work for its findings. Save a collection to share your selection of sources.