Search bioRxivSearch

Biology subjects

Shetty, A. C.

Publications and source records attributed to Shetty, A. C..

2 recordsLinked to original sources

FADU: A Feature Counting Tool for Prokaryotic RNA-Seq Analysis

MotivationThe major algorithms for quantifying transcriptomics data for differential gene expression analysis were designed for analyzing data from human or human-like genomes, specifically those with single gene transcripts and distinct transcriptional boundaries that extend beyond the coding sequence (CDS) as identified through expressed sequence tags (ESTs) or EST-like sequence data. Some eukaryotic genomes and all, or nearly all, bacterial genomes require alternate methods of quantification since they lack annotation of transcriptional boundaries with EST or EST-like data, have overlapping transcriptional boundaries, and/or have polycistronic transcripts.\n\nResultsAn algorithm was developed and tested that better quantifies transcriptomics data for differential gene expression analysis in organisms with overlapping transcriptional units and polycistronic transcripts. Using data from standard libraries originating from Escherichia coli and Ehrlichia chaffeensis, and strand-specific libraries from the Wolbachia endosymbiont wBm, FADU can derive counts for genes that are missed by HTSeq and featurecounts. Using the default parameters with the E. coli data, FADU can detect transcription of 51 more genes than HTSeq in union mode and 21 genes more than featurecounts, with 42 and 18 of these features being <300 bp, respectively. Due to its ability to derive counts for otherwise unrepresented genes without overstating their abundance, we believe FADU to be an improved tool for quantifying transcripts in prokaryotic systems for RNA-Seq analyses.\n\nAvailability and implementationFADU is available at https://github.com/adkinsrs/FADU. FADU was implemented using Python3 and requires the PySAM module (version 0.12.0.1 or later).\n\nContactjdhotopp@som.umaryland.edu

genomics

The Evolutionary Genomic Dynamics of Peruvians Before, During, and After the Inca Empire

Native Americans from the Amazon, Andes, and coast regions of South America have a rich cultural heritage, but have been genetically understudied leading to gaps in our knowledge of their genomic architecture and demographic history. Here, we sequenced 150 high-coverage and genotyped 130 genomes from Native American and mestizo populations in Peru. A majority of our samples possess greater than 90% Native American ancestry and demographic modeling reveals, consistent with a rapid peopling model of the Americas, that most of Peru was peopled approximately 12,000 years ago. While the Native American populations possessed distinct ancestral divisions, the mestizo groups were admixtures of multiple Native American communities which occurred before and during the Inca Empire. The mestizo communities also show Spanish introgression only after Peruvian Independence. Thus, we present a detailed model of the evolutionary dynamics which impacted the genomes of modern day Peruvians.

genomics