bioRxiv · 10.1101/2022.02.15.480535
StrainPanDA: linked reconstruction of strain composition and gene content profiles via pangenome-based decomposition of metagenomic data
Abstract
BackgroundMicrobial strains of variable functional capacities co-exist in microbiomes. Current bioinformatics methods of strain analysis cannot provide the direct linkage between strain composition and their gene contents from metagenomic data. MethodsHere we present StrainPanDA (Strain-level Pangenome Decomposition Analysis), a novel method that uses the pangenome coverage profile of multiple metagenomic samples to simultaneously reconstruct the composition and gene content variation of co-existing strains in microbial communities. ResultsWe systematically validate the accuracy and robustness of StrainPanDA using synthetic datasets. To demonstrate the power of gene-centric strain profiling, we then apply StrainPanDA to analyze the gut microbiome samples of infants, as well as patients treated with fecal microbiota transplantation. We show that the linked reconstruction of strain composition and gene content profiles is critical for understanding the relationship between microbial adaptation and strain-specific functions (e.g., nutrient utilization, pathogenicity). ConclusionsStrainPanDA can be applied to metagenomic datasets to detect association between molecular functions and microbial/host phenotypes to formulate testable hypotheses and gain novel biological insights at the strain or subspecies level.
Source connections
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Hu, H., Tan, Y., Li, C., Chen, J., Kou, Y., Xu, Z., Liu, Y.-Y., Dai, L.. 2022-02-19. StrainPanDA: linked reconstruction of strain composition and gene content profiles via pangenome-based decomposition of metagenomic data. https://doi.org/10.1101/2022.02.15.480535
Cite the original work for its findings. Save a collection to share your selection of sources.