bioRxiv · 10.1101/2021.11.29.470431
PhyloHerb: A phylogenomic pipeline for processing genome skimming data for plants
Abstract
O_LIPremise of the study: The application of high throughput sequencing, especially to herbarium specimens, is greatly accelerating biodiversity research. Among various techniques, low coverage Illumina sequencing of total genomic DNA (genome skimming) can simultaneously recover the plastid, mitochondrial, and nuclear ribosomal regions across hundreds of species. Here, we introduce PhyloHerb -- a bioinformatic pipeline to efficiently and effectively assemble phylogenomic datasets derived from genome skimming. C_LIO_LIMethods and Results: PhyloHerb uses either a built-in database or user-specified references to extract orthologous sequences using BLAST search. It outputs FASTA files and offers a suite of utility functions to assist with alignment, data partitioning, concatenation, and phylogeny inference. The program is freely available at https://github.com/lmcai/PhyloHerb/. C_LIO_LIConclusions: Using published data from Clusiaceae, we demonstrated that PhyloHerb can accurately identify genes using highly fragmented assemblies derived from sequencing older herbarium specimens. Our approach is effective at all taxonomic depths and is scalable to thousands of species. C_LI
Source connections
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Cai, L., Zhang, H., DAVIS, C. C.. 2021-12-01. PhyloHerb: A phylogenomic pipeline for processing genome skimming data for plants. https://doi.org/10.1101/2021.11.29.470431
Cite the original work for its findings. Save a collection to share your selection of sources.