bioRxiv · 10.1101/2020.12.11.422022
An Extensive Sequence Dataset of Gold-Standard Samples for Benchmarking and Development
Abstract
Accurate standards and extensive development datasets are the foundation of technical progress. To facilitate benchmarking and development, we sequence 9 samples, covering the Genome in a Bottle truth sets on multiple instruments (NovaSeq, HiSeqX, HiSeq4000, PacBio Sequel II System) and sample preparations (PCR-Free, PCR-Positive) for both whole genome and multiple exome kits. We benchmark pipelines, quantifying strengths and limitations for sequencing and analysis methods. We identify variability within and between instruments, preparation methods, and analytical pipelines, across various sequencing depths. We discuss the relevance of this variability to downstream analyses, and strategies to reduce variability.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Baid, G., Nattestad, M., Kolesnikov, A., Goel, S., Yang, H., Chang, P.-C., Carroll, A.. 2020-12-11. An Extensive Sequence Dataset of Gold-Standard Samples for Benchmarking and Development. https://doi.org/10.1101/2020.12.11.422022
Cite the original work for its findings. Save a collection to share your selection of sources.