bioRxiv · 10.1101/2022.04.15.488440
Vaeda computationally annotates doublets in single-cell RNA sequencing data
Abstract
MotivationSingle-cell RNA sequencing (scRNA-seq) continues to expand our knowledge by facilitating the study of transcriptional heterogeneity at the level of single cells. Despite this technologys utility and success in biomedical research, technical artifacts are present in scRNA-seq data. Doublets/multiplets are a type of artifact that occurs when two or more cells are tagged by the same barcode, and therefore they appear as a single cell. Because this introduces non-existent transcriptional profiles, doublets can bias and mislead downstream analysis. To address this limitation computational methods to annotate and remove doublets form scRNA-seq datasets are needed. ResultsWe introduce vaeda, a new approach for computational annotation of doublets in scRNA-seq data. Vaeda integrates a variational auto-encoder and Positive-Unlabeled learning to produce doublet scores and binary doublet calls. We apply vaeda, along with seven existing doublet annotation methods, to sixteen benchmark datasets and find that vaeda performs competitively in terms of doublet scores and doublet calls. Notably, vaeda outperforms other python-based methods for doublet annotation. All together, vaeda is a robust and competitive method for scRNA-seq doublet annotation and may be of particular interest in the context of python-based workflows. AvailabilityVaeda is available at https://github.com/kostkalab/vaeda Contactkostka@pitt.edukostka@pitt.edu
Source connections
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Schriever, H., Kostka, D.. 2022-04-15. Vaeda computationally annotates doublets in single-cell RNA sequencing data. https://doi.org/10.1101/2022.04.15.488440
Cite the original work for its findings. Save a collection to share your selection of sources.