bioRxiv · 10.1101/096560
Inference of Multiple-wave Admixtures by Length Distribution of Ancestral Tracks
Abstract
The ancestral tracks in admixed genomes are of valuable information for population history inference. A few methods have been developed to infer admixture history based on ancestral tracks. Nonetheless, these methods suffered the same flaw that only population admixture history under some specific models can be inferred. In addition, the inference of history might be biased or even unreliable if the specific model is deviated from the real situation. To address this problem, we firstly proposed a general discrete admixture model to describe the admixture history with multiple ancestral populations and multiple-wave admixtures. We next deduced the length distribution of ancestral tracks under the general discrete admixture model. We further developed a new method, MultiWaver, to explore the multiple-wave admixture histories. Our method could automatically determine an optimal admixture model based on the length distribution of ancestral tracks, and estimate the corresponding parameters under this optimal model. Specifically, we used a likelihood ratio test (LRT) to determine the number of admixture waves, and implemented an expectation??maximization (EM) algorithm to estimate parameters. We used simulation studies to validate the reliability and effectiveness of our method. Finally, good performance was observed when our method was applied to real datasets of African Americans, Mexicans, Uyghurs, and Hazaras.
Source connections
Explore related subjects
Keep this discovery
Ni, X., Yang, X., Yuan, K., Feng, Q., Guo, W., Ma, Z., Xu, S.. 2016-12-23. Inference of Multiple-wave Admixtures by Length Distribution of Ancestral Tracks. https://doi.org/10.1101/096560
Cite the original work for its findings. Save a collection to share your selection of sources.