Search bioRxivSearch

Biology subjects

van der Lee, S. J.

Publications and source records attributed to van der Lee, S. J..

5 recordsLinked to original sources

Quality Control and Integration of Genotypes from Two Calling Pipelines for Whole Genome Sequence Data in the Alzheimer’s Disease Sequencing Project

The Alzheimers Disease Sequencing Project (ADSP) performed whole genome sequencing (WGS) of 584 subjects from 111 multiplex families at three sequencing centers. Genotype calling of single nucleotide variants (SNVs) and insertion-deletion variants (indels) was performed centrally using GATK-HaplotypeCaller and Atlas V2. The ADSP Quality Control (QC) Working Group applied QC protocols to project-level variant call format files (VCFs) from each pipeline, and developed and implemented a novel protocol, termed \"consensus calling,\" to combine genotype calls from both pipelines into a single high-quality set. QC was applied to autosomal bi-allelic SNVs and indels, and included pipeline-recommended QC filters, variant-level QC, and sample-level QC. Low-quality variants or genotypes were excluded, and sample outliers were noted. Quality was assessed by examining Mendelian inconsistencies (MIs) among 67 parent-offspring pairs, and MIs were used to establish additional genotype-specific filters for GATK calls. After QC, 578 subjects remained. Pipeline-specific QC excluded ~12.0% of GATK and 14.5% of Atlas SNVs. Between pipelines, ~91% of SNV genotypes across all QCed variants were concordant; 4.23% and 4.56% of genotypes were exclusive to Atlas or GATK, respectively; the remaining ~0.01% of discordant genotypes were excluded. For indels, variant-level QC excluded ~36.8% of GATK and 35.3% of Atlas indels. Between pipelines, ~55.6% of indel genotypes were concordant; while 10.3% and 28.3% were exclusive to Atlas or GATK, respectively; and ~0.29% of discordant genotypes were. The final WGS consensus dataset contains 27,896,774 SNVs and 3,133,926 indels and is publicly available.\n\nAbbreviationsAD, Alzheimers disease; QC, Quality Control; LSSAC, Large-Scale Sequencing and Analysis Center; Broad, Broad Institute Genomics Service; Baylor, Baylor College of Medicine Human Genome Sequencing Center; WashU, Washington University-St. Louis McDonnell Genome Institute; WGS, whole genome sequencing; WES, whole exome sequencing; indel, insertion-deletion variants; VCF, variant control format; MI, Mendelian inconsistency; MC, Mendelian consistency; GWAS, genome-wide association study; VR, referent allele read depth; DP, overall read depth; MS, mapping score; GQ, genotype quality score; Ti/Tv, Transition/Transversion; CS, concordance code

genetics

Centenarian Controls Increase Variant Effect-sizes by an average two-fold in an Extreme Case-Extreme Control Analysis of Alzheimer’s Disease

The detection of genetic loci associated with Alzheimers disease (AD) requires large numbers of cases and controls because variant effect-sizes are mostly small. We hypothesized that variant effect-sizes should increase when individuals who represent the extreme ends of a disease spectrum are considered, as their genomes are assumed to be maximally enriched or depleted with disease-associated genetic variants.\n\nWe used 1,073 extensively phenotyped AD cases with relatively young age at onset as extreme cases (66.3{+/-}7.9 years), 1,664 age-matched controls (66.0{+/-}6.5 years) and 255 cognitively healthy centenarians as extreme controls (101.4{+/-}1.3 years). We estimated the effect-size of 29 variants that were previously associated with AD in genome-wide association studies.\n\nComparing extreme AD-cases with centenarian-controls increased the variant effect-size relative to published effect-sizes by on average 1.90-fold (SE=0.29, p=9.0x10-4). The effect-size increase was largest for the rare high-impact TREM2 (R74H) variant (6.5-fold), and significant for variants in/near ECHDC3 (4.6-fold), SLC24A4-RIN3 (4.5-fold), NME8 (3.8-fold), PLCG2 (3.3-fold), APOE-{varepsilon}2 (2.2-fold) and APOE-{varepsilon}4 (2.0-fold). Comparing extreme phenotypes enabled us to replicate the AD association for 10 variants (p<0.05) in relatively small samples. The increase in effect-sizes depended mainly on using centenarians as extreme controls: the average variant effect-size was not increased in a comparison of extreme AD cases and age-matched controls (0.94-fold, p=6.8x10-1), suggesting that on average the tested genetic variants did not explain the extremity of the AD-cases. Concluding, using centenarians as extreme controls in AD case-controls studies boosts the variant effect-size by on average two-fold, allowing the replication of disease-association in relatively small samples.

genetics

Meta-analysis of genetic association with diagnosed Alzheimer’s disease identifies novel risk loci and implicates Abeta, Tau, immunity and lipid processing

Late-onset Alzheimers disease (LOAD, onset age > 60 years) is the most prevalent dementia in the elderly1, and risk is partially driven by genetics2. Many of the loci responsible for this genetic risk were identified by genome-wide association studies (GWAS)3-8. To identify additional LOAD risk loci, the we performed the largest GWAS to date (89,769 individuals), analyzing both common and rare variants. We confirm 20 previous LOAD risk loci and identify four new genome-wide loci (IQCK, ACE, ADAM10, and ADAMTS1). Pathway analysis of these data implicates the immune system and lipid metabolism, and for the first time tau binding proteins and APP metabolism. These findings show that genetic variants affecting APP and A{beta} processing are not only associated with early-onset autosomal dominant AD but also with LOAD. Analysis of AD risk genes and pathways show enrichment for rare variants (P = 1.32 x 10-7) indicating that additional rare variants remain to be identified.

genetics

Genome-wide Association Study Links APOEϵ4 and BACE1 Variants with Plasma Amyloid β Levels

INTRODUCTIONThere is increasing interest in plasma A{beta} as an endophenotype and biomarker of Alzheimers disease (AD). Identifying the genetic determinants of plasma A{beta} levels may elucidate important processes that determine plasma A{beta} measures. METHODSWe included 12,369 non-demented participants derived from eight population-based studies. Imputed genetic data and plasma A{beta}1-40, A{beta}1-42 levels and A{beta}1-42/A{beta}1-40 ratio were used to perform genome-wide association studies, gene-based and pathway analyses. Significant variants and genes were followed-up for the association with PET A{beta} deposition and AD risk. RESULTSSingle-variant analysis identified associations across APOE for A{beta}1-42 and A{beta}1-42/A{beta}1-40 ratio, and BACE1 for A{beta}1-40. Gene-based analysis of A{beta}1-40 additionally identified associations for APP, PSEN2, CCK and ZNF397. There was suggestive interaction between a BACE1 variant and APOE{varepsilon}4 on brain A{beta} deposition. DISCUSSIONIdentification of variants near/in known major A{beta}-processing genes strengthens the relevance of plasma-A{beta} levels both as an endophenotype and a biomarker of AD.

genetics

Genetic Architecture of Subcortical Brain Structures in Over 40,000 Individuals Worldwide

Subcortical brain structures are integral to motion, consciousness, emotions, and learning. We identified common genetic variation related to the volumes of nucleus accumbens, amygdala, brainstem, caudate nucleus, globus pallidus, putamen, and thalamus, using genome-wide association analyses in over 40,000 individuals from CHARGE, ENIGMA and the UK-Biobank. We show that variability in subcortical volumes is heritable, and identify 25 significantly associated loci (20 novel). Annotation of these loci utilizing gene expression, methylation, and neuropathological data identified 62 candidate genes implicated in neurodevelopment, synaptic signaling, axonal transport, apoptosis, and susceptibility to neurological disorders. This set of genes is significantly enriched for Drosophila orthologs associated with neurodevelopmental phenotypes, suggesting evolutionarily conserved mechanisms. Our findings uncover novel biology and potential drug targets underlying brain development and disease.

genetics