Search bioRxivSearch

SEARCH · Search bioRxiv

Results for “Genetics”

Search indexed bioRxiv preprints in genomics, neuroscience, cell biology and bioinformatics. Read source abstracts and check manuscript versions; preprints are not peer reviewed.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 1,099 records · Page 61Linked to original sources

Structural variants and selective sweep foci contribute to insecticide resistance in the Drosophila melanogaster genetic reference panel

Patterns of nucleotide polymorphism within populations of Drosophila melanogaster suggest that insecticides have been the selective agents driving the strongest recent bouts of positive selection. However, there is a need to explicitly link selective sweep loci to the particular insecticide phenotypes that could plausibly account for the drastic selective responses that are observed in these non-target insects. Here, we screen the Drosophila Genetic Reference Panel with two common insecticides; malathion (an organophosphate) and permethrin (a pyrethroid). Genome wide association studies map survival-on-malathion to two of the largest sweeps in the D. melanogaster genome; Ace and Cyp6g1. Malathion survivorship also correlates with lines which have high levels of Cyp12d1 and Jheh1 and Jheh2 transcript abundance. Permethrin phenotypes map to the largest cluster of P450 genes in the Drosophila genome, however in contrast to a selective sweep driven by insecticide use, the derived state seems to be associated with susceptibility. These results underscore previous findings that highlight the importance of structural variation to insecticide phenotypes: Cyp6g1 exhibits copy number variation and transposable element insertions, Cyp12d1 is tandemly duplicated, the Jheh loci are associated with a Bari1 transposable element insertion, and a Cyp6a17 deletion is associated with susceptibility.

genetics

Analysis of the genetic basis of height in large Jewish nuclear families

Despite intensive study, most genetic factors that contribute to variation in human height remain undiscovered. We conducted a family-based linkage study of height in a unique cohort of very large nuclear families from a founder (Jewish) population. This design allowed for increased power to detect linkage, compared to previous family-based studies. We identified loci that together explain an estimated 6% of the variance in height. We showed that these loci are not tagging known common variants associated with height. Rather, we suggest that the observed signals arise from variants with large effects that are rare globally but elevated in frequency in the Jewish population.

genetics

Molecular genetic analysis of Steroid Resistant Nephrotic Syndrome: Detection of a novel mutation

Background: Nephrotic syndrome is one of the most common kidney diseases in childhood. About 20% of children are steroid-resistant NS (SRNS) which progress to end-stage renal disease (ESRD). More than 53 genes are associated with SRNS which represent the genetic heterogeneity of SRNS. This study was aimed to screen disease causing mutations within NPHS1 and NPHS2 and evaluate new potential variants in other genes.\n\nMethod: In first phase of study, 25 patients with SRNS were analyzed for NPHS1 (exon 2, 26) and all exons of NPHS2 genes by Sanger sequencing. In the second phase, whole exome sequencing was performed on 10 patients with no mutations in NPHS1 and NPHS2.\n\nResult: WES analysis revealed a novel mutation in FAT1 (c.10570C>A; Q3524K). We identified 4 pathogenic mutations, located in exon 4 and 5 of NPHS2 gene in 20% of patients (V180M, P118L, R168C and Leu156Phe). Also our study has contributed to the descriptions of previously known pathogenic mutations across WT1 (R205C) and SMARCAL1 (R764Q) and a novel polymorphism in CRB2.\n\nConclusion: Our study concludes that mutations of exon 4 and 5 NPHS2 gene are common in Iranian and some other ethnic groups. We suggest conducting WES after NPHS2 screening and further comprehensive studies to identify the most common genes in the development of SRNS, which might help in Clinical impact on management in patients with SRNS.\n\nDetection of a novel mutation in SRNS

genetics

Genetic diversity in two Plasmodium vivax protein ligands for reticulocyte invasion

The interaction between Plasmodium vivax Duffy binding protein (PvDBP) and Duffy antigen receptor for chemokines (DARC) has been described as critical for the invasion of human reticulocytes, although increasing reports of P. vivax infections in Duffy-negative individuals questions its unique role. To investigate the genetic diversity of the two main protein ligands for reticulocyte invasion, PvDBP and P. vivax Erythrocyte Binding Protein (PvEBP), we analyzed 458 isolates collected in Cambodia and Madagascar. First, we observed a high proportion of isolates with multiple copies PvEBP from Madagascar (56%) where Duffy negative and positive individuals coexist compared to Cambodia (19%) where Duffy-negative population is virtually absent. Whether the gene amplification observed is responsible for alternate invasion pathways remains to be tested. Second, we found that the PvEBP gene was less diverse than PvDBP gene (12 vs. 33 alleles) but provided evidence for an excess of nonsynonymous mutations with the complete absence of synonymous mutations. This finding reveals that PvEBP is under strong diversifying selection, and confirms the importance of this protein ligand in the invasion process of the human reticulocytes and as a target of acquired immunity. These observations highlight how genomic changes in parasite ligands improve the fitness of P. vivax isolates in the face of immune pressure and receptor polymorphisms.

genetics

A promoter interaction map for cardiovascular disease genetics

Over 500 genetic loci have been associated with risk of cardiovascular diseases (CVDs), however most loci are located in gene-distal non-coding regions and their target genes are not known. Here, we generated high-resolution promoter capture Hi-C (PCHi-C) maps in human induced pluripotent stem cells (iPSCs) and iPSC-derived cardiomyocytes (CMs) to provide a resource for identifying and prioritizing the functional targets of CVD associations. We validate these maps by demonstrating that promoters preferentially contact distal sequences enriched for tissue-specific transcription factor motifs and are enriched for chromatin marks that correlate with dynamic changes in gene expression. Using the CM PCHi-C map, we linked 1,999 CVD-associated SNPs to 347 target genes. Remarkably, more than 90% of SNP-target gene interactions did not involve the nearest gene, while 40% of SNPs interacted with at least two genes, demonstrating the importance of considering long-range chromatin interactions when interpreting functional targets of disease loci.

genetics

Mitochondrial genetics of exceptional longevity in multigeneration matrilineages

Some heritable mitochondrial DNA (mtDNA) sequence variants may slow the rate of aging. The European mitochondrial haplogroup K has previously been reported to be increased in frequency in centenarians and nonagenarians relative to its frequency in younger individuals, by standard case/control study designs. To select for mitochondrial genomes likely to carry beneficial genetic variants, we screened a large genealogical database (the Utah Population Database, UPDB) for mitochondrial lineages in which the frequency of survival past 90 years was significantly higher than in the general population, and also significantly higher than in close non-matrilineal relatives. We ranked 14,900 distinct matrilineages by the strength of their association with longevity. Full sequencing of the mtDNAs from a single individual from each of 53 matrilineages in the top longevity ranks and each of 374 control matrilineages from the general Utah population, followed by analyses of the mtDNA haplogroup frequencies, identified haplogroup K2 as the haplogroup most enriched in frequency by the longevity selection (Odds Ratio = 23.05). We then analyzed overall survival and cause-specific mortality in the several thousand individuals aged 40 years or older whose mtDNA genotypes could be imputed from the 374 fully sequenced control mtDNAs. In these control matrilineages Haplogroup K2 individuals (n=332) enjoyed a significantly lower all-cause mortality risk than the general population (HR=0.81), attributable in part to a significantly lower risk of dying from heart disease (HR=0.50), as well as lower (though not significantly lower) risks of dying from cancer (HR=0.72) and diabetes (HR=0.74). Furthermore, K2 was the only haplogroup in which mortality was reduced for all three of these common causes of death.

genetics

A conserved genetic interaction between Spt6 and Set2 regulates H3K36 methylation

The transcription elongation factor Spt6 and the H3K36 methyltransferase Set2 are both required for H3K36 methylation and transcriptional fidelity in Saccharomyces cerevisiae. By selecting for suppressors of a transcriptional defect in an spt6 mutant, we have isolated dominant SET2 mutations (SET2sup mutations) in a region encoding a proposed autoinhibitory domain. The SET2sup mutations suppress the H3K36 methylation defect in the spt6 mutant, as well as in other mutants that impair H3K36 methylation. ChIP-seq studies demonstrate that the H3K36 methylation defect in the spt6 mutant, as well as its suppression by a SET2sup mutation, occur at a step following the recruitment of Set2 to chromatin. Other experiments show that a similar genetic relationship between Spt6 and Set2 exists in Schizosaccharomyces pombe. Taken together, our results suggest a conserved mechanism by which the Set2 autoinhibitory domain requires multiple interactions to ensure that H3K36 methylation occurs specifically on actively transcribed chromatin.

genetics

Genetic data and cognitively-defined late-onset Alzheimer’s disease subgroups

Categorizing people with late-onset Alzheimers disease into biologically coherent subgroups is important for personalized medicine. We evaluated data from five studies (total n=4 050, of whom 2 431 had genome-wide single nucleotide polymorphism (SNP) data). We assigned people to cognitively-defined subgroups on the basis of relative performance in memory, executive functioning, visuospatial functioning, and language at the time of Alzheimers disease diagnosis. We compared genotype frequencies for each subgroup to those from cognitively normal elderly controls. We focused on APOE and on SNPs with p<10-5 and odds ratios more extreme than those previously reported for Alzheimers disease (<0.77 or >1.30). There was substantial variation across studies in the proportions of people in each subgroup. In each study, higher proportions of people with isolated substantial relative memory impairment had [&ge;]1 APOE e4 allele than any other subgroup (overall p= 1.5 x 10-27). Across subgroups, there were 33 novel suggestive loci across the genome with p<10-5 and an extreme OR compared to controls, of which none had statistical evidence of heterogeneity and 30 had ORs in the same direction across all datasets. These data support the biological coherence of cognitively-defined subgroups and nominate novel genetic loci.

genetics

The Subtype Specificity of Genetic Loci Associated with Stroke in 16,664 cases and 32,792 controls

BackgroundGenome-wide association studies have identified multiple loci associated with stroke. However, the specific stroke subtypes affected, and whether loci influence both ischaemic and haemorrhagic stroke, remains unknown. For loci associated with stroke, we aimed to infer the combination of stroke subtypes likely to be affected, and in doing so assess the extent to which such loci have homogeneous effects across stroke subtypes.\n\nMethodsWe performed Bayesian multinomial regression in 16,664 stroke cases and 32,792 controls of European ancestry to determine the most likely combination of stroke subtypes affected for loci with published genome-wide stroke associations, using model selection. Cases were subtyped under two commonly used stroke classification systems, Trial of Org 10172 Acute Stroke Treatment (TOAST) and Causative Classification of Stroke (CCS). All individuals had genotypes imputed to the Haplotype Reference Consortium 1.1 Panel.\n\nResultsSixteen loci were considered for analysis. Seven loci influenced both haemorrhagic and ischaemic stroke, three of which influenced ischaemic and haemorrhagic subtypes under both TOAST and CCS. Under CCS, 4 loci influenced both small vessel stroke and intracerebral haemorrhage. An EDNRA locus demonstrated opposing effects on ischaemic and haemorrhagic stroke. No loci were predicted to influence all stroke subtypes in the same direction and only one locus (12q24) was predicted to influence all ischaemic stroke subtypes.\n\nConclusionsHeterogeneity in the influence of stroke-associated loci on stroke subtypes is pervasive, reflecting differing causal pathways. However, overlap exists between haemorrhagic and ischaemic stroke, which may reflect shared pathobiology predisposing to small vessel arteriopathy. Stroke is a complex, heterogeneous disorder requiring tailored analytic strategies to decipher genetic mechanisms.

genetics

Genetic discovery and translational decision support from exome sequencing of 20,791 type 2 diabetes cases and 24,440 controls from five ancestries

Protein-coding genetic variants that strongly affect disease risk can provide important clues into disease pathogenesis. Here we report an exome sequence analysis of 20,791 type 2 diabetes (T2D) cases and 24,440 controls from five ancestries. We identify rare (minor allele frequency<0.5%) variant gene-level associations in (a) three genes at exome-wide significance, including a T2D-protective series of >30 SLC30A8 alleles, and (b) within 12 gene sets, including those corresponding to T2D drug targets (p=6.1x10-3) and candidate genes from knockout mice (p=5.2x10-3). Within our study, the strongest T2D rare variant gene-level signals explain at most 25% of the heritability of the strongest common single-variant signals, and the rare variant gene-level effect sizes we observe in established T2D drug targets will require 110K-180K sequenced cases to exceed exome-wide significance. To help prioritize genes using associations from current smaller sample sizes, we present a Bayesian framework to recalibrate association p-values as posterior probabilities of association, estimating that reaching p<0.05 (p<0.005) in our study increases the odds of causal T2D association for a nonsynonymous variant by a factor of 1.8 (5.3). To help guide target or gene prioritization efforts, our data are freely available for analysis at www.type2diabetesgenetics.org.

genetics

Mendelian Randomization integrating GWAS and eQTL data reveals genetic determinants of complex and clinical traits

Genome-wide association studies (GWAS) identified thousands of variants associated with complex traits, but their biological interpretation often remains unclear. Most of these variants overlap with expression QTLs (eQTLs), indicating their potential involvement in the regulation of gene expression.\n\nHere, we propose an advanced transcriptome-wide summary statistics-based Mendelian Randomization approach (called TWMR) that uses multiple SNPs jointly as instruments and multiple gene expression traits as exposures, simultaneously.\n\nWhen applied to 43 human phenotypes it uncovered 2,369 genes whose blood expression is putatively associated with at least one phenotype resulting in 3,913 gene-trait associations; of note, 36% of them had no genome-wide significant SNP nearby in previous GWAS analysis. Using independent association summary statistics (UKBiobank), we confirmed that the majority of these loci were missed by conventional GWAS due to power issues. Noteworthy among these novel links is educational attainment-associated BSCL2, known to carry mutations leading to a mendelian form of encephalopathy. We similarly unraveled novel pleiotropic causal effects suggestive of mechanistic connections, e.g. the shared genetic effects of GSDMB in rheumatoid arthritis, ulcerative colitis and Crohns disease.\n\nOur advanced Mendelian Randomization unlocks hidden value from published GWAS through higher power in detecting associations. It better accounts for pleiotropy and unravels new biological mechanisms underlying complex and clinical traits.

genetics

A Key For Hypoxia Genetic Adaptation In Alpaca Could Be A HIF1A Truncated bHLH Protein Domain

Animals exposed to hypoxia, triggers a physiological response via Hypoxia Inducible Factors (HIF1). In this study, we have evidenced the existence of genetic events that caused the loss of most of the bHLH domain in HIF1A proteins borne by Alpaca and other members of the Cetartiodactyla superorder. In these truncate domains, some stop codons are found at identical nucleotide positions in both, Artiodactyls and Cetaceans, indicating that mutations originating the truncated domains occurs before their divergence about 55 million years ago. The relevance of this findings for adaptation of Alpacas to hypoxia of high altitude conditions are discussed.

genetics

C57BL/6 substrain differences in inflammatory and neuropathic nociception and genetic mapping of a major quantitative trait locus underlying acute thermal nociception

Sensitivity to different pain modalities has a genetic basis that remains largely unknown. Employing closely related inbred mouse substrains can facilitate gene mapping of nociceptive behaviors in preclinical pain models. We previously reported enhanced sensitivity to acute thermal nociception in C57BL/6J (B6J) versus C57BL/6N (B6N) substrains. Here, we expanded on nociceptive phenotypes and observed an increase in formalin-induced inflammatory nociceptive behaviors and paw diameter in B6J versus B6N mice (Charles River Laboratories). No strain differences were observed in mechanical or thermal hypersensitivity or in edema following the Complete Freunds Adjuvant (CFA) model of inflammatory pain, indicating specificity in the inflammatory nociceptive stimulus. In the chronic nerve constriction injury (CCI), a model of neuropathic pain, no strain differences were observed in baseline mechanical threshold or in mechanical hypersensitivity up to one month post-CCI. We replicated the enhanced thermal nociception in the 52.5{degrees}C hot plate test in B6J versus B6N mice from The Jackson Laboratory. Using a B6J x B6N-F2 cross (N=164), we mapped a major quantitative trait locus (QTL) underlying hot plate sensitivity to chromosome 7 that peaked at 26 Mb (LOD=3.81, p<0.01; 8.74 Mb-36.50 Mb) that was more pronounced in males. Genes containing expression QTLs (eQTLs) associated with the peak nociceptive marker that are implicated in pain and inflammation include Ryr1, Cyp2a5, Pou2f2, Clip3, Sirt2, Actn4, and Ltbp4 (FDR < 0.05). Future studies involving positional cloning and gene editing will determine the quantitative trait gene(s) and potential pleiotropy of this locus across pain modalities.

genetics

The genetics of situs inversus totalis without primary ciliary dyskinesia

Situs inversus totalis (SIT), a complete left-right mirror reversal of the visceral organs, is usually described as a recessive disorder. SIT can occur with Primary Ciliary Dyskinesia (PCD). However, most people with SIT do not have PCD, and the etiology of their condition remains poorly studied. Those without PCD may have an elevated rate of left-handedness, implying developmental mechanisms linking brain and body laterality. We sequenced the genomes of 15 people with SIT, of which six had PCD, and 15 controls. The SIT subjects with PCD all had likely recessive mutations in genes already known to cause PCD. Two non-PCD SIT cases also had recessive mutations in known PCD genes, suggesting reduced penetrance for PCD in some SIT cases. One non-PCD SIT case had a recessive mutation in PKD1L1, which has previously been linked to SIT without PCD. However, six of the nine non-PCD SIT cases, including most of the left-handers in this dataset, had no obvious candidate genes or significant pathways affected by the mutations that they carried. While we cannot exclude a monogenic basis, more complex genetic models must also be considered, as well as environmental influences or random effects in early development.

genetics

Fast genetic mapping of complex traits in C. elegans using millions of individuals in bulk

Genetic studies of complex traits in animals have been hindered by the need to generate, maintain, and phenotype large panels of recombinant lines. We developed a new method, C. elegans eXtreme Quantitative Trait Locus (ceX-QTL) mapping, that overcomes this obstacle via bulk selection on millions of unique recombinant individuals. We used ceX-QTL to map a drug resistance locus with high resolution. We also mapped differences in gene expression in live worms and discovered a regulatory feedback loop that responds to changes in HSP-90 chaperone activity. Lastly, we used ceX-QTL to map loci that influence fitness and discovered that one such locus is caused by a deletion in a highly conserved chromatin reader in the N2 reference strain. ceX-QTL is fast, powerful and cost-effective, and will accelerate the study of complex traits in animals.

genetics

A high-resolution map of non-crossover events in mice reveals impacts of genetic diversity on meiotic recombination

During meiotic recombination in most mammals, hundreds of programmed DNA Double-Strand Breaks (DSBs) occur across all chromosomes in each cell at sites bound by the protein PRDM9. Faithful DSB repair using the homologous chromosome is essential for fertility, yielding either non-crossovers, which are frequent but difficult to detect, or crossovers. In certain hybrid mice, high sequence divergence causes PRDM9 to bind each homologue at different sites, \"asymmetrically\", and these mice exhibit meiotic failure and infertility, by unknown mechanisms. To investigate the impact of local sequence divergence on recombination, we intercrossed two mouse subspecies over five generations and deep-sequenced 119 offspring, whose high heterozygosity allowed detection of thousands of crossover and non-crossover events with unprecedented power and spatial resolution. Both crossovers and non-crossovers are strongly depleted at individual asymmetric sites, revealing that PRDM9 not only positions DSBs but also promotes their homologous repair by binding to the unbroken homologue at each site. Unexpectedly, we found that non-crossovers containing multiple mismatches repair by a different mechanism than single-mismatch sites, which undergo GC-biased gene conversion. These results demonstrate that local genetic diversity profoundly alters meiotic repair pathway decisions via at least two distinct mechanisms, impacting genome evolution and Prdm9-related hybrid infertility.

genetics

A mixed-model approach for powerful testing of genetic associations with cancer risk incorporating tumor characteristics

AO_SCPLOWBSTRACTC_SCPLOWCancers are routinely classified into subtypes according to various features, including histopathological characteristics and molecular markers. Previous genome-wide association studies have reported heterogeneous associations between loci and cancer subtypes. However, it is not evident what is the optimal modeling strategy for handling correlated tumor features, missing data, and increased degrees-of-freedom in the underlying tests of associations. We propose to test for genetic associations using a mixed-effect two-stage polytomous model score test (MTOP). In the first stage, a standard polytomous model is used to specify all possible sub-types defined by the cross-classification of the tumor characteristics. In the second stage, the subtype-specific case-control odds ratios are specified using a more parsimonious model based on the case-control odds ratio for a baseline subtype, and the case-case parameters associated with tumor markers. Further, to reduce the degrees-of-freedom, we specify case-case parameters for additional exploratory markers using a random-effect model. We use the Expectation-Maximization (EM) algorithm to account for missing data on tumor markers. Through simulations across a range of realistic scenarios and data from the Polish Breast Cancer Study (PBCS), we show MTOP outperforms alternative methods for identifying heterogeneous associations between risk loci and tumor subtypes. The proposed methods have been implemented in a user-friendly and high-speed R statistical package called TOP (https://github.com/andrewhaoyu/TOP).

genetics

Cytologic, Genetic, and Proteomic Analysis of a Yellow Leaf Mutant of Sesame (Sesamum indicum L.), Siyl-1

Leaf color mutation in sesame always affects the growth and development of plantlets, and their yield. To clarify the mechanisms underlying leaf color regulation in sesame, we analyzed a yellow-green leaf mutant. Genetic analysis of the mutant selfing revealed 3 phenotypes--YY, light-yellow (lethal); Yy, yellow-green; and yy, normal green--controlled by an incompletely dominant nuclear gene, Siyl-1. In YY and Yy, the number and morphological structure of the chloroplast changed evidently, with disordered inner matter, and significantly decreased chlorophyll content. To explore the regulation mechanism of leaf color mutation, the proteins expressed among YY, Yy, and yy were analyzed. All 98 differentially expressed proteins (DEPs) were classified into 5 functional groups, in which photosynthesis and energy metabolism (82.7%) occupied a dominant position. Our findings provide the basis for further molecular mechanism and biochemical effect analysis of yellow leaf mutants in plants.

genetics