Search bioRxiv⌕ Search

bioRxiv · 10.1101/2023.09.27.559835

The genetic architecture of the pepper metabolome provides insights into the regulation of capsianoside biosynthesis

Abstract

Capsicum (pepper) is among the most economically important species worldwide, the fruit accumulates specialized metabolites with essential roles in plant environmental interaction and potential health benefits. However, the underlying genetic basis of their biosynthesis remains largely unknown. In this study, we developed and assessed both wild genetic variance and a bespoke mapping population to determine the genetic architecture of the pepper metabolome. The genetic analysis provided over 30 metabolic quantitative trait loci (mQTL) for over 1100 metabolites. We identified 92 candidate genes involved in various mQTL. Among the identified loci, we described and validated by transient overexpression a domestication gene cluster of eleven UDP-glycosyltransferases involved in monomeric capsianoside biosynthesis. We additionally constructed the biosynthetic reactions and annotated the genes involved in capsianoside biosynthesis in pepper. Given that differential glycosylation of acyclic diterpenoid glycosides contributes to plant resistance and acts as anticancer agents in humans, our data provide new insight, and resources for better understanding the biosynthesis of beneficial natural compounds to improve human health.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Nauen, J., Tripodi, P., Wendenberg, R., Tringovska, I., Nankar, A. N., Stoeva, V., Pasev, G., Klemmer, A., Todorova, V., Bulut, M., Tikunov, Y., Bovy, A., Gechev, T., Kostova, D., Alseekh, S., Fernie, A. R.. 2023-09-29. The genetic architecture of the pepper metabolome provides insights into the regulation of capsianoside biosynthesis. https://doi.org/10.1101/2023.09.27.559835

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

Utilizing single-cell data for per-cell type eQTL mapping in the human pancreas

Aims/hypothesis The human pancreas is a central organ for metabolic regulation that is comprised of diverse cell types that uniquely contribute to its function. Previous studies have performed expression quantitative trail loci (eQTL) discovery in either whole pancreas or in pancreatic islets, but due to differences between pancreatic cell types, this approach does not reveal cell type-specific effects. In this study, we sought to either implicate the cell type of action for known eQTLs or identify new eQTLs that may have been masked in bulk studies by performing eQTL discovery in individual pancreatic cell types. Methods We clustered 153,018 single-cell RNA sequencing (scRNA-seq) data from 71 pancreatic islet donors from the Human Pancreas Analysis Program (HPAP). We performed eQTL discovery in six pancreatic cell types using this resource directly. We further utilized this single cell resource as a reference to deconvolute bulk pancreatic RNA sequencing data from 305 Genotype Tissue Expression (GTEx) project donors and performed eQTL discovery in four pancreatic cell types. Finally, we performed fine-mapping and co-localization of pancreatic cell type eQTLs with metabolic GWAS to connect our findings to metabolic disease risk. Results From analyzing 71 individuals with single cell profiles, we identified 112 unique eGenes across six pancreatic cell types, 99 of which had been identified previously and 13 unique to this study. From the deconvoluted eQTLs, we identified 3,134 unique eGenes across four pancreatic cell types, 116 of which were unique to our study. Fine-mapping and co-localization of eQTLs with metabolic GWAS yielded key leads that warrant further investigation, such as the association of rs2168101 with LMO1 expression in alpha cells. Conclusions/interpretation We identified new signals that were previously not found in bulk pancreatic eQTL studies and potential cell type of action for several signals that were identified previously. Although there are limitations to the power, and therefore, discoverability of this study, it provides insights into how individual pancreatic cells differently contribute to metabolic disease.

genetics↗

Snurportin-1 maintains muscle niche integrity and myogenic progenitor homeostasis

Loss-of-function variants in SNUPN, encoding the nuclear import factor Snurportin-1 (SPN1) required for spliceosomal small nuclear ribonucleoprotein (snRNP) transport, cause a recently described form of limb-girdle muscular dystrophy (LGMD). However, the role of SPN1 in skeletal muscle homeostasis remains poorly understood, in part due to the lack of a suitable in vivo model. Here, we generated a zebrafish snupn loss-of-function model that recapitulates key features of the skeletal muscle phenotype observed in patients. Mutant larvae developed severe locomotor impairment by 6 days post-fertilization (dpf), accompanied by sarcomeric disorganization and impaired muscle fiber integrity. Transcriptomic profiling at 6 dpf revealed widespread alternative splicing and transcriptional dysregulation, with prominent alterations in extracellular matrix and basement membrane components, together with upregulation of stress- and inflammation-associated genes. Notably, these late-stage abnormalities were preceded by disruption of the muscle progenitor population at 2 dpf, with reduced Pax7 progenitor abundance and myogenic gene expression together with altered muscle differentiation and organization. Together, these findings identify SPN1 as a key regulator of skeletal muscle homeostasis linking RNA processing to extracellular niche integrity and myogenic progenitor maintenance. This zebrafish model provides an in vivo platform for dissecting LGMD-associated disease mechanisms and developing therapeutic strategies aimed at restoring muscle function and regenerative capacity.

genetics↗

MOD-scTWAS: Leveraging gene co-expression for single-cell transcriptome-wide association studies

Transcriptome-wide association studies (TWAS) provide an effective framework for identifying genes associated with complex traits. Population-scale single-cell transcriptomic data enable genetically regulated expression (GReX) prediction and TWAS analyses at cell-type resolution, but the predictive performance of existing single-cell TWAS methods remains limited. Here, we develop MOD-scTWAS, a module-based method that jointly models GReX for genes within co-expression modules to borrow information across genes. Starting from a generative model for single-cell gene expression, MOD-scTWAS accounts for the heteroscedasticity and cross-gene correlation of individual-level pseudobulk expression in joint GReX prediction. In cross-validation analyses of the OneK1K dataset, MOD-scTWAS achieved higher mean GReX prediction accuracy than scTWAS across all 14 cell types and increased the number of imputable genes. When applied to TWAS analyses of UK Biobank quantitative hematological traits, MOD-scTWAS identified more significant cell type-gene-trait associations than scTWAS. These results demonstrate the potential of leveraging gene co-expression through joint modeling to improve cell-type-specific GReX prediction and TWAS discovery.

genetics↗