Search bioRxiv⌕ Search

bioRxiv · 10.1101/2025.02.07.637054

Computing pathogenicity of mutations in human cytochrome P450 superfamily

Abstract

Cytochrome P450 (CYP) are heme-containing enzymes involved in the metabolism of drugs and endogenous compounds. Humans have 57 known CYPs, with >200 mutations linked to severe disorders, highlighting their role in disease etiology. To investigate their pathogenicity, we performed an in-depth computational study, comparing pathogenic mutations with non-pathogenic ones. First, we evaluated the effects of all possible mutations across 26 known CYPs structures using five structure-and five sequence-based methods, aiming to reproduce pathogenesis. We computed 23,46,410 mutations, categorized into: all possible, non-pathogenic, and pathogenic. Comparisons consistently revealed a meaningful stability pattern: non-pathogenic > all > pathogenic. Second, we found that a significantly higher number of pathogenic mutations were buried in CYP structures, relating to their higher pathogenesis potential. Third, we analyzed mutation allele frequency relative to solvent accessibility. Fourth, we computed change in 48 amino acid properties upon mutation and identified three distinguishing features: Gibbs free energy, isoelectric point and volume. Fifth, the positive residue content was reduced significantly in diseased mutations, with arginine mutations being the main culprit, directly linked to isoelectric point change. Sixth, we found a higher propensity for pathogenic mutations in conserved sites, suggesting disruption of CYP function. Finally, analysis of heme versus substrate sites showed a higher frequency of pathogenic mutations in heme site, with arginine being the main mutating residue, possibly disrupting arginine-heme interactions. We provide the first comprehensive analysis of mutation effects across multiple CYPs. We model the chemical basis of CYP-related pathogenicity, paving way for a semi-quantitative model to predict diseases.

Source connections

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Mondal, S., Shrivastava, P., Mehra, R.. 2025-02-08. Computing pathogenicity of mutations in human cytochrome P450 superfamily. https://doi.org/10.1101/2025.02.07.637054

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

Scaling of structural variability of ecDNA polymer condensates with copy number boosts and stabilises oncogene regulatory contacts

Extrachromosomal DNAs (ecDNAs) form highly heterogeneous condensates in cancer cells that drive oncogene overexpression, yet how structural variability coexists with stable gene regulation remains unclear. Here, we develop a minimal polymer physics model of MYC-harbouring COLO320-DM ecDNAs, where BRD4-like complexes bind and bridge cognate sites along ecDNA rings. Above a critical binder concentration, ecDNAs phase separate into condensates exhibiting diverse conformations because of their thermodynamic folding degeneracy. Despite this variability, condensates retain conserved interaction scaffolds that give rise to reproducible contact patterns, including in-trans associated domains (I-TADs), genomic regions enriched in intermolecular regulatory contacts between distinct ecDNAs. We find that condensate 3D architecture follows universal scaling relations with ecDNA copy number, n, remaining robust to model parameter changes. Regulatory contacts within I TADs increase linearly with n, yet they are one order of magnitude stronger than in size matched control regions outside I TADs, whereas their relative fluctuations are markedly suppressed as n increases. This scaling produces enhanced, low-noise regulatory environments for oncogenes embedded within I-TADs, such as PVT1-MYC fusions, whereas the canonical MYC copy, located outside, is less amplified as experimentally observed. Our findings reveal universal polymer physics principles underlying ecDNA condensate organization, offering a mechanistic basis for selective oncogene amplification and potential advantages in cancer progression.

biophysics↗

High-resolution mapping of RNA structural maturation during Cas9 assembly with ABEL-FRET

The structural flexibility of RNA is essential for forming ribonucleoprotein (RNP) complexes, which regulate diverse biological processes. This intrinsic property permits RNA to act as a dynamic scaffold along the assembly pathway as it folds into a specific structure for initial recognition by protein and undergoes conformational rearrangements for functional maturation as a complex. Yet, RNA flexibility and RNP multicomponent assembly create significant obstacles for traditional structural methods. To overcome these challenges, we applied recently developed ABEL-FRET spectroscopy to measure tether-free single-molecule Forster resonance energy transfer (smFRET) over extended observation times. Furthermore, ABEL-FRET enables the unique ability for simultaneous measurements of ultrahigh resolution smFRET and hydrodynamic size of individual complexes, which offers distinct advantages for studying dynamic RNA molecules that undergo assembly via sequential binding events. Using ABEL-FRET, we explored how the guide RNA (gRNA) of CRISPR genome editing system folds and modulates its structural flexibility to carry out the roles required for each assembly state from its unbound apo form to the functional Cas9 RNP state for target DNA cleavage. Multi-perspective view of gRNA structure gained by probing its two primary functional domains enabled to capture dramatic changes in gRNA flexibility that are highly dependent on its specific structural domains as well as assembly states. Collectively, our work with ABEL-FRET highlights the intrinsic link between the structural flexibility of RNA and its functionality in RNP assembly.

biophysics↗

De novo design of functional RNAs through higher-order interactions

Designing RNA sequences that reliably adopt functional three-dimensional structures remains a central challenge in RNA engineering because folding depends on cooperative interactions beyond canonical base pairing. Here we present DS3dRNA, an interaction-based framework for de novo RNA sequence design that combines a three-body statistical potential with physics-guided sequence sampling and supports design against multiple conformations. Across the evaluated benchmarks, DS3dRNA outperformed representative RNA inverse-design methods in native-sequence recovery and agreement between predicted and target structures. Energy-sequence-quality analyses further showed that lower design energies generally accompanied higher sequence recovery and macro-averaged F1 scores (MacroF1). Experimentally tested Mango II designs retained high-affinity fluorogenic activity, and five twister ribozyme designs yielded mean endpoint cleavage fractions of 37.7-50.6%, compared with 23.5% for the wild type. These results establish explicit higher-order interaction scoring as a complementary approach to emerging data-driven RNA design methods and provide a framework for designing functional RNAs from experimental or predicted structural ensembles.

biophysics↗