Search bioRxiv⌕ Search

bioRxiv · 10.1101/2025.08.06.668254

Q-MOL: High Fidelity Platform for In Silico Drug Discovery and Design

Abstract

For several decades, in silico drug discovery methods have held the promise of cheap and efficient ways of streamlining de novo discovery of small molecule ligands against novel drug targets. The existing computational methodologies proved to be successful in those cases when a protein target is a relatively rigid molecule (e.g., many enzymes), or when an in silico methodology is applied during hit-to-lead optimization of an already potent binder. However, these methods fail when applied to highly flexible or intrinsically disordered protein targets which represent the bulk of pharmacologically important proteins. The major factors contributing to the failure are either incorrect or simplified treatment of protein flexibility, inadequate or overfitted scoring function, absence of correctly annotated druggable pockets, and lack of structural data about potential allosteric sites. The Q-MOL drug discovery platform addresses the problem of treatment of protein flexibility during the course of protein-ligand docking by computationally implementing some of postulates of the energy landscape theory of protein folding. The Q-MOL protein-ligand docking method was thoroughly validated in academic and commercial settings across a panel of more than 60 diverse proteins (of viral, bacterial and animal origin). This methodology, originally developed for high fidelity virtual ligand screening, is also used for the prediction of ligand binding sites. The ability to reliably predict ligand binding sites on the surface of novel and unannotated protein targets is critical for successful drug discovery. The Q-MOL drug discovery platform is agnostic to the nature of a protein target. Rigid, flexible or intrinsically disordered proteins are processed by the same protein-ligand docking protocol. Importantly, the software does not require the presence of structurally well-defined binding pockets or sites. The correct and validated implicit treatment of protein flexibility allows for successful discovery of small molecule ligands against allosteric sites lacking any structural features of a classical druggable pocket. To demonstrate the power of its methodology, Q-MOL was used to detect ligand binding sites on the surface of the bHLHZ domain of the c-Myc protein. Q-MOL virtual ligand screening was then used to dock a human metabolites library against one of the identified binding sites. The top hits were cross-referenced against published data, notably identifying dehydroascorbic acid as one of the top-ranked compounds. In contrast to other docking programs, all of the Q-MOL in silico drug discovery features were thoroughly validated in vitro, in cellulo and in vivo when possible. In this paper, the results of several previously published drug discovery projects are briefly presented and related to relevant computational features of the software. The corresponding protein targets include viral proteinases (West Nile, Hepatitis C, Dengue and Zika viruses), hemopexin domain of membrane type I matrix metalloproteinase, retinoid X receptor , and {beta}-catenin armadillo repeat domain. It has been also shown that contrary to existing belief, highly flexible proteins (or their domains) represent significantly more amenable drug discovery targets than rigid proteins. Flexible and intrinsically disordered proteins exist in a vast array of environment-dependent conformational states, and thus, they can accommodate a significantly larger number of ligand chemotypes ensuring both potency and specificity of interactions with a protein target. It has been also demonstrated that Q-MOL binding sites prediction and ligand docking methodologies, initially developed for protein targets, can be successfully applied "As Is" to polynucleotide-based structures, such as non-coding viral RNA molecules.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Cheltsov, A.. 2025-08-08. Q-MOL: High Fidelity Platform for In Silico Drug Discovery and Design. https://doi.org/10.1101/2025.08.06.668254

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

Conjunctive Targeting Links Drug Synergy to Emergent Proteome Structural States

Combinatorial therapies are widely used in the treatment of acute myeloid leukemia (AML) to address disease heterogeneity, adaptive resistance, and rewired signaling and metabolic states. Yet drug prioritization remains largely guided by clinical or phenotypic evidence, while the molecular mechanisms underlying effective drug combinations remain incompletely defined. To narrow this gap, we developed Combinatorial high-ratio Partial proteolysis with reference PRoteome Analysis (CoPPRA), a structural proteomics workflow based on limited proteolysis of cell lysates that profiles drug-associated changes in regional protein accessibility at peptide-level resolution. Here, we applied CoPPRA to ruxolitinib and ulixertinib, individually and in combination, in AML-related cell lysates. Our findings extend conjunctive targeting (CT), a recently proposed mechanism of combinatorial drug action in which combined exposure produces protein targeting patterns not observed with either drug alone. Previously identified through combination-associated changes in protein solubility/stability, CT is examined here at peptide-level resolution through regional differences in proteolytic accessibility. The ruxolitinib-ulixertinib combination produced broad peptide-level accessibility changes, including a subset meeting the predefined criteria for CT. CT candidates predominantly exhibited regional accessibility changes, with altered peptide regions occurring against comparatively small changes across the remaining quantified peptides from the same proteins. MAP2K1 and ATP6V1G1 showed pronounced differences between overlapping peptide sequences, highlighting localized variation in combination-associated accessibility, including an ATP6V1G1 peptide mapping to an annotated helical region. Combination-associated increases in peptide signals were also observed in PIK3R1, BRD4, and PTPN11, linking regional accessibility changes to signaling and transcriptional regulators relevant to AML. Functional enrichment and network analyses further implicated nucleotide and glucose metabolism, ficolin-1-rich granules, ribosome-associated processes, and phagocytic vesicles. These results extend conjunctive targeting from protein-level solubility/stability changes to regional differences in proteolytic accessibility, showing that combination-associated effects can be concentrated within specific peptide regions rather than distributed uniformly across proteins. More broadly, CoPPRA provides a peptide-resolved approach for investigating the molecular features of combinatorial drug action and prioritizing protein regions for subsequent mechanistic validation.

biochemistry↗

Structural and biochemical characterisation of an iterative GCN5-related N-acetyltransferase required for fungal siderophore tailoring

Siderophore-mediated iron acquisition is essential for fungal survival, particularly under iron-limiting conditions. In Aspergillus fumigatus, SidG, a member of the GCN5-related N-acetyltransferase (GNAT) superfamily, catalyses the final step in the biosynthesis of the extracellular siderophore triacetylfusarinine C (TAFC) through sequential acetylation of the precursor fusarinine C (FsC). However, the timing, catalytic mechanism, and functional significance of this modification are not fully understood. Here, we reconstituted SidG activity in vitro and combined native mass spectrometry, X-ray crystallography, molecular dynamics simulations, and site-directed mutagenesis to investigate its catalytic properties. Our analyses demonstrate that SidG selectively binds acetyl-CoA from the cellular milieu and iteratively acetylates the FsC scaffold prior to iron chelation. Structural, biochemical, and molecular dynamics analyses support a direct transfer mechanism, identify key catalytic residues, and demonstrate the strict selectivity of SidG for short-chain acyl-CoA donors. Together, these findings establish the molecular basis for SidG-dependent siderophore tailoring and expand our understanding of GNAT-catalysed transformations in fungal natural product biosynthesis.

biochemistry↗

Reconstitution of +1 nucleosome transcription reveals coordinated functions of SAGA, Mediator, and TFIIH

The +1 nucleosome has emerged as a key regulator of eukaryotic transcription, but how it controls transcription initiation remains poorly understood. Here we reconstitute transcription through the +1 nucleosome using eleven purified yeast factors: RNA polymerase II (Pol II), the six general transcription factors (GTFs), TFIIS, the activator Pho4, and the SAGA and Mediator complexes. The system recapitulates key features of regulation observed in vivo. SAGA, acting with Pho4, directs pre-initiation complex (PIC) assembly to the correct position through its TBP-loading activity. Mediator stimulates transcription when the +1 nucleosome imposes a barrier to PIC formation, consistent with stabilization of productive TFIIH-DNA engagement. Contrary to the prevailing model, SAGA remains bound to the PIC after TBP loading and acetylates the +1 nucleosome within the assembled complex. The isolated PIC-Mediator-SAGA-nucleosome complex is transcriptionally active, and the repressive effect of the nucleosome is relieved by the DNA translocase activity of Ssl2, the TFIIH subunit that opens promoter DNA. TFIIH thus couples promoter melting to remodeling of the +1 nucleosome.

biochemistry↗