Search bioRxiv⌕ Search

bioRxiv · 10.64898/2026.02.05.704068

Results of a large scale study of the binding of 50 type IIinhibitors to 348 kinases: The role of protein reorganization

Abstract

Kinase family proteins constitute the second largest protein class targeted in drug development efforts, most prominently to treat cancer, but also several other diseases associated with kinase dysfunction. In this work we focus on type II kinase inhibitors which bind to the "classical" inactive conformation of the protein kinase catalytic domain where the DFG motif has a "DFG-out" orientation and the activation loop is folded. Many Tyrosine kinases (TKs) exhibit strong binding affinity with a wide spectrum of type II inhibitors while serine/threonine kinases (STKs) often bind more weakly. Recent work suggests this difference is largely due to differences in the folded to extended conformational equilibrium of the activation loop between TKs vs. STKs. The binding affinity of a type II inhibitor to its kinase target can be decomposed into a sum of two contributions: (1) the free energy cost to reorganize the protein from the active to inactive state, and (2) the binding affinity of the type II inhibitor to the inactive kinase conformation. In previous work we used a Potts statistical energy potential based on sequence co-variation to thread sequences over ensembles of active and inactive kinase structures. The threading function was used to estimate the free energy cost to reorganize kinases from the active to classical inactive conformation, and we showed that this estimator is consistent with the results of molecular dynamics free energy simulations for a small set of STKs and TKs. In the current study, we analyze the results of a large-scale study of the binding affinities of 50 type II inhibitors to 348 kinases, of which the results for 16 of the 50 type II inhibitors were reported in an earlier study (the "Davis dataset"); the binding data for the remaining 34 type II inhibitors to the panel of 348 kinases were recently obtained (the "Schrodinger dataset"). We use the Potts statistical energy model to investigate the contribution of protein reorganization to the selectivity of the large kinase panel against the set of 50 type II inhibitors, and find that protein reorganization makes a significant contribution to the selectivity. The AUC of the receiver-operator characteristic curve is [~]0.8. We report the results of an internal "blind test", that shows how Potts threading energies can provide more accurate estimates of kinase selectivity than corresponding predictions using experimental results of small sample size. We discuss why two STK phylogenetic kinase families, STE and CMGC, appear to contain many outliers, and how to improve the ability to predict kinase selectivity with a more complete analysis of the kinase conformational landscape. We compare the performance of Potts threading for predicting binding properties of the large set of (50) Type II inhibitors to 348 kinases, with those of a sequence-based purely machine learning model, DeepDTAGen, a publicly available machine learning model that was trained on the complete Davis dataset, including both Type I and Type II kinase inhibitors. We observe that DeepDTAGen performs well on binding predictions for the 16 type II inhibitors in the Davis dataset, but performs poorly on binding predictions for the 34 type II inhibitors against 348 kinases in the Schrodinger dataset.

Source connections

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Vardanyan, V. H., Haldane, A., Hwang, H., Coskun, D., Lihan, M., Miller, E. B., Friesner, R. A., Levy, R. M.. 2026-02-08. Results of a large scale study of the binding of 50 type IIinhibitors to 348 kinases: The role of protein reorganization. https://doi.org/10.64898/2026.02.05.704068

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related preprints

Mechanism of molecular recognition revealed through dynamic drug binding pathways to SARS-CoV-2 main protease

Characterization of drug-binding pathways remains experimentally limited by transient intermediates and computationally challenging due to long timescales intractable for conventional molecular dynamics. To address these challenges, we combined solution NMR titrations with weighted ensemble (WE) enhanced sampling simulations to resolve atomistic pathways of nirmatrelvir binding to the SARS-CoV-2 main protease. NMR titration revealed residue-dependent heterogeneity spanning fast, intermediate, and slow exchange regimes. WE simulations complement the NMR by providing insights into unassigned residues and adding time-resolved and three-dimensional structural context. We map key interactions along two distinct binding pathways, provide dynamic explanations for residues involved in resistance, and capture unique backbone conformations compared to those sampled in unbound or bound states. Our comprehensive binding model is consistent with a combined conformational selection and induced fit mechanism in which early transient contacts are made with residues E47 and L50 and allosteric motions are centered around residue V204 of the distal domain. This synergistic application of WE and titration NMR enables a more comprehensive characterization of drug binding than either method alone, providing an integrated framework that may have broader applicability to defining structure-kinetic relationships and guiding design of next-generation inhibitors.

biophysics↗

Discriminating betacoronavirus receptor usage across subgenera using protein structure prediction and molecular dynamics

A critical step in the emergence of a virus is the ability of the viral protein to bind a host receptor and mediate cell entry. For many coronaviruses, this interaction occurs between the Spike S1 subunit and the human ACE2 receptor. Whether this binding interface can be computationally distinguished across unstudied viruses without experimentally resolved protein structures remains an open question. We predicted how 28 emerging coronaviruses may bind to human ACE2 using structural predictions, static interaction prediction programs, and molecular dynamics simulations. To screen the emerging coronaviruses, we predicted a library of S1 structures using AlphaFold. These predicted structures were then used to model the S1-ACE2 interaction with AlphaFold, ClusPro, and HADDOCK. We used known ACE2-binding sarbecoviruses as positive controls and coronaviruses that bind other receptors as negative controls to threshold predicted binding. Contact analysis quantified the predicted binding and revealed that these static interaction prediction methods varied in discriminative power. Less restrained static predictions separated binders from non-binders, whereas heavily restrained docking did not, potentially forcing an interaction where none should exist. This analysis highlighted an emerging coronavirus, Zhejiang2013, as a potential ACE2 binder. We used molecular dynamics simulations to further assess the static predictions and model the interaction over time. Overall, our results indicate that Zhejiang2013 exhibits dynamic interaction patterns consistent with ACE2 binding. Given that two ACE2-binding coronaviruses have caused global pandemics within the past two decades, identifying potential ACE2 binders is critical for early warning and pandemic preparedness.

biophysics↗

De novo design of flexible protein interactions with GuideFlip

De novo design of protein binders requires a target structure. However, for flexible targets, such as intrinsically disordered proteins, this structure does not exist until the binder has stabilized the interaction. Such targets are therefore difficult for methods that separate structure generation from sequence design. We introduce GuideFlip, which co-designs structure and sequence through guided discrete flow matching: binder residues are assigned progressively while the complex is re-predicted at each step, allowing the evolving interface to affect the design process. GuideFlip reduces the hydrophobic bias of direct AlphaFold optimization and improves in silico success rates over existing approaches. We release a database of binder candidates for 177 human disordered proteins. Experimentally, we obtain de novo binders to the C-terminus of -synuclein and the disordered amino terminus of RBX1 with hit rates of 13.5% and 41.7%, respectively, and we confirm the epitopes of selected binders by NMR and mutagenesis. Applying GuideFlip to flexibility on the binder side, we design a nanobody that binds the agonist-bound {beta}1-adrenergic receptor in the active state, but not the receptor in its inactive state, with a 75% hit rate and cryo-EM structure confirming the design. GuideFlip enables protein design where bound structures emerge only upon binding.

biophysics↗