Search bioRxiv⌕ Search

Biology subjects

Tarakanova, A.

Publications and source records attributed to Tarakanova, A..

2 recordsLinked to original sources

Modeling Coronavirus Spike Protein Dynamics: Implications for Immunogenicity and Immune Escape

The ongoing COVID-19 pandemic is a global public health emergency requiring urgent development of efficacious vaccines. While concentrated research efforts are underway to develop antibody-based vaccines that would neutralize SARS-CoV-2, and several first-generation vaccine candidates are currently in Phase III clinical trials or have received emergency use authorization, it is forecasted that COVID-19 will become an endemic disease requiring second-generation vaccines. The SARS-CoV-2 surface Spike (S) glycoprotein represents a prime target for vaccine development because antibodies that block viral attachment and entry, i.e. neutralizing antibodies, bind almost exclusively to the receptor binding domain (RBD). Here, we develop computational models for a large subset of S proteins associated with SARS-CoV-2, implemented through coarse-grained elastic network models and normal mode analysis. We then analyze local protein domain dynamics of the S protein systems and their thermal stability to characterize structural and dynamical variability among them. These results are compared against existing experimental data, and used to elucidate the impact and mechanisms of SARS-CoV-2 S protein mutations and their associated antibody binding behavior. We construct a SARS-CoV-2 antigenic map and offer predictions about the neutralization capabilities of antibody and S mutant combinations based on protein dynamic signatures. We then compare SARS-CoV-2 S protein dynamics to SARS-CoV and MERS-CoV S proteins to investigate differing antibody binding and cellular fusion mechanisms that may explain the high transmissibility of SARS-CoV-2. The outbreaks associated with SARS-CoV, MERS-CoV, and SARS-CoV-2 over the last two decades suggest that the threat presented by coronaviruses is ever-changing and long-term. Our results provide insights into the dynamics-driven mechanisms of immunogenicity associated with coronavirus S proteins, and present a new approach to characterize and screen potential mutant candidates for immunogen design, as well as to characterize emerging natural variants that may escape vaccine-induced antibody responses. STATEMENT OF SIGNIFICANCEWe present novel dynamic mechanisms of coronavirus S proteins that encode antibody binding and cellular fusion properties. These mechanisms may offer an explanation for the widespread nature of SARS-CoV-2 and more limited spread of SARS-CoV and MERS-CoV. A comprehensive computational characterization of SARS-CoV-2 S protein structures and dynamics provides insights into structural and thermal stability associated with a variety of S protein mutants. These findings allow us to make recommendations about the future mutant design of SARS-CoV-2 S protein variants that are optimized to elicit neutralizing antibodies, resist structural rearrangements that aid cellular fusion, and are thermally stabilized. The integrated computational approach can be applied to optimize vaccine immunogen design and predict escape of vaccine-induced antibody responses by SARS-CoV-2 variants.

biophysics↗

DSResSol: A sequence-based solubility predictor created with Dilated Squeeze Excitation Residual Networks

Protein solubility is an important thermodynamic parameter critical for the characterization of a proteins function, and a key determinant for the production yield of a protein in both the research setting and within industrial (e.g. pharmaceutical) applications. Thus, a highly accurate in silico bioinformatics tool for predicting protein solubility from protein sequence is sought. In this study, we developed a deep learning sequence-based solubility predictor, DSResSol, that takes advantage of the integration of squeeze excitation residual networks with dilated convolutional neural networks. The model captures the frequently occurring amino acid k-mers and their local and global interactions, and highlights the importance of identifying long-range interaction information between amino acid k-mers to achieve higher performance in comparison to existing deep learning-based models. DSResSol uses protein sequence as input, outperforming all available sequence-based solubility predictors by at least 5% in accuracy when the performance is evaluated by two different independent test sets. Compared to existing predictors, DSResSol not only reduces prediction bias for insoluble proteins, but also predicts soluble proteins within the test sets with an accuracy that is at least 13% higher. We derive the key amino acids, dipeptides, and tripeptides contributing to protein solubility, identifying glutamic acid and serine as critical amino acids for protein solubility prediction. Overall, DSResSol can be used for fast, reliable, and inexpensive prediction of a proteins solubility to guide experimental design. AvailabilityThe source code, datasets, and web server for this model are available at https://github.com/mahan-fcb/DSResSol

bioinformatics↗