bioRxiv · 10.1101/2024.02.18.580860
Interpretable and explainable predictive machine learning models for data-driven protein engineering
Abstract
Protein engineering using directed evolution and (semi)rational design has emerged as a powerful strategy for optimizing and enhancing enzymes or proteins with desired properties. Integrating artificial intelligence methods has further enhanced and accelerated protein engineering through predictive models developed in data-driven strategies. However, the lack of explainability and interpretability in these models poses challenges. Explainable Artificial Intelligence addresses the interpretability and explainability of machine learning models, providing transparency and insights into predictive processes. Nonetheless, there is a growing need to incorporate explainable techniques in predicting protein properties in machine learning-assisted protein engineering. This work explores incorporating explainable artificial intelligence in predicting protein properties, emphasizing its role in trustworthiness and interpretability. It assesses different machine learning approaches, introduces diverse explainable methodologies, and proposes strategies for seamless integration, improving trust-worthiness. Practical cases demonstrate the explainable models effectiveness in identifying DNA binding proteins and optimizing Green Fluorescent Protein brightness. The study highlights the utility of explainable artificial intelligence in advancing computationally assisted protein design, fostering confidence in model reliability.
Source connections
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Medina-Ortiz, D., Khalifeh, A., Anvari-Kazemabad, H., D. Davari, M.. 2024-02-21. Interpretable and explainable predictive machine learning models for data-driven protein engineering. https://doi.org/10.1101/2024.02.18.580860
Cite the original work for its findings. Save a collection to share your selection of sources.