bioRxiv · 10.1101/2024.07.16.603812
SciMind: A Multimodal Mixture-of-Experts Model for Advancing Pharmaceutical Sciences
Abstract
Large language models (LLMs) have made substantial strides, but their use in reliably tackling issues within specialized domains, particularly in interdisciplinary areas like pharmaceutical sciences, is hindered by data heterogeneity, knowledge complexity, unique objectives, and a spectrum of constraint conditions. In this area, diverse modalities such as nucleic acids, proteins, molecular structures, and natural language are often involved. We designed a specialized token set and introduced a new Mixture-of-Experts (MoEs) pre-training and fine-tuning strategy to unify these modalities in one model. With this strategy, weve created a multi-modal mixture-of-experts foundational model for pharmaceutical sciences, named SciMind. This model has undergone extensive pre-training on publicly accessible datasets including nucleic acid sequences, protein sequences, molecular structure strings, and biomedical texts, and delivers good performance on biomedical text comprehension, promoter prediction, protein function prediction, molecular description, and molecular generation.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Xiong, Z., Fang, X., Chu, H., Wan, X., Liu, L., Xiang, W., Li, Y., Zheng, M.. 2024-07-21. SciMind: A Multimodal Mixture-of-Experts Model for Advancing Pharmaceutical Sciences. https://doi.org/10.1101/2024.07.16.603812
Cite the original work for its findings. Save a collection to share your selection of sources.