bioRxiv · 10.64898/2026.01.07.698308
msmu: a Python toolkit for modular and traceable LC-MS proteomics data analysis based on MuData
Abstract
Computational workflows for MS-based proteomics remain comparatively fragmented, with heterogeneous data formats and analysis pipelines that hinder their reproducibility, interoperability, and reuse of processed data. We present msmu, an open-source Python package that implements a flexible and reproducible end-to-end pipeline for post-search data preprocessing and statistical analysis. At its core, msmu leverages the highly structured MuData format, empowering comprehensive data provenance, transparency in data sharing and reuse, and interoperability with broader Python ecosystem. Together, msmu represents a unique and significant step toward realizing the FAIR (Findable, Accessible, Interoperable, and Reusable) principles in computational proteomics.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Choi, H.-W., Lee, B., Kang, U.-B., Huh, S.. 2026-01-08. msmu: a Python toolkit for modular and traceable LC-MS proteomics data analysis based on MuData. https://doi.org/10.64898/2026.01.07.698308
Cite the original work for its findings. Save a collection to share your selection of sources.