Search bioRxiv⌕ Search

Biology subjects

Preston, J. D.

Publications and source records attributed to Preston, J. D..

2 recordsLinked to original sources

MSMICA: computational metabolite identification in untargeted metabolomics by integrating MS, retention time, and biological evidence

Mass Spectrometry Metabolomics Identification Connection Algorithm (MSMICA) is an algorithm for automated metabolite identification in untargeted liquid chromatography-high-resolution mass spectrometry (LC-HRMS) analyses. Limitations in metabolite identification can occur due to the availability and cost of standards and prevent recognition of metabolic factors impacting human health and disease. MSMICA performs mass-to-charge-ratio matching with chemical structures and clusters of LC-HRMS features for adduct and isotope forms. A local optimization is then used to integrate retention time prediction, metabolite precursor-product and transporter correlations, and biospecimen-specific abundance information for metabolite identification. Applying MSMICA to various internal and external mammalian datasets, validation results showed a 96.2 +- 5.1% correct rate of metabolite identification. When multiple LC-HRMS datasets were used, MSMICA enabled greater metabolite identifications, expanded metabolic pathway coverage, and data harmonization. Thus, MSMICA applies multiple pieces of evidence to substantially improve metabolite identification coverage and accuracy for known metabolites.

bioinformatics↗

TernTables: A Statistical Analysis and Table Generation Web Interface for Clinical and Biomedical Research

Clinical research dissemination is frequently hindered by administrative friction and methodological inconsistency. To address these barriers, we developed TernTables, a freely available, open-source web application (https://www.tern-tables.com/) and R package (https://cran.r-project.org/package=TernTables) that streamlines the transition from raw data to formatted results for descriptive and univariate clinical reporting. The system integrates a client-side screening protocol for protected health information (PHI) with a rule-based decision tree that selects and executes appropriate frequency-based, parametric, or non-parametric statistical tests based on data distribution and class. TernTables generates publication-ready summary tables in Microsoft Word format, complemented by dynamically generated methods text and the underlying R code to ensure complete transparency and reproducibility. Validation using a landmark clinical trial dataset demonstrated concordance with established biostatistical approaches for descriptive and univariate analyses. TernTables is designed to supplement, not replace, formal statistical consultation by standardizing routine descriptive and univariate workflows, allowing biostatistical expertise to be focused on complex analyses and study design. By lowering technical and financial barriers, the platform democratizes access to rigorous statistical workflows while maintaining methodological excellence and reducing "researcher degrees of freedom."

bioinformatics↗