bioRxiv · 10.64898/2025.12.04.691622
Analyzing associations and higher-order effects in multi-omics data with double machine learning
Abstract
MotivationIntegrative omics analyses enhance our understanding of disease mechanisms and biomarkers by investigating relationships among traits, omics measurements, genetic variants, and epidemiological factors. Statistically, these analyses are challenging and require robust, flexible methodologies due to high dimensionality, non-standard data distributions, and potentially complex, non-linear confounding effects. ResultsTo facilitate the integration and analysis of multi-omics data, we introduce the Robust Omics MethodologY (ROMY) framework and its corresponding R implementation, the romy package. ROMY enables users to (a) perform robust association testing between two target variables while incorporating flexible covariate adjustments, (b) examine effects on measurement variances and covariances (e.g., co-expression, co-abundance), and (c) conduct rigorous interaction-effect testing. ROMY builds on recent advances in theoretical statistics and double machine learning to ensure robustness and statistical validity. We illustrate the performance of our framework through simulation studies. Availability and implementationromy is an R package available under the GNU GPLv3 license on GitHub: https://github.com/julianhecker/romy.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Hecker, J., Prokopenko, D., Hahn, G., Lee, S., Lutz, S. M., Lange, C.. 2025-12-08. Analyzing associations and higher-order effects in multi-omics data with double machine learning. https://doi.org/10.64898/2025.12.04.691622
Cite the original work for its findings. Save a collection to share your selection of sources.