Lemonite: identification of regulatory metabolites through data-driven, interpretable integration of transcriptomics and metabolomics data
Vandemoortele, B.; Devlies, H.; Michoel, T.; Vanhaecke, L.; Vandenbroucke, R. E.; Laukens, D.; Vermeirssen, V.
Show abstract
Biological regulation emerges from coordinated interactions between genes, proteins, and metabolites; yet, despite their central regulatory potential, metabolites remain largely absent from genome-wide gene regulatory network inference. Current transcriptomics-metabolomics integration approaches are either limited by poor interpretability or constrained by incomplete prior knowledge, preventing the systematic identification of regulatory metabolites. Here, we present Lemonite, a data-driven and interpretable framework for integrating bulk transcriptomics and metabolomics data to uncover regulatory metabolites acting on gene modules. Lemonite extends module network inference to jointly associate transcription factors and metabolites with coexpressed gene programs, without requiring prior differential analysis or complete metabolome annotation. To contextualize predictions, we constructed a comprehensive gene-metabolite knowledge graph integrating over 370 000 metabolite-gene and 2.1 million protein-protein interactions. Applied to glioblastoma and inflammatory bowel disease cohorts (n=99 and n=75 resp), Lemonite identified over 50 functionally coherent gene modules per disease, revealing established and previously uncharacterized metabolite-gene regulatory relationships. In glioblastoma, myo-inositol and phosphatidylcholines, together with IRF6, regulate mesenchymal-like immune programs, which upon integration with single-cell transcriptomics are primarily expressed in tumor-associated macrophages and monocytes. In inflammatory bowel disease, regulatory metabolites were prioritized that change the expression of their predicted target genes in colonic epithelial cells in vitro. Overall, Lemonite provides a principled framework to explore the genome-wide regulatory potential of the metabolome and to generate biologically interpretable, experimentally testable hypotheses from multi-omics data.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Multi-omic integration of microbiome data for identifying disease-associated modules 96%
- Genetic analysis of blood molecular phenotypes reveals regulatory networks affecting complex traits: a DIRECT study 96%
- TidyMass2: Advancing LC-MS Untargeted Metabolomics Through Metabolite Origin Inference and Metabolic Feature-based Functional Module Analysis 96%
Similar papers in this journal
- Multi-resolution characterization of molecular taxonomies in bulk and single-cell transcriptomics data 96%
- Disentangling single-cell omics representation with a power spectral density-based feature extraction 94%
- NetActivity enhances transcriptional signals by combining gene expression into robust gene set activity scores through interpretable autoencoders 94%
Similar papers in this journal
Similar papers in this journal
- AutoFocus: A hierarchical framework to explore multi-omic disease associations spanning multiple scales of biomolecular interaction 95%
- Gene Function Revealed at the Moment of Stochastic Gene Silencing 94%
- The pharmacoepigenomic landscape of cancer cell lines reveals the epigenetic component of drug sensitivity 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.