Unfolding and De-confounding: Biologically meaningful causal inference from longitudinal multi-omic networks using METALICA
Ruiz-Perez, D.; Gimon, I.; Sazal, M.; Mathee, K.; Narasimhan, G.
Show abstract
A key challenge in the analysis of microbiome data is the integration of multi-omic datasets and the discovery of interactions between microbial taxa, their expressed genes, and the metabolites they consume and/or produce. In an effort to improve the state-of-the-art in inferring biologically meaningful multi-omic interactions, we sought to address some of the most fundamental issues in causal inference from longitudinal multi-omics microbiome data sets. We developed METALICA, a suite of tools and techniques that can infer interactions between microbiome entities. METALICA introduces novel unrolling and de-confounding techniques used to uncover multi-omic entities that are believed to act as confounders for some of the relationships that may be inferred using standard causal inferencing tools. The results lend support to predictions about biological models and processes by which microbial taxa interact with each other in a microbiome. The unrolling process helps to identify putative intermediaries (genes and/or metabolites) to explain the interactions between microbes; the de-confounding process identifies putative common causes that may lead to spurious relationships to be inferred. METALICA was applied to the networks inferred by existing causal discovery and network inference algorithms applied to a multi-omics data set resulting from a longitudinal study of IBD microbiomes. The most significant unrollings and de-confoundings were manually validated using the existing literature and databases. ImportanceWe have developed a suite of tools and techniques capable of inferring interactions between microbiome entities. METALICAintroduces novel techniques called unrolling and de-confounding that are employed to uncover multi-omic entities considered to be confounders for some of the relationships that may be inferred using standard causal inferencing tools. To evaluate our method, we conducted tests on the Inflammatory Bowel Disease (IBD) dataset from the iHMP longitudinal study, which we pre-processed in accordance with our previous work.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Decoding the Language of Microbiomes: Leveraging Patterns in 16S Public Data using Word-Embedding Techniques and Applications in Inflammatory Bowel Disease 96%
- Pathway-based and phylogenetically adjusted quantification of metabolic interaction between microbial species 95%
- MiMeNet: Exploring Microbiome-Metabolome Relationships using Neural Networks 95%
Similar papers in this journal
- Assembling bacterial puzzles: piecing together functions into microbial pathways 94%
- Integrated de novo Gene Prediction and Peptide Assembly of Metagenomic Sequencing Data 94%
- CORGIAS: identifying correlated gene pairs by considering evolutionary history in a large-scale prokaryotic genome dataset 93%
Similar papers in this journal
Similar papers in this journal
- iModulonDB: a knowledgebase of microbial transcriptional regulation derived from machine learning 94%
- Unveiling the Microbial Realm with VEBA 2.0: A modular bioinformatics suite for end-to-end genome-resolved prokaryotic, (micro)eukaryotic, and viral multi-omics from either short- or long-read sequencing 94%
- MetaOmGraph: a workbench for interactive exploratory data analysis of large expression datasets 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.