Weighted variance component test for the integrative multi-omics analysis of microbiome data
Zhang, A.; Ling, W.; Little, A.; Williams-Nguyen, J. S.; Moon, J.-Y.; Burk, R. D.; Knight, R.; Wang, D. D.; Qi, Q.; Kaplan, R.; Zhao, N.; Wu, M.
Show abstract
Metabolic dysregulation and alterations have been linked to various diseases and conditions. Innovations in high-throughput technology now allow rapid profiling of the metabolome and metagenome -- often the gene content of bacterial populations -- for characterizing metabolism. Due to the small sample sizes and high dimensionality of the data, pathway analysis (wherein the effect of multiple genes or metabolites on an outcome is cumulatively assessed) of metabolomic data is commonly conducted and also represents a standard for metagenomic analysis. However, how to integrate both data types remains unclear. Recognizing that a metabolic pathway can be complementarily characterized by both metagenomics and metabolomics, we propose a weighted variance components framework to test if the joint effect of genes and metabolites in a biological pathway is associated with outcomes. The approach allows analytic p-value calculation, correlation between data types, and optimal weighting. Power simulations show that our approach often outperforms other strategies while maintaining type I error. The approach is illustrated on real data.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Pathway Analysis Through Mutual Information 96%
- Joint Modeling of Longitudinal Biomarker and Survival Outcomes with the Presence of Competing Risk in Nested Case-Control Studies with Application to the TEDDY Microbiome Dataset 96%
- BAGSE: a Bayesian hierarchical model approach for gene set enrichment analysis 95%
Similar papers in this journal
- Dirichlet distribution parameter estimation with application in microbiome analyses 95%
- A robust and fast two-sample test of equal correlations with an application to differential co-expression 95%
- Using a supervised principal components analysis for variable selection in high-dimensional datasets reduces false discovery rates 94%
Similar papers in this journal
- PaIRKAT: A pathway integrated regression-based kernel association test with applications to metabolomics and COPD phenotypes 96%
- Addressing Erroneous Scale Assumptions in Microbe and Gene Set Enrichment Analysis 95%
- Reconstruction Set Test (RESET): a computationally efficient method for single sample gene set testing based on randomized reduced rank reconstruction error 95%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.