Uncovering Multi-Omics Profiles of Population Heterogeneity: A Cluster-Based Bayesian Approach
Shi, R.; Xiong, Z.; Banaschewski, T.; Barker, G. J.; Bokde, A. L. W.; Flor, H.; Garavan, H.; Gowland, P.; Grigis, A.; Heinz, A.; Martinot, J.-L.; Martinot, M.-L. P.; Artiges, E.; Nees, F.; Orfanos, D. P.; Poustka, L.; Smolka, M. N.; Hohmann, S.; Vaidya, N.; Walter, H.; Whelan, R.; Schumann, G.; Desrivieres, S.; Lin, X.; Feng, J.
Show abstract
Understanding the heterogeneous nature of genetic effects is critical for advancing our knowledge of the genetic architecture of complex traits and developing personalized management strategies. However, existing methods often rely on pre-specified modifying variables to model this heterogeneity, limiting their ability to capture effects driven by complex or unobserved factors. Here, we propose MOCHA (Multi-Omics Clustering for Heterogeneous Association), a novel Bayesian analytical paradigm that identifies latent population subgroups with distinct genetic effects directly from multi-omics data, without requiring a priori variable specification. Extensive simulations confirm that MOCHA accurately identifies the underlying clustering structure, demonstrates superior performance in identifying and ranking features with cluster-specific effects, and provides reliable parameter estimates. Applying MOCHA to genomic and transcriptomic data from the IMAGEN study, we identified two distinct neurodevelopmental clusters associated with adolescent inhibitory control. Post-hoc characterization of these clusters provided novel insights into the mechanisms of brain plasticity, demonstrating the methods practical utility and interpretability.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Testing and controlling for horizontal pleiotropy with the probabilistic Mendelian randomization in transcriptome-wide association studies 95%
- Genetic analysis of blood molecular phenotypes reveals regulatory networks affecting complex traits: a DIRECT study 94%
- Projecting genetic associations through gene expression patterns highlights disease etiology and drug mechanisms 94%
Similar papers in this journal
Similar papers in this journal
- Correspondence analysis for dimension reduction, batch integration, and visualization of single-cell RNA-seq data 94%
- Major cell-types in multiomic single-nucleus datasets impact statistical modeling of links between regulatory sequences and target genes 93%
- Controlling for background genetic effects using polygenic scores improves the power of genome-wide association studies 93%
Similar papers in this journal
- RAMEN: Dissecting individual, additive and interactive gene-environment contributions to DNA methylome variability in cord blood 95%
- A fast non-parametric test of association for multiple traits 95%
- GeneWalk identifies relevant gene functions for a biological context using network representation learning 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.