Back

HILAMA: High-dimensional multi-omic mediation analysis with latent confounding

Wang, X.; Liu, J.; Hu, S. S.; Liu, Z.; Lu, H.; Liu, L.

2023-09-15 bioinformatics
10.1101/2023.09.15.557839 bioRxiv
Show abstract

MotivationThe increasingly available multi-omic datasets have posed both new opportunities and challenges to the development of quantitative methods for discovering novel mechanisms in biomedical research. One natural approach to analyzing such datasets is mediation analysis originated from the causal inference literature. Mediation analysis can help unravel the mechanisms through which exposure(s) exert the effect on outcome(s). However, existing methods fail to consider the case where (1) both exposures and mediators are potentially high-dimensional and (2) it is very likely that some important confounding variables are unmeasured or latent; both issues are quite common in practice. To the best of our knowledge, however, no methods have been developed to address these challenges with statistical guarantees. ResultsIn this article, we propose a new method for HIgh-dimensional LAtent-confounding Mediation Analysis, abbreviated as "HILAMA", that considers both high-dimensional exposures and mediators, and more importantly, the possible existence of latent confounding variables. HILAMA achieves false discovery rate (FDR) control under finite sample size for multiple mediation effect testing. The proposed method is evaluated through extensive simulation experiments, demonstrating its improved stability in FDR control and superior power in finite sample size compared to existing competitive methods. Furthermore, our method is applied to the proteomics-radiomics data from ADNI, identifying some key proteins and brain regions relating to Alzheimers disease. The results show that HILAMA can effectively control FDR and provide valid statistical inference for high dimensional mediation analysis with latent confounding variables. AvailabilityThe R package HILAMA is publicly available at https://github.com/Cinbo-Wang/HILAMA. Contactcinbo_w@sjtu.edu.cn

Published in BMC Medical Research Methodology · training set

Matching journals

The top 4 journals account for 50% of the predicted probability mass.

1
Bioinformatics
1204 papers in training set
Top 0.7%
27.3%
2
Briefings in Bioinformatics
354 papers in training set
Top 0.3%
13.5%
3
The Annals of Applied Statistics
19 papers in training set
Top 0.1%
8.1%
4
Biostatistics
24 papers in training set
Top 0.1%
6.9%
50% of probability mass above
5
Genetic Epidemiology
55 papers in training set
Top 0.1%
5.6%
6
BMC Bioinformatics
457 papers in training set
Top 2%
4.4%
7
Biometrics
23 papers in training set
Top 0.1%
3.6%
8
Statistics in Medicine
40 papers in training set
Top 0.2%
2.8%
9
eLife
5828 papers in training set
Top 38%
2.7%
10
Patterns
78 papers in training set
Top 1%
1.8%
11
PLOS Computational Biology
1863 papers in training set
Top 14%
1.8%
12
NeuroImage
903 papers in training set
Top 4%
1.8%
13
Nature Communications
5641 papers in training set
Top 44%
1.8%
14
PLOS Genetics
862 papers in training set
Top 8%
1.4%
15
Journal of Biomedical Informatics
47 papers in training set
Top 0.9%
1.2%
16
Human Brain Mapping
329 papers in training set
Top 3%
1.2%
17
Statistical Methods in Medical Research
11 papers in training set
Top 0.2%
1.0%
18
PLOS ONE
5266 papers in training set
Top 60%
0.9%
19
Genome Biology
637 papers in training set
Top 8%
0.9%
20
Frontiers in Genetics
230 papers in training set
Top 6%
0.6%
21
Computers in Biology and Medicine
128 papers in training set
Top 5%
0.6%
22
European Journal of Epidemiology
43 papers in training set
Top 0.8%
0.6%