A Simple Strategy for Identifying Conserved Features across Non-independent Omics Studies
Reed, E.; Sebastiani, P.
Show abstract
False discovery is an ever-present concern in omics research, especially for burgeoning technologies with unvetted specificity of their biomolecular measurements, as such unknowns obscure the ability to characterize biologically informative features from studies performed with any single platform. Accordingly, performing replication studies of the same samples using different omics platforms is a viable strategy for identifying high-confidence molecular associations that are conserved across studies. However, an important caveat of replication studies that include the same samples is that they are inherently non-independent, leading to overestimating conservation if studies are treated otherwise. Strategies for accounting for such inter-study dependencies have been proposed for meta-analysis methods devised to increase statistical power to detect molecular associations in one or more studies. Still, they are not immediately suited for identifying conserved molecular associations across multiple studies. Here, we present a unifying strategy for performing inter-study conservation analysis as an alternative to meta-analysis strategies for aggregating summary statistical results of shared features across complementary studies while accounting for inter-study dependency. This method, which we call "adjusted maximum p-value" (AdjMaxP), is easy to implement with inter-study dependency and conservation estimated directly from the p-values from each studys molecular feature-level association testing results. Through simulation-based assessment, we demonstrate AdjMaxPs improved performance for accurately identifying conserved features over a related meta-analysis strategy for non-independent studies. AdjMaxP offers an easily implementable strategy for improving the precision of analyses for biomarker discovery from cross-platform omics study designs, thereby facilitating the adoption of such protocols for robust inference from emerging omics technologies.
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- TRIAGE: A web-based iterative analysis platform integrating pathway and network approaches optimizes hit selection from high- throughput assays. 93%
- An efficient not-only-linear correlation coefficient based on machine learning 91%
- Differential Allele-Specific Expression Uncovers Breast Cancer Genes Dysregulated By Cis Noncoding Mutations 91%
Similar papers in this journal
- Modelling group heteroscedasticity in single-cellRNA-seq pseudo-bulk data 93%
- Assessment of statistical methods from single cell, bulk RNA-seq and metagenomics applied to microbiome data 93%
- Primo: integration of multiple GWAS and omics QTL summary statistics for elucidation of molecular mechanisms of trait-associated SNPs and detection of pleiotropy in complex traits 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.