Constructing genotype and phenotype network helps reveal disease heritability and phenome-wide association studies
Cao, X.; Zhu, L.; Liang, X.; Zhang, S.; Sha, Q.
Show abstract
Analyses of a bipartite Genotype and Phenotype Network (GPN), linking the genetic variants and phenotypes based on statistical associations, provide an integrative approach to elucidate the complexities of genetic relationships across diseases and identify pleiotropic loci. In this study, we first assess contributions to constructing a well-defined GPN with a clear representation of genetic associations by comparing the network properties with a random network, including connectivity, centrality, and community structure. Next, we construct network topology annotations of genetic variants that quantify the possibility of pleiotropy and apply stratified linkage disequilibrium (LD) score regression to 12 highly genetically correlated phenotypes to identify enriched annotations. The constructed network topology annotations are informative for disease heritability after conditioning on a broad set of functional annotations from the baseline-LD model. Finally, we extend our discussion to include an application of bipartite GPN in phenome-wide association studies (PheWAS). The community detection method can be used to obtain a priori grouping of phenotypes detected from GPN based on the shared genetic architecture, then jointly test the association between multiple phenotypes in each network module and one genetic variant to discover the cross-phenotype associations and pleiotropy. Significance thresholds for PheWAS are adjusted for multiple testing by applying the false discovery rate (FDR) control approach. Extensive simulation studies and analyses of 633 electronic health record (EHR)-derived phenotypes in the UK Biobank GWAS summary dataset reveal that most multiple phenotype association tests based on GPN can well-control FDR and identify more significant genetic variants compared with the tests based on UK Biobank categories.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Identification of putative causal loci in whole-genome sequencing data via knockoff statistics 96%
- OTTERS: A powerful TWAS framework leveraging summary-level reference data 96%
- Testing and controlling for horizontal pleiotropy with the probabilistic Mendelian randomization in transcriptome-wide association studies 95%
Similar papers in this journal
- A Poisson binomial based statistical testing framework for comprehensive comorbidity discovery across massive Electronic Health Record datasets 94%
- Adversarial domain translation networks for fast and accurate integration of large-scale atlas-level single-cell datasets 92%
- Automated customization of large-scale spiking network models to neuronal population activity 92%
Similar papers in this journal
- Uncovering genetic associations in the human diseasome using an endophenotype-augmented disease network 95%
- Exploiting deep transfer learning for the prediction of functional noncoding variants using genomic sequence 94%
- NetSHy: Network Summarization via a Hybrid Approach Leveraging Topological Properties 94%
Similar papers in this journal
- Efficient mixed model approach for large-scale genome-wide association studies of ordinal categorical phenotypes 96%
- mBAT-combo: a more powerful test to detect gene-trait associations from GWAS data 95%
- An integrative multi-context Mendelian randomization method for identifying risk genes across human tissues 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.