Pleiotropic Variability Score: A Genome Interpretation Metric to Quantify Phenomic Associations of Genomic Variants
Khader, S.; Glicksberg, B.; Johnson, K. W.; Badgeley, M.; Dudley, J.
Show abstract
A more complete understanding of phenomic space is critical for elucidating genome-phenome relationships and for assessing disease risk from genome sequencing. To incorporate knowledge of how related a variants associations are, we developed a new genome interpretation metric called Pleiotropic Variability Score (PVS). PVS uses semantic reasoning to score the relatedness of a genetic variants associated phenotypes based on those phenotypes relationships in the human phenotype ontology (HPO) and disease ontology (DO). We tested 78 unique semantic similarity methods and integrated six robust metrics to define the pleiotropy score of SNPs. We computed PVS for 12,541 SNPs which were mapped to 382 HPO and 317 DO unique phenotype terms in a genotype-phenotype catalog (10,021 SNPs mapped to DO phenotypes and 8,569 SNPs mapped to HPO phenotypes). We validated the utility of PVS by computing pleiotropy using an electronic health record linked genomic database (BioME, n=11,210). Further we demonstrate the application of PVS in personalized medicine using "personalized pleiotropy score" reports for individuals with genomic data that could potentially aid in variant interpretation. We further developed a software framework to incorporate PVS into VCF files and to consolidate pleiotropy assessment as part of genome interpretation pipelines. As the genome-phenome catalogs are growing, PVS will be a useful metric to assess genetic variation to find SNPs with highly pleiotropic effects. Additionally, variants with varying degree of pleiotropy can be prioritized for explorative studies to understand specific roles of SNPs and pleiotropic hubs in mediating novel phenotypes and drug development.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- GA4GH Phenopacket-Driven Characterization of Genotype-Phenotype Correlations in Mendelian Disorders 95%
- UK-Biobank Whole Exome Sequence Binary Phenome Analysis with Robust Region-based Rare Variant Test 95%
- Identifying digenic disease genes using machine learning in the undiagnosed diseases network 95%
Similar papers in this journal
- Towards a standard benchmark for phenotype-driven variant and gene prioritisation algorithms: PhEval - Phenotypic inference Evaluation framework 94%
- VarSight: Prioritizing Clinically Reported Variants with Binary Classification Algorithms 94%
- ATAV: a comprehensive platform for population-scale genomic analyses 93%
Similar papers in this journal
- Phen2Gene: Rapid Phenotype-Driven Gene Prioritization for Rare Diseases 96%
- DeepCOMBI: Explainable artificial intelligence for the analysis and discovery in genome-wide association studies 94%
- Predicting Gene Disease Associations With Knowledge Graph Embeddings For Diseases With Curtailed Information 94%
Similar papers in this journal
- Accelerate the discovery of genetic variants in mitochondrial diseases with VIOLA: Variant PrIOritization using Latent space 95%
- kTWAS: integrating kernel-machine with transcriptome-wide association studies improves statistical power and reveals novel genes 94%
- WEVar: a novel statistical learning framework for predicting noncoding regulatory variants 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.