Evaluating Aggregated Gene Level eQTL Scores
Meyer, D.; Popko, N.; Laub, D.; Schofield, P.; Amariuta, T.; Alexandrov, L. B.; Carter, H.
Show abstract
Genetic feature engineering, used in methods such as transcriptome-wide association study, supports gene-trait association testing by aggregating single variants into gene-level features predictive of expression. To evaluate how different model architectures, LD filtering thresholds, and variant prioritization methods affect expression prediction quality, we trained over 3 million models and evaluated their performance in independent cohorts. Using the best performing models to impute expression and immunotherapy response as an example trait, we found a significant association with the reactive oxygen species pathway (p=0.032). Our model training workflow will support genetic feature engineering towards improved complex trait modeling.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- DAESC+: High-performance, integrated software for single-cell allele-specific expression data 92%
- An elastic-net logistic regression approach to generate classifiers and gene signatures for types of immune cells and T helper cell subsets 91%
- Deconvolution of bulk blood eQTL effects into immune cell subpopulations 91%
Similar papers in this journal
Similar papers in this journal
- Bioinformatics workflows for genomic analysis of tumors from Patient Derived Xenografts (PDX): challenges and guidelines 90%
- GeneTerpret: a customizable multilayer approach to genomic variant prioritization and interpretation 90%
- Evaluating single-subject study methods for personal transcriptomic interpretations to advance precision medicine 90%
Similar papers in this journal
- A novel machine learning-based algorithm for eQTL identification reveals complex pleiotropic effects in the MHC region 92%
- WEVar: a novel statistical learning framework for predicting noncoding regulatory variants 92%
- kTWAS: integrating kernel-machine with transcriptome-wide association studies improves statistical power and reveals novel genes 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.