Comparative Evaluation of Cross-Ancestry Polygenic Risk Scoring of Type 1 Diabetes in the All of Us Cohort
Ahangari, M.; Moore, S.; Davidson, I.; Anomaly, J.; Maier, R.; Li, J. H.; Christensen, M.; Stern, D.; Wolfram, T.
Show abstract
Type 1 diabetes is a highly heritable autoimmune condition characterized by the destruction of pancreatic beta cells, resulting in insulin deficiency. Here, we developed a novel polygenic score we call the HLA-Augmented SBayesRC Framework (HLA-ARC). HLA-ARC integrates direct modeling of HLA haplotypes, with a Bayesian regression approach for the non-HLA component. SBayesRC leverages extensive functional genomic annotations and linkage disequilibrium patterns across approximately 7.4 million variants, substantially enhancing predictive accuracy. We systematically compared HLA-ARC to three existing T1D polygenic scores (Polygenic Risk Score extension for Diabetes Mellitus [PRSedm], Trans-Ancestry Polygenic Score for Diabetes [TA-PS], and Type 1 Diabetes Multi-Ancestry Polygenic Score [T1D-MAPS]) using data from the ancestrally-diverse All of Us cohort. Among the three existing methods, T1D-MAPS showed superior performance in all ancestry groups. However, HLA-ARC consistently outperformed the existing methods, achieving AUROC values exceeding 0.91 in European individuals and 0.89 in non-European groups. Our results demonstrate that integrating HLA haplotype modeling with genomic annotation and ancestry-informed linkage disequilibrium methods significantly improves polygenic risk prediction for autoimmune diseases characterized by major genetic risk loci.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Calibrated rare variant genetic risk scores for complex disease prediction using large exome sequence repositories 95%
- A multi-task convolutional deep learning method for HLA allelic imputation and its application to trans-ethnic MHC fine-mapping of type 1 diabetes 95%
- Improved Allele Frequencies in gnomAD through Local Ancestry Inference 95%
Similar papers in this journal
- A combined polygenic score of 21,293 rare and 22 common variants significantly improves diabetes diagnosis based on hemoglobin A1C levels 95%
- Combining case-control status and family history of disease increases association power 95%
- Scalable generalized linear mixed model for region-based association tests in large biobanks and cohorts 95%
Similar papers in this journal
- Inclusion of Variants Discovered from Diverse Populations Improves Polygenic Risk Score Transferability 95%
- Multivariate adaptive shrinkage improves cross-population transcriptome prediction for transcriptome-wide association studies in underrepresented populations 95%
- Polygenic risk score prediction accuracy convergence 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.