Genetic Prediction of Circulating Lipoprotein(a) Levels in Diverse Populations
Levin, M.; Selvaraj, M. S.; Vy, H. M.; Judy, R.; Honigberg, M. C.; Bajaj, A.; Nadkarni, G. C.; Do, R.; Denny, J. C.; Loh, P.-R.; Penn Medicine Biobank, ; Natarajan, P.
Show abstract
BackgroundCirculating lipoprotein(a) [Lp(a)] levels are highly heritable and linked to atherosclerotic cardiovascular disease, yet clinical measurement rates remain low (<1%) in the United States. The high heritability of Lp(a) across populations makes genetic prediction an attractive approach for closing this testing gap, but existing polygenic scores transfer poorly across populations. Haplotype-based prediction models, which use standard genome-wide genotype data to capture common-, rare-, and structural-variation at the LPA locus, could bridge this gap, enabling opportunistic identification of individuals with elevated Lp(a) levels across diverse populations within existing large, genotyped cohorts. ObjectivesThis study sought to develop and validate a haplotype-based prediction model using genome-wide genotype data to identify individuals with elevated Lp(a) levels across diverse populations. MethodsWe developed an LPA-haplotype model using data from the All of Us Research Program and validated it in the Penn Medicine BioBank (PMBB), Mass General Brigham Biobank (MGBB), and Mount Sinai BioMe cohorts. Primary outcomes included model performance for predicting continuous Lp(a) concentrations (r{superscript 2}) and identifying elevated Lp(a) levels (>125 nmol/L) through positive predictive value (PPV) and number needed to test (NNT). ResultsAmong PMBB (n = 1856), MGBB (n = 1401), and BioMe (n = 1686) participants with available genotype and Lp(a) measurements, average age was 60 years, and 51% were female. Overall r{superscript 2} of the haplotype model was 0.46 (95% Credible Interval [CrI] 0.32 to 0.6), with similar performance across genetically inferred ancestries and cohorts. For identifying elevated Lp(a) levels >125 nmol/L the overall PPV was 0.81 (95% CrI 0.6 to 0.89), corresponding to a NNT of 1.2 (95% CrI 1.1 to 1.7) individuals predicted to have elevated levels needing to undergo clinical testing to identify one true elevation. In the full PMBB cohort (n = 49310), the haplotype model identified elevated Lp(a) at a rate of 128 per 1000 (95% CrI 125 to 130), corresponding to an estimated 14.4-fold improvement (95% CrI 13.1 to 15.9; P(improvement) = 1) in identification rate compared with the existing rate of clinical assessment. ConclusionsA haplotype-based genetic model effectively identified individuals with elevated Lp(a) levels across diverse populations, with potential utility for opportunistic screening among cohorts where genotype data is available, but Lp(a) testing rates are low.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Population genomic screening leads to improved lipid management in patients with familial hypercholesterolemia 96%
- Coronary Artery Disease Risk of Familial Hypercholesterolemia Genetic Variants Independent of Historical Cholesterol Exposure 95%
- Validation of genome-wide polygenic risk scores for coronary artery disease in French Canadians 93%
Similar papers in this journal
- Adverse pregnancy outcomes and coronary artery disease risk: A negative control Mendelian randomization study 92%
- Genetically Predicted IL-18 Inhibition and Risk of Cardiovascular Events: A Mendelian Randomization Study 92%
- Genetically downregulated interleukin-6 signaling is associated with a favorable cardiometabolic profile: a phenome-wide association study 92%
Similar papers in this journal
- Polygenic Hyperlipidemias and Coronary Artery Disease Risk 94%
- LDLR Variant Classification for Improved Cardiovascular Risk Prediction in Familial Hypercholesterolemia 94%
- Markers of systemic iron status show sex-specific differences in peripheral artery disease: a cross-sectional analysis of HEIST-DiC and NHANES participants 92%
Similar papers in this journal
- Investigation of an LPA KIV-2 nonsense mutation in 11,000 individuals: the importance of linkage disequilibrium structure in LPA genetics. 94%
- Novel protein-altering variants associated with serum apolipoprotein and lipid levels 94%
- Validation of a Trans-Ancestry Polygenic Risk Score for Type 2 Diabetes in Diverse Populations 91%
Similar papers in this journal
- Causally Associations of Blood Lipids Levels with COVID-19 Risk: Mendelian Randomization Study 92%
- Identification of endothelial-derived proteins in plasma associated with cardiovascular risk factors 92%
- Antithrombin, protein C and protein S: Genome and transcriptome wide association studies identify 7 novel loci regulating plasma levels 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.