Back

An Integrated Large-Scale Atlas of Protein Quantitative Trait Loci across Olink and SomaScan platforms

Khunsriraksakul, C.; Zhang, F.; Wang, L.; Chen, S.; Markus, H.; Chen, D.; Shenoy, G.; Lopez-Silva, C.; Jabboure, F.; Kesaf, A. E.; Zhan, X.; Melia, J.; Hui, K.; Jiang, B.; Liu, D.

2025-10-07 genetic and genomic medicine
10.1101/2025.10.06.25336803 medRxiv
Show abstract

Protein quantitative trait loci (pQTLs) provide insight into the genetic regulation of protein expression and disease biology. We performed a large-scale cross-platform meta-analysis of plasma proteomics, integrating data from over 90,000 individuals across Olink and SomaScan platforms. This effort identified >30,000 sentinel pQTLs, with multi-trait approach (MTAG) markedly boosting both discovery and replication, despite the potential differences in the protein measurements of the two platforms. While both platforms captured well-known pleiotropic loci (HLA, ABO, SH2B3), we also uncovered platform-specific associations, reflecting unique assay designs. We further established a comprehensive resource for transcriptome- and proteome-wide association studies (TWAS, PWAS), identifying >100,000 upstream regulators with strong replication. As a proof of concept, we applied this framework to inflammatory bowel disease (IBD), generating the largest to date GWAS directly comparing Crohns disease (CD) and ulcerative colitis (UC). Integrative multi-omics analyses revealed divergent loci enriched in the NF-{kappa}B signaling pathway and improved CD vs. UC classification when proteomic features were combined with polygenic risk scores (PRS). Our findings provide a comprehensive resource for plasma protein genetics and demonstrate the value of integrative multi-omics for disease subtype classification, with broad implications for precision medicine.

Matching journals

The top 3 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.