Back

Functional and Computational Interrogation of the Juvenile Idiopathic Arthritis Risk Loci Identifies Candidate Causal SNPs and Target Genes in CD4+ T cells

Jiang, K.; Haley, E. K.; Barshad, G.; He, A.; Rogic, A.; Rice, E. J.; Sudman, M.; Thompson, S. D.; Danko, C. G.; Jarvis, J. N.

2025-12-16 genetic and genomic medicine
10.64898/2025.12.15.25342296 medRxiv
Show abstract

GWAS have identified multiple genetic regions that confer risk for juvenile idiopathic arthritis (JIA). However, identifying the single nucleotide polymorphisms (SNPs) that drive disease risk has been impeded by the fact that the SNPs used to identify risk loci are in linkage disequilibrium (LD) with hundreds of other SNPs. Since the causal SNPs remain unknown, it is difficult to identify target genes and thus use genetic information to elucidate disease biology and inform patient care. We next used existing genotyping data from 3,939 children with JIA and 14,412 healthy controls to identify SNPs on JIA risk haplotypes that: present within open chromatin in multiple immune cell types and more common in children with JIA than the controls (p<0.05) in the genotyping data sets. We identified SNPs within cis-regulatory regions (CREs) using precision run-on sequencing data, and identified likely target genes using MicroC in both resting and activated CD4+ T cells. We identified 138 SNPs within the PROseq-identified CREs, and n=41 genes with which these CREs physically interacted. Data from GTEx corroborated these analyses by showing allelic effects for SNPs within the CREs in the ERAP2 and IRF1 risk loci. We further corroborated IRF1 allelic effects using a luciferase reporter assay. Our findings significantly reduce the genomic search space for risk-driving variants and target genes and support the roles of IRF1, ERAP2 and LNPEP in driving risk for JIA.

Matching journals

The top 7 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.