Back

Genetic integration with cell-specific nucleosome positioning resolves causal relationships underlying chromatin accessibility profiles

Wang, X.; Robertson, C. C.; Varshney, A.; Manickam, N.; Orchard, P. C.; Laakso, M.; Tuomilehto, J.; Lakka, T. A.; Mohlke, K. L.; Boehnke, M.; Scott, L. J.; Koistinen, H. A.; Collins, F. S.; Parker, S. C.

2025-09-13 genomics
10.1101/2025.09.09.674883 bioRxiv
Show abstract

Cell type-specific chromatin accessibility QTL (caQTL) mapping is a promising approach to understand genetic control of chromatin landscapes and identify regulatory mechanisms underlying GWAS associations. However, current caQTL studies lack resolution and do not distinguish nucleosome-free regions (NFR) from positioned nucleosomes. Here, we leverage statistical modeling of fragment position and length to decompose ATAC-seq profiles into NFRs and phased nucleosomes. With single nucleus (sn)ATAC-seq from 281 human muscle biopsies, we map cell type-specific genetic effects on NFRs (76,027 nfrQTLs) and nucleosome occupancy (24,623 nucQTLs) across skeletal muscle cell types. Colocalization and causal inference between nucQTLs and nearby nfrQTLs and show that nfrQTLs are substantially more likely to causally influence nucQTLs and phase adjacent nucleosomes, indicating a causal relationship in shaping chromatin profiles. Hundreds of nfrQTLs colocalize with GWAS signals for muscle-related traits, including grip strength, atrial fibrillation, and fasting insulin, and the majority of colocalizing signals mapped to credible sets overlapping the corresponding nfrPeak. This approach adds mechanistic insights for how variants underlying caQTLs and GWAS signals exert their cis regulatory effects by initially modifying NFR accessibility and subsequently shaping broader chromatin landscapes, nucleosome positioning, gene expression, and ultimately higher-level traits and disease.

Matching journals

The top 4 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.