Back

Biophysical Modeling Uncovers Transcription Factor and Nucleosome Binding on Single DNA Molecules

Dalakishvili, L.; Managori, H.; Bardet, A.; Slaninova, V.; Bertrand, E.; Molina, N.

2025-05-16 genomics
10.1101/2025.05.13.653852 bioRxiv
Show abstract

Gene regulation in eukaryotes emerges from a dynamic interplay between transcription factors (TFs), nucleosomes, and RNA Polymerase II (Pol II), whose competitive and cooperative binding shapes DNA accessibility and transcriptional output. Single-molecule footprinting (SMF) and long-read chromatin accessibility assays such as Fiber-seq now capture these interactions at nucleotide resolution on individual DNA molecules. However, existing computational tools remain insufficient to decode complex binding events from sparse methylation data. Here, we introduce HiddenFoot, a probabilistic modeling framework based on statistical mechanics that quantitatively infers TF, nucleosome, and Pol II occupancy profiles on single DNA molecules by systematically evaluating all thermodynamically plausible binding configurations. Applying HiddenFoot to SMF and Fiber-seq data from mouse, Drosophila, and human cells, we recovered known TF footprints, precisely resolved Pol II pausing, and identified extensive heterogeneity in nucleosome positioning driven by TF binding. HiddenFoot further distinguishes direct TF-TF cooperativity from nucleosome-mediated co-dependency by estimating pairwise interaction energies and comparing to null models under equilibrium. By integrating biophysical modeling with high-resolution single-molecule data, HiddenFoot offers a general, interpretable framework for dissecting regulatory logic in native chromatin with base-pair precision. O_FIG O_LINKSMALLFIG WIDTH=200 HEIGHT=149 SRC="FIGDIR/small/653852v1_ufig1.gif" ALT="Figure 1"> View larger version (34K): org.highwire.dtl.DTLVardef@46f3b5org.highwire.dtl.DTLVardef@2a2cbdorg.highwire.dtl.DTLVardef@df42f2org.highwire.dtl.DTLVardef@1a4205d_HPS_FORMAT_FIGEXP M_FIG C_FIG

Matching journals

The top 4 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.