Back

A unified framework for causal gene regulatory network inference grounded in orthogonal molecular evidence

Pugh, F. B.; Li, R.; Lai, W.

2026-05-01 systems biology
10.64898/2026.04.28.721354 bioRxiv
Show abstract

Gene regulatory networks (GRNs) govern gene expression, cellular differentiation, and stable transcriptional states. Yet inferring GRNs that integrate molecular regulatory mechanisms and reproduce transcriptional states as stable outcomes remains a central challenge. Here we present SETIA, a framework that infers GRNs whose explicit dynamical models reproduce transcriptional profiles as one or more stable states across conditions. Applied to RNA-seq data from wild-type and transcription factor knockout strains in Saccharomyces cerevisiae, SETIA infers GRNs that accurately reproduce held-out transcriptional states in cross-validation experiments. Incorporating TF-promoter binding and protein-protein interaction priors, SETIA yields GRNs ranging from mechanistically grounded architectures to flexible models that capture indirect regulatory influences. SETIA reveals that gene expression organizes into discrete stable states that represent distinct transcriptional programs, all emerging as stable attractors of a single underlying GRN whose dynamics are predominantly explained by TF-DNA binding and protein-protein interactions from orthogonal molecular evidence. HighlightsO_LISETIA infers causal gene regulatory networks whose dynamics reproduce and generalize to held-out transcriptional profiles as stable attractors C_LIO_LIGenes occupy multiple discrete, reproducible expression states C_LIO_LIA ChIP-exo-derived protein-protein/DNA interaction network provides structural priors that ground the GRN in molecular mechanisms C_LIO_LIMolecular structural priors improve mechanistic interpretability while maintaining dynamical performance C_LIO_LISETIA generalizes across bulk and single-cell data and scales to a semi-genome-scale regulatory network C_LI

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.