Back

A Context-Specific, Literature-Supported Framework for Validating Stress Response Models in Mammals

Frishman, B. A.; Gonzalez, J. L.; Forbes, V. E.

2025-12-12 bioinformatics
10.64898/2025.12.10.693541 bioRxiv
Show abstract

Computational models of stress responses can highlight genes underlying physiological adaptation, but their utility depends on rigorous validation. Using existing biological databases, we tested a novel approach to identify and group differentially expressed genes (DEGs) from RNA-seq data. Key-Response, Treatment-Specific, Support, and Noisy groups of genes were previously identified as representing the Principal Response of the species to the stress. The Support Group was also suspected to represent housekeeping genes based on variability patterns. To validate these assumptions, we built protein-protein interaction (PPI) networks using the Human Protein Atlas and STRING, incorporating both direct and second-order connections. Crucially, second-order connections were restricted to those made via DEGs, ensuring that connectivity reflected condition-specific stress responses rather than generic hubs. Across two conditions, >75% of Principal Response genes assembled into largest connected components (LCC) that were significantly larger than random networks. Support Group genes also showed strong connectivity and enriched overlap with a housekeeping gene database, supporting their distinct classification. STRING confirmed PPI enrichment but produced less reliable results than our DEG-restricted framework. Overall, this study demonstrates that the novel method identifies biologically meaningful, interconnected stress-response networks. By emphasizing DEG-restricted second-order connections, our framework addresses key limitations of context-free enrichment methods and advances validation strategies for computational models of gene regulation, with implications for stress physiology and computational model validation.

Matching journals

The top 8 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.