Predicting which genes will respond to perturbations of a TF: TF-independent properties of genes are major determinants of their responsiveness
Kang, Y.; Brent, M.
Show abstract
BackgroundThe ability to predict which genes will respond to perturbation of a TFs activity serves as a benchmark for our systems-level understanding of transcriptional regulatory networks. In previous work, machine learning models have been trained to predict static gene expression levels in a given sample by using data from the same or similar conditions, including data on TF binding locations, histone marks, or DNA sequence. We report on a different challenge - training machine learning models that can predict which genes will respond to perturbation of a TF without using any data from the perturbed cells. ResultsExisting TF location data (ChIP-Seq) from human K562 cells have no detectable utility for predicting which genes will respond to perturbation of the TF, but data obtained by newer methods in yeast cells are useful. TF-independent features of genes, including their pre-perturbation expression level and expression variation, are very useful for predicting responses to TF perturbations. This shows that some genes are poised to respond to TF perturbations and others are resistant, shedding significant light on why it has been so difficult to predict responses from binding locations. Certain histone marks (HMs), including H3K4me1 and H3K4me3, have some predictive power, especially when downstream of the transcription start site. In human, the predictive power of HMs is much less than that of gene expression level and variation. Code is available at https://github.com/yiming-kang/TFPertRespExplainer. ConclusionsSequence-based or epigenetic properties of genes strongly influence their tendency to respond to direct TF perturbations, partially explaining the oft-noted difficulty of predicting responsiveness from TF binding location data. These molecular features are largely reflected in and summarized by the genes expression level and expression variation.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Genome-wide nucleosome and transcription factor responses to genetic perturbations reveal chromatin-mediated mechanisms of transcriptional regulation 96%
- Domain adaptive neural networks improvecross-species prediction of transcription factor binding 96%
- Independent evolution of transcript abundance and gene regulatory dynamics 95%
Similar papers in this journal
- Model-X knockoffs reveal data-dependent limits on regulatory network identification 96%
- Transcriptional kinetic synergy: a complex landscape revealed by integrating modelling and synthetic biology 95%
- Comprehensive prediction of robust synthetic lethality between paralog pairs in cancer cell lines 94%
Similar papers in this journal
- Predictive features of gene expression variation reveal a mechanistic link between expression variation and differential expression 95%
- Transcription factor expression is the main determinant of variability in gene co-activity 94%
- Unique features of transcription termination and initiation at closely spaced tandem human genes 93%
Similar papers in this journal
Similar papers in this journal
- BET family members Bdf1/2 modulate global transcription initiation and elongation in Saccharomyces cerevisiae 95%
- Promoter sequence and architecture determine expression variability and confer robustness to genetic variants 95%
- Unraveling the influences of sequence and position on yeast uORF activity using massively parallel reporter systems and machine learning. 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.