Pervasive noise in human splice site selection
Khokhar, E. S.; Brokaw, K.; Kartje, Z. J.; Sanabria, V.; Javeed, N.; Kumar, A.; Watts, J. K.; Pai, A. A.
Show abstract
RNA splicing has historically been thought to be highly efficient and accurate, with little opportunity for deviation from regulated alternative splicing decisions. This dogma has been challenged by recent observations that suggest that biological noise may contribute substantially to transcriptome diversity. However, quantitative understanding of stochastic variations in splicing is challenging because these transcripts are likely subject to rapid degradation. Here, we use ultra-deep sequencing across RNA compartments to track splicing intermediates in human cells and see abundant cryptic splicing associated with genomic features that promote splicing noise. We observe pervasive usage of low-fidelity splice sites, likely due to stochasticity in recruitment or binding of the spliceosome. These sites are most likely degraded in the nucleus rather than targeted by translation-dependent degradation processes, suggesting widespread surveillance and rapid quality control of non-productive RNA transcripts. Our findings provide unprecedented insights into the propensity for error in RNA processing mechanisms and the regulation of alternative splice sites across a gene.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Empirical prediction of variant-associated cryptic-donors with 87% sensitivity and 95% specificity 97%
- Integrative analysis reveals RNA G-Quadruplexes in UTRs are selectively constrained and enriched for functional associations 97%
- Expanded palette of RNA base editors for comprehensive RBP-RNA interactome studies 97%
Similar papers in this journal
- Machine learning-optimized targeted detection of alternative splicing 97%
- Using single-cell perturbation screens to decode the regulatory architecture of splicing factor programs 97%
- Insplico: Effective computational tool for studying intron splicing order genome-wide with short and long RNA-seq reads 96%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.