ASTA-P: a pipeline for the detection, quantification and statistical analysis of complex alternative splicing events
Tiwari, K.; Nielsen, L. K.
Show abstract
Alternative splicing dramatically increases the repertoire of the human transcriptome and plays a critical role in cellular differentiation. Long read sequencing has dramatically improved our ability to explore isoform diversity directly. However, short read sequencing still provides advantages in terms of sequencing depth at low cost, which is important in comparative quantitative studies. Here, we present a pipeline called ASTA-P for profiling, quantification, and differential splicing analysis of tissue-specific, arbitrarily complex alternative splicing patterns. We discover novel events by supplementing existing annotation with reconstructed transcripts and use spliced RNA-seq reads to quantify splicing changes accurately based on their unique assignments. We used simulated RNA-seq data to demonstrate that ASTA-P provides a good trade-off between discovery and accuracy compared with several popular methods. Further, we applied ASTA-P to analyse AS patterns in real data from hiPSC derived cranial neural crest cells capturing the transition from primary neural cells into migratory cranial neural crest cells, differentiated by their expression of the transcription factor, SOX10. Our analysis revealed a significant splicing complexity, i.e., numerous AS events that cannot be described using the conventionally analysed 2D splicing event patterns. Such events are misclassified when analysed using current differential splicing analysis methods. Thus, ASTA-P provides a new approach for studying both conventional and complex splicing across different cellular conditions and the dynamic regulation of AS. The pipeline is available at https://github.com/uqktiwar/ASTAP/tree/main
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Dividing out quantification uncertainty allows efficient assessment of differential transcript expression with edgeR 96%
- Dividing out quantification uncertainty enables assessment of differential transcript usage with limma and edgeR 96%
- Conservation assessment of human splice site annotation based on a 470-genome alignment 96%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.