GLASS: A Graph Learning Algorithm for Screening Splice-Aware Alignments of Long-Read RNA-seq
Li, J.; Li, G.; Yu, T.
Show abstract
With the continuous development of third-generation RNA-seq, obtaining accurate splice-aware alignments for the long RNA-seq reads to reference genomes has become one of the major challenges in transcriptomic analysis. To mitigate the erroneous alignments arising from existing long-read RNA-seq aligners, we propose GLASS, a novel splice-aware alignments filtering approach based on third-generation transcriptome data. GLASS introduces a newly designed Read-AS Map model and integrates graph machine learning techniques for detecting and removing falsely spliced aligned reads from alignment files. Experimental results demonstrate that GLASS significantly reduces the error rate of spliced alignmnt and enhance the accuracy of subsequent transcriptome reconstruction. Additionally, GLASS demonstrates strong generalization ability in data from species such as Mus musculus and Arabidopsis thaliana, providing a new approach for the efficient processing of third-generation transcriptome data and offering more reliable data support for subsequent gene expression analysis and functional studies.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Detection of alternative splicing: deep sequencing or deep learning? 95%
- A Robust and Scalable Graph Neural Network for Accurate Single Cell Classification 95%
- A novel splicing graph allows a direct comparison between exon-based and splice junction-based approaches to alternative splicing detection 95%
Similar papers in this journal
- Comprehensive benchmark of differential transcript usage analysis for static and dynamic conditions 96%
- FLYNC: A Machine Learning-Driven Framework for Discovering Long Non-Coding RNAs in Drosophila melanogaster 95%
- Informative RNA-base embedding for functional RNA structural alignment and clustering by deep representation learning 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.