Assembly Arena: Benchmarking RNA isoform reconstruction algorithms for nanopore sequencing
Sagniez, M.; Budhraja, A.; Pare, B.; Simpson, S. M.; Vinet-Ouellette, C.; Rozendaal, M.; Smith, M. A.
Show abstract
Resolving the transcriptomes of higher eukaryotes is more tangible with the advent of long read sequencing, which greatly facilitates the identification of new transcripts and their splicing isoforms. However, the computational analysis of long read RNA sequencing data remains challenging as it is difficult to disentangle technical artifacts from bona fide biological information. To address this, we evaluated the performance of multiple leading transcriptome assembly algorithms on their ability to accurately reconstruct RNA transcript isoforms. We specifically focused on deep nanopore sequencing of synthetic RNA spike-in controls (Sequins and SIRVs) across different chemistries, including cDNA and direct RNA protocols. Our systematic comparative benchmarking exposes the strengths and limitations of the different surveyed strategies. We also highlight conceptual and technical challenges with the annotation of transcriptomes and the formalization of assembly quality metrics. Our results complement similar recent endeavors, helping forge a path towards a gold standard analytical pipeline for long read transcriptome assembly.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Identifying and quantifying isoforms from accurate full-length transcriptome sequencing reads with Mandalorion 97%
- Transcriptome assembly from long-read RNA-seq alignments with StringTie2 96%
- Enhancing transcriptome expression quantification through accurate assignment of long RNA sequencing reads with TranSigner 96%
Similar papers in this journal
- Improved characterization of single-cell RNA-seq libraries with paired-end avidity sequencing 95%
- Kmerator Suite: design of specific k-mer signatures andautomatic metadata discovery in large RNA-Seq datasets. 94%
- Tailored machine learning models for functional RNA detection in genome-wide screens 94%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.