I-CONVEX: Fast and Accurate de Novo Transcriptome Recovery from Long Reads
Baharlouei, S.; Razaviyayn, M.; Tseng, E.; Tse, D.
Show abstract
Long-read sequencing technologies demonstrate high potential for de novo discovery of complex transcript isoforms, but high error rates pose a significant challenge. Existing error correction methods rely on clustering reads based on isoform-level alignment and cannot be efficiently scaled. We propose a new method, I-CONVEX, that performs fast, alignment-free isoform clustering with almost linear computational complexity, and leads to better consensus accuracy on simulated, synthetic, and real datasets.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Juggling offsets unlocks RNA-seq tools for fast scalable differential usage, aberrant splicing and expression analyses. 95%
- Enhancing transcriptome expression quantification through accurate assignment of long RNA sequencing reads with TranSigner 95%
- DeepSAP: Improved RNA-Seq Alignment by Integrating Transcriptome Guidance with Transformer-Based Splice Junction Scoring 95%
Similar papers in this journal
- bayNorm: Bayesian gene expression recovery, imputation and normalisation for single cell RNA-sequencing data 95%
- Souporcell3: Robust Demultiplexing for High-Donor Single-Cell RNA-seq Datasets 94%
- Resolving single-cell heterogeneity from hundreds of thousands of cells through sequential hybrid clustering and NMF 94%
Similar papers in this journal
- Single-Cell Omics for Transcriptome CHaracterization (SCOTCH): isoform-level characterization of gene expression through long-read single-cell RNA sequencing 95%
- Constructing Ensemble Gene Functional Networks Capturing Tissue/condition-specific Co-expression from Unlabled Transcriptomic Data with TEA-GCN 94%
- Error correction enables use of Oxford Nanopore technology for reference-free transcriptome analysis 94%
Similar papers in this journal
- Sensitive detection of circular DNA at single-nucleotide resolution using guided realignment of partially aligned reads 94%
- SpliceRead: Improving Canonical and Non-Canonical Splice Site Prediction with Residual Blocks and Synthetic Data Augmentation 93%
- cDNA-detector: Detection and removal of cDNA contamination in DNA sequencing libraries 93%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.