Back

I-CONVEX: Fast and Accurate de Novo Transcriptome Recovery from Long Reads

Baharlouei, S.; Razaviyayn, M.; Tseng, E.; Tse, D.

2020-10-01 bioinformatics
10.1101/2020.09.28.317594 bioRxiv
Show abstract

Long-read sequencing technologies demonstrate high potential for de novo discovery of complex transcript isoforms, but high error rates pose a significant challenge. Existing error correction methods rely on clustering reads based on isoform-level alignment and cannot be efficiently scaled. We propose a new method, I-CONVEX, that performs fast, alignment-free isoform clustering with almost linear computational complexity, and leads to better consensus accuracy on simulated, synthetic, and real datasets.

Matching journals

The top 3 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.