HyPo: Super Fast & Accurate Polisher for Long Read Genome Assemblies
Kundu, R.; Casey, J.; Sung, W.-K.
Show abstract
Efforts towards making population-scale long read genome assemblies (especially human genomes) viable have intensified recently with the emergence of many fast assemblers. The reliance of these fast assemblers on polishing for the accuracy of assemblies makes it crucial. We present HyPo-a Hybrid Polisher-that utilises short as well as long reads within a single run to polish a long read assembly of small and large genomes. It exploits unique genomic kmers to selectively polish segments of contigs using partial order alignment of selective read-segments. As demonstrated on human genome assemblies, Hypo generates significantly more accurate polished assemblies in about one-third time with about half the memory requirements in comparison to Racon (the widely used polisher currently).
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Vulcan: Improved long-read mapping and structural variant calling via dual-mode alignment 96%
- ntsm: an alignment-free, ultra low coverage, sequencing technology agnostic, intraspecies sample comparison tool for sample swap detection 96%
- A graph clustering algorithm for detection and genotyping of structural variants from long reads 96%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.