Pangenomic read mapping
Sheikhizadeh Anari, S.; de Ridder, D.; Schranz, M. E.; Smit, S.
Show abstract
In modern genomics, mapping reads to a single reference genome is common practice. However, a reference genome does not necessarily accurately represent a population or species and as a result a substantial percentage of reads often cannot be mapped. A number of graph-based variation-aware mapping methods have recently been proposed to remedy this. Here, we propose an alternative multi-reference approach, which aligns reads to large collections of genomes simultaneously. Our approach, an extension to our pangenomics suite PanTools (https://git.wur.nl/bioinformatics/pantools), is as accurate as state-of the-art tools but more efficient on large numbers of genomes. We successfully applied PanTools to map genomic and metagenomic reads to large collections of viral, archaeal, bacterial, fungal and plant genomes.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Read-SpaM: assembly-free and alignment-free comparison of bacterial genomes with low sequencing coverage 96%
- HapSolo: An optimization approach for removing secondary haplotigs during diploid genome assembly and scaffolding. 96%
- MTG-Link: leveraging barcode information from linked-reads to assemble specific loci 96%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.