Snappy: de novo identification of DNA methylation sites based on Oxford Nanopore reads
Konanov, D. N.; Krivonos, D. V.; Babenko, V. V.; Ilina, E. N.
Show abstract
SummaryNowadays, search for methylation sites in bacteria is usually performed by direct detection of nucleotide motifs over-represented in modified contexts, using classical motif enrichment approaches oriented only on context sequences themselves. Herein, we present a new algorithm Snappy, which is actually rethinking of the original Snapper algorithm but does not use any enrichment heuristics and does not require control sample sequencing. Opposite to previous methods, Snappy uses raw basecalling data simultaneously with the motif enrichment process, thus significantly enhancing the enrichment sensitivity and accuracy compared with other enrichment algorithms. The versatility of the method was shown on both our and external data, representing different bacterial taxa with complex and simple methylome. Availability and implementationSource code and documentation is hosted on GitHub (https://github.com/DNKonanov/ont-snappy) and Zenodo (zenodo.org/records/16731817). For accessibility, Snappy is installable from PyPi using pip install ont-snappy command.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Snapper: a high-sensitive algorithm to detect methylation motifs based on Oxford Nanopore reads 97%
- AdDeam: A Fast and Scalable Tool for Estimating and Clustering Reference-Level Damage Profiles 94%
- MethPanel: a parallel pipeline and interactive analysis tool for multiplex bisulphite PCR sequencing to assess DNA methylation biomarker panels for disease detection 94%
Similar papers in this journal
Similar papers in this journal
- kASA: Taxonomic Analysis of Metagenomic Data on a Notebook 94%
- LINbase: A Web service for genome-based identification of microbes as members of crowdsourced taxa 93%
- Rapid Identification of Methylase Specificity (RIMS-seq) jointly identifies methylated motifs and generates shotgun sequencing of bacterial genomes 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.