Leveraging Basecaller's Move Table to Generate a Lightweight k-mer Model
Samarakoon, H.; Kei Wan, Y.; Parameswaran, S.; Göke, J.; Gamaarachchi, H.; Deveson, I. W.
Show abstract
Nanopore sequencing by Oxford Nanopore Technologies (ONT) enables direct analysis of DNA and RNA by capturing raw electrical signals. Different nanopore chemistries have varied k-mer lengths, current levels, and standard deviations, which are stored in k-mer models. Particularly in cases where official models are lacking or unsuitable for specific sequencing conditions, tailored k-mer models are crucial to ensure precise signal-to-sequence alignment and interpretation. The process of transforming raw signals into nucleotide sequences, known as basecalling, is a fundamental step in nanopore sequencing. In this study, we leverage the basecallers move table to create a lightweight denovo k-mer model for RNA004 chemistry. We showcase the effectiveness of our custom k-mer model through high alignment rates (97.48%) compared to larger default models. Additionally, our 5-mer model exhibits similar performance as the default 9-mer models in m6A methylation detection.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- LongReadSum: A fast and flexible quality control and signal summarization tool for long-read sequencing data 96%
- UMI-Gen: a UMI-based reads simulator for variant calling evaluation in paired-end sequencing NGS libraries 95%
- Giraffe: a tool for comprehensive processing and visualization of multiple long-read sequencing data 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.