ViralConsensus: A fast and memory-efficient tool for calling viral consensus genome sequences directly from read alignment data
Moshiri, N.
Show abstract
MotivationIn viral molecular epidemiology, reconstruction of consensus genomes from sequence data is critical for tracking mutations and variants of concern. However, as the number of samples that are sequenced grows rapidly, compute resources needed to reconstruct consensus genomes can become prohibitively large. ResultsViralConsensus is a fast and memory-efficient tool for calling viral consensus genome sequences directly from read alignment data. ViralConsensus is orders of magnitude faster and more memory-efficient than existing methods. Further, unlike existing methods, ViralConsensus can pipe data directly from a read mapper via standard input and performs viral consensus calling on-the-fly, making it an ideal tool for viral sequencing pipelines. AvailabilityViralConsensus is freely available at https://github.com/niemasd/ViralConsensus as an open-source software project. Contactniema@ucsd.edu Supplementary informationSupplementary data are available online.
Matching journals
The top 1 journal accounts for 50% of the predicted probability mass.
Similar papers in this journal
- V-pipe: a computational pipeline for assessing viral genetic diversity from high-throughput sequencing data 97%
- ViralMSA: Massively scalable reference-guided multiple sequence alignment of viral genomes 96%
- ONTdeCIPHER: An amplicon-based nanopore sequencing pipeline for tracking pathogen variants 95%
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- Choice of assemblers has a critical impact on de novo assembly of SARS-CoV-2 genome and characterizing variants 92%
- A Computational Toolset for Rapid Identification of SARS-CoV-2, other Viruses, and Microorganisms from Sequencing Data 92%
- DeepSSV: detecting somatic small variants in paired tumor and normal sequencing data with convolutional neural network 91%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.