MAGqual: A standalone pipeline to assess the quality of metagenome-assembled genomes
Cansdale, A.; Chong, J. P. J.
Show abstract
Metagenomics, the whole genome sequencing of microbial communities, has provided insight into complex ecosystems. It has facilitated the discovery of novel microorganisms, explained community interactions, and found applications in various fields. Advances in high-throughput and third-generation sequencing technologies have further fuelled its popularity. Nevertheless, managing the vast data produced and addressing variable dataset quality remain ongoing challenges. Another challenge arises from the number of assembly and binning strategies used across studies. Comparing datasets and analysis tools is complex as it requires a measure of metagenome quality. The inherent limitations of metagenomic sequencing, which often involves sequencing complex communities means community members are challenging to interrogate with traditional culturing methods leading to many lacking reference sequences. The MIMAG standards (Bowers et al., 2017) aim to provide a method to assess metagenome quality for comparison but have not been widely adopted. To bridge this gap, the MAGqual pipeline outlined here offers an accessible way to evaluate metagenome quality and generate metadata on a large scale. MAGqual is built in Snakemake to ensure readability and scalability and its open-source nature promotes accessibility, community development, and ease of updates. Here, we introduce the pipeline MAGqual (metagenome-assembled genome qualifier) and demonstrate its effectiveness at determining metagenomic dataset quality when compared to the MIMAG standards. MAGqual is built in Snakemake, R, and Python and is available under the MIT License on GitHub at https://github.com/ac1513/MAGqual.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- MerCat2: a versatile k-mer counter and diversity estimator for database-independent property analysis obtained from omics data 97%
- BugBuster: A novel automatic and reproducible workflow for metagenomic data analysis 96%
- ExplorePipolin: reconstruction and annotation of bacterial mobile elements from draft genomes 94%
Similar papers in this journal
- SQMtools: automated processing and visual analysis of 'omics data with R and anvi'o 97%
- Persistent Memory as an Effective Alternative to Random Access Memory in Metagenome Assembly 97%
- Natrix: A Snakemake-based workflow for processing, clustering, and taxonomically assigning amplicon sequencing reads 95%
Similar papers in this journal
Similar papers in this journal
- TaxonTableTools - A comprehensive, platform-independent graphical user interface software to explore and visualise DNA metabarcoding data 95%
- In-situ metagenomics: A platform for rapid sequencing and analysis of metagenomes in less than one day 94%
- debar, a sequence-by-sequence denoiser for COI-5P DNA barcode data 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.