Back

Snaq: A Dynamic Snakemake Pipeline for Microbiome data analysis with QIIME2

Mohsen, A.; Chen, Y.-A.; Allendes Osorio, R. S.; Higuchi, C.; Mizuguchi, K.

2022-03-12 bioinformatics
10.1101/2022.03.10.483866 bioRxiv
Show abstract

Optimizing a protocol for 16S microbiome data analysis with QIIME2 is a challenging task for biologists with no programming experience. It involves a multi-step process, and multiple parameters and options that need to be tested and determined. In this article, we describe Snaq, a snakemake pipeline that helps automate and optimize 16S data analysis using QIIME2. Snaq offers an informative file naming system and automatically performs the analysis of a data set by downloading and installing the required databases and classifiers, all through a single command-line instruction. It works natively on Linux and Mac and on Windows through the use of containers, and is potentially extendable by adding new rules. This pipeline will substantially reduce the efforts in sending commands and prevent the confusion caused by the accumulation of analysis results due to testing multiple parameters.

Matching journals

The top 3 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.