Mandatory use of mock communities highlighted by the descriptive comparison of Epi2Me 16S and EMU bioinformatic workflows for full-length 16S rRNA Nanopore sequencing.
Shedleur-Bourguignon, F.; Theriault, W. P.; Thibodeau, A.
Show abstract
Full-length 16S rRNA gene sequencing using Oxford Nanopore Technologies has emerged as a promising approach to improve species-level resolution in microbiota studies. However, the accuracy of taxonomic assignment remains highly dependent on the bioinformatics s used to process Nanopore long-read data. Therefore, the only way to ensure a good level of certainty in obtained results is to use positive controls in the form of mock communities in the experimental designs. In this study, we compared the performance of Epi2Me 16S (using Minimap2 or Kraken2) workflows provided by Oxford Nanopore Technologies and an EMU workflow for full-length 16S rRNA gene analysis. Using a commercial mock community sequenced across multiple Nanopore runs, taxonomic assignment accuracy and reproducibility was evaluated. Epi2Me-Kraken2 exhibited 18 % of incorrect genus-level assignments and failed to identify 3 species present in the mock community. While Epi2Me-Minimap2 achieved an excellent genus-level classification, reporting 9 % of sequences assigned to a genus not in the mock community, species-level assignments were inconsistent for several community members such as Listeria. In contrast, EMU provided accurate and consistent species-level taxonomic profiles, with all species correctly identified while keeping the number of genus absent from the mock community at 1.2%. ImportanceThese results highlight that Epi2Me integrated workflows are not the best option for specie-level taxonomic assignation. More importantly, this paper underscores the importance of routine inclusion of positive controls for microbiota studies, in the form of mock communities, as a critical safeguard for accurate data interpretation. Without the use of a mock community, a paper published would be at risk of reporting wrong observations and inaccurate conclusions.
Matching journals
The top 8 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- centriflaken: an automated data analysis pipeline for assembly and in silico analyses of foodborne pathogens from metagenomic samples 95%
- Omnicrobe, an open-access database of microbial habitats and phenotypes using a comprehensive text mining and data fusion approach 94%
- Precision long-read metagenomics sequencing for food safety by detection and assembly of Shiga toxin-producing Escherichia coli in irrigation water 94%
Similar papers in this journal
- Exploring Taxonomic and Functional Microbiome of Hawaiian Stream and Spring Irrigation Water Systems Using Illumina and Oxford Nanopore Sequencing Platforms 95%
- Improved microbial community characterization of 16S rRNA via metagenome hybridization capture enrichment 95%
- SARS-CoV-2 virus in Raw Wastewater from Student Residence Halls with concomitant 16S rRNA Bacterial Community Structure changes 94%
Similar papers in this journal
Similar papers in this journal
- Full-length 16S rRNA gene amplicon analysis of human gut microbiota using MinION™ nanopore sequencing confers species-level resolution 94%
- Multi-factorial examination of amplicon sequencing workflows from sample preparation to bioinformatic analysis 93%
- rpoB, a promising marker for analyzing the diversity of bacterial communities by amplicon sequencing 93%
Similar papers in this journal
- rRNA Operon Improves Species-Level Classification of Bacteria and Microbial Community Analysis Compared to 16S rRNA 95%
- Machine-learning based detection of adventitious microbes in T-cell therapy cultures using long read sequencing 95%
- Library Preparation and Sequencing Platform Introduce Bias in Metagenomic-Based Characterizations of Microbiomes 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.