CuBi-MeAn Customized Pipeline for Metagenomic Data Analysis
Keshani-Langroodi, S.; Sales, C. M.
Show abstract
1.Whole genome shotgun sequencing is a powerful to study microbial community is a given environment. Metagenomic binning offers a genome centric approach to study microbiomes. There are several tools available to process metagenomic data from raw reads to the interpretation there is still lack of standard approach that can be used to process the metagenomic data step by step. In this study CuBi-MeAn (Customizable Binning and Metagenomic Analysis) create a customizable and flexible processing pipeline, to process the metagenomic data and generate results for further interpretation. This study aims to perform metagenomic binning to enhance taxonomical classification, functional potentials, and interactions among microbial populations in environmental systems. This customized pipeline which is comprised of a series of genomic/metagenomic tools designed to recover better quality results and reliable interpretation of the system dynamics for the given systems. For this reason, a metagenomic data processing pipeline is developed to evaluate metagenomic data from three environmental engineering projects. The use of our pipeline was demonstrated and compared on three different datasets that were of different sizes, from different sequencing platforms, and generated from three different environmental sources. By designing and developing a flexible and customized pipeline, this study has showed how to process large metagenomic data sets with limited resources. This result not only would help to uncover new information from environmental samples, but also, could be applicable to any other metagenomic studies across various disciplines.
Matching journals
The top 9 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Comprehensive benchmarking of metagenomic classification tools for long-read sequencing data 95%
- Persistent Memory as an Effective Alternative to Random Access Memory in Metagenome Assembly 94%
- Natrix: A Snakemake-based workflow for processing, clustering, and taxonomically assigning amplicon sequencing reads 94%
Similar papers in this journal
- Interpretations of microbial community studies are biased by the selected 16S rRNA gene amplicon sequencing pipeline. 95%
- Using QC-Blind for quality control and contamination screening of bacteria DNA sequencing data without reference genome 93%
- SARS-CoV-2 virus in Raw Wastewater from Student Residence Halls with concomitant 16S rRNA Bacterial Community Structure changes 92%
Similar papers in this journal
- Omnicrobe, an open-access database of microbial habitats and phenotypes using a comprehensive text mining and data fusion approach 93%
- Insights on aquatic microbiome of the Indian Sundarbans mangrove areas 93%
- metaVaR: introducing metavariant species models for reference-free metagenomic-based population genomics 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.