TAMPA: interpretable analysis and visualization of metagenomics-based taxon abundance profiles
Sarwal, V.; Brito, J.; Mangul, S.; Koslicki, D. J.
Show abstract
Metagenomic taxonomic profiling aims to predict the identity and relative abundance of taxa in a given whole genome sequencing metagenomic sample. A recent surge in computational methods that aim to accurately estimate taxonomic profiles, called taxonomic profilers, have motivated community driven efforts to create standardized benchmarking datasets and platforms, standardized taxonomic profile formats, as well as a benchmarking platform to assess tool performance. While this standardization is essential, there is currently a lack of tools to visualize the standardized output of the many existing taxonomic profilers, and benchmarking studies rely on a single value metrics to compare performance of tools and compare to benchmarking datasets. Here we report the development of TAMPA (Taxonomic metagenome profiling evaluation), a robust and easy-to-use method that allows scientists to easily interpret and interact with taxonomic profiles produced by the many different taxonomic profiler methods beyond the standard metrics used by the scientific community. We demonstrate the unique ability of TAMPA to provide important biological insights into the taxonomic differences between samples otherwise missed by commonly utilized metrics. Additionally, we show that TAMPA can help visualize the output of taxonomic profilers, enabling biologists to effectively choose the most appropriate profiling method to use on their metagenomics data. TAMPA is available on GitHub, Bioconda and Galaxy Toolshed at https://github.com/dkoslicki/TAMPA, and is released under the MIT license.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- getphylo: rapid and automatic generation of multi-locus phylogenetic trees 94%
- Evaluation of taxonomic classification and profiling methods for long-read shotgun metagenomic sequencing datasets 93%
- Functional Analysis of Metagenomes by Likelihood Inference (FAMLI) Successfully Compensates for Multi-Mapping Short Reads from Metagenomic Samples 93%
Similar papers in this journal
- MADRe: Strain-Level Metagenomic Classification Through Assembly-Driven Database Reduction 94%
- To assemble or not to resemble -- A validated Comparative Metatranscriptomics Workflow (CoMW) 93%
- HVRLocator: A Computationally Efficient Tool for Identifying Hypervariable Regions in 16S rRNA Big Datasets 93%
Similar papers in this journal
- zAMP and zAMPExplorer: Reproducible Scalable Amplicon-based Metagenomics Analysis and Visualization 94%
- MerCat2: a versatile k-mer counter and diversity estimator for database-independent property analysis obtained from omics data 94%
- baseLess: Lightweight detection of sequences in raw MinION data 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.