Sarand: Exploring Antimicrobial Resistance Gene Neighborhoods in Complex Metagenomic Assembly Graphs
Kafaie, S.; Beiko, R. G.; Maguire, F.
Show abstract
Antimicrobial resistance (AMR) is a major global challenge to human and animal health. The genomic element (e.g., chromosome, plasmid, and genomic islands) and neighbouring genes associated with an AMR gene play a major role in its function, regulation, evolution, and propensity to undergo lateral gene transfer. Therefore, characterising these genomic contexts is vital to effective AMR surveillance, risk assessment, and stewardship. Metagenomic sequencing is widely used to identify AMR genes in microbial communities, but analysis of short-read data offers fragmentary information that lacks this critical contextual information. Alternatively, metagenomic assembly, in which a complex assembly graph is generated and condensed into contigs, provides some contextual information but systematically fails to recover many mobile genetic elements. Here we introduce Sarand, a method that combines the sensitivity of read-based methods with the genomic context offered by assemblies by extracting AMR genes and their associated context directly from metagenomic assembly graphs. Sarand combines BLAST-based homology searches with coverage statistics to sensitively identify and visualise AMR gene contexts while minimising inference of chimeric contexts. Using both real and simulated metagenomic data, we show that Sarand outperforms metagenomic assembly and recently developed graph-based tools in terms of precision and sensitivity for this problem. Sarand (https://github.com/beiko-lab/sarand) enables effective extraction of metagenomic AMR gene contexts to better characterize AMR evolutionary dynamics within complex microbial communities.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Targeted genome mining with GATOR-GC maps the evolutionary landscape of biosynthetic diversity 96%
- Unveiling the Microbial Realm with VEBA 2.0: A modular bioinformatics suite for end-to-end genome-resolved prokaryotic, (micro)eukaryotic, and viral multi-omics from either short- or long-read sequencing 96%
- BGCFlow: Systematic pangenome workflow for the analysis of biosynthetic gene clusters across large genomic datasets 95%
Similar papers in this journal
Similar papers in this journal
- Evaluation of taxonomic classification and profiling methods for long-read shotgun metagenomic sequencing datasets 95%
- Functional Analysis of Metagenomes by Likelihood Inference (FAMLI) Successfully Compensates for Multi-Mapping Short Reads from Metagenomic Samples 95%
- PlasForest: a homology-based random forest classifier for plasmid detection in genomic datasets 95%
Similar papers in this journal
- Whokaryote: distinguishing eukaryotic and prokaryotic contigs in metagenomes based on gene structure 96%
- From defaults to databases: parameter and database choice dramatically impact the performance of metagenomic taxonomic classification tools 96%
- Coinfinder: Detecting Significant Associations and Dissociations in Pangenomes 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.