Afanc: a Metagenomics Tool for Variant Level Disambiguation of NGS Datasets
Morris, A. V.; Price, A.; Connor, T.
Show abstract
Genomics is amongst the most powerful tools available for mounting a clinical response to infectious disease. The accurate and precise taxonomic evaluation of pathogens is essential when building a picture of pathogenicity, virulence, transmission, and drug resistance. Carrying out such profiling in a high throughput manner necessitates the development of reliable bioinformatic tools. Here we present Afanc, a novel metagenomic profiler which is sensitive down to species and strain level taxa, and capable of elucidating the complex pathogen profile of compound datasets. We compared Afanc against currently available cutting edge profilers using 3 datasets: single species read sets simulated from the full Mycobacteriaceae taxonomic landscape; compound read sets containing multiple Mycobacteriaceae species and variants; and real data covering the majority of the M. tuberculosis lineage taxonomic space. Afanc outperformed all profilers, both generic and Mycobacteriaceae specific, across all tested fields. As a species agnostic profiler, we predict that Afanc will be of great utility when carrying out highly specific and sensitive pathogen profiling of clinical datasets. Such analyses are essential in advising both the clinical response to an individual disease case, and in forming the foundation of epidemiological surveys.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- ganon2: up-to-date and scalable metagenomics analysis 95%
- Metagenomics-Toolkit: The Flexible and Efficient Cloud-Based Metagenomics Workflow featuring Machine Learning-Enabled Resource Allocation 94%
- iLoci: Robust evaluation of genome content and organization for provisional and mature genome assemblies 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.