SaLTy: a novel Staphylococcus aureus Lineage Typer
Cheney, L.; Payne, M.; Kaur, S.; Lan, R.
Show abstract
Staphylococcus aureus asymptomatically colonises 30% of humans and in 2017 was associated with 20,000 deaths in the USA alone. Dividing S. aureus into smaller sub-groups can reveal the emergence of distinct sub-populations with varying potential to cause infections. Despite multiple molecular typing methods categorising such sub-groups, they do not take full advantage of S. aureus WGS when describing the fundamental population structure of the species. In this study, we developed Staphylococcus aureus Lineage Typing (SaLTy), which rapidly divides the species into 61 phylogenetically congruent lineages. Alleles of three core genes were identified that uniquely define the 61 lineages and were used for SaLTy typing. SaLTy was validated on 5,000 genomes and 99.12% (4,956/5,000) of isolates were assigned the correct lineage. We compared SaLTy lineages to previously calculated clonal complexes (CCs) from BIGSdb (n=21,173). SALTy improves on CCs by grouping isolates congruently with phylogenetic structure. SaLTy lineages were further used to describe the carriage of Staphylococcal chromosomal cassette containing mecA (SCCmec) which is carried by methicillin-resistant S. aureus (MRSA). Most lineages had isolates lacking SCCmec and the four largest lineages varied in SCCmec over time. Classifying isolates into SaLTy lineages, which were further SCCmec typed, allowed SaLTy to describe high-level MRSA epidemiology We provide SALTy as a simple typing method that defines phylogenetic lineages (https://github.com/LanLab/SaLTy). SALTy is highly accurate and can quickly analyse large amounts of S. aureus WGS. SALTy will aid the characterisation of S. aureus populations and the ongoing surveillance of sub-groups that threaten human health.
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Large-scale comparative genomics of Salmonella enterica to refine the organization of the global Salmonella population structure 95%
- Diverse Genetic Determinants of Nitrofurantoin Resistance in UK Escherichia coli 95%
- Resolving plasmid-encoded carbapenem resistance dynamics and reservoirs in a hospital setting through nanopore sequencing 95%
Similar papers in this journal
- Small pangenome of Candida parapsilosis reflects overall low intraspecific diversity 96%
- The emergence of successful Streptococcus pyogenes lineages through convergent pathways of capsule loss and recombination directing high toxin expression 96%
- Transposon-sequencing across multiple Mycobacterium abscessus isolates reveals significant functional genomic diversity among strains 95%
Similar papers in this journal
- Species-wide phylogenomics of the Staphylococcus aureus agr operon reveals convergent evolution of frameshift mutations 97%
- Development of an amplicon nanopore sequencing strategy for detection of mutations conferring intermediate resistance to vancomycin in Staphylococcus aureus strains 95%
- Genetic Determinants Underlying the Progressive Phenotype of Beta-lactam/Beta-lactamase Inhibitor Resistance in Escherichia coli 95%
Similar papers in this journal
- Achromobacter xylosoxidans isolates exhibit genome diversity, variable virulence, high levels of antibiotic resistance and potential intrahost evolution. 95%
- Diversity and prevalence of Clostridium innocuum in the human gut microbiota 95%
- Genomic epidemiology and evolution of Escherichia coli in wild animals 94%
Similar papers in this journal
- Transmission and antibiotic resistance of Achromobacter in cystic fibrosis 96%
- Population genomic molecular epidemiological study of macrolide resistant Streptococcus pyogenes in Iceland,1995-2016: Identification of a large clonal population with a pbp2x mutation conferring reduced in vitro beta-lactam susceptibility 95%
- Accurate and Reproducible Whole-Genome Genotyping for Bacterial Genomic Surveillance with Nanopore Sequencing Data 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.