Pre-epidemic evolution of the USA300 clade and a molecular key for classification
Bianco, C.; Moustafa, A.; OBrien, K.; Martin, M.; Read, T.; Kreiswirth, B. N.; Planet, P.
Show abstract
USA300 has remained the dominant community and healthcare associated methicillin-resistant Staphylococcus aureus (MRSA) clone in the United States and in northern South America for at least the past 20 years. In this time, it has experienced epidemic spread in both of these locations. However, its pre-epidemic evolutionary history and origins are incompletely understood. Large sequencing databases, such as NCBI, PATRIC, and Staphopia, contain clues to the early evolution of USA300 in the form of sequenced genomes of USA300 isolates that are representative of lineages that diverged prior to the establishment of the South American (SAE) and North American (NAE) epidemics. In addition, historical isolates collected prior to the emergence of epidemics can help reconstruct early events in the history of this lineage. Here, we take advantage of the accrued, publicly available data, as well as two newly sequenced pre-epidemic historical isolates from 1996, and a very early diverging ACME-negative NAE genome to understand the pre-epidemic evolution of USA300. We use database mining techniques to emphasize genomes similar to pre-epidemic isolates, with the goal of reconstructing the early molecular evolution of the USA300 lineage. Phylogenetic analysis with these genomes confirms that the North American Epidemic and South American Epidemic USA300 lineages diverged from a most recent common ancestor around 1970 with high confidence, and it also pinpoints the independent acquisition events of the of the ACME and COMER loci with greater precision than in previous studies. We solidify evidence for a North American origin of the USA300 lineage and identify multiple introductions of USA300 into South America from North America. Notably, we describe a third major USA300 clade (the pre-epidemic branching clade; PEB1) consisting of both MSSA and MRSA isolates circulating around the world that diverged from the USA300 lineage prior to the establishment of the South American and North American epidemics. We present a detailed analysis of specific sequence characteristics of each of the major clades, and present diagnostic positions that can be used to classify new genomes.
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Adaptative and ancient co-evolution of integrons with Xanthomonas genomes 96%
- Large contribution of repeats to genetic variation in a transmission cluster of Mycobacterium tuberculosis 96%
- Genomic Insights into the Diversity, Antimicrobial Resistance, and Zoonotic Potential of Campylobacter fetus Across Diverse Hosts and Geographies 96%
Similar papers in this journal
- Global emergence and dissemination of Neisseria gonorrhoeae ST-9363 isolates with reduced susceptibility to azithromycin 96%
- Comparative genomics reveals multipartite genomes undergoing loss in the fungal endosymbiotic genus Mycetohabitans 96%
- Evolutionary route of resistant genes in Staphylococcus aureus 95%
Similar papers in this journal
- No Assembly Required: Using BTyper3 to Assess the Congruency of a Proposed Taxonomic Framework for the Bacillus cereus group with Historical Typing Methods 96%
- High-throughput nanopore sequencing of Treponema pallidum tandem repeat genes arp and tp0470 reveals clade-specific patterns and recapitulates global whole genome phylogeny 96%
- Monitoring the Antimicrobial Resistance Dynamics of Salmonella enterica in Healthy Dairy Cattle Populations at the Individual Farm Level Using Whole-Genome Sequencing 95%
Similar papers in this journal
- A co-speciation dilemma and a lifestyle transition with genomic consequences in Wolbachia of Neotropical Drosophila 94%
- In silico secretome characterization of clinical Mycobacterium abscessus isolates provides insights into antigenic differences 94%
- Pango lineage designation and assignment using SARS-CoV-2 spike gene nucleotide sequences 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.