Back

Genomic network analysis of an environmental and livestock IncF plasmid population

Matlock, W.; Chau, K. K.; AbuOun, M.; Stubberfield, E.; Barker, L.; Kavanagh, J.; Pickford, H.; Gilson, D.; Smith, R. P.; Gweon, H. S.; Hoosdally, S. J.; Swann, J.; Sebra, R.; Bailey, M. J.; Peto, T. E. A.; Crook, D. W.; Anjum, M. F.; Read, D. S.; Walker, A. S.; Stoesser, N.; Shaw, L. P.; REHAB Consortium,

2020-07-24 bioinformatics
10.1101/2020.07.24.215889 bioRxiv
Show abstract

IncF plasmids are diverse and of great clinical significance, often carrying genes conferring antimicrobial resistance (AMR) such as extended-spectrum {beta}-lactamases, particularly in Enterobacteriaceae. Organising this plasmid diversity is challenging, and current knowledge is largely based on plasmids from clinical settings. Here, we present a network community analysis of a large survey of IncF plasmids from environmental (influent, effluent, and upstream/downstream waterways surrounding wastewater treatment works) and livestock settings. We use a tractable and scalable methodology to examine the relationship between plasmid metadata and network communities. This reveals how niche (sampling compartment and host genera) partition and shape plasmid diversity. We also perform pangenome-style analyses on network communities. We show that such communities define unique combinations of core genes, with limited overlap. Building plasmid phylogenies based on alignments of these core genes, we demonstrate that plasmid accessory function is closely linked to core gene content. Taken together, our results suggest that stable IncF plasmid backbone structures can persist in environmental settings while allowing dramatic variation in accessory gene content that may be linked to niche adaptation. The recent association of IncF plasmids with AMR likely reflects their suitability for rapid niche adaptation.

Matching journals

The top 6 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.