Back

Time-series sewage metagenomics can separate the seasonal, human-derived and environmental microbial communities, holding promise for source-attributed surveillance

Becsei, A.; Fuschi, A.; Otani, S.; Kant, R.; Weinstein, I.; Alba, P.; Steger, J.; Visontai, D.; Brinch, C.; de Graaf, M.; Schapendonk, C. M. E.; Battisti, A.; De Cesare, A.; Oliveri, C.; Troja, F.; Vapalahti, O.; Pasquali, F.; Banyai, K.; Mako, M.; Pollner, P.; Merlotti, A.; Koopmans, M.; Csabai, I.; Remondini, D.; Aarestrup, F. M.; Munk, P.

2024-05-31 microbiology
10.1101/2024.05.30.596588 bioRxiv
Show abstract

Sewage metagenomics has risen to prominence in urban population surveillance of pathogens and antimicrobial resistance (AMR). Unknown species with similarity to known genomes cause database bias in reference-based metagenomics. To improve surveillance, we designed this study to recover sewage genomes and develop a quantification and correlation workflow for these genomes and AMR over time. We used longitudinal sewage sampling in seven treatment plants from five major European cities to explore the utility of catch-all sequencing of these population-level samples. Using metagenomic assembly methods, we recovered 2,332 metagenome-assembled genomes (MAGs) from prokaryotic species, 1,334 of which were previously undescribed. These genomes account for [~]69% of sequenced DNA and provide insight into sewage microbial dynamics. Rotterdam (Netherlands) and Copenhagen (Denmark) showed strong seasonal microbial community shifts, while Bologna, Rome, (Italy) and Budapest (Hungary) had occasional blooms of Pseudomonas-dominated communities, accounting for up to [~]95% of sample DNA. Seasonal shifts and blooms present challenges for effective sewage surveillance. We find that bacteria of known shared origin, like human gut microbiota, form communities, suggesting the potential for source-attributing novel species and their ARGs through network community analysis. This could significantly improve AMR tracking in urban environments.

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.