Spatiotemporal structure of SARS-CoV-2 mutational frequencies in wastewater samples from Ontario
Magbor, P.; Wang, W. Z.; Gugan, G.; Olabode, A. S.; Becker, D. G.; Parreira, V. R.; Lawal, O. U.; Fedynak, A.; Zhang, L.; Rizvi, F.; Precious, M.; DeGroot, C. T.; Goodridge, L.; Poon, A. F. Y.
Show abstract
Starting October 2021, the Ontario wastewater surveillance initiative has used next-generation sequencing (NGS) to monitor SARS-CoV-2 RNA in wastewater samples. The fragmented and heterogeneous nature of these data precludes using comparative methods that require full-length genome sequences. In this study, we investigate the utility of the inner product of the vectors of mutation frequencies to quantify the temporal and spatial structure of these data. Raw sequence data were trimmed and mapped to the SARS-CoV-2 reference genome to extract mutation frequencies and coverage statistics. These data were filtered for samples with incomplete metadata, positions with insufficient coverage (>100 reads), or mutations with frequencies below 1%. For every pair of samples, we calculated the inner product D(x, y) of the respective mutation frequency vectors x and y, and normalized by [Formula]. In total, we processed 1,619 samples from October 2021 to June 2023. The average depth was 7,693 reads, with mean coverage of 24,853 nt. A total of 241,078 mutations were detected in these samples. We restricted our analysis to 20 consecutive months with samples from at least one health region per month. A projection of the resulting distance matrix revealed substantial temporal structure largely driven by the rapid spread of variants of concern. Genetic similarity, as quantified by the normalized dot product of mutation frequencies, was significantly negatively correlated with the geographic distance between sampling locations. These results suggest that spatial differentiation in the genomic variation of SARS-CoV-2 among wastewater samples can be measured, even at the relatively small scale of a single province.
Matching journals
The top 8 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Longitudinal metatranscriptomic sequencing of Southern California wastewater representing 16 million people from August 2020-21 reveals widespread transcription of antibiotic resistance genes. 96%
- Combining individual and wastewater whole genome sequencing improves SARS-CoV-2 surveillance 96%
- Metrics to relate COVID-19 wastewater data to clinical testing dynamics 95%
Similar papers in this journal
- Assessing multiplex tiling PCR sequencing approaches for detecting genomic variants of SARS-CoV-2 in municipal wastewater 94%
- Advantages and limits of metagenomic assembly and binning of a giant virus 94%
- Rapid, large-scale wastewater surveillance and automated reporting system enabled early detection of nearly 85% of COVID-19 cases on a University campus 94%
Similar papers in this journal
- Systematic SARS-CoV-2 S Gene Sequencing in Wastewater Samples Enables Early Lineage Detection and Uncovers Rare Mutations in Portugal 96%
- COVID-19 infection dynamics revealed by SARS-CoV-2 wastewater sequencing analysis and deconvolution 95%
- Development of a COVID-19 warning system for neighborhood-scale wastewater-based epidemiology in low incidence situations 95%
Similar papers in this journal
- A resistome survey across hundreds of freshwater bacterial communities reveals the impacts of veterinary and human antibiotics use 94%
- Secrets of the hospital underbelly: patterns of abundance of antimicrobial resistance genes in hospital wastewater vary by specific antimicrobial and bacterial family 94%
- An improved hgcAB primer set and direct high throughput sequencing expand Hg-methylator diversity in nature 93%
Similar papers in this journal
- Differential Decay of Multiple eNA Components from a Cetacean 93%
- Optimization of subsampling, decontamination, and DNA extraction of difficult peat and silt permafrost samples 93%
- Wastewater sequencing from a rural community enables identification of widespread adaptive mutations in a SARS-CoV-2 Alpha variant 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.