Back

Targeted tiled amplicon based protocol for sequencing the Hemagglutinin (HA) gene segment of seasonal influenza A and influenza B virus from wastewater at high depth of coverage

Hetherington-Rauth, M. C.; Nguyen, V.; Lequia, G.; Pysnack, N.; Bankers, L. A.; Rossheim, A. E.; Matzinger, S. R.

2025-10-17 public and global health
10.1101/2025.10.15.25338105 medRxiv
Show abstract

Wastewater based epidemiology has emerged as a compelling tool to monitor the spread and evolution of pathogens of public health concern. Next generation sequencing (NGS) of pathogens detected in wastewater enables sequence characterization which is essential for monitoring the changing genomic landscape of pathogens. Influenza virus, with its potential to cause epidemics and pandemics, poses a serious risk to human health. Genomic surveillance of the virus is essential for safeguarding public health by monitoring the viruss evolution, and developing preventive vaccines and therapeutics. As such, methods for successfully sequencing influenza virus at a consistent high depth of coverage across the genome can aid in this endeavor. Here, we present a novel targeted tiled amplicon based sequencing protocol that uses short tiled amplicons (<250 bp in length) to successfully capture the Hemagglutinin (HA) gene segment of seasonal influenza A subtypes (H1 and H3) and Influenza B at high depth of coverage. We observed near consistent coverage across the HA gene segment for wastewater samples that had influenza viral target dPCR detections of at least 103 copies/L. We were able to successfully detect low frequency single nucleotide variants (SNVs) at high depth of coverage demonstrating the utility of the data to characterize the diversity of circulating influenza A and B viruses at the community level. Our approach is flexible and future directions include expanding this approach to sequence additional influenza virus HA subtypes and gene segments.

Matching journals

The top 3 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.