Back

PrimalScheme: open-source community resources for low-cost viral genome sequencing

Kent, C.; Smith, A. D.; Tyson, J.; Stepniak, D.; Kinganda-Lusamaki, E.; Lee, T.; Weaver, M.; Sparks, N.; Landsdowne, L.; Wilkinson, S. A.; Brier, T.; Colquhoun, R.; O'Toole, A. N.; Kingebeni, P. M.; Goodfellow, I. G.; Rambaut, A.; Loman, N. J.; Quick, J.

2024-12-22 genomics
10.1101/2024.12.20.629611 bioRxiv
Show abstract

Viral genome sequencing using the ARTIC protocol has been a vital tool for understanding the spread of epidemics including Ebola, Zika, Covid-19 and Mpox and has seen widespread adoption due to its low cost and high sensitivity. Here, we describe PrimalScheme, an open-source toolkit and website that allows users to easily design primer schemes for amplicon sequencing of viruses, and has generated over 67,000 primer schemes for the global community since 2017. In January 2020, PrimalScheme was used to rapidly generate a primer scheme for SARS-CoV-2, with primer pools distributed to researchers from 44 countries to help scale-up genomic surveillance efforts. Overall, these primers were used to generate an estimated 18M genome sequences and the protocols were viewed online [~]250K times. To complement PrimalScheme, we have built PrimalScheme Labs, a scheme repository which allows users to find and share primer schemes as well as establishing a set of data standards. Through improvements to the primer design process, including the use of discrete primer clouds, we have expanded the use of amplicon sequencing to include diverse virus species. We demonstrate the utility of this approach through a high diversity pan-genotype Measles virus (MeV) scheme. We also demonstrate its use on a high sensitivity, short amplicon Monkeypox virus (MPXV) scheme with over 1000 primers, showing high genome recovery on low-titre clinical samples. These developments have implications for sequencing from samples such as wastewater, for genomic surveillance of endemic pathogens and in preparing for future pandemics.

Matching journals

The top 8 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.