Back

SituSeq: An offline protocol for rapid and remote Nanopore amplicon sequence analysis

Zorz, J.; Li, C.; Chakraborty, A.; Gittins, D.; Surcon, T.; Morrison, N.; Bennett, R.; MacDonald, A.; Hubert, C. R.

2022-10-18 microbiology
10.1101/2022.10.18.512610 bioRxiv
Show abstract

Microbiome analysis through 16S rRNA gene sequencing is a crucial tool for understanding the microbial ecology of any habitat or ecosystem. However, workflows require large equipment, stable internet, and extensive computing power such that most of the work is performed far away from sample collection in both space and time. Performing amplicon sequencing and analysis at sample collection would have positive implications in many instances including remote fieldwork and point-of-care medical diagnoses. Here we present SituSeq, an offline and portable workflow for the sequencing and analysis of 16S rRNA gene amplicons using the Nanopore MinION and a standard laptop computer. SituSeq was validated using the same environmental DNA to sequence Nanopore 16S rRNA gene amplicons, Illumina 16S rRNA gene amplicons, and Illumina metagenomes. Comparisons revealed consistent community composition, ecological trends, and sequence identity across platforms. Correlation between the abundance of phyla in Illumina and Nanopore data sets was high (Pearsons r = 0.9), and over 70% of Illumina 16S rRNA gene sequences matched a Nanopore sequence with greater than 97% sequence identity. On board a research vessel on the open ocean, SituSeq was used to analyze amplicon sequences from deep sea sediments less than two hours after sequencing, and eight hours after sample collection. The rapidly available results informed decisions about subsequent sampling in near real-time while the offshore expedition was still underway. SituSeq is a portable and robust workflow that helps to bring the power of microbial genomics and diagnostics to many more researchers and situations.

Matching journals

The top 6 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.