Accurate Microbiome Sequencing with Synthetic Long Read Sequencing
Chung, N.; Van Goethem, M. W.; Preston, M. A.; Lhota, F.; Cerna, L.; Garcia-Pichel, F.; Fernandes, V. M. C.; Giraldo-Silva, A.; Kim, H. S.; Hurowitz, E.; Balamotis, M.; Wu, I.; Ben Yehezkel, T.
Show abstract
The microbiome plays a central role in biochemical cycling and nutrient turnover of most ecosystems. Because it can comprise myriad microbial prokaryotes, eukaryotes and viruses, microbiome characterization requires high-throughput sequencing to attain an accurate identification and quantification of such co-existing microbial populations. Short-read next-generation-sequencing (srNGS) revolutionized the study of microbiomes and remains the most widely used approach, yet read lengths spanning only a few of the nine hypervariable regions of the 16S rRNA gene limit phylogenetic resolution leading to misclassification or failure to classify in a high percentage of cases. Here we evaluate a synthetic long-read (SLR) NGS approach for full-length 16S rRNA gene sequencing that is high-throughput, highly accurate and low-cost. The sequencing approach is amenable to highly multiplexed sequencing and provides microbiome sequence data that surpasses existing short and long-read modalities in terms of accuracy and phylogenetic resolution. We validated this commercially-available technology, termed LoopSeq, by characterizing the microbial composition of well-established mock microbiome communities and diverse real-world samples. SLR sequencing revealed differences in aquatic community complexity associated with environmental gradients, resolved species-level community composition of uterine lavage from subjects with histories of misconception and accurately detected strain differences, multiple copies of the 16S rRNA in a single strains genome, as well as low-level contamination in soil cyanobacterial cultures. This approach has implications for widespread adoption of high-resolution, accurate long-read microbiome sequencing as it is generated on popular short read sequencing platforms without the need for additional infrastructure.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Decomposing a San Francisco Estuary microbiome using long read metagenomics reveals species and species- and strain-level dominance from picoeukaryotes to viruses 96%
- Evaluation of the effects of library preparation procedure and sample characteristics on the accuracy of metagenomic profiles 96%
- Carbon Assimilation Strategies in Ultrabasic Groundwater: Clues from the Integrated Study of a Serpentinization-Influenced Aquifer 96%
Similar papers in this journal
- Unlocking River Biofilm Microbial Diversity: A Comparative Analysis of Sequencing Technologies 97%
- Benchmarking long-read sequencing strategies for obtaining ASV-resolved rRNA operons from environmental microeukaryotes 97%
- High molecular weight DNA extraction strategies for long-read sequencing of complex metagenomes 97%
Similar papers in this journal
- An improved hgcAB primer set and direct high throughput sequencing expand Hg-methylator diversity in nature 96%
- Strain-level variation influences Methanothermococcus fitness in marine hydrothermal systems 96%
- Gut Microbiome Dynamics and Predictive Value in Hospitalized COVID-19 Patients: A Comparative Analysis of Shallow and Deep Shotgun Sequencing 96%
Similar papers in this journal
- Ultra-accurate Microbial Amplicon Sequencing with Synthetic Long Reads 98%
- Machine learning algorithm to characterize antimicrobial resistance associated with the International Space Station surface microbiome 96%
- Time-series metagenomics reveals changing protistan ecology of a temperate dimictic lake 95%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.