Back

Long-range PCR amplification and nanopore sequencing of 8-10 kb mitochondrial fragments from environmental DNA

Matthews, S.; Scott, O. M.; Ip, A. Y.; Allan, E. A.; Kelly, R. P.

2026-01-14 bioinformatics
10.64898/2026.01.13.699327 bioRxiv
Show abstract

O_LILong-read sequencing data can provide increased taxonomic resolution and genomic linkage information that is otherwise impossible to obtain from short-read amplicon or shotgun sequencing. However, the use of long-read sequencing for environmental DNA (eDNA) analysis has thus far been limited by both the apparent rarity of long DNA molecules in eDNA samples and the lack of established bioinformatics workflows for long-read metabarcoding, particularly from mixed template samples. C_LIO_LIHere, we report nanopore sequencing of 8.0 - 9.5 kb mitochondrial fragments obtained from mesocosm and field eDNA samples, amplified with long-range PCR (LR-PCR) using primers designed to preferentially amplify teleost mitogenomes. C_LIO_LIWe recovered half-mitochondria from 13 fish species in field-collected samples (Puget Sound in Washington State, USA), as well as from approximately half of the fish species inhabiting the positive control mesocosm (the Seattle Aquarium). Among biological replicates, we observed consistent detection of read-abundant species, while there was greater stochasticity in the presence of rarer species. C_LIO_LIWe demonstrate that long fragments can be obtained from standard eDNA samples and successfully amplified and sequenced to obtain species identifications despite higher error rates characteristic of nanopore sequencing. We present both laboratory methods and an accessible bioinformatic pipeline for obtaining and analyzing LR-PCR amplified fragments from eDNA, providing a framework for future long-read metabarcoding studies. C_LI

Matching journals

The top 3 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.