A reference library for the identification of Canadian invertebrates: 1.5 million DNA barcodes, voucher specimens, and genomic samples
deWaard, J. R.; Ratnasingham, S.; Zakharov, E. V.; Borisenko, A. V.; Steinke, D.; Telfer, A. C.; Perez, K. H. J.; Sones, J. E.; Young, M. R.; Levesque-Beaudin, V.; Sobel, C. N.; Abrahamyan, A.; Bessonov, K.; Blagoev, G.; deWaard, S. L.; Ho, C.; Ivanova, N. V.; Layton, K. K. S.; Lu, L.; Manjunath, R.; McKeown, J. T. A.; Milton, M. A.; Miskie, R.; Monkhouse, N.; Naik, S.; Nikolova, N.; Pentinsaari, M.; Prosser, S. W. J.; Radulovici, A. E.; Steinke, C.; Warne, C. P.; Hebert, P. D. N.
Show abstract
The reliable taxonomic identification of organisms through DNA sequence data requires a well parameterized library of curated reference sequences. However, it is estimated that just 15% of described animal species are represented in public sequence repositories. To begin to address this deficiency, we provide DNA barcodes for 1,500,003 animal specimens collected from 23 terrestrial and aquatic ecozones at sites across Canada, a nation that comprises 7% of the planets land surface. In total, 14 phyla, 43 classes, 163 orders, 1123 families, 6186 genera, and 64,264 Barcode Index Numbers (BINs; a proxy for species) are represented. Species-level taxonomy was available for 38% of the specimens, but higher proportions were assigned to a genus (69.5%) and a family (99.9%). Voucher specimens and DNA extracts are archived at the Centre for Biodiversity Genomics where they are available for further research. The corresponding sequence and taxonomic data can be accessed through the Barcode of Life Data System, GenBank, the Global Biodiversity Information Facility, and the Global Genome Biodiversity Network Data Portal.\n\n\n\nO_FIG O_LINKSMALLFIG WIDTH=200 HEIGHT=61 SRC=\"FIGDIR/small/701805v1_ufig1.gif\" ALT=\"Figure 1\">\nView larger version (17K):\norg.highwire.dtl.DTLVardef@5c6449org.highwire.dtl.DTLVardef@1bc0224org.highwire.dtl.DTLVardef@309dddorg.highwire.dtl.DTLVardef@1cc2cc8_HPS_FORMAT_FIGEXP M_FIG C_FIG
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Diversity-disease relationships in natural microscopic nematode communities 91%
- Circadian Activity Predicts Breeding Phenology in the Asian Burying Beetle Nicrophorus nepalensis 91%
- Gut microbial community in proboscis monkeys (Nasalis larvatus): implications for effects of geographical and social factors 91%
Similar papers in this journal
Similar papers in this journal
- Message in a Bottle, Metabarcoding Enables Biodiversity Comparisons Across Ecoregions 97%
- Chromosome-level genome assembly for the Aldabra giant tortoise enables insights into the genetic health of a threatened population 94%
- Draft genome assemblies using sequencing reads from Oxford Nanopore Technology and Illumina platforms for four species of North American killifish from the Fundulus genus 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.