WIO-ReefFish: A High-Resolution Dataset for Taxon-Aware Coral Reef Fish Detection in the Western Indian Ocean
Gerard, J.; Branger, L.; Huyghe, F.; Kochzius, M.; Otwoma, L.; Bergacker, S.; op't Roodt, L.; Rumisha, c.; Di Bella, L.
Show abstract
Coral reef fish assemblages are widely used as indicators of ecosystem condition, yet manual annotation of underwater video remains a major bottleneck for scalable biodiversity monitoring. Despite rapid progress in automated detection, ecologically realistic and publicly available datasets remain scarce, particularly for the Western Indian Ocean. Here, we present WIO-ReefFish, a reef fish detection dataset derived from diver-operated line-intercept transects and designed for ecological monitoring under natural survey conditions. WIO-ReefFish comprises 1,000 ultra-high-definition images (3840 $\times$ 2160 pixels) and 6,768 exhaustive bounding-box annotations spanning 24 taxonomic categories, thereby preserving full-frame assemblage structure in complex reef scenes. We also establish a standardized benchmark across nine object detection models under two complementary protocols: class-aware detection and class-agnostic fish localization. Detection performance was consistently higher under the class-agnostic protocol. The best-performing model (RT-DETR) improved from 0.48 mAP50 in the class-aware setting to 0.70 mAP50 when taxonomic constraints were removed, indicating that taxonomic discrimination remains substantially more challenging than fish localisation in reef imagery. Spatially independent evaluation revealed a pronounced generalisation gap, particularly for taxonomic detection, whereas class-agnostic fish localisation remained substantially more robust across transects and countries. Together, these results establish WIO-ReefFish as a realistic benchmark for automated reef fish detection and provide a foundation for more robust computer-vision tools in coral reef biodiversity monitoring. The WIO-ReefFish dataset and associated benchmarking resources are publicly available.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Semi-Automated Seal Detection on the Western Antarctic Peninsula: An Unsupervised Machine Learning Approach for Detecting Ice Seals in Aerial Survey Data 95%
- Using very high-resolution satellite imagery and deep learning to detect and count African elephants in heterogeneous landscapes 92%
- Robotics-assisted acoustic surveys could deliver reliable, landscape-level biodiversity insights 92%
Similar papers in this journal
- Using machine learning to count Antarctic shag (Leucocarbo bransfieldensis) nests on images captured by Remotely Piloted Aircraft Systems 94%
- Exploiting facial side similarities to improve AI-driven sea turtle photo-identification systems 93%
- Comparison of Two Individual Identification Algorithms for Snow Leopards (Panthera uncia) after Automated Detection 92%
Similar papers in this journal
- AquaX: An enhanced and revised AquaMaps framework to model marine species distributions and biodiversity 93%
- Accuracy and precision of citizen scientist animal counts from drone imagery 93%
- Improving deep learning-based segmentation of diatoms in gigapixel-sized virtual slides by object-based tile positioning and object integrity constraint 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.