Back

Optimizing automated classification for zooplankton in coastal conditions: the impact of model selection, imaging instruments, and colour information

Hovenkamp, P. D. L.; van Walraven, L.; Ollevier, A.; van Oevelen, D.; van der Stappen, A. F.

2026-07-13 bioinformatics
10.64898/2026.07.09.733739 bioRxiv
Show abstract

The advancement in deep learning techniques has made Convolutional Neural Networks (CNNs) a powerful tool for the fully automated classification of zooplankton images. In this study, we systematically investigate how network selection, colour information and differences in imaging instruments affect the classification of zooplankton images by comparing multiple state-of-the-art CNNs on images of zooplankton and marine snow from the in situ Continuous Particle Imaging and Classification Sensor (CPICS), Video Plankton Recorder (VPR), In Situ Ichtyoplankton Imaging System (ISIIS), and the on-board Plankton Imager (Pi-10). With differences between models of 7.8 to 19% in F1-score, we find that model selection strongly affects the classification performance, with EfficientNetV2S showing the most reliable overall performance. Moreover, differences between model architectures are largest for the least abundant classes (<100 labeled images), which implies that when these are present, careful model selection is most beneficial. The high image quality of the Pi-10 strongly increases the performance for the least abundant classes compared to the other instruments. In addition, we find a significant correlation (r = 0.597) between ImageNet the performance and F1-score on zooplankton images, which implies that more generally, a model that performs well on ImageNet will perform well for zooplankton classification. Colour information increases the F1-score of the best performing classifier with 2.8%, but provides a stronger benefit (25% F1-score) for classes with <100 images. The overall performance increase of colour information is less than expected and questions the advantage of recording colour information for zooplankton.

Matching journals

The top 5 journals account for 50% of the predicted probability mass.

1
Scientific Reports
3612 papers in training set
Top 0.6%
19.0%
2
PLOS ONE
5266 papers in training set
Top 9%
19.0%
3
Limnology and Oceanography: Methods
11 papers in training set
Top 0.1%
5.7%
4
Animals
23 papers in training set
Top 0.1%
4.4%
5
Royal Society Open Science
214 papers in training set
Top 1%
3.3%
50% of probability mass above
6
Remote Sensing
10 papers in training set
Top 0.1%
2.5%
7
BMC Bioinformatics
457 papers in training set
Top 4%
1.8%
8
PLOS Computational Biology
1863 papers in training set
Top 14%
1.8%
9
Bioinformatics
1204 papers in training set
Top 7%
1.8%
10
Bioengineering
29 papers in training set
Top 0.4%
1.7%
11
Ecological Informatics
33 papers in training set
Top 0.4%
1.5%
12
PeerJ
308 papers in training set
Top 7%
1.4%
13
IEEE Access
35 papers in training set
Top 0.8%
1.4%
14
Frontiers in Microbiology
427 papers in training set
Top 6%
1.2%
15
Water Research
79 papers in training set
Top 0.9%
1.1%
16
Remote Sensing in Ecology and Conservation
14 papers in training set
Top 0.2%
1.1%
17
Applied Sciences
25 papers in training set
Top 0.6%
1.1%
18
The Journal of the Acoustical Society of America
35 papers in training set
Top 0.3%
1.1%
19
Sensors
43 papers in training set
Top 1%
1.1%
20
Frontiers in Marine Science
62 papers in training set
Top 0.9%
1.0%
21
Environmental Science & Technology
64 papers in training set
Top 1%
0.9%
22
Biomedical Optics Express
95 papers in training set
Top 1.0%
0.9%
23
Scientific Data
209 papers in training set
Top 3%
0.9%
24
GigaScience
212 papers in training set
Top 4%
0.9%
25
Journal of Experimental Botany
219 papers in training set
Top 3%
0.9%
26
Journal of Animal Ecology
75 papers in training set
Top 2%
0.9%
27
Computational and Structural Biotechnology Journal
242 papers in training set
Top 6%
0.9%
28
Light: Science & Applications
16 papers in training set
Top 0.3%
0.9%