Self-supervised Component Segmentation To Improve Object Detection and Classification For Bumblebee Identification
Choton, J. C.; Margapuri, V.; Grijalva, I.; Spiesman, B.; Hsu, W. H.
Show abstract
The performance of computer vision models for object detection and classification is heavily influenced by the number of classes and quality of input images, particularly in biological applications such as species-level identification of bumblebees. Bee identification is time-consuming, costly, and requires specialized taxonomic training. Different deep learning based computer vision models have been proven to overcome this methodological bottleneck through automated identification of bee species from captured images. However, accurate identification of bee species in images containing multiple objects of various classes poses significant challenges due to ambiguity, poor image quality, and noisy backgrounds. Existing pipelines (baselines) primarily rely on object detection to crop bees from images and classify the species for each cropped instance. This approach is limited by the inclusion of noisy backgrounds, low resolution, and poor image quality. To address these limitations, we propose an enhanced pipeline that integrates object detection with segmentation to generate body masks for bees and remove background noise. This process is complemented by a classification model that identifies the top k species for each masked image. The proposed methodology significantly improves both detection and classification performance in most cases, demonstrating its potential to advance automated identification of bee species in complex image datasets. For the cases where the baselines performed much better, we investigated using a state-of-the-art explainable AI model (Grad-CAM) to explain the reason.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- SN-FPN: Self-attention Nested Feature Pyramid Network for Digital Pathology Image Segmentation 94%
- Cell segmentation without annotation by unsupervised domain adaptation based on cooperative self-learning 94%
- The tempest in a cubic millimeter: Image-based refinements necessitate the reconstruction of 3D microvasculature from a large series of damaged alternately-stained histological sections 92%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.