Back

DINOSim: Zero-Shot Object Detection and Semantic Segmentation on Electron Microscopy Images

Gonzalez-Marfil, A.; Gomez-de-Mariscal, E.; Arganda-Carreras, I.

2025-03-13 bioinformatics
10.1101/2025.03.09.642092 bioRxiv
Show abstract

We present DINOSim, a novel method for detecting and segmenting objects in microscopy images without the need for large annotated datasets or additional training. DINOSim builds on the pretrained DINOv2 image encoder, which captures semantic information from images. By comparing the encoders features of images patches to those of a user-selected reference, DINOSim generates pseudo-labels that guide object detection and segmentation. Subsequently, a k-nearest neighbors framework is then used to refine predictions across new images. Our experiments show that DINOSim can effectively identify and segment previously unseen objects in diverse microscopy datasets, offering performance comparable to supervised approaches while avoiding the need for costly manual labeling. We also investigate how different choices of user prompts selection and model size affect accuracy and generalization. To make the method widely accessible, we provide an open-source Napari plugin (github.com/AAitorG/napari-DINOSim), enabling researchers to easily apply DINOSim to their own data. Overall, DINOSim offers a fast, flexible and practical solution for bioimage analysis, particularly valuable in resource-constrained settings.

Matching journals

The top 7 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.