Temporal coding enables hyperacuity in event based vision
Assa, E.; Rivkind, A.; Kreiserman, M.; Khan, F. S.; Khan, S.; Ahissar, E.
Show abstract
The fact that the eyes are constantly in motion, even during fixation, entails that the spike times of retinal outputs carry information about the visual scene even when the scene is static. Moreover, this motion implies that fine details of the visual scene could not be decoded from pure spatial retinal representations due to smearing. Understanding the interplay of temporal and spatial information in visual processing is thus pivotal for both biological research and bio-inspired computer-vision applications. In this study, we consider data from a popular event-based camera that was designed to emulate the function of a biological retina in hardware. Similarly to biological eye, and in contrast to standard frame-based cameras, this camera outputs an asynchronous sequence of "spike" events. We used this camera to obtain dataset of event streams of tiny images, i.e., images whose recognition is impaired by photosensors pixelization and thus their recognition requires hyperacuity. Using these datasets we demonstrate here the superiority of event-based spatio-temporal coding over frame-based spatial coding in the recognition of tiny images by artificial neural networks (ANNs). We further demonstrate the benefits of event sequences for unsupervised learning. Interestingly, Vernier hyperacuity, which is a standard measure of shape hyperacuity, emerged in ANNs following training on tiny images, resembling the natural hyperacuity observed in humans. Our findings underscore the essential role of precise temporal information in visual processing, offering insights for advancing both biological understanding and bio-inspired engineering of visual perception.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- LiftPose3D, a deep learning-based approach for transforming 2D to 3D pose in laboratory animals 96%
- Lightning Pose: improved animal pose estimation via semi-supervised learning, Bayesian ensembling, and cloud-native open-source tools 96%
- A-SOiD, an active learning platform for expert-guided, data efficient discovery of behavior 95%
Similar papers in this journal
- DAMM for the detection and tracking of multiple animals within complex social and environmental settings 96%
- DeepAction: A MATLAB toolbox for automated classification of animal behavior in video 96%
- Generalising uncertainty improves accuracy and safety of deep learning analytics applied to oncology 94%
Similar papers in this journal
- Emergence of brain-like mirror-symmetric viewpoint tuning in convolutional neural networks 96%
- Deep Neural Networks to Register and Annotate Cells in Moving and Deforming Nervous Systems 95%
- CEM500K - A large-scale heterogeneous unlabeled cellular electron microscopy image dataset for deep learning. 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.