Noise-robust recognition of objects by humans and deep neural networks
Jang, H.; McCormack, D. R.; Tong, F.
Show abstract
Deep neural networks (DNNs) for object classification have been argued to provide the most promising model of the visual system, accompanied by claims that they have attained or even surpassed human-level performance. Here, we evaluated whether DNNs provide a viable model of human vision when tested with challenging noisy images of objects, sometimes presented at the very limits of visibility. We show that popular state-of-the-art DNNs perform in a qualitatively different manner than humans - they are unusually susceptible to spatially uncorrelated white noise and less impaired by spatially correlated noise. We implemented a noise-training procedure to determine whether noise-trained DNNs exhibit more robust responses that better match human behavioral and neural performance. We found that noise-trained DNNs provide a better qualitative match to human performance; moreover, they reliably predict human recognition thresholds on an image-by-image basis. Functional neuroimaging revealed that noise-trained DNNs provide a better correspondence to the pattern-specific neural representations found in both early visual areas and high-level object areas. A layer-specific analysis of the DNNs indicated that noise training led to broad-ranging modifications throughout the network, with greater benefits of noise robustness accruing in progressively higher layers. Our findings demonstrate that noise-trained DNNs provide a viable model to account for human behavioral and neural responses to objects in challenging noisy viewing conditions. Further, they suggest that robustness to noise may be acquired through a process of visual learning.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- The representational hierarchy in human and artificial visual systems in the presence of object-scene regularities 96%
- Depth in convolutional neural networks solves scene segmentation 96%
- Increasing neural network robustness improves match to macaque V1 eigenspectrum, spatial frequency preference and predictivity 96%
Similar papers in this journal
- Brain-Guided Convolutional Neural Networks Reveal Task-Specific Representations in Scene Processing 97%
- Orthogonal Representations of Object Shape and Category in Deep Convolutional Neural Networks and Human Visual Cortex 95%
- Orthogonal neural representations support perceptual judgements of natural stimuli 94%
Similar papers in this journal
- Understanding transformation tolerant visual object representations in the human brain and convolutional neural networks 94%
- Voxel-to-voxel predictive models reveal unexpected structure in unexplained variance 94%
- Predicting the retinotopic organization of human visual cortex from anatomy using geometric deep learning 94%
Similar papers in this journal
- Small hand-designed convolutional neural networks outperform transfer learning in automated cell shape detection in confluent tissues 94%
- Deep Learning Classification of Lipid Droplets in Quantitative Phase Images 94%
- Fine-tuning TrailMap: The utility of transfer learning to improve the performance of deep learning in axon segmentation of light-sheet microscopy images 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.