Perceptual Expertise and Attention: An Exploration using Deep Neural Networks
Das, S.; Mangun, G.; Ding, M.
Show abstract
Perceptual expertise and attention are two important factors that enable superior object recognition and task performance. While expertise enhances knowledge and provides a holistic understanding of the environment, attention allows us to selectively focus on task-related information and suppress distraction. It has been suggested that attention operates differently in experts and in novices, but much remains unknown. This study investigates the relationship between perceptual expertise and attention using convolutional neural networks (CNNs), which are shown to be good models of primate visual pathways. Two CNN models were trained to become experts in either face or scene recognition, and the effect of attention on performance was evaluated in tasks involving complex stimuli, such as superimposed images containing superimposed faces and scenes. The goal was to explore how feature-based attention (FBA) influences recognition within and outside the domain of expertise of the models. We found that each model performed better in its area of expertise--and that FBA further enhanced task performance, but only within the domain of expertise, increasing performance by up to 35% in scene recognition, and 15% in face recognition. However, attention had reduced or negative effects when applied outside the models expertise domain. Neural unit-level analysis revealed that expertise led to stronger tuning towards category-specific features and sharper tuning curves, as reflected in greater representational dissimilarity between targets and distractors, which, in line with the biased competition model of attention, leads to enhanced performance by reducing competition. These findings highlight the critical role of neural tuning at single as well as network level neural in distinguishing the effects of attention in experts and in novices and demonstrate that CNNs can be used fruitfully as computational models for addressing neuroscience questions not practical with the empirical methods.
Matching journals
The top 2 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Emergence of Emotion Selectivity in A Deep Neural Network Trained to Recognize Visual Objects 97%
- How well do models of visual cortex generalize to out of distribution samples? 97%
- Arousal state affects perceptual decision-making by modulating hierarchical sensory processing in a large-scale visual system model 96%
Similar papers in this journal
- Leveraging prior concept learning improves ability to generalize from few examples in computational models of human object recognition 94%
- The face module emerged in a deep convolutional neural network selectively deprived of face experience 94%
- Invariance of object detection in untrained deep neural networks 94%
Similar papers in this journal
- Analysis of convolutional neural networks reveals the computational properties essential for subcortical processing of facial expression 96%
- Brain-Guided Convolutional Neural Networks Reveal Task-Specific Representations in Scene Processing 95%
- Orthogonal neural representations support perceptual judgements of natural stimuli 95%
Similar papers in this journal
- Unsupervised alignment reveals structural commonalities and differences in neural representations of natural scenes across individuals and brain areas 95%
- Distinct rich and diverse clubs regulate coarse and fine binocular disparity processing: Evidence from stereoscopic task-based fMRI 94%
- Face familiarity detection with complex synapses 94%
Similar papers in this journal
- Deep neural networks and visuo-semantic models explain complementary components of human ventral-stream representational dynamics 95%
- Sharpened visual memory representations are reflected in inferotemporal cortex 95%
- Detecting object boundaries in natural images requires 'incitatory' cell-cell interactions 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.