Face-selective responses correlate with deep networks that learn from environment feedback
Zhou, M.; Schwartz, E.; Alreja, A.; Richardson, M.; Ghuman, A.; Anzellotti, S.
Show abstract
AO_SCPLOWBSTRACTC_SCPLOWDeep neural networks have shown high accuracy in modeling neural responses in the visual system, but most models rely on supervised learning, which requires training on ground-truth labels that are typically unavailable in real-world settings. While unsupervised models can address this limitation, they miss another key aspect: visual representations are shaped by feedback from the environment. We introduce a reinforcement learning (RL) model of face perception that incorporates both input stimuli and feedback from the environment. Inspired by human interactions, we train the model to approach faces yielding positive interactions and avoid faces yielding negative interactions. Using intracortical electroencephalography (iEEG) data and Representational Dissimilarity Matrices (RDMs), we evaluate the models ability to account for neural responses. Our RL model performs at the same level as supervised and unsupervised models, capturing neural responses to complex visual stimuli. The findings suggest that RL models are a promising approach for understanding perception. Significance StatementUnderstanding how the brain encodes faces is central to vision science. Existing models rely on supervised learning, which requires ground-truth labels that are often unavailable in real-world settings, or on unsupervised learning, which ignores the role of environmental-feedback in shaping visual representations. We introduce a reinforcement learning (RL) model that learns through environmental feedback, simulating human interactions by associating approaching faces with positive interactions and avoiding faces with negative interactions. Using intracortical electroencephalography (iEEG) data from face-selective regions, we show that an RL model with a variational DenseNet encoder accounts for neural representations comparably to supervised and unsupervised models. Task and architecture jointly shaped representational geometry, highlighting the importance of both learning objective and encoder design. These findings suggest the potential of RL-based approaches to understand neural representations of naturalistic faces.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- A large and rich EEG dataset for modeling human visual object recognition 96%
- Evidence for dimensional representations and anticipatory dynamics in facial expression perception 95%
- Understanding transformation tolerant visual object representations in the human brain and convolutional neural networks 94%
Similar papers in this journal
- The face module emerged in a deep convolutional neural network selectively deprived of face experience 97%
- Multidimensional face representation in deep convolutional neural network reveals the mechanism underlying AI racism 96%
- Hierarchical sparse coding of objects in deep convolutional neural networks 95%
Similar papers in this journal
- Analysis of convolutional neural networks reveals the computational properties essential for subcortical processing of facial expression 96%
- Brain-Guided Convolutional Neural Networks Reveal Task-Specific Representations in Scene Processing 96%
- Tracking cortical representations of facial attractiveness using time-resolved representational similarity analysis 94%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.