Temporal misalignment in scene perception: Divergent representations of locomotive action affordances in human brain responses and DNNs
Bartnik, C. G.; Fraats, E. I. C.; Groen, I. I. A.
Show abstract
The human visual system processes scenes with remarkable speed, enabling the extraction of essential information to navigate our surroundings in a single glance. To elucidate how the brain transforms visual inputs into neural representations of navigationally relevant information, we collected electroencephalography (EEG) responses to diverse indoor and outdoor scenes along with behavioral annotations of locomotive action affordances (e.g., walking, cycling), object annotations, and low-level image features to model distinct types of scene information. Using representational similarity analysis, we examined the neural representation of locomotive action affordances over time, their co-localization within scene-selective cortex, and their computational alignment with deep neural networks (DNNs). Our results show that locomotive action affordance representations emerge within 200 ms of visual processing, showing unique contributions to EEG responses at temporally distinct time-points from objects and low-level properties. Spatiotemporal fusion with functional magnetic resonance imaging (fMRI) recordings in scene-selective brain regions reveals that both the parahippocampal (PPA) and occipital place region (OPA), but not the medial place region (MPA), contribute to locomotive action affordance representations, with a distinct temporal hierarchy between them. While DNNs exhibit good predictivity of early EEG responses, they primarily capture low-level features and show limited alignment with affordance processing. These findings reveal a temporally distinct neural representation of action affordances and highlight a limitation of current DNNs in modeling affordance perception.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Novel verbal instructions recruit abstract neural patterns of time-variable information dimensionality 97%
- The relative coding strength of object identity and nonidentity features in human occipito-temporal cortex and convolutional neural networks 96%
- Object selection by automatic spreading of top-down attentional signals in V1 96%
Similar papers in this journal
Similar papers in this journal
- Rapid contextualization of fragmented scene information in the human visual system 97%
- Independent spatiotemporal effects of spatial attention and background clutter on human object location representations 97%
- The contribution of object size, manipulability, and stability on neural responses to inanimate objects 96%
Similar papers in this journal
- Representation of locomotive action affordances in human behavior, brains and deep neural networks 97%
- Large-scale dissociations between views of objects, scenes, and reachable-scale environments in visual cortex 96%
- High-level cognition is supported by information-rich but compressible brain activity patterns 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.