Aligning transformer circuit mechanisms to neural representations in relational reasoning
Hearne, L. J.; Robinson, C. N.; Cocchi, L.; Ito, T.
Show abstract
Relational reasoning--the capacity to understand how elements relate to one another--is a defining feature of human intelligence, yet its computational basis remains unclear. Here, we combined human neuroimaging (7T fMRI) with artificial neural network modeling to identify circuit-level analogues of human reasoning computations. Using the Latin Square Task, we found that humans and transformers were able to generalize the task reliably, while standard architectures used in cognitive neuroscience could not. Analysing the transformer components revealed distinct computational roles: positional encoding captured the spatial structure of the task and aligned with representations in visual cortex, whereas attention encoded relational structure and mapped onto frontoparietal and default-mode networks. Attention weights tracked the relational complexity of the task, providing a computational analogue of working-memory demands. These results advance knowledge on the core computations supporting complex reasoning, highlighting attention-based architectures as powerful models for investigating the neural basis of higher cognition.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Human hippocampus and dorsomedial prefrontal cortex infer and update latent causes during social interaction 97%
- Entorhinal and ventromedial prefrontal cortices abstract and generalise the structure of reinforcement learning problems 97%
- Joint representation of working memory and uncertainty in human cortex 96%
Similar papers in this journal
- The lexical categorization model: A computational model of left-ventral occipito-temporal cortex activation in visual word recognition 97%
- Meta-Reinforcement Learning reconciles surprise, value and control in the anterior cingulate cortex. 96%
- Diverse and flexible behavioral strategies arise in recurrent neural networks trained on multisensory decision making 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.