Attention weights accurately predict language representations in the brain
Lamarre, M.; Chen, C.; Deniz, F.
Show abstract
In Transformer-based language models (LMs) the attention mechanism converts token embeddings into contextual embeddings that incorporate information from neighboring words. The resulting contextual hidden state embeddings have enabled highly accurate models of brain responses, suggesting that the attention mechanism constructs contextual embeddings that carry information reflected in language-related brain representations. However, it is unclear whether the attention weights that are used to integrate information across words are themselves related to language representations in the brain. To address this question we analyzed functional magnetic resonance imaging (fMRI) recordings of participants reading English language narratives. We provided the narrative text as input to two LMs (BERT and GPT-2) and extracted their corresponding attention weights. We then used encoding models to determine how well attention weights can predict recorded brain responses. We find that attention weights accurately predict brain responses in much of the frontal and temporal cortices. Our results suggest that the attention mechanism itself carries information that is reflected in brain representations. Moreover, these results indicate cortical areas in which context integration may occur.
Matching journals
The top 5 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Artificial neural network language models predict human brain responses to language even after a developmentally realistic amount of training 97%
- Lexical semantic content, not syntactic structure, is the main contributor to ANN-brain similarity of fMRI responses in the language network 96%
- Beyond Letters: Optimal Transport as a Model for Sub-Letter Orthographic Processing 95%
Similar papers in this journal
- Encoding neural representations of time-continuous stimulus-response transformations in the human brain with advanced deep neural networks 97%
- Alignment massive of auditory individual artificial networks with fMRI brain data leads to generalizable improvements in brain encoding and downstream tasks 96%
- Representational similarity learning reveals a graded multi-dimensional semantic space in the human anterior temporal cortex 96%
Similar papers in this journal
Similar papers in this journal
- Simultaneous Modeling of Reaction Times and Brain Dynamics in a Spatial Cuing Task 95%
- Prediction of individual melodic contour processing in sensory association cortices from resting state functional connectivity 95%
- DeepComBat: A Statistically Motivated, Hyperparameter-Robust, Deep Learning Approach to Harmonization of Neuroimaging Data 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.