Attention is all you need (in the brain): semantic contextualization in human hippocampus
Katlowitz, K.; Belanger, J. L.; Ismail, T.; Chavez, A. G.; Chericoni, A.; Franch, M. C.; Mickiewicz, E. A.; Mathura, R. K.; Paulo, D.; Bartoli, E.; Piantadosi, S. T.; Provenza, N. R.; Watrous, A. J.; Sheth, S. A.; Hayden, B. Y.
Show abstract
In natural language, word meanings are contextualized, that is, modified by meanings of nearby words. Inspired by self-attention mechanisms in transformer-based large language models (LLMs), we hypothesized that contextualization in the brain results from a weighted summation of canonical neural population responses to words with those of the words that contextualize them. We examined single unit responses in the human hippocampus while participants listened to podcasts. We first find that neurons encode the position of words within a clause, that they do so at multiple scales, and that they make use of both ordinal and frequency-domain positional encoding (which are used in some transformer models). Critically, neural responses to specific words correspond to a weighted sum of that words non-contextual embedding and the embedding of the words that contextualize it. Moreover, the relative weighting of the contextualizing words is correlated with the magnitude of the LLM-derived estimates of self-attention weighting. Finally, we show that contextualization is aligned with next-word prediction, which includes prediction of multiple possible words simultaneously. Together these results support the idea that the principles of self-attention used in LLMs overlap with the mechanisms of language processing within the human hippocampus, possibly due to similar prediction-oriented computational goals.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Gated recurrence enables simple and accurate sequence prediction in stochastic, changing, and structured environments 96%
- When to retrieve and encode episodic memories: a neural network model of hippocampal-cortical interaction 94%
- A neural network model of hippocampal contributions to category learning 94%
Similar papers in this journal
- Lexical semantic content, not syntactic structure, is the main contributor to ANN-brain similarity of fMRI responses in the language network 95%
- Artificial neural network language models predict human brain responses to language even after a developmentally realistic amount of training 95%
- Beyond Letters: Optimal Transport as a Model for Sub-Letter Orthographic Processing 94%
Similar papers in this journal
- A shared linguistic space for transmitting our thoughts from brain to brain in natural conversations 95%
- Neural state space alignment for magnitude generalization in humans and recurrent networks 94%
- Mental compression of spatial sequences in human working memory using numerical and geometrical primitives 94%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.