Back

Semantic axes in the brain support analogical representations

Zhu, H.; Franch, M.; Mickiewicz, E.; Belanger, J.; Cowan, R. L.; Katlowitz, K.; Chavez, A. G. L.; Chericoni, A.; Paulo, D.; Yan, X.; Bartoli, E.; Hennig, J.; Provenza, N.; Smith, E. H.; Piantadosi, S.; Sheth, S.; Hayden, B. Y.

2026-01-28 neuroscience
10.64898/2026.01.28.702241 bioRxiv
Show abstract

In vectorial embeddings of word meaning, semantic features often reflect consistent directions--or axes--within a representational space. A classic example is gender: the vector spanning "boy"[->]"girl" can be added to the embedding for "king" to predict "queen." Here we show that the same principle governs semantically driven neural responses in the human brain. We recorded single-neuron activity in three brain regions while participants listened to podcasts. Across fifteen analogical categories --including gender, number, and antonymy--we observed consistent vectorial directions, resulting in parallelogram structure within the neural manifold. Among pronouns, vectors corresponding to grammatical case, number, person, and possession obeyed the principle of commutativity, resulting in a prismatic structure, demonstrating that semantic axes can be factorized. We found parallelogram structure in all three regions, with some regional specialization: noun pluralization was more robust in the hippocampus than in the anterior cingulate cortex (ACC) or orbitofrontal cortex, whereas verbal conjugation effects were more robust in the ACC. Finally, we found evidence for partial functional specialization at the neuron level: neurons most strongly involved in one analogy type were less involved in others. These principles parallel representational structure observed in large language models. Together, these findings support a geometric basis for analogical reasoning.

Matching journals

The top 3 journals account for 50% of the predicted probability mass.

50% of probability mass above

"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.