Building Goal-Directed Cognitive Graphs
Gungi, A.; Sepulveda Delgado, P.; Aitsahalia, I.; Blanco-Pozo, M.; Iigaya, K.
Show abstract
Flexible behavior requires building internal structured models from experience to support goal-directed action. Although environmental transition statistics accumulate gradually, how they are used to construct a compact goal-directed graph remains unclear. We introduce the Sparse Cognitive Graph (SCG), a reinforcement-learning framework that separates gradual transition learning from the sparse directed graph that governs valuation and action selection. In the SCG, transition statistics accumulate in a dense predictive representation, while nonlinear selection determines which transitions are expressed as graph edges. Consequently, gradual strengthening can trigger discrete reorganization of graph topology and abrupt behavioral shifts. Across human reward and transition revaluation tasks, the SCG explains bimodal and trimodal behavioral regimes as emergent consequences of distinct graph configurations. Across human and mouse two-step tasks, dynamic graph reconfiguration captures canonical reward-by-transition interactions without requiring mixtures of control systems. In mice, transitions preceding reward strengthened more rapidly, biasing graph topology toward reward-directed paths. Temporally precise optogenetic dopamine stimulation produced behavioral effects consistent with accelerated graph edge formation predicted by the SCG. The model further generates a testable prediction: graph topology determines the geometry of low-dimensional population activity. Directed acyclic graphs yield activity concentrated at graph entry and goal states, whereas cyclic graphs produce periodic, grid-like structure. Together, these findings identify reward-dependent graph reorganization as a computational principle that reconciles stable predictive learning with efficient goal-directed control.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Linear reinforcement learning: Flexible reuse of computation in planning, grid fields, and cognitive control 97%
- Tonic dopamine and biases in value learning linked through a biologically inspired reinforcement learning model 96%
- Adaptive mechanisms of social and asocial learning in immersive collective foraging 96%
Similar papers in this journal
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.