Change point estimation by the mouse medial frontal cortex during probabilistic reward learning
Atilgan, H.; Murphy, C. E.; Wang, H.; Ortega, H. K.; Pinto, L.; Kwan, A. C.
Show abstract
There are often sudden changes in the state of environment. For a decision maker, accurate prediction and detection of change points are crucial for optimizing performance. Still unclear, however, is whether rodents are simply reactive to reinforcements, or if they can be proactive to estimate future change points during value-based decision making. In this study, we characterize head-fixed mice performing a two-armed bandit task with probabilistic reward reversals. Choice behavior deviates from classic reinforcement learning, but instead suggests a strategy involving belief updating, consistent with the anticipation of change points to exploit the task structure. Excitotoxic lesion and optogenetic inactivation implicate the anterior cingulate and premotor regions of medial frontal cortex. Specifically, over-estimation of hazard rate arises from imbalance across frontal hemispheres during the time window before the choice is made. Collectively, the results demonstrate that mice can capitalize on their knowledge of task regularities, and this estimation of future changes in the environment may be a main computational function of the rodent dorsal medial frontal cortex.
Matching journals
The top 3 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- New information triggers prospective codes to adapt for flexible navigation 96%
- Dopamine Release Plateau and Outcome Signals in Dorsal Striatum Contrast with Classic Reinforcement Learning Formulations 96%
- Functional Dichotomy of Orbitofrontal Cortex and Anterior Cingulate Cortex Subregions in Decision-Making and Brain-Body Regulation 96%
Similar papers in this journal
- Behavioral strategy shapes activation of the Vip-Sst disinhibitory circuit in visual cortex 95%
- Hippocampal replay reflects specific past experiences rather than a plan for subsequent choice 95%
- Prefrontal cortical dynorphin peptidergic transmission constrains threat-driven behavioral and network states 95%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.