Difficulty in Cessation of Undesired Habits: Goal-Based Reduced Successor Representation and Reward Prediction Errors
Shimomura, K.; Kato, A.; Morita, K.
Show abstract
Difficulty in cessation of drinking, smoking, or gambling has been widely recognized. Conventional theories proposed relative dominance of habitual over goal-directed control, but human studies have not convincingly supported them. Referring to the recently suggested "successor representation" of states that enables partially goal-directed control, we propose a dopamine-related mechanism potentially underlying the difficulty in resisting habitual reward-seeking, common to substance and non-substance reward. Consider that a person has long been taking a series of actions leading to a certain reward without resisting temptation. Given the suggestions of the successor representation and the dimension reduction in the brain, we assumed that the person has acquired a dimension-reduced successor representation of states based on the goal state under the established non-resistant policy. Then, we show that if the person changes the policy to resist temptation, a large positive reward prediction error (RPE) becomes generated upon eventually reaching the goal, and it sustains given that the acquired state representation is so rigid that it does not change. Inspired by the anatomically suggested spiral striatum-midbrain circuit and the theoretically proposed spiraling accumulation of RPE bias in addiction, we further simulated the influence of RPEs generated in the goal-based representation system on another system representing individual actions. We then found that such an influence could potentially enhance the propensity of non-resistant choice. These results suggest that the inaccurate value estimation in the reduced successor representation system and its influence through the spiral striatum-midbrain circuit might contribute to the difficulty in cessation of habitual reward-seeking.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- A Computational Model of Learning Flexible Navigation in a Maze by Layout-Conforming Replay of Place Cells 95%
- A Multiscale, Systems-level, Neuropharmacological Model of Cortico-Basal Ganglia System for Arm Reaching under Normal, Parkinsonian and Levodopa Medication Conditions 95%
- An active inference approach to modeling structure learning: concept learning as an example case 94%
Similar papers in this journal
- Opponent Learning with Different Representations in the Cortico-Basal Ganglia Pathways Can Develop Obsession-Compulsion Cycle 98%
- A nonlinear relationship between prediction errors and learning rates in human reinforcement learning 95%
- Vigilance, arousal, and acetylcholine: Optimal control of attention in a simple detection task 95%
Similar papers in this journal
- Reward prediction-errors weighted by cue salience produces addictive behaviors in simulations, with asymmetrical learning and steeper delay discounting 96%
- On the Boundary Conditions of Avoidance Memory Reconsolidation: An Attractor Network Perspective 94%
- A novel density-based neural mass model for simulating neuronal network dynamics with conductance-based synapses and membrane current adaptation 93%
Similar papers in this journal
- Bipolar oscillations between positive and negative mood states in a computational model of Basal Ganglia 93%
- Modeling Time to Visual Insight in Mooney Image Recognition with a Chaotic Recurrent Neural Network 92%
- Influence of Various Temporal Recoding on Pavlovian Eyeblink Conditioning in The Cerebellum 92%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.