A meta reinforcement learning account of behavioral adaptation to volatility in recurrent neural networks
Tuzsus, D.; Brands, A. M.; Pappas, I.; Peters, J.
Show abstract
Natural environments often exhibit various degrees of volatility, ranging from slowly changing to rapidly changing contingencies. How learners adapt to changing environments is a central issue in both reinforcement learning theory and psychology. For example, learners may adapt to changes in volatility by increasing learning if volatility increases, and reducing it if volatility decreases. Computational neuroscience and neural network modeling work suggests that this adaptive capacity may result from a meta-reinforcement learning process (implemented for example in the prefrontal cortex), where past experience endows the system with the capacity to rapidly adapt to environmental conditions in new situations. Here we provide direct evidence for a meta-reinforcement learning account of adaptation to environmental volatility. Recurrent neural networks (RNNs) were trained on a restless four-armed bandit reinforcement learning problem under three different training regimes (low volatility training only, medium volatility training only, or meta-volatility training across a range of volatilities). We show that, in contrast to RNNs trained in the low volatility regime, RNNs trained under the medium or meta-volatility regimes exhibited a superior adaption to environmental volatility. This was reflected in a better performance, and computational modeling of the networks behavior revealed a more adaptive adjustment of learning and exploration to varying levels of volatility during test. Results extend the meta-RL account to volatility adaptation and confirm past experience as a crucial factor of adaptive decision-making.
Matching journals
The top 7 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- On the normative advantages of dopamine and striatal opponency for learning and choice 96%
- When to retrieve and encode episodic memories: a neural network model of hippocampal-cortical interaction 95%
- Adaptive chunking improves effective working memory capacity in a prefrontal cortex and basal ganglia circuit 95%
Similar papers in this journal
- Artificial neural networks for model identification and parameter estimation in computational cognitive models 96%
- Asymmetric and adaptive reward coding via normalized reinforcement learning 95%
- A Recurrent Neural Network Model for Flexible and Adaptive Decision Making based on Sequence Learning 95%
Similar papers in this journal
- Using top-down modulation to optimally balance shared versus separated task representations 94%
- On the Boundary Conditions of Avoidance Memory Reconsolidation: An Attractor Network Perspective 94%
- Reward prediction-errors weighted by cue salience produces addictive behaviors in simulations, with asymmetrical learning and steeper delay discounting 93%
Similar papers in this journal
Similar papers in this journal
- Unsupervised learning and clustered connectivity enhance reinforcement learning in spiking neural networks 95%
- Global remapping emerges as the mechanism for renewal of context-dependent behavior in a reinforcement learning model 94%
- An active inference approach to modeling structure learning: concept learning as an example case 94%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.