Biologically plausible gated recurrent neural networks for working memory and learning-to-learn
van den Berg, A. R.; Roelfsema, P. R.; Bohte, S. M.
Show abstract
The acquisition of knowledge does not occur in isolation; rather, learning experiences in the same or similar domains amalgamate. This process through which learning can accelerate over time is referred to as learning-to-learn or meta-learning. While meta-learning can be implemented in recurrent neural networks, these networks tend to be trained with architectures that are not easily interpretable or mappable to the brain and with learning rules that are biologically implausible. Specifically, these rules employ backpropagation-through-time for learning, which relies on information that is unavailable at synapses that are undergoing plasticity in the brain. While memory models that exclusively use local information for their weight updates have been developed, they have limited capacity to integrate information over long timespans and therefore cannot easily learn-to-learn. Here, we propose a novel gated recurrent network named RECOLLECT, which can flexibly retain or forget information by means of a single memory gate and biologically plausible trial-and-error-learning that requires only local information. We demonstrate that RECOLLECT successfully learns to represent task-relevant information over increasingly long memory delays in a pro-/anti-saccade task, and that it learns to flush its memory at the end of a trial. Moreover, we show that RECOLLECT can learn-to-learn an effective policy on a reversal bandit task. Finally, we show that the solutions acquired by RECOLLECT resemble how animals learn similar tasks.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- A Recurrent Neural Network Model for Flexible and Adaptive Decision Making based on Sequence Learning 97%
- Learning compositional sequences with multiple time scales through a hierarchical network of spiking neurons 97%
- Learning spatiotemporal signals using a recurrent spiking network that discretizes time 96%
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- Unsupervised learning and clustered connectivity enhance reinforcement learning in spiking neural networks 98%
- Global remapping emerges as the mechanism for renewal of context-dependent behavior in a reinforcement learning model 96%
- An active inference approach to modeling structure learning: concept learning as an example case 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.