Combining Backpropagation with Equilibrium Propagation to improve an Actor-Critic RL framework
Kubo, Y.; Chalmers, E.; Luczak, A.
Show abstract
Backpropagation has been used to train neural networks for many years, allowing them to solve a wide variety of tasks like image classification, speech recognition, and reinforcement learning tasks. But the biological plausibility of backpropagation as a mechanism of neural learning has been questioned. Equilibrium Propagation (EP) has been proposed as a more biologically plausible alternative and achieves comparable accuracy on the CIFAR-10 image classification task. This study proposes the first EP-based reinforcement learning architecture: an actor-critic architecture with the actor network trained by EP. We show that this model can solve the basic control tasks often used as benchmarks for BP-based models. Interestingly, our trained model demonstrates more consistent high-reward behavior than a comparable model trained exclusively by backpropagation.
Matching journals
The top 4 journals account for 50% of the predicted probability mass.
Similar papers in this journal
Similar papers in this journal
- Predicting human decision making in psychological tasks with recurrent neural networks 97%
- Training a spiking neuronal network model of visual-motor cortex to play a virtual racket-ball game using reinforcement learning 95%
- Collective Evolution Learning Model for Vision-Based Collective Motion with Collision Avoidance 95%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.