Paraphernalia
PPubMed23 Aug 2022Cited 9×

Combining backpropagation with Equilibrium Propagation to improve an Actor-Critic reinforcement learning framework

Yoshimasa Kubo, Eric Chalmers, Artur Luczak

Abstract

Backpropagation (BP) has been used to train neural networks for many years, allowing them to solve a wide variety of tasks like image classification, speech recognition, and reinforcement learning tasks. But the biological plausibility of BP as a mechanism of neural learning has been questioned. Equilibrium Propagation (EP) has been proposed as a more biologically plausible alternative and achieves comparable accuracy on the CIFAR-10 image classification task. This study proposes the first EP-based reinforcement learning architecture: an Actor-Critic architecture with the actor network trained by EP. We show that this model can solve the basic control tasks often used as benchmarks for BP-based models. Interestingly, our trained model demonstrates more consistent high-reward behavior than a comparable model trained exclusively by BP.

A figure from Combining backpropagation with Equilibrium Propagation to improve an Actor-Critic reinforcement learning framework
fig. from the paper

§ The Valyu brief

Reading the full paper and taking notes. This takes a few seconds…

§ Ask this paper

Ask a question about this paper

Valyu reads the full text and answers from what the paper actually says.

Q.

Searching the other archives…

Combining backpropagation with Equilibrium Propagation to improve an Actor-Critic reinforcement learning framework · Paraphernalia