arXiv · 1210.3569
Autonomous Reinforcement of Behavioral Sequences in Neural Dynamics
Abstract
We introduce a dynamic neural algorithm called Dynamic Neural (DN) SARSA(λ) for learning a behavioral sequence from delayed reward. DN-SARSA(λ) combines Dynamic Field Theory models of behavioral sequence representation, classical reinforcement learning, and a computational neuroscience model of working memory, called Item and Order working memory, which serves as an eligibility trace. DN-SARSA(λ) is implemented on both a simulated and real robot that must learn a specific rewarding sequence of elementary behaviors from exploration. Results show DN-SARSA(λ) performs on the level of the discrete SARSA(λ), validating the feasibility of general reinforcement learning without compromising neural dynamics.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Sohrob Kazerounian, Matthew Luciw, Mathis Richter, Yulia Sandamirskaya. 2013-05-14. Autonomous Reinforcement of Behavioral Sequences in Neural Dynamics. https://arxiv.org/abs/1210.3569
Cite the original work for its findings. Save a collection to share your selection of sources.