arXiv · 1210.0077
Optimistic Agents are Asymptotically Optimal
Abstract
We use optimism to introduce generic asymptotically optimal reinforcement learning agents. They achieve, with an arbitrary finite or compact class of environments, asymptotically optimal behavior. Furthermore, in the finite deterministic case we provide finite error bounds.
Explore related subjects
Keep this discovery
Peter Sunehag, Marcus Hutter. 2012-09-29. Optimistic Agents are Asymptotically Optimal. https://arxiv.org/abs/1210.0077
Cite the original work for its findings. Save a collection to share your selection of sources.