arXiv · 2312.12325
Optimizing Local Satisfaction of Long-Run Average Objectives in Markov Decision Processes
Abstract
Long-run average optimization problems for Markov decision processes (MDPs) require constructing policies with optimal steady-state behavior, i.e., optimal limit frequency of visits to the states. However, such policies may suffer from local instability, i.e., the frequency of states visited in a bounded time horizon along a run differs significantly from the limit frequency. In this work, we propose an efficient algorithmic solution to this problem.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
David Klaška, Antonín Kučera, Vojtěch Kůr, Vít Musil, Vojtěch Řehák. 2023-12-19. Optimizing Local Satisfaction of Long-Run Average Objectives in Markov Decision Processes. https://arxiv.org/abs/2312.12325
Cite the original work for its findings. Save a collection to share your selection of sources.