arXiv · 2103.15901
Distributed learning in congested environments with partial information
Abstract
How can non-communicating agents learn to share congested resources efficiently? This is a challenging task when the agents can access the same resource simultaneously (in contrast to multi-agent multi-armed bandit problems) and the resource valuations differ among agents. We present a fully distributed algorithm for learning to share in congested environments and prove that the agents' regret with respect to the optimal allocation is poly-logarithmic in the time horizon. Performance in the non-asymptotic regime is illustrated in numerical simulations. The distributed algorithm has applications in cloud computing and spectrum sharing. Keywords: Distributed learning, congestion games, poly-logarithmic regret.
Explore related subjects
Keep this discovery
Tomer Boyarski, Amir Leshem, Vikram Krishnamurthy. 2021-03-29. Distributed learning in congested environments with partial information. https://arxiv.org/abs/2103.15901
Cite the original work for its findings. Save a collection to share your selection of sources.