arXiv · 2609.26783
A Decentralized Partially Observable Team Decision Methodology with Delayed Information Sharing
Abstract
We study decentralized partially observable team decision problems with low-rank latent dynamics and unknown system models. The proposed framework combines team-theoretic equivalence with low-rank model representations to address cooperative decision-making in partially observable Markov decision processes without prior knowledge of the transition model. Each team member makes decisions based on local private information and delayed common information shared across the team. Using only this available information, each member learns an approximate low-rank Markov decision process and applies least-squares value iteration to compute its policy. This yields a fully decentralized learning and planning algorithm that requires neither a centralized coordinator nor centralized training. We show that the resulting member-side solutions approximate the centralized team solution: despite partial observability, unknown dynamics, and delayed common information, each member recovers the corresponding component of an approximate team-optimal policy. We further establish finite-sample performance guarantees and derive a corresponding sample-complexity bound for the proposed algorithm.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Xiaoxing Ren, Thomas Parisini, Andreas A. Malikopoulos. 2026-09-22. A Decentralized Partially Observable Team Decision Methodology with Delayed Information Sharing. https://arxiv.org/abs/2609.26783
Cite the original work for its findings. Save a collection to share your selection of sources.