Search arXiv⌕ Search

arXiv · 2009.04336

Polynomial-Time Computation of Optimal Correlated Equilibria in Two-Player Extensive-Form Games with Public Chance Moves and Beyond

Abstract

Unlike normal-form games, where correlated equilibria have been studied for more than 45 years, extensive-form correlation is still generally not well understood. Part of the reason for this gap is that the sequential nature of extensive-form games allows for a richness of behaviors and incentives that are not possible in normal-form settings. This richness translates to a significantly different complexity landscape surrounding extensive-form correlated equilibria. As of today, it is known that finding an optimal extensive-form correlated equilibrium (EFCE), extensive-form coarse correlated equilibrium (EFCCE), or normal-form coarse correlated equilibrium (NFCCE) in a two-player extensive-form game is computationally tractable when the game does not include chance moves, and intractable when the game involves chance moves. In this paper we significantly refine this complexity threshold by showing that, in two-player games, an optimal correlated equilibrium can be computed in polynomial time, provided that a certain condition is satisfied. We show that the condition holds, for example, when all chance moves are public, that is, both players observe all chance moves. This implies that an optimal EFCE, EFCCE and NFCCE can be computed in polynomial time in the game size in two-player games with public chance moves, providing the biggest positive complexity result surrounding extensive-form correlation in more than a decade.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Gabriele Farina, Tuomas Sandholm. 2020-09-09. Polynomial-Time Computation of Optimal Correlated Equilibria in Two-Player Extensive-Form Games with Public Chance Moves and Beyond. https://arxiv.org/abs/2009.04336

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Dynamic Welfare-Maximizing Pooled Testing

Pooled testing uses one test to certify several agents as healthy when the pooled result is negative. We study a budget-constrained welfare problem in which agents have heterogeneous utilities and independent prior probabilities of being healthy. Welfare is earned when an agent is certified healthy, and a dynamic policy may choose each pool after observing earlier test outcomes. We ask how much such adaptation can improve over a static allocation that fixes all pools in advance. Our main result proves that the optimal dynamic policy has value at most twice that of the optimal static overlapping allocation, for every population, test budget, and pool-size cap. The proof samples a path through the dynamic policy using an independent health profile, randomly thins the selected pools, and compares the resulting static allocation with the dynamic policy one agent at a time. A Boolean-cube argument proves the comparison on uniform subcubes, and a complementary-profile coupling lifts the result to arbitrary heterogeneous product priors. We also identify regimes in which adaptivity has no value, show that strict adaptive gains require re-pooling agents after positive tests, and give a three-agent instance in which adaptation is strictly beneficial. Exact-Joint Greedy obtains a $1/(e+1)$ fraction of optimal static overlapping welfare. The static non-overlapping Greedy algorithm of Finster et al. has the same guarantee, hence our factor-two theorem newly implies that each is a $2(e+1)$-approximation to the optimal dynamic policy. Finally, a reproducible exact small-instance study compares the dynamic and static benchmarks and greedy policies. The appendix records separate exploratory results for Gibbs-marginal and reinforcement-learning approaches at larger scales.

cs.GT↗

Induced Representations in Cooperative Games with Homogeneous Groups of Players

Oftentimes, the Shapley value, a measure of the contribution of a player to a game, becomes infeasible for games with many players. However, establishing symmetry allows for polynomial-time computation. To examine this reduction, we identify the spectrum of a homogeneous group game by using an induced representation from a Young subgroup. We prove that the depth of interaction of a two-group game is limited by the size of the minority group. Therefore, the algebraic structure of the game filters out a large space of irrelevant complexities. We then show that this filtration constrains any symmetric linear value to a specific subspace. This recovers the Shapley value uniquely for games consisting of exactly two homogeneous groups under standard axioms. Finally, we explore applications to the UN Security Council and complementary goods markets to illustrate the practical power of this approach.

cs.GT↗

From Bilateral Trade to Matching Markets: Sharp Gains from Trade

We study gains from trade in matching markets with independent private values and costs, Bayesian incentive compatibility, interim individual rationality, and no expected budget deficit. A second-best guarantee for finite bilateral trade extends without loss to matching markets with independent Borel priors, arbitrary downward-closed feasibility, and finite expected first-best gains. For bounded buyers with monotone hazard rates and arbitrary bounded sellers, we determine the exact worst-case ratio of second-best to first-best gains, approximately $0.72490721$. For binary buyers and sellers with at most $m$ types, we determine the exact ratio for every $m$, including $8/9$ when $m=2$ and a limit of $4/5$ as $m$ grows. Both families of bounds are tight already in bilateral trade.

cs.GT↗