English

Near-Optimal Policy Optimization for Correlated Equilibrium in General-Sum Markov Games

Machine Learning 2024-05-03 v2 Computer Science and Game Theory Optimization and Control

Abstract

We study policy optimization algorithms for computing correlated equilibria in multi-player general-sum Markov Games. Previous results achieve O(T1/2)O(T^{-1/2}) convergence rate to a correlated equilibrium and an accelerated O(T3/4)O(T^{-3/4}) convergence rate to the weaker notion of coarse correlated equilibrium. In this paper, we improve both results significantly by providing an uncoupled policy optimization algorithm that attains a near-optimal O~(T1)\tilde{O}(T^{-1}) convergence rate for computing a correlated equilibrium. Our algorithm is constructed by combining two main elements (i) smooth value updates and (ii) the optimistic-follow-the-regularized-leader algorithm with the log barrier regularizer.

Keywords

Cite

@article{arxiv.2401.15240,
  title  = {Near-Optimal Policy Optimization for Correlated Equilibrium in General-Sum Markov Games},
  author = {Yang Cai and Haipeng Luo and Chen-Yu Wei and Weiqiang Zheng},
  journal= {arXiv preprint arXiv:2401.15240},
  year   = {2024}
}

Comments

AISTATS 2024 Oral