English
Related papers

Related papers: Offline Two-Player Zero-Sum Markov Games with KL R…

200 papers

Establishing the existence of exact or near Markov or stationary perfect Nash equilibria in nonzero-sum Markov games over Borel spaces is a challenging problem with limited positive results. Motivated by problems in multi-agent and Bayesian…

Systems and Control · Electrical Eng. & Systems 2025-07-22 Naci Saldi , Gurdal Arslan , Serdar Yuksel

In this work, we study stochastic non-cooperative games, where only noisy black-box function evaluations are available to estimate the cost function for each player. Since each player's cost function depends on both its own decision…

Computer Science and Game Theory · Computer Science 2025-11-18 Haidong Li , Anzhi Sheng , Yijie Peng , Long Wang

Model-based algorithms -- algorithms that explore the environment through building and utilizing an estimated model -- are widely used in reinforcement learning practice and theoretically shown to achieve optimal sample efficiency for…

Machine Learning · Computer Science 2021-02-09 Qinghua Liu , Tiancheng Yu , Yu Bai , Chi Jin

Worst-case hardness results for most equilibrium computation problems have raised the need for beyond-worst-case analysis. To this end, we study the smoothed complexity of finding pure Nash equilibria in Network Coordination Games, a…

Computational Complexity · Computer Science 2019-02-27 Shant Boodaghians , Rucha Kulkarni , Ruta Mehta

Mean-field games have been used as a theoretical tool to obtain an approximate Nash equilibrium for symmetric and anonymous $N$-player games. However, limiting applicability, existing theoretical results assume variations of a "population…

Optimization and Control · Mathematics 2023-06-12 Batuhan Yardim , Semih Cayci , Matthieu Geist , Niao He

We derive the rate of convergence to Nash equilibria for the payoff-based algorithm proposed in \cite{tat_kam_TAC}. These rates are achieved under the standard assumption of convexity of the game, strong monotonicity and differentiability…

Optimization and Control · Mathematics 2022-02-24 Tatiana Tatarenko , Maryam Kamgarpour

We explore the use of policy approximations to reduce the computational cost of learning Nash equilibria in zero-sum stochastic games. We propose a new Q-learning type algorithm that uses a sequence of entropy-regularized soft policies to…

Machine Learning · Computer Science 2021-06-29 Yue Guan , Qifan Zhang , Panagiotis Tsiotras

We consider the problem of finding stationary Nash equilibria (NE) in a finite discounted general-sum stochastic game. We first generalize a non-linear optimization problem from Filar and Vrieze [2004] to a $N$-player setting and break down…

Computer Science and Game Theory · Computer Science 2015-07-06 H. L Prasad , L. A. Prashanth , Shalabh Bhatnagar

Self-play (SP) is a popular multi-agent reinforcement learning (MARL) framework for solving competitive games, where each agent optimizes policy by treating others as part of the environment. Despite the empirical successes, the theoretical…

Artificial Intelligence · Computer Science 2023-10-06 Zelai Xu , Yancheng Liang , Chao Yu , Yu Wang , Yi Wu

This paper addresses the problem of learning a Nash equilibrium in $\gamma$-discounted multiplayer general-sum Markov Games (MG). A key component of this model is the possibility for the players to either collaborate or team apart to…

Computer Science and Game Theory · Computer Science 2017-03-07 Julien Pérolat , Florian Strub , Bilal Piot , Olivier Pietquin

We examine global non-asymptotic convergence properties of policy gradient methods for multi-agent reinforcement learning (RL) problems in Markov potential games (MPG). To learn a Nash equilibrium of an MPG in which the size of state space…

Machine Learning · Computer Science 2022-08-08 Dongsheng Ding , Chen-Yu Wei , Kaiqing Zhang , Mihailo R. Jovanović

Mean Field Games (MFGs) offer a powerful framework for studying large-scale multi-agent systems. Yet, learning Nash equilibria in MFGs remains a challenging problem, particularly when the initial distribution is unknown or when the…

Machine Learning · Computer Science 2025-09-04 Zida Wu , Mathieu Lauriere , Matthieu Geist , Olivier Pietquin , Ankur Mehta

$ $This paper addresses the inverse problem for Linear-Quadratic (LQ) nonzero-sum $N$-player differential games, where the goal is to learn parameters of an unknown cost function for the game, called observed, given the demonstrated…

Optimization and Control · Mathematics 2024-10-28 Emin Martirosyan , Ming Cao

We study generalized Nash equilibrium (GNE) problems in games with quadratic costs and individual linear equality constraints. Departing from approaches that require strong monotonicity and/or shared constraints, we reformulate the KKT…

Optimization and Control · Mathematics 2025-12-23 Tatiana Tatarenko , Lucas Wey Hacker

Optimization under uncertainty is a fundamental problem in learning and decision-making, particularly in multi-agent systems. Previously, Feldman, Kalai, and Tennenholtz [2010] demonstrated the ability to efficiently compete in repeated…

Computer Science and Game Theory · Computer Science 2026-01-29 Daniel Ablin , Alon Cohen

In this paper we consider the problem of computing an $\epsilon$-approximate Nash Equilibrium of a zero-sum game in a payoff matrix $A \in \mathbb{R}^{m \times n}$ with $O(1)$-bounded entries given access to a matrix-vector product oracle…

Optimization and Control · Mathematics 2025-09-05 Ishani Karmarkar , Liam O'Carroll , Aaron Sidford

Many large-scale platforms and networked control systems have a centralized decision maker interacting with a massive population of agents under strict observability constraints. Motivated by such applications, we study a cooperative Markov…

Multiagent Systems · Computer Science 2026-05-12 Emile Anand , Ishani Karmarkar

This paper proposes a novel approach for local convergence to Nash equilibrium in quadratic noncooperative games based on a distributed Lie-bracket extremum seeking control scheme. This is the first instance of noncooperative games being…

Optimization and Control · Mathematics 2025-01-22 Victor Hugo Pereira Rodrigues , Tiago Roux Oliveira , Miroslav Krstic , Tamer Basar

This work studies an algorithm, which we call magnetic mirror descent, that is inspired by mirror descent and the non-Euclidean proximal gradient algorithm. Our contribution is demonstrating the virtues of magnetic mirror descent as both an…

This work studies Nash equilibrium seeking for a class of stochastic aggregative games, where each player has an expectation-valued objective function depending on its local strategy and the aggregate of all players' strategies. We propose…

Optimization and Control · Mathematics 2022-05-17 Tongyu Wang , Peng Yi , Jie Chen
‹ Prev 1 4 5 6 7 8 10 Next ›