中文
相关论文

相关论文: Offline Two-Player Zero-Sum Markov Games with KL R…

200 篇论文

We study dynamic finite-player and mean-field stochastic games within the framework of Markov perfect equilibria (MPE). Our focus is on discrete time and space structures without monotonicity. Unlike their continuous-time analogues,…

最优化与控制 · 数学 2025-09-29 Felix Höfer , H. Mete Soner , Atilla Yılmaz

We consider a class of N-player stochastic games of multi-dimensional singular control, in which each player faces a minimization problem of monotone-follower type with submodular costs. We call these games "monotone-follower games". In a…

最优化与控制 · 数学 2019-02-05 Jodi Dianetti , Giorgio Ferrari

In this paper, we present a novel consensus-based zeroth-order algorithm tailored for non-convex multiplayer games. The proposed method leverages a metaheuristic approach using concepts from swarm intelligence to reliably identify global…

动力系统 · 数学 2024-07-30 Enis Chenchene , Hui Huang , Jinniao Qiu

Finding Nash equilibria in two-player zero-sum imperfect-information games remains a central challenge in multi-agent reinforcement learning. Recent multi-round regularization methods offer a promising direction, yet existing approaches…

机器学习 · 计算机科学 2026-05-01 Eason Yu , Tzu Hao Liu , Clément L. Canonne , Yunke Wang , Chang Xu , Nguyen H. Tran , Stefano V. Albrecht

In an inverse game problem, one needs to infer the cost function of the players in a game such that a desired joint strategy is a Nash equilibrium. We study the inverse game problem for a class of multiplayer matrix games, where the cost…

计算机科学与博弈论 · 计算机科学 2022-10-17 Yue Yu , Jonathan Salfity , David Fridovich-Keil , Ufuk Topcu

Mean Field Games (MFGs) have the ability to handle large-scale multi-agent systems, but learning Nash equilibria in MFGs remains a challenging task. In this paper, we propose a deep reinforcement learning (DRL) algorithm that achieves…

计算机科学与博弈论 · 计算机科学 2024-03-07 Zida Wu , Mathieu Lauriere , Samuel Jia Cong Chua , Matthieu Geist , Olivier Pietquin , Ankur Mehta

Real-world games, which concern imperfect information, multiple players, and simultaneous moves, are less frequently discussed in the existing literature of game theory. While reinforcement learning (RL) provides a general framework to…

计算机科学与博弈论 · 计算机科学 2023-06-02 Runyu Lu , Yuanheng Zhu , Dongbin Zhao

In this paper, we propose a second-order extension of the continuous-time game-theoretic mirror descent (MD) dynamics, referred to as MD2, which provably converges to mere (but not necessarily strict) variationally stable states (VSS)…

最优化与控制 · 数学 2024-10-28 Bolin Gao , Lacra Pavel

We study decentralized policy learning in Markov games where we control a single agent to play with nonstationary and possibly adversarial opponents. Our goal is to develop a no-regret online learning algorithm that (i) takes actions based…

机器学习 · 计算机科学 2022-06-06 Wenhao Zhan , Jason D. Lee , Zhuoran Yang

This work proposes an algorithm for seeking generalised feedback Nash equilibria (GFNE) in noncooperative dynamic games. The focus is on cyber-physical systems with dynamics which are linear, stochastic, potentially unstable, and partially…

最优化与控制 · 数学 2025-04-01 Otacilio B. L. Neto , Michela Mulas , Francesco Corona

Finding approximate Nash equilibria in zero-sum imperfect-information games is challenging when the number of information states is large. Policy Space Response Oracles (PSRO) is a deep reinforcement learning algorithm grounded in game…

计算机科学与博弈论 · 计算机科学 2021-02-22 Stephen McAleer , John Lanier , Roy Fox , Pierre Baldi

We study Markov perfect equilibria in continuous-time dynamic games with finitely many symmetric players. The corresponding Nash system reduces to the Nash-Lasry-Lions equation for the common value function, also known as the master…

最优化与控制 · 数学 2025-07-29 Felix Höfer , Mathieu Laurière , H. Mete Soner , Qinxin Yan

Decentralized online learning for seeking generalized Nash equilibrium (GNE) of noncooperative games in dynamic environments is studied in this paper. Each player aims at selfishly minimizing its own time-varying cost function subject to…

最优化与控制 · 数学 2021-05-14 Min Meng , Xiuxian Li , Yiguang Hong , Jie Chen , Long Wang

We consider the problem of computing an equilibrium in a class of \textit{nonlinear generalized Nash equilibrium problems (NGNEPs)} in which the strategy sets for each player are defined by equality and inequality constraints that may…

最优化与控制 · 数学 2023-02-07 Michael I. Jordan , Tianyi Lin , Manolis Zampetakis

We consider the problem of computing mixed Nash equilibria of two-player zero-sum games with continuous sets of pure strategies and with first-order access to the payoff function. This problem arises for example in game-theory-inspired…

最优化与控制 · 数学 2025-09-04 Guillaume Wang , Lénaïc Chizat

Zero-sum stochastic games have found important applications in a variety of fields, from machine learning to economics. Work on this model has primarily focused on the computation of Nash equilibrium due to its effectiveness in solving…

计算机科学与博弈论 · 计算机科学 2022-11-28 Denizalp Goktas , Jiayi Zhao , Amy Greenwald

We derive the rate of convergence to the strongly variationally stable Nash equilibrium in a convex game, for a zeroth-order learning algorithm. Though we do not assume strong monotonicity of the game, our rates for the one-point feedback…

最优化与控制 · 数学 2024-03-12 Tatiana Tatarenko , Maryam Kamgarpour

We study the convergence to local Nash equilibria of gradient methods for two-player zero-sum differentiable games. It is well-known that such dynamics converge locally when $S \succ 0$ and may diverge when $S=0$, where $S\succeq 0$ is the…

最优化与控制 · 数学 2023-11-08 Guillaume Wang , Lénaïc Chizat

Model-free learning for multi-agent stochastic games is an active area of research. Existing reinforcement learning algorithms, however, are often restricted to zero-sum games, and are applicable only in small state-action spaces or other…

机器学习 · 计算机科学 2022-10-25 Philippe Casgrain , Brian Ning , Sebastian Jaimungal

We investigate the convergence of symmetric stochastic differential games with interactions via control, where the volatility terms of both idiosyncratic and common noises are controlled. We apply the stochastic maximum principle, following…

概率论 · 数学 2026-02-19 Erhan Bayraktar , Hiroaki Horikawa