中文
相关论文

相关论文: Fictitious Play in Markov Games with Single Contro…

200 篇论文

We show that an N-person non-cooperative semi-Markov game under limiting ratio average pay-off has a pure semi-stationary Nash equilibrium. In an earlier paper, the zero-sum two person case has been dealt with. The proof follows by reducing…

计算机科学与博弈论 · 计算机科学 2024-02-27 K. G. Bakshi , S. Sinha

We consider two-player stochastic games played on a finite graph for infinitely many rounds. Stochastic games generalize both Markov decision processes (MDP) by adding an adversary player, and two-player deterministic games by adding…

计算机科学与博弈论 · 计算机科学 2022-02-28 Laurent Doyen

Mean field games (MFGs) model equilibria in games with a continuum of weakly interacting players as limiting systems of symmetric $n$-player games. We consider the finite-state, infinite-horizon problem with ergodic cost. Assuming Markovian…

最优化与控制 · 数学 2025-03-25 Asaf Cohen , Ethan Zell

We develop provably efficient reinforcement learning algorithms for two-player zero-sum finite-horizon Markov games with simultaneous moves. To incorporate function approximation, we consider a family of Markov games where the reward…

机器学习 · 计算机科学 2020-06-25 Qiaomin Xie , Yudong Chen , Zhaoran Wang , Zhuoran Yang

The works of (Daskalakis et al., 2009, 2022; Jin et al., 2022; Deng et al., 2023) indicate that computing Nash equilibria in multi-player Markov games is a computationally hard task. This fact raises the question of whether or not…

计算机科学与博弈论 · 计算机科学 2023-05-30 Fivos Kalogiannis , Ioannis Panageas

We consider a class of N-player stochastic games of multi-dimensional singular control, in which each player faces a minimization problem of monotone-follower type with submodular costs. We call these games "monotone-follower games". In a…

最优化与控制 · 数学 2019-02-05 Jodi Dianetti , Giorgio Ferrari

Markov games with coupling constraints model constrained dynamical decision-making involving self-interested agents, where the feasibility of an individual agent's strategy depends on the joint strategies of the others. Such games arise in…

计算机科学与博弈论 · 计算机科学 2026-05-27 Tingting Ni , Anna Maddux , Maryam Kamgarpour

In this paper, we study finite-agent linear-quadratic games on graphs. Specifically, we propose a comprehensive framework that extends the existing literature by incorporating heterogeneous and interpretable player interactions. Compared to…

最优化与控制 · 数学 2025-11-19 Ruimeng Hu , Jihao Long , Haosheng Zhou

A real-valued game has the finite improvement property (FIP), if starting from an arbitrary strategy profile and letting the players change strategies to increase their individual payoffs in a sequential but non-deterministic order always…

计算机科学与博弈论 · 计算机科学 2014-10-17 Stephane Le Roux

This paper studies a multi-player, general-sum stochastic game characterized by a dual-stage temporal structure per period. The agents face uncertainty regarding the time-evolving state that is realized at the beginning of each period.…

计算机科学与博弈论 · 计算机科学 2023-10-09 Tao Zhang , Quanyan Zhu

Independent learners are agents that employ single-agent algorithms in multi-agent systems, intentionally ignoring the effect of other strategic agents. This paper studies mean-field games from a decentralized learning perspective, with two…

计算机科学与博弈论 · 计算机科学 2025-02-04 Bora Yongacoglu , Gürdal Arslan , Serdar Yüksel

Stochastic games generalize Markov decision processes (MDPs) to a multiagent setting by allowing the state transitions to depend jointly on all player actions, and having rewards determined by multiplayer matrix games at each state. We…

计算机科学与博弈论 · 计算机科学 2013-01-18 Michael Kearns , Yishay Mansour , Satinder Singh

We consider a general type of non-Markovian impulse control problems under adverse non-linear expectation or, more specifically, the zero-sum game problem where the adversary player decides the probability measure. We show that the upper…

最优化与控制 · 数学 2022-06-30 Magnus Perninge

We show that computing approximate stationary Markov coarse correlated equilibria (CCE) in general-sum stochastic games is computationally intractable, even when there are two players, the game is turn-based, the discount factor is an…

机器学习 · 计算机科学 2022-04-11 Constantinos Daskalakis , Noah Golowich , Kaiqing Zhang

In this paper, we deepen the analysis of continuous time Fictitious Play learning algorithm to the consideration of various finite state Mean Field Game settings (finite horizon, $\gamma$-discounted), allowing in particular for the…

最优化与控制 · 数学 2020-10-27 Sarah Perrin , Julien Perolat , Mathieu Laurière , Matthieu Geist , Romuald Elie , Olivier Pietquin

We study discrete-time mean-field Markov games with infinite numbers of agents where each agent aims to minimize its ergodic cost. We consider the setting where the agents have identical linear state transitions and quadratic cost…

最优化与控制 · 数学 2019-10-17 Zuyue Fu , Zhuoran Yang , Yongxin Chen , Zhaoran Wang

It is now well known that decentralised optimisation can be formulated as a potential game, and game-theoretical learning algorithms can be used to find an optimum. One of the most common learning techniques in game theory is fictitious…

机器学习 · 统计学 2011-12-13 Michalis Smyrnakis , David S. Leslie

We study a multi-agent reinforcement learning dynamics, and analyze its asymptotic behavior in infinite-horizon discounted Markov potential games. We focus on the independent and decentralized setting, where players do not know the game…

机器学习 · 计算机科学 2025-04-02 Chinmay Maheshwari , Manxi Wu , Druv Pai , Shankar Sastry

Researchers on artificial intelligence have achieved human-level intelligence in large-scale perfect-information games, but it is still a challenge to achieve (nearly) optimal results (in other words, an approximate Nash Equilibrium) in…

人工智能 · 计算机科学 2019-04-09 Li Zhang , Wei Wang , Shijian Li , Gang Pan

We consider a stochastic differential game in the context of forward-backward stochastic differential equations, where one player implements an impulse control while the opponent controls the system continuously. Utilizing the notion of…

最优化与控制 · 数学 2021-12-20 Magnus Perninge