中文
相关论文

相关论文: Fictitious Play in Markov Games with Single Contro…

200 篇论文

While multi-agent reinforcement learning (MARL) has produced numerous algorithms that converge to Nash or related equilibria, such equilibria are often non-unique and can exhibit widely varying efficiency. This raises a fundamental…

计算机科学与博弈论 · 计算机科学 2026-01-29 Runyu Zhang , Gioele Zardini , Asuman Ozdaglar , Jeff Shamma , Na Li

We study the long-term behavior of the fictitious play process in repeated extensive-form games of imperfect information with perfect recall. Each player maintains incorrect beliefs that the moves at all information sets, except the one at…

计算机科学与博弈论 · 计算机科学 2025-04-28 Jason Castiglione , Gürdal Arslan

Decentralised optimisation tasks are important components of multi-agent systems. These tasks can be interpreted as n-player potential games: therefore game-theoretic learning algorithms can be used to solve decentralised optimisation…

多智能体系统 · 计算机科学 2013-01-16 Michalis Smyrnakis

We propose a deep neural network-based algorithm to identify the Markovian Nash equilibrium of general large $N$-player stochastic differential games. Following the idea of fictitious play, we recast the $N$-player game into $N$ decoupled…

最优化与控制 · 数学 2020-06-08 Jiequn Han , Ruimeng Hu

Fictitious play (FP) is a canonical game-theoretic learning algorithm which has been deployed extensively in decentralized control scenarios. However standard treatments of FP, and of many other game-theoretic models, assume rather…

最优化与控制 · 数学 2016-09-29 Brian Swenson , Soummya Kar , João Xavier , David S. Leslie

Multi-agent reinforcement learning has been successfully applied to fully-cooperative and fully-competitive environments, but little is currently known about mixed cooperative/competitive environments. In this paper, we focus on a…

机器学习 · 计算机科学 2021-10-22 Roy Fox , Stephen McAleer , Will Overman , Ioannis Panageas

This letter studies multi-agent reinforcement learning in partially observable Markov potential games. Solving this problem is challenging due to partial observability, decentralized information, and the curse of dimensionality. First, to…

多智能体系统 · 计算机科学 2026-04-02 Wonseok Yang , Thinh T. Doan

We consider a class of mean field games in which the agents interact through both their states and controls, and we focus on situations in which a generic agent tries to adjust her speed (control) to an average speed (the average is made in…

偏微分方程分析 · 数学 2020-03-10 Y Achdou , Z Kobeissi

We consider a class of two-player dynamic stochastic nonzero-sum games where the state transition and observation equations are linear, and the primitive random variables are Gaussian. Each controller acquires possibly different dynamic…

系统与控制 · 计算机科学 2014-01-21 Abhishek Gupta , Ashutosh Nayyar , Cedric Langbort , Tamer Basar

We consider two classes of constrained finite state-action stochastic games. First, we consider a two player nonzero sum single controller constrained stochastic game with both average and discounted cost criterion. We consider the same…

最优化与控制 · 数学 2012-06-11 Vikas Vikram Singh , N. Hemachandra

In this paper, a Nash-type fictitious game framework is introduced to handle a time-inconsistent linear-quadratic optimal control. The Nash-type game in this framework is called fictitious as it is between the decision maker (called real…

最优化与控制 · 数学 2021-10-04 Yuan-Hua Ni , Binbin Si , Xinzhen Zhang

Fictitious play (FP) is a natural learning dynamic in two-player zero-sum games. Samuel Karlin conjectured in 1959 that FP converges at a rate of $O(t^{-1/2})$ to Nash equilibrium, where $t$ is the number of steps played. However,…

计算机科学与博弈论 · 计算机科学 2025-07-15 Yuanhao Wang

This paper studies mean field game (MFG) of controls by featuring the joint distribution of the state and the control with the reflected state process along an exogenous stochastic reflection boundary. We contribute to the literature with a…

最优化与控制 · 数学 2025-11-10 Lijun Bo , Jingfei Wang , Xiang Yu

A classic model to study strategic decision making in multi-agent systems is the normal-form game. This model can be generalised to allow for an infinite number of pure strategies leading to continuous games. Multi-objective normal-form…

计算机科学与博弈论 · 计算机科学 2023-03-02 Willem Röpke , Carla Groenland , Roxana Rădulescu , Ann Nowé , Diederik M. Roijers

We study finite-player dynamic stochastic games with heterogeneous interactions and non-Markovian linear-quadratic objective functionals. We derive the Nash equilibrium explicitly by converting the first-order conditions into a coupled…

最优化与控制 · 数学 2024-11-12 Eyal Neuman , Sturmius Tuschmann

We consider static finite-player network games and their continuum analogs, graphon games. Existence and uniqueness results are provided, as well as convergence of the finite-player network game optimal strategy profiles to their analogs…

最优化与控制 · 数学 2019-11-26 Rene Carmona , Daniel Cooney , Christy Graves , Mathieu Lauriere

This paper studies policy optimization algorithms for multi-agent reinforcement learning. We begin by proposing an algorithm framework for two-player zero-sum Markov Games in the full-information setting, where each iteration consists of a…

机器学习 · 计算机科学 2022-07-26 Runyu Zhang , Qinghua Liu , Huan Wang , Caiming Xiong , Na Li , Yu Bai

A mean-field game (MFG) seeks the Nash Equilibrium of a game involving a continuum of players, where the Nash Equilibrium corresponds to a fixed point of the best-response mapping. However, simple fixed-point iterations do not always…

最优化与控制 · 数学 2025-07-15 Jiajia Yu , Xiuyuan Cheng , Jian-Guo Liu , Hongkai Zhao

In this paper we present a scalable deep learning framework for finding Markovian Nash Equilibria in multi-agent stochastic games using fictitious play. The motivation is inspired by theoretical analysis of Forward Backward Stochastic…

人工智能 · 计算机科学 2021-05-24 Tianrong Chen , Ziyi Wang , Ioannis Exarchos , Evangelos A. Theodorou

We study the convergence properties of decentralized fictitious play (DFP) for the class of near-potential games where the incentives of agents are nearly aligned with a potential function. In DFP, agents share information only with their…

最优化与控制 · 数学 2021-03-19 Sarper Aydın , Sina Arefizadeh , Ceyhun Eksin