中文
相关论文

相关论文: Dynamics of Boltzmann Q-Learning in Two-Player Two…

200 篇论文

The framework of multi-agent learning explores the dynamics of how individual agent strategies evolve in response to the evolving strategies of other agents. Of particular interest is whether or not agent strategies converge to well known…

计算机科学与博弈论 · 计算机科学 2023-11-21 Sarah A. Toonsi , Jeff S. Shamma

Consider a two-player zero-sum stochastic game where the transition function can be embedded in a given feature space. We propose a two-player Q-learning algorithm for approximating the Nash equilibrium strategy via sampling. The algorithm…

机器学习 · 计算机科学 2019-06-04 Zeyu Jia , Lin F. Yang , Mengdi Wang

Learning in games discusses the processes where multiple players learn their optimal strategies through the repetition of game plays. The dynamics of learning between two players in zero-sum games, such as Matching Pennies, where their…

计算机科学与博弈论 · 计算机科学 2025-03-06 Yuma Fujimoto , Kaito Ariu , Kenshi Abe

Recent studies in the spatial prisoner's dilemma games with reinforcement learning have shown that static agents can learn to cooperate through a diverse sort of mechanisms, including noise injection, different types of learning algorithms…

人工智能 · 计算机科学 2025-07-08 Gustavo C. Mangold , Heitor C. M. Fernandes , Mendeli H. Vainstein

Learning in games considers how multiple agents maximize their own rewards through repeated games. Memory, an ability that an agent changes his/her action depending on the history of actions in previous games, is often introduced into…

计算机科学与博弈论 · 计算机科学 2024-02-19 Yuma Fujimoto , Kaito Ariu , Kenshi Abe

When deployed in the world, a learning agent such as a recommender system or a chatbot often repeatedly interacts with another learning agent (such as a user) over time. In many such two-agent systems, each agent learns separately and the…

机器学习 · 计算机科学 2024-06-24 Kate Donahue , Nicole Immorlica , Meena Jagadeesan , Brendan Lucier , Aleksandrs Slivkins

We study strategic interaction in linear-quadratic network games where agents act on subjective, misspecified models of their environment. Agents observe noisy aggregate signals generated by local network externalities and interpret them…

计算机科学与博弈论 · 计算机科学 2026-03-19 Quanyan Zhu , Zhengye Han

In this paper, we investigate the seeking of Nash equilibrium (NE) in a non-cooperative quadratic game where all agents exchange their delayed strategy information with their neighbors. To extend best-response algorithms to the delayed…

系统与控制 · 电气工程与系统科学 2026-02-24 Kaichen Jiang , Yuyue Yan , Mingda Yue , Yuhu Wu

We present a novel variant of fictitious play dynamics combining classical fictitious play with Q-learning for stochastic games and analyze its convergence properties in two-player zero-sum stochastic games. Our dynamics involves players…

计算机科学与博弈论 · 计算机科学 2022-06-03 Muhammed O. Sayin , Francesca Parise , Asuman Ozdaglar

Several multiagent reinforcement learning (MARL) algorithms have been proposed to optimize agents decisions. Due to the complexity of the problem, the majority of the previously developed MARL algorithms assumed agents either had some…

机器学习 · 计算机科学 2014-01-16 Sherief Abdallah , Victor Lesser

We discuss similarities and differencies between systems of many interacting players maximizing their individual payoffs and particles minimizing their interaction energy. We analyze long-run behavior of stochastic dynamics of many…

统计力学 · 物理学 2007-05-23 Jacek Miekisz

In this paper, we examine the convergence landscape of multi-agent learning under uncertainty. Specifically, we analyze two stochastic models of regularized learning in continuous games -- one in continuous and one in discrete time with the…

计算机科学与博弈论 · 计算机科学 2025-12-10 Kyriakos Lotidis , Panayotis Mertikopoulos , Nicholas Bambos , Jose Blanchet

We consider the problem of finding stationary Nash equilibria (NE) in a finite discounted general-sum stochastic game. We first generalize a non-linear optimization problem from Filar and Vrieze [2004] to a $N$-player setting and break down…

计算机科学与博弈论 · 计算机科学 2015-07-06 H. L Prasad , L. A. Prashanth , Shalabh Bhatnagar

Beyond specific settings, many multi-agent learning algorithms fail to converge to an equilibrium solution, instead displaying complex, non-stationary behaviours such as recurrent or chaotic orbits. In fact, recent literature suggests that…

多智能体系统 · 计算机科学 2026-02-12 Dan Leonte , Aamal Hussain , Raphael Huser , Francesco Belardinelli , Dario Paccagnan

We develop a method based on computer algebra systems to represent the mutual pure strategy best-response dynamics of symmetric two-player, two-action repeated games played by players with a one-period memory. We apply this method to the…

动力系统 · 数学 2022-10-04 Janusz M Meylahn , Lars Janssen

We study a dynamic game with a large population of players who choose actions from a finite set in continuous time. Each player has a state in a finite state space that evolves stochastically with their actions. A player's reward depends…

系统与控制 · 电气工程与系统科学 2025-11-06 Leonardo Pedroso , Andrea Agazzi , W. P. M. H. Heemels , Mauro Salazar

We provide a classification of symmetric three-player games with two strategies and investigate evolutionary and asymptotic stability (in the replicator dynamics) of their Nash equilibria. We discuss similarities and differences between…

种群与进化 · 定量生物学 2007-05-23 Maciej Bukowski , Jacek Miekisz

In the theory of multi-agent systems, deception refers to the strategic manipulation of information to influence the behavior of other agents, ultimately altering the long-term dynamics of the entire system. Recently, this concept has been…

系统与控制 · 电气工程与系统科学 2025-08-27 Michael Tang , Miroslav Krstic , Jorge Poveda

Learning problems commonly exhibit an interesting feedback mechanism wherein the population data reacts to competing decision makers' actions. This paper formulates a new game theoretic framework for this phenomenon, called "multi-player…

计算机科学与博弈论 · 计算机科学 2022-04-08 Adhyyan Narang , Evan Faulkner , Dmitriy Drusvyatskiy , Maryam Fazel , Lillian J. Ratliff

Learning processes in games explain how players grapple with one another in seeking an equilibrium. We study a natural model of learning based on individual gradients in two-player continuous games. In such games, the arguably natural…

计算机科学与博弈论 · 计算机科学 2020-11-10 Benjamin J. Chasnov , Daniel Calderone , Behçet Açıkmeşe , Samuel A. Burden , Lillian J. Ratliff