中文
相关论文

相关论文: Dynamics of Boltzmann Q-Learning in Two-Player Two…

200 篇论文

We study analytically and by computer simulations a complex system of adaptive agents with finite memory. Borrowing the framework of the Minority Game and using the replica formalism we show the existence of an equilibrium phase transition…

统计力学 · 物理学 2009-11-07 M. Marsili , R. Mulet , F. Ricci-Tersenghi , R. Zecchina

Distributed Nash equilibrium (NE) seeking problem for multi-coalition games has attracted increasing attention in recent years, but the research mainly focuses on the case without agreement demand within coalitions. This paper considers a…

最优化与控制 · 数学 2021-12-10 Jialing Zhou , Yuezu Lv , Guanghui Wen , Jinhu Lv , Dezhi Zheng

Distributed optimization and Nash equilibrium (NE) seeking problems have drawn much attention in the control community recently. This paper studies a class of non-cooperative games, known as N-cluster game, which subsumes both cooperative…

最优化与控制 · 数学 2023-03-01 Yipeng Pang , Guoqiang Hu

Consider a strongly monotone game where the players' utility functions include a reward function and a linear term for each dimension, with coefficients that are controlled by the manager. Gradient play converges to a unique Nash…

多智能体系统 · 计算机科学 2026-02-25 Siddharth Chandak , Ilai Bistritz , Nicholas Bambos

Power system operators and electric utility companies often impose a coincident peak demand charge on customers when the aggregate system demand reaches its maximum. This charge incentivizes customers to strategically shift their peak usage…

系统与控制 · 电气工程与系统科学 2025-05-16 Liudong Chen , Jay Sethuraman , Bolun Xu

We study stochastic effects on the lagging anchor dynamics, a reinforcement learning algorithm used to learn successful strategies in iterated games, which is known to converge to Nash points in the absence of noise. The dynamics is…

适应与自组织系统 · 物理学 2012-04-20 James B. T. Sanders , Tobias Galla , Jonathan Shapiro

Evolutionary Prisoner's Dilemma games with quenched inhomogeneities in the spatial dynamical rules are considered. The players following one of the two pure strategies (cooperation or defection) are distributed on a two-dimensional lattice.…

种群与进化 · 定量生物学 2007-05-23 Attila Szolnoki , Gyorgy Szabo

In this work, we study the sample complexity of obtaining a Nash equilibrium (NE) estimate in two-player zero-sum matrix games with noisy feedback. Specifically, we propose a novel algorithm that repeatedly solves linear programs (LPs) to…

最优化与控制 · 数学 2026-02-16 Jiashuo Jiang , Mengxiao Zhang

Boltzmann exploration is a classic strategy for sequential decision-making under uncertainty, and is one of the most standard tools in Reinforcement Learning (RL). Despite its widespread use, there is virtually no theoretical understanding…

机器学习 · 计算机科学 2017-11-08 Nicolò Cesa-Bianchi , Claudio Gentile , Gábor Lugosi , Gergely Neu

We study the problem of learning a Nash equilibrium (NE) in an imperfect information game (IIG) through self-play. Precisely, we focus on two-player, zero-sum, episodic, tabular IIG under the perfect-recall assumption where the only…

机器学习 · 统计学 2021-06-14 Tadashi Kozuno , Pierre Ménard , Rémi Munos , Michal Valko

This paper considers data-based solutions of linear-quadratic nonzero-sum differential games. Two cases are considered. First, the deterministic game is solved and Nash equilibrium strategies are obtained by using persistently excited data…

系统与控制 · 电气工程与系统科学 2026-05-15 Victor G. Lopez , Matthias A. Müller

We study the global convergence of policy optimization for finding the Nash equilibria (NE) in zero-sum linear quadratic (LQ) games. To this end, we first investigate the landscape of LQ games, viewing it as a nonconvex-nonconcave…

机器学习 · 计算机科学 2021-02-12 Kaiqing Zhang , Zhuoran Yang , Tamer Başar

This paper investigates Nash equilibrium (NE) seeking problems for noncooperative games over multi-players networks with finite bandwidth communication. A distributed quantized algorithm is presented, which consists of local gradient play,…

分布式、并行与集群计算 · 计算机科学 2021-11-16 Ziqin Chen , Ji Ma , Shu Liang , Li Li

Recently, the eco-evolutionary game theory which describes the coupled dynamics of strategies and environment have attracted great attention. At the same time, most of the current work is focused on the classic two-player two-strategy game.…

物理与社会 · 物理学 2021-11-22 Bin-Quan Li , Cong Liu , Zhi-Xi Wu , Jian-Yue Guan

We study the connection between the evolutionary replicator dynamics and the number of Nash equilibria in large random bi-matrix games. Using techniques of disordered systems theory we compute the statistical properties of both, the fixed…

种群与进化 · 定量生物学 2015-06-26 Tobias Galla

Quantum games with incomplete information can be studied within a Bayesian framework. We analyze games quantized within the EWL framework [Eisert, Wilkens, and Lewenstein, Phys Rev. Lett. 83, 3077 (1999)]. We solve for the Nash equilibria…

量子物理 · 物理学 2017-03-10 Neal Solmeyer , Radhakrishnan Balu

Whilst network coordination games and network anti-coordination games have received a considerable amount of attention in the literature, network games with coexisting coordinating and anti-coordinating players are known to exhibit more…

计算机科学与博弈论 · 计算机科学 2021-10-26 Laura Arditti , Giacomo Como , Fabio Fagnani , Martina Vanelli

We study the problem of repeated play in a zero-sum game in which the payoff matrix may change, in a possibly adversarial fashion, on each round; we call these Online Matrix Games. Finding the Nash Equilibrium (NE) of a two player zero-sum…

机器学习 · 计算机科学 2020-04-06 Adrian Rivera Cardoso , Jacob Abernethy , He Wang , Huan Xu

We discuss long-run behavior of stochastic dynamics of many interacting agents. In particular, three-player spatial games are studied. The effect of the number of players and the noise level on the stochastic stability of Nash equilibria is…

其他凝聚态物理 · 物理学 2009-11-10 Jacek Miekisz

We propose an adaptive incentive mechanism that learns the optimal incentives in environments where players continuously update their strategies. Our mechanism updates incentives based on each player's externality, defined as the difference…

计算机科学与博弈论 · 计算机科学 2025-03-04 Chinmay Maheshwari , Kshitij Kulkarni , Manxi Wu , Shankar Sastry