中文
相关论文

相关论文: Dynamics of Boltzmann Q-Learning in Two-Player Two…

200 篇论文

We show that under some general conditions the finite memory determinacy of a class of two-player win/lose games played on finite graphs implies the existence of a Nash equilibrium built from finite memory strategies for the corresponding…

计算机科学与博弈论 · 计算机科学 2016-07-13 Stéphane Le Roux , Arno Pauly

Existing settings of decentralized learning either require players to have full information or the system to have certain special structure that may be hard to check and hinder their applicability to practical systems. To overcome this, we…

系统与控制 · 电气工程与系统科学 2023-05-17 Yan Jiang , Wenqi Cui , Baosen Zhang , Jorge Cortés

When learning in strategic environments, a key question is whether agents can overcome uncertainty about their preferences to achieve outcomes they could have achieved absent any uncertainty. Can they do this solely through interactions…

计算机科学与博弈论 · 计算机科学 2024-11-21 Nivasini Ananthakrishnan , Nika Haghtalab , Chara Podimata , Kunhe Yang

The long-run behavior of multi-agent learning - and, in particular, no-regret learning - is relatively well-understood in potential games, where players have aligned interests. By contrast, in harmonic games - the strategic counterpart of…

计算机科学与博弈论 · 计算机科学 2024-12-31 Davide Legacci , Panayotis Mertikopoulos , Christos H. Papadimitriou , Georgios Piliouras , Bary S. R. Pradelski

This paper studies two important signal processing aspects of equilibrium behavior in non-cooperative games arising in social networks, namely, reinforcement learning and detection of equilibrium play. The first part of the paper presents a…

计算机科学与博弈论 · 计算机科学 2015-01-07 Omid Namvar Gharehshiran , William Hoiles , Vikram Krishnamurthy

To achieve an optimal outcome in many situations, agents need to choose distinct actions from one another. This is the case notably in many resource allocation problems, where a single resource can only be used by one agent at a time. How…

计算机科学与博弈论 · 计算机科学 2014-02-05 Ludek Cigler , Boi Faltings

In this paper, we investigate a class of nonzero-sum dynamic stochastic games, where players have linear dynamics and quadratic cost functions. The players are coupled in both dynamics and cost through a linear regression (weighted average)…

最优化与控制 · 数学 2020-10-20 Jalal Arabneydi , Amir G. Aghdam , Roland P. Malhamé

In this tutorial article, we give an overview of new challenges and representative results on distributed no-regret learning in multi-agent systems modeled as repeated unknown games. Four emerging game characteristics---dynamicity,…

计算机科学与博弈论 · 计算机科学 2020-02-24 Xiao Xu , Qing Zhao

We study a class of stochastic dynamic games that exhibit strategic complementarities between players; formally, in the games we consider, the payoff of a player has increasing differences between her own state and the empirical…

计算机科学与博弈论 · 计算机科学 2010-12-13 Sachin Adlakha , Ramesh Johari

We study the behavior of a stochastic variant of replicator dynamics in two-agent zero-sum games. We characterize the statistics of such systems by their invariant measures which can be shown to be entirely supported on the boundary of the…

动力系统 · 数学 2023-10-26 Maximilian Engel , Georgios Piliouras

In this paper, $2\times2$ zero-sum games are studied under the following assumptions: $(1)$ One of the players (the leader) commits to choose its actions by sampling a given probability measure (strategy); $(2)$ The leader announces its…

计算机科学与博弈论 · 计算机科学 2023-05-12 Ke Sun , Samir M. Perlaza , Alain Jean-Marie

This paper investigates the convergence of learning dynamics in Stackelberg games. In the class of games we consider, there is a hierarchical game being played between a leader and a follower with continuous action spaces. We establish a…

计算机科学与博弈论 · 计算机科学 2024-12-07 Tanner Fiez , Benjamin Chasnov , Lillian J. Ratliff

This paper is concerned with an overlapping information linear-quadratic (LQ) Stackelberg stochastic differential game with two leaders and two followers, where the diffusion terms of the state equation contain both the control and state…

最优化与控制 · 数学 2024-01-17 Yu Si , Jingtao Shi

Quantum games with incomplete information can be studied within a Bayesian framework. We consider a version of prisoner's dilemma (PD) in this framework with three players and characterize the Nash equilibria. A variation of the standard PD…

量子物理 · 物理学 2017-03-10 Neal Solmeyer , Ricky Dixon , Radhakrishnan Balu

This paper studies the complexity of solving two classes of non-cooperative games in a distributed manner in which the players communicate with a set of system nodes over noisy communication channels. The complexity of solving each game…

信息论 · 计算机科学 2017-01-25 Ehsan Nekouei , Girish N. Nair , Tansu Alpcan , Robin J. Evans

In this paper, we propose a passivity-based methodology for analysis and design of reinforcement learning in multi-agent finite games. Starting from a known exponentially-discounted reinforcement learning scheme, we show that convergence to…

最优化与控制 · 数学 2024-10-30 Bolin Gao , Lacra Pavel

We have studied an evolutionary prisoner's dilemma game with players located on two types of random regular graphs with a degree of 4. The analysis is focused on the effects of payoffs and noise (temperature) on the maintenance of…

统计力学 · 物理学 2009-11-11 Jeromos Vukov , György Szabó , Attila Szolnoki

We study the asymptotic behavior of deterministic, continuous-time imitation dynamics for population games over networks. The basic assumption of this learning mechanism -- encompassing the replicator dynamics -- is that players belonging…

系统与控制 · 电气工程与系统科学 2020-10-23 Giacomo Como , Fabio Fagnani , Lorenzo Zino

The behavior of no-regret learning algorithms is well understood in two-player min-max (i.e, zero-sum) games. In this paper, we investigate the behavior of no-regret learning in min-max games with dependent strategy sets, where the strategy…

计算机科学与博弈论 · 计算机科学 2022-04-15 Denizalp Goktas , Jiayi Zhao , Amy Greenwald

Risk-aversion and bounded rationality are two key characteristics of human decision-making. Risk-averse quantal-response equilibrium (RQE) is a solution concept that incorporates these features, providing a more realistic depiction of human…

计算机科学与博弈论 · 计算机科学 2025-08-13 Yizhou Zhang , Eric Mazumdar
‹ 上一页 1 8 9 10 下一页 ›