中文
相关论文

相关论文: Effects of noise on convergent game learning dynam…

200 篇论文

We establish finite-time last-iterate guarantees for vanilla stochastic gradient descent in co-coercive games under noisy feedback. This is a broad class of games that is more general than strongly monotone games, allows for multiple Nash…

计算机科学与博弈论 · 计算机科学 2026-04-22 Siddharth Chandak , Ramanan Tamizholi , Nicholas Bambos

We discuss long-run behavior of stochastic dynamics of many interacting agents. In particular, three-player spatial games are studied. The effect of the number of players and the noise level on the stochastic stability of Nash equilibria is…

其他凝聚态物理 · 物理学 2009-11-10 Jacek Miekisz

We use analytical techniques based on an expansion in the inverse system size to study the stochastic evolutionary dynamics of finite populations of players interacting in a repeated prisoner's dilemma game. We show that a mechanism of…

种群与进化 · 定量生物学 2012-04-20 Alex J. Bladon , Tobias Galla , Alan J. McKane

Evolutionary game theory has traditionally employed deterministic models to describe population dynamics. These models, due to their inherent nonlinearities, can exhibit deterministic chaos, where population fluctuations follow complex,…

种群与进化 · 定量生物学 2025-04-02 Maria Alejandra Ramirez , George Datseris , Arne Traulsen

We study the limiting behavior of the mixed strategies that result from optimal no-regret learning strategies in a repeated game setting where the stage game is any 2 by 2 competitive game. We consider optimal no-regret algorithms that are…

计算机科学与博弈论 · 计算机科学 2022-03-03 Vidya Muthukumar , Soham Phade , Anant Sahai

This paper proposes a payoff perturbation technique for the Mirror Descent (MD) algorithm in games where the gradient of the payoff functions is monotone in the strategy profile space, potentially containing additive noise. The optimistic…

计算机科学与博弈论 · 计算机科学 2024-06-25 Kenshi Abe , Kaito Ariu , Mitsuki Sakamoto , Atsushi Iwasaki

We consider the influence of stochastic perturbations on stability of a unique positive equilibrium of a difference equation subject to prediction-based control. These perturbations may be multiplicative $$x_{n+1}=f(x_n)-\left( \alpha +…

动力系统 · 数学 2016-06-08 Elena Braverman , Conall Kelly , Alexandra Rodkina

We study the problem of computing an approximate Nash equilibrium of continuous-action game without access to gradients. Such game access is common in reinforcement learning settings, where the environment is typically treated as a black…

计算机科学与博弈论 · 计算机科学 2023-08-30 Carlos Martin , Tuomas Sandholm

Fluctuations and noise may alter the behavior of dynamical systems considerably. For example, oscillations may be sustained by demographic fluctuations in biological systems where a stable fixed point is found in the absence of noise. We…

适应与自组织系统 · 物理学 2009-11-13 Richard P. Boland , Tobias Galla , Alan J. McKane

Prediction is a well-studied machine learning task, and prediction algorithms are core ingredients in online products and services. Despite their centrality in the competition between online companies who offer prediction-based products,…

计算机科学与博弈论 · 计算机科学 2019-05-09 Omer Ben-Porat , Moshe Tennenholtz

We address learning Nash equilibria in convex games under the payoff information setting. We consider the case in which the game pseudo-gradient is monotone but not necessarily strictly monotone. This relaxation of strict monotonicity…

最优化与控制 · 数学 2023-08-17 Tatiana Tatarenko , Maryam Kamgarpour

Reinforcement learning from self-play has recently reported many successes. Self-play, where the agents compete with themselves, is often used to generate training data for iterative policy improvement. In previous work, heuristic rules are…

机器学习 · 计算机科学 2020-09-15 Yuanyi Zhong , Yuan Zhou , Jian Peng

Many important real-world settings contain multiple players interacting over an unknown duration with probabilistic state transitions, and are naturally modeled as stochastic games. Prior research on algorithms for stochastic games has…

计算机科学与博弈论 · 计算机科学 2021-02-19 Sam Ganzfried

In order to better understand the impact of environmental stochastic fluctuations on the evolution of animal behavior, we introduce the concept of a stochastic Nash equilibrium (SNE) that extends the classical concept of a Nash equilibrium…

种群与进化 · 定量生物学 2023-10-26 Cong Li , Tianjiao Feng , Xiudeng Zheng , Sabin Lessard , Yi Tao

In this paper, we investigate how randomness and uncertainty influence learning in games. Specifically, we examine a perturbed variant of the dynamics of "follow-the-regularized-leader" (FTRL), where the players' payoff observations and…

计算机科学与博弈论 · 计算机科学 2025-06-17 Pierre-Louis Cauvin , Davide Legacci , Panayotis Mertikopoulos

The effect of entanglement and correlated noise in a four-player quantum Minority game is investigated. Different time correlated quantum memory channels are considered to analyze the Nash equilibrium payoff of the 1st player. It is seen…

量子物理 · 物理学 2013-05-10 M. Ramzan , M. K. Khan

We formulate a general framework for competitive gradient-based learning that encompasses a wide breadth of multi-agent learning algorithms, and analyze the limiting behavior of competitive gradient-based learning algorithms using dynamical…

机器学习 · 计算机科学 2020-02-21 Eric Mazumdar , Lillian J. Ratliff , S. Shankar Sastry

Consider a two-player zero-sum stochastic game where the transition function can be embedded in a given feature space. We propose a two-player Q-learning algorithm for approximating the Nash equilibrium strategy via sampling. The algorithm…

机器学习 · 计算机科学 2019-06-04 Zeyu Jia , Lin F. Yang , Mengdi Wang

In this paper, we examine the long-run behavior of regularized, no-regret learning in finite games. A well-known result in the field states that the empirical frequencies of no-regret play converge to the game's set of coarse correlated…

计算机科学与博弈论 · 计算机科学 2023-11-07 Victor Boone , Panayotis Mertikopoulos

We consider multi-agent decision making where each agent's cost function depends on all agents' strategies. We propose a distributed algorithm to learn a Nash equilibrium, whereby each agent uses only obtained values of her cost function at…

多智能体系统 · 计算机科学 2019-04-04 Tatiana Tatarenko , Maryam Kamgarpour