中文
相关论文

相关论文: Poincar\'{e}-Bendixson Limit Sets in Multi-Agent L…

200 篇论文

We consider the Reinforcement Learning problem of controlling an unknown dynamical system to maximise the long-term average reward along a single trajectory. Most of the literature considers system interactions that occur in discrete time…

人工智能 · 计算机科学 2023-09-07 Lorenzo Croissant , Marc Abeille , Bruno Bouchard

It is known that there are uncoupled learning heuristics leading to Nash equilibrium in all finite games. Why should players use such learning heuristics and where could they come from? We show that there is no uncoupled learning heuristic…

计算机科学与博弈论 · 计算机科学 2015-04-27 Burkhard C. Schipper

Repeated games consider a situation where multiple agents are motivated by their independent rewards throughout learning. In general, the dynamics of their learning become complex. Especially when their rewards compete with each other like…

计算机科学与博弈论 · 计算机科学 2023-05-23 Yuma Fujimoto , Kaito Ariu , Kenshi Abe

Rock-scissors-paper game, as the simplest model of intransitive relation between competing agents, is a frequently quoted model to explain the stable diversity of competitors in the race of surviving. When increasing the number of…

物理与社会 · 物理学 2019-05-21 D. Bazeia , B. F. de Oliveira , A. Szolnoki

A local agglomeration of cooperators can support the survival or spreading of cooperation, even when cooperation is predicted to die out according to the replicator equation, which is often used in evolutionary game theory to study the…

物理与社会 · 物理学 2009-03-06 Dirk Helbing

In multi-agent problems requiring a high degree of cooperation, success often depends on the ability of the agents to adapt to each other's behavior. A natural solution concept in such settings is the Stackelberg equilibrium, in which the…

机器学习 · 计算机科学 2024-06-14 Robert Loftin , Mustafa Mert Çelikok , Herke van Hoof , Samuel Kaski , Frans A. Oliehoek

Evolutionary game theory classically investigates which behavioral patterns are evolutionarily successful in a single game. More recently, a number of contributions have studied the evolution of preferences instead: which subjective…

计算机科学与博弈论 · 计算机科学 2015-05-27 Paolo Galeazzi , Michael Franke

In view of the complexity of the dynamics of learning in games, we seek to decompose a game into simpler components where the dynamics' long-run behavior is well understood. A natural starting point for this is Helmholtz's theorem, which…

计算机科学与博弈论 · 计算机科学 2024-05-21 Davide Legacci , Panayotis Mertikopoulos , Bary Pradelski

A nonlinear dynamical system is called eventually competitive (or cooperative) provided that it preserves a partial order in backward (or forward) time only after some reasonable initial transient. We presented in this paper the…

动力系统 · 数学 2018-09-27 Lin Niu , Yi Wang

Under what conditions do the behaviors of players, who play a game repeatedly, converge to a Nash equilibrium? If one assumes that the players' behavior is a discrete-time or continuous-time rule whereby the current mixed strategy profile…

计算机科学与博弈论 · 计算机科学 2022-03-29 Jason Milionis , Christos Papadimitriou , Georgios Piliouras , Kelly Spendlove

Self-play is a technique for machine learning in multi-agent systems where a learning algorithm learns by interacting with copies of itself. Self-play is useful for generating large quantities of data for learning, but has the drawback that…

计算机科学与博弈论 · 计算机科学 2023-11-30 Revan MacQueen , James R. Wright

In this paper, we examine the long-run behavior of regularized, no-regret learning in finite games. A well-known result in the field states that the empirical frequencies of no-regret play converge to the game's set of coarse correlated…

计算机科学与博弈论 · 计算机科学 2023-11-07 Victor Boone , Panayotis Mertikopoulos

When people play a repeated game they usually try to anticipate their opponents' moves based on past observations, and then decide what action to take next. Behavioural economics studies the mechanisms by which strategic decisions are taken…

物理与社会 · 物理学 2012-04-20 Tobias Galla

Reinforcement learning in multiagent systems has been studied in the fields of economic game theory, artificial intelligence and statistical physics by developing an analytical understanding of the learning dynamics (often in relation to…

多智能体系统 · 计算机科学 2019-06-25 Wolfram Barfuss , Jonathan F. Donges , Jürgen Kurths

We study the limiting behavior of the mixed strategies that result from optimal no-regret learning strategies in a repeated game setting where the stage game is any 2 by 2 competitive game. We consider optimal no-regret algorithms that are…

计算机科学与博弈论 · 计算机科学 2022-03-03 Vidya Muthukumar , Soham Phade , Anant Sahai

This paper investigates the impact of feedback quantization on multi-agent learning. In particular, we analyze the equilibrium convergence properties of the well-known "follow the regularized leader" (FTRL) class of algorithms when players…

计算机科学与博弈论 · 计算机科学 2022-09-13 Kyriakos Lotidis , Panayotis Mertikopoulos , Nicholas Bambos

Game dynamics in which three or more strategies are cyclically competitive, as represented by the rock-scissors-paper game, have attracted practical and theoretical interests. In evolutionary dynamics, cyclic competition results in…

种群与进化 · 定量生物学 2008-03-11 Naoki Masuda

We consider an example of cyclic competition bimatrix game which is a Rock-Scissors-Paper game with assumption about perfect memory of the playing agents. At first we investigate the dynamics in the neighbourhood of the Nash equilibrium as…

动力系统 · 数学 2016-09-06 Cezary Olszowiec

We introduce novel multi-agent interaction models of entropic spatially inhomogeneous evolutionary undisclosed games and their quasi-static limits. These evolutions vastly generalize first and second order dynamics. Besides the…

最优化与控制 · 数学 2022-03-10 Mauro Bonafini , Massimo Fornasier , Bernhard Schmitzer

Evolutionary dynamics in finite populations is known to fixate eventually in the absence of mutation. We here show that a similar phenomenon can be found in stochastic game dynamical batch learning, and investigate fixation in learning…

物理与社会 · 物理学 2015-05-27 John Realpe-Gomez , Bartosz Szczesny , Luca Dall'Asta , Tobias Galla