中文
相关论文

相关论文: Adaptive Learning in Large Populations

200 篇论文

In this work we propose a kinetic formulation for evolutionary game theory for zero sum games when the agents use mixed strategies. We start with a simple adaptive rule, where after an encounter each agent increases the probability of play…

偏微分方程分析 · 数学 2020-06-16 Juan Pablo Pinasco , Mauro Rodriguez-Cartabia , Nicolas Saintier

We investigate an evolutionary prisoner's dilemma game among self-driven agents, where collective motion of biological flocks is imitated through averaging directions of neighbors. Depending on the temptation to defect and the velocity at…

物理与社会 · 物理学 2015-03-13 Zhuo Chen , Jian-Xi Gao , Yun-Ze Cai , Xiao-Ming Xu

This work introduces an online Bayesian game-theoretic method for behavior identification in multi-agent dynamical systems. By casting Hamilton-Jacobi-Bellman optimality conditions as linear-in-parameter residuals, the method enables fast…

系统与控制 · 电气工程与系统科学 2026-01-09 Francesco Bianchin , Robert Lefringhausen , Sandra Hirche

Spatial evolutionary games model individuals who are distributed in a spatial domain and update their strategies upon playing a normal form game with their neighbors. We derive integro-differential equations as deterministic approximations…

概率论 · 数学 2010-07-06 Sung-Ha Hwang , Markos Katsoulakis , Luc Rey-Bellet

This chapter focuses on variable maturation delay or, more precisely, on the mathematical description of a size-structured population consuming an unstructured resource. When the resource concentration is a known function of time, we can…

种群与进化 · 定量生物学 2025-10-21 Odo Diekmann , Francesca Scarabel

We study the evolution of behavior under reinforcement learning in a Prisoner's Dilemma where agents interact in a regular network and can learn about whether they play one-shot or repeatedly by incurring a cost of deliberation. With…

物理与社会 · 物理学 2024-03-28 Rossana Mastrandrea , Leonardo Boncinelli , Ennio Bilancini

Holding on to one's strategy is natural and common if the later warrants success and satisfaction. This goes against widespread simulation practices of evolutionary games, where players frequently consider changing their strategy even…

种群与进化 · 定量生物学 2012-05-04 Yongkui Liu , Xiaojie Chen , Lin Zhang , Long Wang , Matjaz Perc

An adaptive agent predicting the future state of an environment must weigh trust in new observations against prior experiences. In this light, we propose a view of the adaptive immune system as a dynamic Bayesian machinery that updates its…

种群与进化 · 定量生物学 2019-05-14 Andreas Mayer , Vijay Balasubramanian , Aleksandra M. Walczak , Thierry Mora

We introduce an analytical model to study the evolution towards equilibrium in spatial games, with `memory-aware' agents, i.e., agents that accumulate their payoff over time. In particular, we focus our attention on the spatial Prisoner's…

物理与社会 · 物理学 2016-03-23 Marco Alberto Javarone

The Kelly or proportional allocation mechanism is a simple and efficient auction-based scheme that distributes an infinitely divisible resource proportionally to the agents bids. When agents are aware of the allocation rule, their…

计算机科学与博弈论 · 计算机科学 2026-03-27 Younes Ben Mazziane , Cleque-Marlain Mboulou Moutoubi , Eitan Altman , Francesco De Pellegrini

Multi-agent learning is intrinsically harder, more unstable and unpredictable than single agent optimization. For this reason, numerous specialized heuristics and techniques have been designed towards the goal of achieving convergence to…

机器学习 · 计算机科学 2023-06-05 Emmanouil-Vasileios Vlatakis-Gkaragkounis , Lampros Flokas , Georgios Piliouras

This paper discusses the role of opportunistic punisher who may act selfishly to free-ride cooperators or not to be exploited by defectors. To consider opportunistic punisher, we make a change to the sequence of one-shot public good game;…

种群与进化 · 定量生物学 2010-08-10 Jun-Sok Huhh

Evolutionary dynamics in finite populations is known to fixate eventually in the absence of mutation. We here show that a similar phenomenon can be found in stochastic game dynamical batch learning, and investigate fixation in learning…

物理与社会 · 物理学 2015-05-27 John Realpe-Gomez , Bartosz Szczesny , Luca Dall'Asta , Tobias Galla

According to the standard imitation protocol, a less successful player adopts the strategy of the more successful one faithfully for future success. This is the cornerstone of evolutionary game theory that explores the vitality of competing…

物理与社会 · 物理学 2019-09-27 Attila Szolnoki , Xiaojie Chen

We use analytical techniques based on an expansion in the inverse system size to study the stochastic evolutionary dynamics of finite populations of players interacting in a repeated prisoner's dilemma game. We show that a mechanism of…

种群与进化 · 定量生物学 2012-04-20 Alex J. Bladon , Tobias Galla , Alan J. McKane

We develop a flexible stochastic approximation framework for analyzing the long-run behavior of learning in games (both continuous and finite). The proposed analysis template incorporates a wide array of popular learning algorithms,…

计算机科学与博弈论 · 计算机科学 2023-07-04 Panayotis Mertikopoulos , Ya-Ping Hsieh , Volkan Cevher

In this paper we present results and analyses of a class of games in which heterogeneous agents are rewarded for being in a minority group. Each agent possesses a number of fixed strategies each of which are predictors of the next minority…

adap-org · 物理学 2007-05-23 Radu Manuca , Yi Li , Rick Riolo , Robert Savit

Although learning has found wide application in multi-agent systems, its effects on the temporal evolution of a system are far from understood. This paper focuses on the dynamics of Q-learning in large-scale multi-agent systems modeled as…

多智能体系统 · 计算机科学 2022-03-04 Shuyue Hu , Chin-Wing Leung , Ho-fung Leung , Harold Soh

Consider a 2-player normal-form game repeated over time. We introduce an adaptive learning procedure, where the players only observe their own realized payoff at each stage. We assume that agents do not know their own payoff function, and…

计算机科学与博弈论 · 计算机科学 2013-06-13 Mario Bravo , Mathieu Faure

Given a large population of players, each player has three possible choices between option 1 or 2 or no option. The two options are equally favorable and the population has to reach consensus on one of the two options quickly and in a…

系统与控制 · 计算机科学 2017-06-06 Leonardo Stella , Dario Bauso