中文
相关论文

相关论文: The State-Action-Reward-State-Action Algorithm in …

200 篇论文

Repeated interaction between individuals is the main mechanism for maintaining cooperation in social dilemma situations. Variants of tit-for-tat (repeating the previous action of the opponent) and the win-stay lose-shift strategy are known…

种群与进化 · 定量生物学 2011-11-08 Shoma Tanabe , Naoki Masuda

Human decision behaviour is quite diverse. In many games humans on average do not achieve maximal payoff and the behaviour of individual players remains inhomogeneous even after playing many rounds. For instance, in repeated prisoner…

物理与社会 · 物理学 2015-11-11 Martin Spanknebel , Klaus Pawelzik

Reinforcement learning (RL) is a general framework for adaptive control, which has proven to be efficient in many domains, e.g., board games, video games or autonomous vehicles. In such problems, an agent faces a sequential decision-making…

机器学习 · 计算机科学 2020-06-16 Olivier Buffet , Olivier Pietquin , Paul Weng

A canonical social dilemma arises when finite resources are allocated to a group of people, who can choose to either reciprocate with interest, or keep the proceeds for themselves. What resource allocation mechanisms will encourage levels…

This paper presents a novel approach to Multi-Agent Reinforcement Learning (MARL) that combines cooperative task decomposition with the learning of reward machines (RMs) encoding the structure of the sub-tasks. The proposed method helps…

人工智能 · 计算机科学 2025-02-17 Leo Ardon , Daniel Furelos-Blanco , Alessandra Russo

The spatial Prisoner's Dilemma is a prototype model to show the emergence of cooperation in very competitive environments. It considers players, at site of lattices, that can either cooperate or defect when playing the Prisoner's Dilemma…

计算物理 · 物理学 2007-09-18 Marcelo Alves Pereira , Alexandre Souto Martinez , Aquino Lauri Espindola

Individual cooperative strategy influences the surrounding dynamic population, which in turn affects cooperative strategy. To better model this phenomenon, we develop a Markov decision chain based game transitions model and examine the…

物理与社会 · 物理学 2025-12-30 Chaoyang Luo , Yuji Zhang , Minyu Feng , Attila Szolnoki

To investigate the origin of cooperative behaviors, we developed an evolutionary model of sequential strategies and tested our model with computer simulations. The sequential strategies represented by stochastic machines were evaluated…

计算机科学与博弈论 · 计算机科学 2021-01-12 Jin Hong Kuan , Aadesh Salecha

We study effects of strategy-dependent time delays on equilibria of evolving populations. It is well known that time delays may cause oscillations in dynamical systems. Here we report a novel behavior. We show that microscopic models of…

种群与进化 · 定量生物学 2021-04-13 Jacek Miȩkisz , Marek Bodnar

The conflict between individual and collective interests is in the heart of every social dilemmas established by evolutionary game theory. We cannot avoid these conflicts but sometimes we may choose which interaction framework to use as a…

物理与社会 · 物理学 2021-06-09 Attila Szolnoki , Xiaojie Chen

The self-organization in cooperative regimes in a simple mean-field version of a model based on "selfish" agents which play the Prisoner's Dilemma (PD) game is studied. The agents have no memory and use strategies not based on direct…

适应与自组织系统 · 物理学 2009-11-07 H. Fort

We study an evolutionary spatial prisoner's dilemma game where the fitness of the players is determined by both the payoffs from the current interaction and their history. We consider the situation where the selection timescale is slower…

物理与社会 · 物理学 2009-09-15 Zhi-Xi Wu , Zhihai Rong , Petter Holme

With the development of artificial intelligence, human beings are increasingly interested in human-agent collaboration, which generates a series of problems about the relationship between agents and humans, such as trust and cooperation.…

物理与社会 · 物理学 2025-04-30 Danyang Jia , Xiangfeng Dai , Junliang Xing , Pin Tao , Yuanchun Shi , Zhen Wang

Deep reinforcement learning has become an important paradigm for constructing agents that can enter complex multi-agent situations and improve their policies through experience. One commonly used technique is reactive training - applying…

人工智能 · 计算机科学 2017-12-11 Alexander Peysakhovich , Adam Lerer

We present a collaboration ring model -- a network of players playing the prisoner's dilemma game and collaborating among the nearest neighbours by forming coalitions. The microscopic stochastic updating of the players' strategies are…

物理与社会 · 物理学 2026-02-13 Joy Das Bairagya , Jonathan Newton , Sagar Chakraborty

We explore the evolutionary dynamics of two games - the Prisoner's Dilemma and the Snowdrift Game - played within distinct networks (layers) of interdependent networks. In these networks imitation and interaction between individuals of…

物理与社会 · 物理学 2014-04-10 M. D. Santos , S. N. Dorogovtsev , J. F. F. Mendes

Understanding cooperative behavior in biological and social systems constitutes a scientific challenge, being the object of intense research over the past decades. Many mechanisms have been proposed to explain the presence and persistence…

物理与社会 · 物理学 2022-08-24 Hugo Perez-Martinez , Carlos Gracia-Lazaro , Fabio Dercole , Yamir Moreno

Multi-agent reinforcement learning has received significant interest in recent years notably due to the advancements made in deep reinforcement learning which have allowed for the developments of new architectures and learning algorithms.…

多智能体系统 · 计算机科学 2018-12-27 Nicolas Anastassacos , Mirco Musolesi

The Prisoner's dilemma is the main game theoretical framework in which the onset and maintainance of cooperation in biological populations is studied. In the spatial version of the model, we study the robustness of cooperation in…

无序系统与神经网络 · 物理学 2009-11-07 Mendeli H. Vainstein , Jeferson J. Arenzon

In spatial games players typically alter their strategy by imitating the most successful or one randomly selected neighbor. Since a single neighbor is taken as reference, the information stemming from other neighbors is neglected, which…

物理与社会 · 物理学 2012-11-01 Xiaofeng Wang , Matjaz Perc , Yongkui Liu , Xiaojie Chen , Long Wang