English
Related papers

Related papers: The State-Action-Reward-State-Action Algorithm in …

200 papers

Reinforcement Learning (RL) in various decision-making tasks of machine learning provides effective results with an agent learning from a stand-alone reward function. However, it presents unique challenges with large amounts of environment…

Machine Learning · Computer Science 2020-03-10 Neda Navidi

Cooperative behavior lies at the very basis of human societies, yet its evolutionary origin remains a key unsolved puzzle. Whereas reciprocity or conditional cooperation is one of the most prominent mechanisms proposed to explain the…

Physics and Society · Physics 2016-03-22 Giulio Cimini , Angel Sánchez

We consider the coupled dynamics of the adaption of network structure and the evolution of strategies played by individuals occupying the network vertices. We propose a computational model in which each agent plays a $n$-round Prisoner's…

Physics and Society · Physics 2007-11-05 Feng Fu , Xiaojie Chen , Lianghuan Liu , Long Wang

Reward shaping (RS) is a powerful method in reinforcement learning (RL) for overcoming the problem of sparse or uninformative rewards. However, RS typically relies on manually engineered shaping-reward functions whose construction is…

Cooperation, fairness, trust, and resource coordination are cornerstones of modern civilization, yet their emergence remains inadequately explained by the persistent discrepancies between theoretical predictions and behavioral experiments.…

Populations and Evolution · Quantitative Biology 2026-05-21 Guozhong Zheng , Xin Ou , Shengfeng Deng , Jiqiang Zhang , Li Chen

We study a model for switching strategies in the Prisoner's Dilemma game on adaptive networks of player pairings that coevolve as players attempt to maximize their return. We use a node-based strategy model wherein each player follows one…

Social and Information Networks · Computer Science 2017-07-25 Hsuan-Wei Lee , Nishant Malik , Peter J. Mucha

Prisoner's dilemma game is the most commonly used model of spatial evolutionary game which is considered as a paradigm to portray competition among selfish individuals. In recent years, Win-Stay-Lose-Learn, a strategy updating rule base on…

Physics and Society · Physics 2021-01-27 Zhenyu Shi , Wei Wei , Xiangnan Feng , Xing Li , Zhiming Zheng

Evolutionary Prisoner's Dilemma games with quenched inhomogeneities in the spatial dynamical rules are considered. The players following one of the two pure strategies (cooperation or defection) are distributed on a two-dimensional lattice.…

Populations and Evolution · Quantitative Biology 2007-05-23 Attila Szolnoki , Gyorgy Szabo

We investigate the challenge of multi-agent deep reinforcement learning in partially competitive environments, where traditional methods struggle to foster reciprocity-based cooperation. LOLA and POLA agents learn reciprocity-based…

Computer Science and Game Theory · Computer Science 2024-04-11 Milad Aghajohari , Tim Cooijmans , Juan Agustin Duque , Shunichi Akatsuka , Aaron Courville

We propose a model for demonstrating spontaneous emergence of collective intelligent behavior from selfish individual agents. Agents' behavior is modeled using our proposed selfish algorithm ($SA$) with three learning mechanisms: reinforced…

Adaptation and Self-Organizing Systems · Physics 2020-01-06 Korosh Mahmoodi , Bruce J. West , Cleotilde Gonzalez

One of the most direct human mechanisms of promoting cooperation is rewarding it. We study the effect of sharing a reward among cooperators in the most stringent form of social dilemma, namely the Prisoner's Dilemma. Specifically, for a…

Populations and Evolution · Quantitative Biology 2012-02-02 J. A. Cuesta , R. Jimenez , H. Lugo , A. Sanchez

In this paper, the Optional Prisoner's Dilemma game in a spatial environment, with coevolutionary rules for both the strategy and network links between agents, is studied. Using a Monte Carlo simulation approach, a number of experiments are…

Neural and Evolutionary Computing · Computer Science 2017-12-04 Marcos Cardinot , Colm O'Riordan , Josephine Griffith

In this work, we ask for and answer what makes classical temporal-difference reinforcement learning with epsilon-greedy strategies cooperative. Cooperating in social dilemma situations is vital for animals, humans, and machines. While…

Machine Learning · Computer Science 2023-02-22 Wolfram Barfuss , Janusz Meylahn

In this paper we study the cooperative behavior of agents playing the Prisoner's Dilemma game in random scale-free networks. We show that the survival of cooperation is enhanced with respect to random homogeneous graphs but, on the other…

Physics and Society · Physics 2015-05-13 J. Poncela , J. Gomez-Gardenes , Y. Moreno , L. M. Floria

When an individual's behavior has rational characteristics, this may lead to irrational collective actions for the group. A wide range of organisms from animals to humans often evolve the social attribute of cooperation to meet this…

Multiagent Systems · Computer Science 2021-11-18 Zhenbo Cheng , Xingguang Liu , Leilei Zhang , Hangcheng Meng , Qin Li , Xiao Gang

The Prisoner's Dilemma has been a subject of extensive research due to its importance in understanding the ever-present tension between individual self-interest and social benefit. A strictly dominant strategy in a Prisoner's Dilemma…

Computer Science and Game Theory · Computer Science 2016-05-17 John J. Nay , Yevgeniy Vorobeychik

Cooperation is the foundation of ecosystems and the human society, and the reinforcement learning provides crucial insight into the mechanism for its emergence. However, most previous work has mostly focused on the self-organization at the…

Physics and Society · Physics 2024-05-17 Zhen-Wei Ding , Guo-Zhong Zheng , Chao-Ran Cai , Wei-Ran Cai , Li Chen , Ji-Qiang Zhang , Xu-Ming Wang

We introduce an analytical model to study the evolution towards equilibrium in spatial games, with `memory-aware' agents, i.e., agents that accumulate their payoff over time. In particular, we focus our attention on the spatial Prisoner's…

Physics and Society · Physics 2016-03-23 Marco Alberto Javarone

According to the standard imitation protocol, a less successful player adopts the strategy of the more successful one faithfully for future success. This is the cornerstone of evolutionary game theory that explores the vitality of competing…

Physics and Society · Physics 2019-09-27 Attila Szolnoki , Xiaojie Chen

As an important psychological and social experiment, the Iterated Prisoner's Dilemma (IPD) treats the choice to cooperate or defect as an atomic action. We propose to study the behaviors of online learning algorithms in the Iterated…

Computer Science and Game Theory · Computer Science 2022-08-30 Baihan Lin , Djallel Bouneffouf , Guillermo Cecchi