中文
相关论文

相关论文: Forgiveness is an Adaptation in Iterated Prisoner'…

200 篇论文

Language models are increasingly deployed in interactive online environments, from personal chat assistants to domain-specific agents, raising questions about their cooperative and competitive behavior in multi-party settings. While prior…

人工智能 · 计算机科学 2025-09-08 Mukul Singh , Arjun Radhakrishna , Sumit Gulwani

In this work we study a weak Prisoner\^as Dilemma game in which both strategies and update rules are subjected to evolutionary pressure. Interactions among agents are specified by complex topologies, and we consider both homogeneous and…

物理与社会 · 物理学 2010-11-13 A. Cardillo , J. Gomez-Gardenes , D. Vilone , A. Sanchez

Promoting cooperation is an intellectual challenge in the social sciences, for which the iterated Prisoners' Dilemma (IPD) is a fundamental framework. The traditional view that there exists no simple ultimatum strategy whereby one player…

物理与社会 · 物理学 2015-05-12 Bin Xu , Yanran Zhou , Jaimie W. Lien , Jie Zheng , Zhijian Wang

Evolutionary game theory assumes that players replicate a highly scored player's strategy through genetic inheritance. However, when learning occurs culturally, it is often difficult to recognize someone's strategy just by observing the…

种群与进化 · 定量生物学 2021-07-01 Minjae Kim , Jung-Kyoo Choi , Seung Ki Baek

The self-organization in cooperative regimes in a simple mean-field version of a model based on "selfish" agents which play the Prisoner's Dilemma (PD) game is studied. The agents have no memory and use strategies not based on direct…

适应与自组织系统 · 物理学 2009-11-07 H. Fort

One of the most direct human mechanisms of promoting cooperation is rewarding it. We study the effect of sharing a reward among cooperators in the most stringent form of social dilemma, namely the Prisoner's Dilemma. Specifically, for a…

种群与进化 · 定量生物学 2012-02-02 J. A. Cuesta , R. Jimenez , H. Lugo , A. Sanchez

In this work, we investigate the application of Reinforcement Learning to two well known decision dilemmas, namely Newcomb's Problem and Prisoner's Dilemma. These problems are exemplary for dilemmas that autonomous agents are faced with…

人工智能 · 计算机科学 2016-10-25 Dominik Meyer , Johannes Feldmaier , Hao Shen

In this paper the results of a simulation of a prisoner's dilemma robin-round tournament are presented. In the tournament each participating strategy plays an iterated prisoner's dilemma against each other strategy (round-robin) and as a…

计算机科学与博弈论 · 计算机科学 2014-02-10 Tobias Kretz

We study the evolution of behavior under reinforcement learning in a Prisoner's Dilemma where agents interact in a regular network and can learn about whether they play one-shot or repeatedly by incurring a cost of deliberation. With…

物理与社会 · 物理学 2024-03-28 Rossana Mastrandrea , Leonardo Boncinelli , Ennio Bilancini

We investigate an evolutionary prisoner's dilemma game among self-driven agents, where collective motion of biological flocks is imitated through averaging directions of neighbors. Depending on the temptation to defect and the velocity at…

物理与社会 · 物理学 2015-03-13 Zhuo Chen , Jian-Xi Gao , Yun-Ze Cai , Xiao-Ming Xu

In this paper, the Optional Prisoner's Dilemma game in a spatial environment, with coevolutionary rules for both the strategy and network links between agents, is studied. Using a Monte Carlo simulation approach, a number of experiments are…

神经与进化计算 · 计算机科学 2017-12-04 Marcos Cardinot , Colm O'Riordan , Josephine Griffith

In the evolutionary Prisoner's Dilemma (PD) game, agents play with each other and update their strategies in every generation according to some microscopic dynamical rule. In its spatial version, agents do not play with every other but,…

种群与进化 · 定量生物学 2009-03-08 Luis G. Moyano , Angel Sánchez

Partner selection is an important process in many social interactions, permitting individuals to decrease the risks associated with cooperation. In large populations, defectors may escape punishment by roving from partner to partner, but…

adap-org · 物理学 2008-02-03 Dan Ashlock , Mark D. Smucker , E. Ann Stanley , Leigh Tesfatsion

We present a detailed study of prisoner's dilemma game with stochastic modifications on a two-dimensional lattice, in presence of evolutionary dynamics. By very nature of the rules, the cooperators have incentive to cheat and the fear to…

统计力学 · 物理学 2015-05-14 M. Ali Saif , Prashant M. Gade

While many theoretical studies have revealed the strategies that could lead to and maintain cooperation in the Iterated Prisoner's Dilemma, less is known about what human participants actually do in this game and how strategies change when…

计算机科学与博弈论 · 计算机科学 2024-03-12 Eladio Montero-Porras , Jelena Grujic , Elias Fernandez-Domingos , Tom Lenaerts

An adaptive agent predicting the future state of an environment must weigh trust in new observations against prior experiences. In this light, we propose a view of the adaptive immune system as a dynamic Bayesian machinery that updates its…

种群与进化 · 定量生物学 2019-05-14 Andreas Mayer , Vijay Balasubramanian , Aleksandra M. Walczak , Thierry Mora

So far, the theory of equilibrium selection in the infinitely repeated prisoner's dilemma is insensitive to communication possibilities. To address this issue, we incorporate the assumption that communication reduces -- but does not…

理论经济学 · 经济学 2023-04-25 Maximilian Andres

David Gauthier in his article, Maximization constrained: the rationality of cooperation, tries to defend the joint strategy in situations in which no outcome is both equilibrium and optimal. Prisoner Dilemma is the most familiar example of…

综合经济学 · 经济学 2021-09-07 Shahin Esmaeili

Growing concerns about safety and alignment of AI systems highlight the importance of embedding moral capabilities in artificial agents: a promising solution is the use of learning from experience, i.e., Reinforcement Learning. In…

多智能体系统 · 计算机科学 2026-02-11 Elizaveta Tennant , Stephen Hailes , Mirco Musolesi

Imitation learning is an effective alternative approach to learn a policy when the reward function is sparse. In this paper, we consider a challenging setting where an agent and an expert use different actions from each other. We assume…

机器学习 · 计算机科学 2019-08-27 Konrad Zolna , Negar Rostamzadeh , Yoshua Bengio , Sungjin Ahn , Pedro O. Pinheiro