中文
相关论文

相关论文: Forgiveness is an Adaptation in Iterated Prisoner'…

200 篇论文

We consider a setting for Inverse Reinforcement Learning (IRL) where the learner is extended with the ability to actively select multiple environments, observing an agent's behavior on each environment. We first demonstrate that if the…

人工智能 · 计算机科学 2016-01-26 Kareem Amin , Satinder Singh

Large language model (LLM) agents achieve impressive single-task performance but commonly exhibit repeated failures, inefficient exploration, and limited cross-task adaptability. Existing reflective strategies (e.g., Reflexion, ReAct)…

人工智能 · 计算机科学 2025-09-09 Chunlong Wu , Ye Luo , Zhibo Qu , Min Wang

We present tournament results and several powerful strategies for the Iterated Prisoner's Dilemma created using reinforcement learning techniques (evolutionary and particle swarm algorithms). These strategies are trained to perform well…

计算机科学与博弈论 · 计算机科学 2018-02-07 Marc Harper , Vincent Knight , Martin Jones , Georgios Koutsovoulos , Nikoleta E. Glynatsi , Owen Campbell

We study a modified prisoner's dilemma game taking place on two-dimensional disordered square lattices. The players are pure strategists and can either cooperate or defect with their immediate neighbors. In the generations each player…

物理与社会 · 物理学 2007-05-23 Zhi-Xi Wu , Xin-Jian Xu , Zi-Gang Huang , Sheng-Jun Wang , Ying-Hai Wang

The exploration of whether agents can align with their environment without relying on human-labeled data presents an intriguing research topic. Drawing inspiration from the alignment process observed in intelligent organisms, where…

计算与语言 · 计算机科学 2024-03-06 Bo Wang , Tianxiang Sun , Hang Yan , Siyin Wang , Qingyuan Cheng , Xipeng Qiu

We will study a population of individuals playing the infinitely repeated Prisoner's Dilemma under replicator dynamics. The population consists of three kinds of individuals using the following reactive strategies: ALLD (individuals which…

种群与进化 · 定量生物学 2015-09-04 Irene Núñez Rodríguez , Armando G. M. Neves

Real-world applications of reinforcement learning for recommendation and experimentation faces a practical challenge: the relative reward of different bandit arms can evolve over the lifetime of the learning agent. To deal with these…

机器学习 · 计算机科学 2022-06-29 Srivas Chennu , Andrew Maher , Jamie Martin , Subash Prabanantham

In spatial games players typically alter their strategy by imitating the most successful or one randomly selected neighbor. Since a single neighbor is taken as reference, the information stemming from other neighbors is neglected, which…

物理与社会 · 物理学 2012-11-01 Xiaofeng Wang , Matjaz Perc , Yongkui Liu , Xiaojie Chen , Long Wang

The field of Game Theory provides a useful mechanism for modeling many decision-making scenarios. In participating in these scenarios individuals and groups adopt particular strategies, which generally perform with varying levels of…

多智能体系统 · 计算机科学 2018-07-24 Francis Lawlor , Rem Collier , Vivek Nallur

Exploiting others is beneficial individually but it could also be detrimental globally. The reverse is also true: a higher cooperation level may change the environment in a way that is beneficial for all competitors. To explore the possible…

物理与社会 · 物理学 2018-02-23 Attila Szolnoki , Xiaojie Chen

The Prisoner's Dilemma is a simple model that captures the essential contradiction between individual rationality and global rationality. Although the one-shot Prisoner's Dilemma is usually viewed simple, in this paper we will categorize it…

计算机科学与博弈论 · 计算机科学 2015-03-17 Haoyang Wu

Here we study the effects of adopting different strategies against different opponent instead of adopting the same strategy against all of them in the prisoner dilemma structured in well-mixed populations. We consider an evolutionary…

物理与社会 · 物理学 2015-05-14 Lucas Wardil , Jafferson K. L. da Silva

We analyze the performance of heterogeneous learning agents in asset markets with stochastic payoffs. Our main focus is on comparing Bayesian learners and no-regret learners who compete in markets and identifying the conditions under which…

计算机科学与博弈论 · 计算机科学 2026-05-04 David Easley , Yoav Kolumbus , Eva Tardos

We study the evolution of behavioral rules in environments with multiple contexts. Agents copy rules used by better-performing peers in the same context and apply them across contexts. Multiple contexts turn discrete-time imitation dynamics…

理论经济学 · 经济学 2026-05-08 Enrique Urbano Arellano , Xinyang Wang

We investigate the repeated prisoner's dilemma game where both players alternately use reinforcement learning to obtain their optimal memory-one strategies. We theoretically solve the simultaneous Bellman optimality equations of…

计算机科学与博弈论 · 计算机科学 2021-06-02 Yuki Usui , Masahiko Ueda

Evolution of cooperation in the prisoner's dilemma game is studied where initially all players are linked via a regular graph, having four neighbors each. Simultaneously with the strategy evolution, players are allowed to make new…

物理与社会 · 物理学 2009-03-27 Attila Szolnoki , Matjaz Perc , Zsuzsa Danku

The adaptive immune system provides a diverse set of molecules that can mount specific responses against a multitude of pathogens. Memory is a key feature of adaptive immunity, which allows organisms to respond more readily upon…

种群与进化 · 定量生物学 2021-04-14 Oskar H Schnaack , Armita Nourmohammad

Anonymous online business environments have a social dilemma situation in it. A dilemma on whether to cooperate or Defect. Defection by a buyer to seller and/or seller to buyer might give each a better profit at the cost of the loss of…

计算机科学与博弈论 · 计算机科学 2013-05-15 Sanat Kumar Bista , Keshav P. Dahal , Peter I. Cowling , Bhadra Man Tuladhar

In this paper we address the cooperation problem in structured populations by considering the prisoner's dilemma game as metaphor of the social interactions between individuals with imitation capacity. We present a new strategy update rule…

计算机科学与博弈论 · 计算机科学 2013-03-19 Ignacio Gomez Portillo

We propose to reformulate the payoff matrix structure of Prisoner's Dilemma Game, by introducing threat and greed factors, and show their effect on the co-evolution of memory and cooperation. Our findings are as follows. (i) Memory protects…

物理与社会 · 物理学 2017-04-20 Uzay Cetin , Haluk O. Bingol