中文
相关论文

相关论文: Forgiveness is an Adaptation in Iterated Prisoner'…

200 篇论文

This paper studies a preference evolution model in which a population of agents are matched to play a sequential prisoner's dilemma in an incomplete information environment. An institution can design an incentive-compatible screening…

综合经济学 · 经济学 2023-11-07 Ethan Holdahl , Jiabin Wu

Previous studies suggest that punishment is a useful way to promote cooperation in the well-mixed public goods game, whereas it still lacks specific evidence that punishment maintains cooperation in spatial prisoner's dilemma game as well.…

物理与社会 · 物理学 2011-03-08 Qing Jin , Zhen Wang , Zhen Wang , Yi-Ling Wang

This paper characterizes how different incentive instruments shape cooperation in a repeated Prisoner`s Dilemma with a continuum of players. A simple tit-for-tat strategy competes against unconditional defection, and the long-run outcome is…

理论经济学 · 经济学 2025-11-14 Alexander Kangas

The prisoner's dilemma game is the most known contribution of game theory into social sciences. Here we describe new implications of this game for transactional and transformative leadership. While the autocratic (Stackelberg's) leadership…

物理与社会 · 物理学 2019-10-22 S. G. Babajanyan , A. V. Melkikh , A. E. Allahverdyan

Continual learning is often motivated by the idea, known as the big world hypothesis, that "the world is bigger" than the agent. Recent problem formulations capture this idea by explicitly constraining an agent relative to the environment.…

人工智能 · 计算机科学 2025-12-30 Alex Lewandowski , Adtiya A. Ramesh , Edan Meyer , Dale Schuurmans , Marlos C. Machado

Stochastic games have become a prevalent framework for studying long-term multi-agent interactions, especially in the context of multi-agent reinforcement learning. In this work, we comprehensively investigate the concept of constant-memory…

计算机科学与博弈论 · 计算机科学 2025-10-16 Fengming Zhu , Fangzhen Lin

Are Large Language Models (LLMs) a new form of strategic intelligence, able to reason about goals in competitive settings? We present compelling supporting evidence. The Iterated Prisoner's Dilemma (IPD) has long served as a model for…

人工智能 · 计算机科学 2025-07-04 Kenneth Payne , Baptiste Alloui-Cros

This work considers a repeated principal-agent bandit game, where the principal can only interact with her environment through the agent. The principal and the agent have misaligned objectives and the choice of action is only left to the…

Cooperation and competition are fundamental forces shaping both natural and human systems, yet their interplay remains poorly understood. The Prisoner's Dilemma Game (PDG) has long served as a foundational framework in Game Theory for…

物理与社会 · 物理学 2025-08-01 Alfonso de Miguel-Arribas , Chengbin Sun , Carlos Gracia-Lázaro , Yamir Moreno

Today's high-stakes adversarial interactions feature attackers who constantly breach the ever-improving security measures. Deception mitigates the defender's loss by misleading the attacker to make suboptimal decisions. In order to formally…

We address the problem of reinforcement learning in which observations may exhibit an arbitrary form of stochastic dependence on past observations and actions, i.e. environments more general than (PO)MDPs. The task for an agent is to attain…

机器学习 · 计算机科学 2009-12-30 Daniil Ryabko , Marcus Hutter

In dynamic settings each economic agent's choices can be revealing of her private information. This elicitation via the rationalization of observable behavior depends each agent's perception of which payoff-relevant contingencies other…

理论经济学 · 经济学 2021-05-17 Evan Piermont , Peio Zuazo-Garin

Cooperation underlies many natural and artificial systems. While voluntary participation can sustain cooperation without informational assumptions, real interactions are rarely anonymous, leaving the joint effects of participation and…

物理与社会 · 物理学 2026-01-27 Chen Shen , Zhao Song , Xinyu Wang , Lei Shi , Matjaž Perc , Zhen Wang , Jun Tanimoto

The paper studies the emergence and stability of cooperative behavior in populations of agents who interact among themselves in Prisoner's Dilemma games and who are allowed to choose their partners. The population is then subject to…

无序系统与神经网络 · 物理学 2007-05-23 Pawel Sobkowicz

To learn directed behaviors in complex environments, intelligent agents need to optimize objective functions. Various objectives are known for designing artificial agents, including task rewards and intrinsic motivation. However, it is…

人工智能 · 计算机科学 2022-02-15 Danijar Hafner , Pedro A. Ortega , Jimmy Ba , Thomas Parr , Karl Friston , Nicolas Heess

The inherent complexity of human beings manifests in a remarkable diversity of responses to intricate environments, enabling us to approach problems from varied perspectives. However, in the study of cooperation, existing research within…

种群与进化 · 定量生物学 2026-02-04 Guozhong Zheng , Zhenwei Ding , Jiqiang Zhang , Shengfeng Deng , Weiran Cai , Li Chen

Fast adapting to unknown peers (partners or opponents) with different strategies is a key challenge in multi-agent games. To do so, it is crucial for the agent to probe and identify the peer's strategy efficiently, as this is the…

人工智能 · 计算机科学 2024-08-12 Long Ma , Yuanfei Wang , Fangwei Zhong , Song-Chun Zhu , Yizhou Wang

The problem of reinforcement learning is considered where the environment or the model undergoes a change. An algorithm is proposed that an agent can apply in such a problem to achieve the optimal long-time discounted reward. The algorithm…

系统与控制 · 电气工程与系统科学 2023-04-25 Wuxia Chen , Taposh Banerjee , Jemin George , Carl Busart

We study the effect of imperfect memory on decision making in the context of a stochastic sequential action-reward problem. An agent chooses a sequence of actions which generate discrete rewards at different rates. She is allowed to make…

概率论 · 数学 2019-09-20 Kuang Xu , Se-Young Yun

We propose an extension of the evolutionary Prisoner's Dilemma cellular automata, introduced by Nowak and May \cite{nm92}, in which the pressure of the environment is taken into account. This is implemented by requiring that individuals…

物理与社会 · 物理学 2009-11-11 Julia Alonso , Ariel Fernandez , Hugo Fort