中文
相关论文

相关论文: Using a theory of mind to find best responses to m…

200 篇论文

Since the introduction of zero-determinant strategies, extortionate strategies have received considerable interest. While an interesting class of strategies, the definitions of extortionate strategies are algebraically rigid, apply only to…

计算机科学与博弈论 · 计算机科学 2019-04-02 Vincent A. Knight , Marc Harper , Nikoleta E. Glynatsi , Jonathan Gillard

Mutual relationships, such as cooperation and exploitation, are the basis of human and other biological societies. The foundations of these relationships are rooted in the decision making of individuals, and whether they choose to be…

最优化与控制 · 数学 2021-09-29 Yuma Fujimoto , Kunihiko Kaneko

Repeated games are difficult to analyze, especially when agents play mixed strategies. We study one-memory strategies in iterated prisoner's dilemma, then generalize the result to k-memory strategies in repeated games. Our result shows that…

计算机科学与博弈论 · 计算机科学 2019-02-26 Shiheng Wang , Fangzhen Lin

We use replicator dynamics to study an iterated prisoners' dilemma game with memory. In this study, we investigate the characteristics of all 32 possible strategies with a single-step memory by observing the results when each strategy…

物理与社会 · 物理学 2014-03-06 Young Jin Kim , Myungkyoon Roh , Seung-Woo Son

Evolutionary game theory assumes that players replicate a highly scored player's strategy through genetic inheritance. However, when learning occurs culturally, it is often difficult to recognize someone's strategy just by observing the…

种群与进化 · 定量生物学 2021-07-01 Minjae Kim , Jung-Kyoo Choi , Seung Ki Baek

Stochastic games have become a prevalent framework for studying long-term multi-agent interactions, especially in the context of multi-agent reinforcement learning. In this work, we comprehensively investigate the concept of constant-memory…

计算机科学与博弈论 · 计算机科学 2025-10-16 Fengming Zhu , Fangzhen Lin

The iterated prisoner's dilemma is a game that produces many counter-intuitive and complex behaviors in a social environment, based on very simple basic rules. It illustrates that cooperation can be a good thing even in a competitive world,…

计算机科学与博弈论 · 计算机科学 2020-09-07 Robert Prentner

We present tournament results and several powerful strategies for the Iterated Prisoner's Dilemma created using reinforcement learning techniques (evolutionary and particle swarm algorithms). These strategies are trained to perform well…

计算机科学与博弈论 · 计算机科学 2018-02-07 Marc Harper , Vincent Knight , Martin Jones , Georgios Koutsovoulos , Nikoleta E. Glynatsi , Owen Campbell

We investigate symmetric equilibria of mutual reinforcement learning when both players alternately learn the optimal memory-two strategies against the opponent in the repeated prisoners' dilemma game. We provide a necessary condition for…

物理与社会 · 物理学 2023-01-03 Masahiko Ueda

In this paper the results of a simulation of a prisoner's dilemma robin-round tournament are presented. In the tournament each participating strategy plays an iterated prisoner's dilemma against each other strategy (round-robin) and as a…

计算机科学与博弈论 · 计算机科学 2014-02-10 Tobias Kretz

We investigate the repeated prisoner's dilemma game where both players alternately use reinforcement learning to obtain their optimal memory-one strategies. We theoretically solve the simultaneous Bellman optimality equations of…

计算机科学与博弈论 · 计算机科学 2021-06-02 Yuki Usui , Masahiko Ueda

Repeated games have provided an explanation how mutual cooperation can be achieved even if defection is more favorable in a one-shot game in prisoner's dilemma situation. Recently found zero-determinant strategies have substantially been…

计算机科学与博弈论 · 计算机科学 2021-05-27 Masahiko Ueda

We introduce a two-player model of reinforcement learning with memory. Past actions of an iterated game are stored in a memory and used to determine player's next action. To examine the behaviour of the model some approximate methods are…

统计力学 · 物理学 2009-11-13 Adam Lipowski , Krzysztof Gontarek , Marcel Ausloos

Game theory is fundamental to understanding cooperation between agents. Mainly, the Prisoner's Dilemma is a well-known model that has been extensively studied in complex networks. However, although the emergence of cooperation has been…

物理与社会 · 物理学 2023-01-04 Nastaran Lotfi , Francisco A. Rodrigues

We present insights and empirical results from an extensive numerical study of the evolutionary dynamics of the iterated prisoner's dilemma. Fixation probabilities for Moran processes are obtained for all pairs of 164 different strategies…

计算机科学与博弈论 · 计算机科学 2020-02-11 Vincent Knight , Marc Harper , Nikoleta E. Glynatsi , Owen Campbell

In an iterated game between two players, there is much interest in characterizing the set of feasible payoffs for both players when one player uses a fixed strategy and the other player is free to switch. Such characterizations have led to…

种群与进化 · 定量生物学 2022-02-18 Alex McAvoy , Martin A. Nowak

This paper examines the integration of computational complexity into game theoretic models. The example focused on is the Prisoner's Dilemma, repeated for a finite length of time. We show that a minimal bound on the players' computational…

计算机科学与博弈论 · 计算机科学 2007-05-23 Yishay Mor , Jeffrey S. Rosenschein

Exploration of mechanisms underlying the emergence of collective cooperation remains a focal point in field of evolution of cooperation. Prevailing studies often neglect historical information, relying on the latest rewards as the primary…

物理与社会 · 物理学 2024-02-07 Changyan Di , Jianyue Guan , Qingguo Zhou , Jingqiang Wang , Xiangyang Li

The Prisoner's Dilemma game has a long history stretching across the social, biological, and physical sciences. In 2012, Press and Dyson developed a method for analyzing the mapping of the 8-dimensional strategy profile onto the…

物理与社会 · 物理学 2019-11-28 Robert D. Young

Researchers have explored the performance of Iterated Prisoner's Dilemma strategies for decades, from the celebrated performance of Tit for Tat to the introduction of the zero-determinant strategies and the use of sophisticated learning…

计算机科学与博弈论 · 计算机科学 2024-01-25 Nikoleta E. Glynatsi , Vincent Knight , Marc Harper
‹ 上一页 1 2 3 10 下一页 ›