中文
相关论文

相关论文: Memory-two strategies forming symmetric mutual rei…

200 篇论文

In an iterated game between two players, there is much interest in characterizing the set of feasible payoffs for both players when one player uses a fixed strategy and the other player is free to switch. Such characterizations have led to…

种群与进化 · 定量生物学 2022-02-18 Alex McAvoy , Martin A. Nowak

We consider turn-based stochastic two-player games with a combination of a parity condition that must hold surely, that is in all possible outcomes, and of a parity condition that must hold almost-surely, that is with probability 1. The…

计算机科学与博弈论 · 计算机科学 2026-01-08 Laurent Doyen , Shibashis Guha

Reciprocity is an important feature of human social interaction and underpins our cooperative nature. What is more, simple forms of reciprocity have proved remarkably resilient in matrix game social dilemmas. Most famously, the tit-for-tat…

多智能体系统 · 计算机科学 2019-03-20 Tom Eccles , Edward Hughes , János Kramár , Steven Wheelwright , Joel Z. Leibo

Learning problems commonly exhibit an interesting feedback mechanism wherein the population data reacts to competing decision makers' actions. This paper formulates a new game theoretic framework for this phenomenon, called "multi-player…

计算机科学与博弈论 · 计算机科学 2022-04-08 Adhyyan Narang , Evan Faulkner , Dmitriy Drusvyatskiy , Maryam Fazel , Lillian J. Ratliff

In repeated interactions between individuals, we do not expect that exactly the same situation will occur from one time to another. Contrary to what is common in models of repeated games in the literature, most real situations may differ a…

种群与进化 · 定量生物学 2007-05-23 Anders Eriksson , Kristian Lindgren

Do boundedly rational players learn to choose equilibrium strategies as they play a game repeatedly? A large literature in behavioral game theory has proposed and experimentally tested various learning algorithms, but a comparative analysis…

经济学 · 定量金融 2021-09-03 Marco Pangallo , James Sanders , Tobias Galla , Doyne Farmer

Across many domains of interaction, both natural and artificial, individuals use past experience to shape future behaviors. The results of such learning processes depend on what individuals wish to maximize. A natural objective is one's own…

种群与进化 · 定量生物学 2022-09-02 Alex McAvoy , Julian Kates-Harbeck , Krishnendu Chatterjee , Christian Hilbe

Punishment is a common tactic to sustain cooperation and has been extensively studied for a long time. While most of previous game-theoretic work adopt the imitation learning where players imitate the strategies who are better off, the…

种群与进化 · 定量生物学 2024-12-20 Chenyang Zhao , Guozhong Zheng , Chun Zhang , Jiqiang Zhang , Li Chen

For the iterated Prisoner's Dilemma, there exist Markov strategies which solve the problem when we restrict attention to the long term average payoff. When used by both players these assure the cooperative payoff for each of them. Neither…

动力系统 · 数学 2017-04-27 Ethan Akin

This paper addresses a mathematically tractable model of the Prisoner's Dilemma using the framework of active inference. In this work, we design pairs of Bayesian agents that are tracking the joint game state of their and their opponent's…

物理与社会 · 物理学 2023-08-31 Daphne Demekas , Conor Heins , Brennan Klein

We examine the effects of memory and different updating paradigms in a game-theoretic model of competitive learning, where agents are influenced in their choice of strategy by both the choices made by, and the consequent success rates of,…

物理与社会 · 物理学 2012-01-23 Ajaz Ahmad Bhat , Anita Mehta

Zero-determinant strategies are memory-one strategies in repeated games which unilaterally enforce linear relations between expected payoffs of players. Recently, the concept of zero-determinant strategies was extended to the class of…

最优化与控制 · 数学 2022-09-07 Masahiko Ueda

We study the limiting behavior of the mixed strategies that result from optimal no-regret learning strategies in a repeated game setting where the stage game is any 2 by 2 competitive game. We consider optimal no-regret algorithms that are…

计算机科学与博弈论 · 计算机科学 2022-03-03 Vidya Muthukumar , Soham Phade , Anant Sahai

Learning algorithm design for state-based games is investigated. A heuristic uncoupled learning algorithm, which is a two memory better reply with inertia dynamics, is proposed. Under certain reasonable conditions it is proved that for any…

最优化与控制 · 数学 2018-09-18 Changxi Li , Yu Xing , Fenghua He , Daizhan Cheng

Repeated games consider a situation where multiple agents are motivated by their independent rewards throughout learning. In general, the dynamics of their learning become complex. Especially when their rewards compete with each other like…

计算机科学与博弈论 · 计算机科学 2023-05-23 Yuma Fujimoto , Kaito Ariu , Kenshi Abe

This paper presents an algorithmic framework for learning robust policies in asymmetric imperfect-information games, where the joint reward could depend on the uncertain opponent type (a private information known only to the opponent itself…

人工智能 · 计算机科学 2020-03-05 Macheng Shen , Jonathan P. How

Colonel Blotto games with discrete strategy spaces effectively illustrate the intricate nature of multidimensional strategic reasoning. This paper studies the equilibrium set of such games where, in line with prior experimental work, the…

计算机科学与博弈论 · 计算机科学 2024-03-28 Christian Ewerhart , Stanisław Kaźmierowski

Two-player games have had a long and fruitful history of applications stretching across the social, biological, and physical sciences. Most applications of two-player games assume synchronous decisions or moves even when the games are…

生物物理 · 物理学 2017-12-15 Robert D. Young

In real-world scenarios, individuals often cooperate for mutual benefit. However, differences in wealth can lead to varying outcomes for similar actions. In complex social networks, individuals' choices are also influenced by their…

物理与社会 · 物理学 2024-07-08 Yunhao Ding , Chunyan Zhang , Jianlei Zhang

The Prisoner's Dilemma game has a long history stretching across the social, biological, and physical sciences. In 2012, Press and Dyson developed a method for analyzing the mapping of the 8-dimensional strategy profile onto the…

物理与社会 · 物理学 2019-11-28 Robert D. Young