中文
相关论文

相关论文: Cooperation under Incomplete Information on the Di…

200 篇论文

We examine sequential equilibrium in the context of computational games, where agents are charged for computation. In such games, an agent can rationally choose to forget, so issues of imperfect recall arise. In this setting, we consider…

计算机科学与博弈论 · 计算机科学 2014-12-22 Joseph Y. Halpern , Rafael Pass

We consider the one-shot Prisoner's Dilemma between algorithms with read-access to one anothers' source codes, and we use the modal logic of provability to build agents that can achieve mutual cooperation in a manner that is robust, in that…

计算机科学与博弈论 · 计算机科学 2021-04-06 Mihaly Barasz , Paul Christiano , Benja Fallenstein , Marcello Herreshoff , Patrick LaVictoire , Eliezer Yudkowsky

We study a two-player, zero-sum, stochastic game with incomplete information on one side in which the players are allowed to play more and more frequently. The informed player observes the realization of a Markov chain on which the payoffs…

最优化与控制 · 数学 2013-07-15 Pierre Cardaliaguet , Catherine Rainer , Dinah Rosenberg , Nicolas Vieille

Self-serving, rational agents sometimes cooperate to their mutual benefit. The two-player iterated prisoner's dilemma game is a model for including the emergence of cooperation. It is generally believed that there is no simple ultimatum…

计算机科学与博弈论 · 计算机科学 2024-11-08 Jin-Li Guo

We study a model of games that combines concurrency, imperfect information and stochastic aspects. Those are finite states games in which, at each round, the two players choose, simultaneously and independently, an action. Then a successor…

形式语言与自动机理论 · 计算机科学 2011-08-31 Vincent Gripon , Olivier Serre

Social dilemmas, where mutual cooperation can lead to high payoffs but participants face incentives to cheat, are ubiquitous in multi-agent interaction. We wish to construct agents that cooperate with pure cooperators, avoid exploitation by…

人工智能 · 计算机科学 2019-05-27 Alexander Peysakhovich , Adam Lerer

We study multi-player games with perfect information and general payoff function, where the set of stages is the set of non-positive integers $\{\ldots,-2,-1,0\}$. We define two related equilibrium concepts: one considering only deviations…

最优化与控制 · 数学 2025-12-02 Galit Ashkenazi-Golan , János Flesch , Eilon Solan

Games with incomplete preferences are an important model for studying rational decision-making in scenarios where players face incomplete information about their preferences and must contend with incomparable outcomes. We study the problem…

计算机科学与博弈论 · 计算机科学 2024-08-13 Abhishek N. Kulkarni , Jie Fu , Ufuk Topcu

We obtain global, non-asymptotic convergence guarantees for independent learning algorithms in competitive reinforcement learning settings with two agents (i.e., zero-sum stochastic games). We consider an episodic setting where in each…

机器学习 · 计算机科学 2021-01-13 Constantinos Daskalakis , Dylan J. Foster , Noah Golowich

Game theory provides a quantitative framework for analyzing the behavior of rational agents. The Iterated Prisoner's Dilemma in particular has become a standard model for studying cooperation and cheating, with cooperation often emerging as…

种群与进化 · 定量生物学 2015-06-18 Alexander J. Stewart , Joshua B. Plotkin

The effects of payoffs and noise on the maintenance of cooperative behavior are studied in an evolutionary Prisoner's Dilemma game with players located on the sites of different two-dimensional lattices. This system exhibits a phase…

统计力学 · 物理学 2009-11-11 Gyorgy Szabo , Jeromos Vukov , Attila Szolnoki

We investigate symmetric equilibria of mutual reinforcement learning when both players alternately learn the optimal memory-two strategies against the opponent in the repeated prisoners' dilemma game. We provide a necessary condition for…

物理与社会 · 物理学 2023-01-03 Masahiko Ueda

We consider a network of coupled agents playing the Prisoner's Dilemma game, in which players are allowed to pick a strategy in the interval [0,1], with 0 corresponding to defection, 1 to cooperation, and intermediate values representing…

适应与自组织系统 · 物理学 2015-05-28 Francesco Sorrentino , Nicholas Mecholsky

Direct reciprocity and conditional cooperation are important mechanisms to prevent free riding in social dilemmas. But in large groups these mechanisms may become ineffective, because they require single individuals to have a substantial…

种群与进化 · 定量生物学 2014-11-05 Christian Hilbe , Arne Traulsen , Bin Wu , Martin A. Nowak

In this paper, we study cooperation in distributed games under network-constrained communication. Building on the framework of Monderer and Tennenholtz (1999), we derive a sufficient condition for cooperative equilibrium in settings where…

计算机科学与博弈论 · 计算机科学 2025-11-10 Tommy Mordo , Omer Madmon , Moshe Tennenholtz

This paper studies a nonzero-sum Dynkin game in discrete time under non-exponential discounting. For both players, there are two levels of game-theoretic reasoning intertwined. First, each player looks for an intra-personal equilibrium…

最优化与控制 · 数学 2022-05-09 Yu-Jui Huang , Zhou Zhou

Game theory formalizes certain interactions between physical particles or between living beings in biology, sociology, and economics, and quantifies the outcomes by payoffs. The prisoner's dilemma (PD) describes situations in which it is…

物理与社会 · 物理学 2015-05-13 Dirk Helbing , Sergi Lozano

We study best-response type learning dynamics for zero-sum polymatrix games under two information settings. The two settings are distinguished by the type of information that each player has about the game and their opponents' strategy. The…

最优化与控制 · 数学 2025-08-13 Fathima Zarin Faizal , Asuman Ozdaglar , Martin J. Wainwright

In this paper the results of a simulation of a prisoner's dilemma robin-round tournament are presented. In the tournament each participating strategy plays an iterated prisoner's dilemma against each other strategy (round-robin) and as a…

计算机科学与博弈论 · 计算机科学 2014-02-10 Tobias Kretz

Using simulations between pairs of $\epsilon$-greedy q-learners with one-period memory, this article demonstrates that the potential function of the stochastic replicator dynamics (Foster and Young, 1990) allows it to predict the emergence…

理论经济学 · 经济学 2023-02-23 Maximilian Schaefer