中文
相关论文

相关论文: Cooperation under Incomplete Information on the Di…

200 篇论文

Zero-determinant strategies are a class of memory-one strategies in repeated games which unilaterally enforce linear relationships between payoffs. It has long been unclear for what stage games zero-determinant strategies exist. We provide…

物理与社会 · 物理学 2022-07-12 Masahiko Ueda

We study a simple model of algorithmic collusion in which Q-learning algorithms are designed in a strategic fashion. We let players (\textit{designers}) choose their exploration policy simultaneously prior to letting their algorithms…

理论经济学 · 经济学 2024-09-13 Ivan Conjeaud

It is generally believed that in a situation where individual and collective interests are in conflict, the availability of optional participation is a key mechanism to maintain cooperation. Surprisingly, this effect is sensitive to the use…

物理与社会 · 物理学 2023-06-21 Marcos Cardinot , Colm O'Riordan , Josephine Griffith , Attila Szolnoki

This paper studies repeated games where two players play multiple duopolistic games simultaneously (multimarket contact). A key assumption is that each player receives a noisy and private signal about the other's actions (private monitoring…

计算机科学与博弈论 · 计算机科学 2019-11-26 Atsushi Iwasaki , Tadashi Sekiguchi , Shun Yamamoto , Makoto Yokoo

Semi-Markov model is one of the most general models for stochastic dynamic systems. This paper deals with a two-person zero-sum game for semi-Markov processes. We focus on the expected discounted payoff criterion with state-action-dependent…

计算机科学与博弈论 · 计算机科学 2021-03-09 Zhihui Yu , Xianping Guo , Li Xia

It is well-known that for infinitely repeated games, there are computable strategies that have best responses, but no computable best responses. These results were originally proved for either specific games (e.g., Prisoner's dilemma), or…

计算机科学与博弈论 · 计算机科学 2020-06-11 Jakub Dargaj , Jakob Grue Simonsen

Pricing decisions are increasingly made by AI. Thanks to their ability to train with live market data while making decisions on the fly, deep reinforcement learning algorithms are especially effective in taking such pricing decisions. In…

人工智能 · 计算机科学 2021-07-06 Michael Schlechtinger , Damaris Kosack , Heiko Paulheim , Thomas Fetzer

Recently, Press and Dyson have proposed a new class of probabilistic and conditional strategies for the two-player iterated Prisoner's Dilemma, so-called zero-determinant strategies. A player adopting zero-determinant strategies is able to…

计算机科学与博弈论 · 计算机科学 2014-02-17 Liming Pan , Dong Hao , Zhihai Rong , Tao Zhou

In spatial games players typically alter their strategy by imitating the most successful or one randomly selected neighbor. Since a single neighbor is taken as reference, the information stemming from other neighbors is neglected, which…

物理与社会 · 物理学 2012-11-01 Xiaofeng Wang , Matjaz Perc , Yongkui Liu , Xiaojie Chen , Long Wang

We study a spatial two-strategy (cooperation and defection) Prisoner's Dilemma game with two types ($A$ and $B$) of players located on the sites of a square lattice. The evolution of strategy distribution is governed by iterated strategy…

物理与社会 · 物理学 2009-01-15 Gyorgy Szabo , Attila Szolnoki

The conditional commitment abilities of mutually transparent computer agents have been studied in previous work on commitment games and program equilibrium. This literature has shown how these abilities can help resolve Prisoner's Dilemmas…

计算机科学与博弈论 · 计算机科学 2022-12-06 Anthony DiGiovanni , Jesse Clifton

We study zero-sum differential games with state constraints and one-sided information, where the informed player (Player 1) has a categorical payoff type unknown to the uninformed player (Player 2). The goal of Player 1 is to minimize his…

计算机科学与博弈论 · 计算机科学 2024-06-05 Mukesh Ghimire , Lei Zhang , Zhe Xu , Yi Ren

Consider a very simple class of (finite) games: after an initial move by nature, each player makes one move. Moreover, the players have common interests: at each node, all the players get the same payoff. We show that the problem of…

计算机科学与博弈论 · 计算机科学 2007-05-23 Francis Chu , Joseph Y. Halpern

This paper characterizes differentiable subgame perfect equilibria in a continuous time intertemporal decision optimization problem with non-constant discounting. The equilibrium equation takes two different forms, one of which is…

最优化与控制 · 数学 2007-05-23 Ivar Ekeland , Ali Lazrak

We investigate the mechanism design problem faced by a principal who hires \emph{multiple} agents to gather and report costly information. Then, the principal exploits the information to make an informed decision. We model this problem as a…

计算机科学与博弈论 · 计算机科学 2023-07-13 Federico Cacciamani , Matteo Castiglioni , Nicola Gatti

As machine learning agents act more autonomously in the world, they will increasingly interact with each other. Unfortunately, in many social dilemmas like the one-shot Prisoner's Dilemma, standard game theory predicts that ML agents will…

计算机科学与博弈论 · 计算机科学 2023-11-14 Caspar Oesterheld , Johannes Treutlein , Roger Grosse , Vincent Conitzer , Jakob Foerster

We investigate the spatial distribution and the global frequency of agents who can either cooperate or defect. The agent interaction is described by a deterministic, non-iterated prisoner's dilemma game, further each agent only locally…

统计力学 · 物理学 2007-05-23 Frank Schweitzer , Laxmidhar Behera , Heinz Muehlenbein

Information uncertainty is one of the major challenges facing applications of game theory. In the context of Stackelberg games, various approaches have been proposed to deal with the leader's incomplete knowledge about the follower's…

计算机科学与博弈论 · 计算机科学 2019-05-21 Jiarui Gan , Haifeng Xu , Qingyu Guo , Long Tran-Thanh , Zinovi Rabinovich , Michael Wooldridge

We study a system in which N agents have to decide between two strategies \theta_i (i \in 1... N), for defection or cooperation, when interacting with other n agents (either spatial neighbors or randomly chosen ones). After each round, they…

物理与社会 · 物理学 2012-11-07 Frank Schweitzer , Pavlin Mavrodiev , Claudio J. Tessone

We study a two-player, zero-sum, dynamic game with incomplete information where one of the players is more informed than his opponent. We analyze the limit value as the players play more and more frequently. The more informed player…

最优化与控制 · 数学 2015-09-14 Fabien Gensbittel