中文
相关论文

相关论文: Winning Without Observing Payoffs: Exploiting Beha…

200 篇论文

Human cooperation depends on how accurately we infer others' motives--how much they value fairness, generosity, or self-interest from the choices they make. We model that process in binary dictator games, which isolate moral trade-offs…

神经元与认知 · 定量生物学 2025-11-12 Gregory Stanley , Jun Zhang , Rick Lewis

We consider infinite-state turn-based stochastic games of two players, Box and Diamond, who aim at maximizing and minimizing the expected total reward accumulated along a run, respectively. Since the total accumulated reward is unbounded,…

计算机科学与博弈论 · 计算机科学 2012-08-09 Tomáš Brázdil , Antonín Kučera , Petr Novotný

Overcoming the impact of selfish behavior of rational players in multiagent systems is a fundamental problem in game theory. Without any intervention from a central agent, strategic users take actions in order to maximize their personal…

计算机科学与博弈论 · 计算机科学 2024-09-06 Maria-Florina Balcan , Matteo Pozzi , Dravyansh Sharma

Path planning plays an essential role in many areas of robotics. Various planning techniques have been presented, either focusing on learning a specific task from demonstrations or retrieving trajectories by optimizing for hand-crafted cost…

机器人学 · 计算机科学 2018-09-26 Salvatore Virga , Christian Rupprecht , Nassir Navab , Christoph Hennersperger

Evolutionary game theory provides a mathematical foundation for cross-disciplinary fertilization, especially for integrating ideas from artificial intelligence and game theory. Such integration offers a transparent and rigorous approach to…

物理与社会 · 物理学 2023-05-31 Xingru Chen , Feng Fu

Limited lookahead has been studied for decades in perfect-information games. We initiate a new direction via two simultaneous deviation points: generalization to imperfect-information games and a game-theoretic approach. We study how one…

计算机科学与博弈论 · 计算机科学 2020-03-20 Christian Kroer , Tuomas Sandholm

We consider the problem of learning to exploit learning algorithms through repeated interactions in games. Specifically, we focus on the case of repeated two player, finite-action games, in which an optimizer aims to steer a no-regret…

计算机科学与博弈论 · 计算机科学 2025-05-29 Yizhou Zhang , Yi-An Ma , Eric Mazumdar

Bargaining can be used to resolve mixed-motive games in multi-agent systems. Although there is an abundance of negotiation strategies implemented in automated negotiating agents, most agents are based on single fixed strategies, while it is…

多智能体系统 · 计算机科学 2022-12-21 Bram M. Renting , Holger H. Hoos , Catholijn M. Jonker

Adversarial training, a special case of multi-objective optimization, is an increasingly prevalent machine learning technique: some of its most notable applications include GAN-based generative modeling and self-play techniques in…

Imitation sometimes achieves success in multi-agent situations even though it is very simple. In game theory, success of imitation has been characterized by unbeatability against other agents. Previous studies specified conditions under…

物理与社会 · 物理学 2025-07-23 Masahiko Ueda

This paper examines the convergence of no-regret learning in games with continuous action sets. For concreteness, we focus on learning via "dual averaging", a widely used class of no-regret learning schemes where players take small steps…

最优化与控制 · 数学 2018-01-17 Panayotis Mertikopoulos , Zhengyuan Zhou

We present a general framework for evolutionary learning to emergent unbiased state representation without any supervision. Evolutionary frameworks such as self-play converge to bad local optima in case of multi-agent reinforcement learning…

机器学习 · 统计学 2023-02-03 Shohei Ohsawa

Various social dilemma games that follow different strategy updating rules have been studied on many networks.The reported results span the entire spectrum, from significantly boosting,to marginally affecting,to seriously decreasing the…

物理与社会 · 物理学 2015-06-17 Qiang Zhang , Tianxiao Qi , Keqiang Li , Zengru Di , Jinshan Wu

In this work, we proposed a new $N$-person game in which the players can bet on two options, for example represented by two boxers. Some of the players have privileged information about the boxers and part of them can provide this…

统计力学 · 物理学 2018-04-04 Roberto da Silva , Henrique A. Fernandes

For zero-sum two-player continuous-time games with integral payoff and incomplete information on one side, one shows that the optimal strategy of the informed player can be computed through an auxiliary optimization problem over some…

概率论 · 数学 2008-10-02 Pierre Cardaliaguet , Catherine Rainer

We present an approach for systematically anticipating the actions and policies employed by \emph{oblivious} environments in concurrent stochastic games, while maximizing a reward function. Our main contribution lies in the synthesis of a…

人工智能 · 计算机科学 2024-09-19 Shadi Tasdighi Kalat , Sriram Sankaranarayanan , Ashutosh Trivedi

We study a simple adaptive model in the framework of an N -player normal form game. The model consists of a repeated game where the players only know their own action space and their own payoff scored at each stage, not those of the other…

计算机科学与博弈论 · 计算机科学 2017-06-12 Mario Bravo

We consider imperfect information stochastic games where we require the players to use pure (i.e. non randomised) strategies. We consider reachability, safety, B\"uchi and co-B\"uchi objectives, and investigate the existence of…

形式语言与自动机理论 · 计算机科学 2018-03-28 Arnaud Carayol , Christof Löding , Olivier Serre

In this paper, we consider zero-sum repeated games in which the maximizer is restricted to strategies requiring no more than a limited amount of randomness. Particularly, we analyze the maxmin payoff of the maximizer in two models: the…

信息论 · 计算机科学 2018-10-11 Mehrdad Valizadeh , Amin Gohari

Long-term cooperation, competition, or exploitation among individuals can be modeled through repeated games. In repeated games, Press and Dyson discovered zero-determinant (ZD) strategies that enforce a special relationship between two…

计算机科学与博弈论 · 计算机科学 2022-07-18 Daiki Miyagawa , Azumi Mamiya , Genki Ichinose