中文
相关论文

相关论文: Winning Without Observing Payoffs: Exploiting Beha…

200 篇论文

In repeated interactions between individuals, we do not expect that exactly the same situation will occur from one time to another. Contrary to what is common in models of repeated games in the literature, most real situations may differ a…

种群与进化 · 定量生物学 2007-05-23 Anders Eriksson , Kristian Lindgren

Zero Determinant (ZD) strategies are a new class of probabilistic and conditional strategies that are able to unilaterally set the expected payoff of an opponent in iterated plays of the Prisoner's Dilemma irrespective of the opponent's…

种群与进化 · 定量生物学 2013-08-07 Christoph Adami , Arend Hintze

Multi-round competitions often double or triple the points awarded in the final round, calling it a bonus, to maximize spectators' excitement. In a two-player competition with $n$ rounds, we aim to derive the optimal bonus size to maximize…

计算机科学与博弈论 · 计算机科学 2024-06-10 Zhihuan Huang , Yuqing Kong , Tracy Xiao Liu , Grant Schoenebeck , Shengwei Xu

We consider the problem of learning by demonstration from agents acting in unknown stochastic Markov environments or games. Our aim is to estimate agent preferences in order to construct improved policies for the same task that the agents…

机器学习 · 计算机科学 2014-08-12 Aristide Tossou , Christos Dimitrakakis

We consider the problem of learning by demonstration from agents acting in unknown stochastic Markov environments or games. Our aim is to estimate agent preferences in order to construct improved policies for the same task that the agents…

机器学习 · 统计学 2013-07-16 Aristide C. Y. Tossou , Christos Dimitrakakis

Auctions in which agents' payoffs are random variables have received increased attention in recent years. In particular, recent work in algorithmic mechanism design has produced mechanisms employing internal randomization, partly in…

计算机科学与博弈论 · 计算机科学 2012-06-15 Shaddin Dughmi , Yuval Peres

We introduce and study the problem of detecting whether an agent is updating their prior beliefs given new evidence in an optimal way that is Bayesian, or whether they are biased towards their own prior. In our model, biased agents form…

计算机科学与博弈论 · 计算机科学 2024-10-31 Yiling Chen , Tao Lin , Ariel D. Procaccia , Aaditya Ramdas , Itai Shapira

This paper considers information sharing in a multi-player repeated game. Every round, each player observes a subset of components of a random vector and then takes a control action. The utility earned by each player depends on the full…

最优化与控制 · 数学 2014-12-31 Michael J. Neely

Repeated games have a long tradition in the behavioral sciences and evolutionary biology. Recently, strategies were discovered that permit an unprecedented level of control over repeated interactions by enabling a player to unilaterally…

种群与进化 · 定量生物学 2016-10-25 Alex McAvoy , Christoph Hauert

While discounted payoff games and classic games that reduce to them, like parity and mean-payoff games, are symmetric, their solutions are not. We have taken a fresh view on the properties that optimal solutions need to have, and devised a…

数据结构与算法 · 计算机科学 2026-03-11 Daniele Dell'Erba , Arthur Dumas , Sven Schewe

Von Neumann's Min-Max Theorem guarantees that each player of a zero-sum matrix game has an optimal mixed strategy. This paper gives an elementary proof that each player has a near-optimal mixed strategy that chooses uniformly at random from…

计算复杂性 · 计算机科学 2015-06-02 Richard Lipton , Neal E. Young

How can one detect friendly and adversarial behavior from raw data? Detecting whether an environment is a friend, a foe, or anything in between, remains a poorly understood yet desirable ability for safe and robust agents. This paper…

人工智能 · 计算机科学 2018-07-03 Pedro A. Ortega , Shane Legg

Which classes can be learned properly in the online model? -- that is, by an algorithm that at each round uses a predictor from the concept class. While there are simple and natural cases where improper learning is necessary, it is natural…

机器学习 · 计算机科学 2021-02-03 Steve Hanneke , Roi Livni , Shay Moran

Algorithms for playing in Stackelberg games have been deployed in real-world domains including airport security, anti-poaching efforts, and cyber-crime prevention. However, these algorithms often fail to take into consideration the…

计算机科学与博弈论 · 计算机科学 2024-10-15 Keegan Harris , Zhiwei Steven Wu , Maria-Florina Balcan

We are witnessing an increasing use of data-driven predictive models to inform decisions. As decisions have implications for individuals and society, there is increasing pressure on decision makers to be transparent about their decision…

The integrity of democratic elections depends on voters' access to accurate information. However, modern media environments, which are dominated by social media, provide malicious actors with unprecedented ability to manipulate elections…

计算机科学与博弈论 · 计算机科学 2018-11-22 Bryan Wilder , Yevgeniy Vorobeychik

We study decision-making with rational inattention in settings where agents have perception constraints. In such settings, inaccurate prior beliefs or models of others may lead to inattention blindness, where an agent is unaware of its…

计算机科学与博弈论 · 计算机科学 2025-10-06 Mustafa O. Karabag , Jesse Milzman , Ufuk Topcu

We study online learning in Bayesian Stackelberg games, where a leader repeatedly interacts with a follower whose unknown private type is independently drawn at each round from an unknown probability distribution. The goal is to design…

计算机科学与博弈论 · 计算机科学 2026-02-03 Matteo Bollini , Francesco Bacchiocchi , Samuel Coutts , Matteo Castiglioni , Alberto Marchesi

We investigate how effective an attacker can be when it only learns from its victim's actions, without access to the victim's reward. In this work, we are motivated by the scenario where the attacker wants to behave strategically when the…

机器学习 · 计算机科学 2021-12-03 Ted Fujimoto , Timothy Doster , Adam Attarian , Jill Brandenberger , Nathan Hodas

Many auction settings implicitly or explicitly require that bidders are treated equally ex-ante. This may be because discrimination is philosophically or legally impermissible, or because it is practically difficult to implement or…

计算机科学与博弈论 · 计算机科学 2014-11-06 Christos Tzamos , Christopher A. Wilkens