中文
相关论文

相关论文: No-Regret Learning in Bayesian Games

200 篇论文

A Bayesian game is a game of incomplete information in which the rules of the game are not fully known to all players. We consider the Bayesian game of Battle of Sexes that has several Bayesian Nash equilibria and investigate its outcome…

量子物理 · 物理学 2014-11-19 Azhar Iqbal , James M. Chappell , Qiang Li , Charles E. M. Pearce , Derek Abbott

In this paper, we extend the Descent framework, which enables learning and planning in the context of two-player games with perfect information, to the framework of stochastic games. We propose two ways of doing this, the first way…

人工智能 · 计算机科学 2023-02-10 Quentin Cohen-Solal , Tristan Cazenave

Bayesian optimal experiments that maximize the information gained from collected data are critical to efficiently identify behavioral models. We extend a seminal method for designing Bayesian optimal experiments by introducing two…

应用统计 · 统计学 2025-03-19 Stefano Balietti , Brennan Klein , Christoph Riedl

We study the Bayesian coarse correlated equilibrium (BCCE) of continuous and discretised first-price and all-pay auctions under the standard symmetric independent private-values model. Our study is motivated by the question of how the…

计算机科学与博弈论 · 计算机科学 2024-11-19 Mete Şeref Ahunbay , Martin Bichler

In stochastic games with incomplete information, the uncertainty is evoked by the lack of knowledge about a player's own and the other players' types, i.e. the utility function and the policy space, and also the inherent stochasticity of…

机器学习 · 计算机科学 2022-03-21 Hannes Eriksson , Debabrota Basu , Mina Alibeigi , Christos Dimitrakakis

Driven by recent successes in two-player, zero-sum game solving and playing, artificial intelligence work on games has increasingly focused on algorithms that produce equilibrium-based strategies. However, this approach has been less…

计算机科学与博弈论 · 计算机科学 2022-06-24 Dustin Morrill , Ryan D'Orazio , Reca Sarfati , Marc Lanctot , James R. Wright , Amy Greenwald , Michael Bowling

Claude Shannon's zero-error communication paradigm reshaped our understanding of fault-tolerant information transfer. Here, we adapt this notion into game theory with incomplete information. We ask: can players with private information…

Best-response (BR) schemes represent an important avenue for learning equilibria in noncooperative games. However, extant rate guarantees for BR schemes generally necessitate stringent smoothness requirements on player objectives and the…

最优化与控制 · 数学 2026-03-03 Zhuoyu Xiao , Uday V. Shanbhag

We show that learning algorithms satisfying a $\textit{low approximate regret}$ property experience fast convergence to approximate optimality in a large class of repeated games. Our property, which simply requires that each learner has…

计算机科学与博弈论 · 计算机科学 2016-12-19 Dylan J. Foster , Zhiyuan Li , Thodoris Lykouris , Karthik Sridharan , Eva Tardos

We consider extensive games with perfect information with well-founded game trees and study the problems of existence and of characterization of the sets of subgame perfect equilibria in these games. We also provide such characterizations…

计算机科学与博弈论 · 计算机科学 2021-06-23 Krzysztof R. Apt , Sunil Simon

An ideal strategy in zero-sum games should not only grant the player an average reward no less than the value of Nash equilibrium, but also exploit the (adaptive) opponents when they are suboptimal. While most existing works in Markov games…

机器学习 · 计算机科学 2022-06-15 Qinghua Liu , Yuanhao Wang , Chi Jin

We consider a class of two-player dynamic stochastic nonzero-sum games where the state transition and observation equations are linear, and the primitive random variables are Gaussian. Each controller acquires possibly different dynamic…

系统与控制 · 计算机科学 2014-01-21 Abhishek Gupta , Ashutosh Nayyar , Cedric Langbort , Tamer Basar

Deception plays a critical role in many interactions in communication and network security. Game-theoretic models called "cheap talk signaling games" capture the dynamic and information asymmetric nature of deceptive interactions. But…

密码学与安全 · 计算机科学 2017-10-17 Jeffrey Pawlick , Quanyan Zhu

Understanding the convergence landscape of multi-agent learning is a fundamental problem of great practical relevance in many applications of artificial intelligence and machine learning. While it is known that learning dynamics converge to…

计算机科学与博弈论 · 计算机科学 2025-03-21 Martin Bichler , Davide Legacci , Panayotis Mertikopoulos , Matthias Oberlechner , Bary Pradelski

The computational study of equilibria involving constraints on players' strategies has been largely neglected. However, in real-world applications, players are usually subject to constraints ruling out the feasibility of some of their…

计算机科学与博弈论 · 计算机科学 2024-08-08 Martino Bernasconi , Matteo Castiglioni , Alberto Marchesi , Francesco Trovò , Nicola Gatti

We study Bayesian coordination games where agents receive noisy private information over the game's payoffs, and over each others' actions. If private information over actions is of low quality, equilibrium uniqueness obtains in a manner…

综合经济学 · 经济学 2021-11-23 Dominik Grafenhofer , Wolfgang Kuhle

Quantum entanglement has been recently demonstrated as a useful resource in conflicting interest games of incomplete information between two players, Alice and Bob [Pappa et al., Phys. Rev. Lett. 114, 020401 (2015)]. General setting for…

量子物理 · 物理学 2017-11-17 Ashutosh Rai , Goutam Paul

In this work, we study the system of interacting non-cooperative two Q-learning agents, where one agent has the privilege of observing the other's actions. We show that this information asymmetry can lead to a stable outcome of population…

机器学习 · 计算机科学 2021-01-26 Ezra Tampubolon , Haris Ceribasic , Holger Boche

We study online learning settings in which experts act strategically to maximize their influence on the learning algorithm's predictions by potentially misreporting their beliefs about a sequence of binary events. Our goal is twofold.…

机器学习 · 计算机科学 2020-07-02 Rupert Freeman , David M. Pennock , Chara Podimata , Jennifer Wortman Vaughan

How should a player who repeatedly plays a game against a no-regret learner strategize to maximize his utility? We study this question and show that under some mild assumptions, the player can always guarantee himself a utility of at least…

计算机科学与博弈论 · 计算机科学 2025-11-12 Yuan Deng , Jon Schneider , Balusubramanian Sivan