English
Related papers

Related papers: Learning to Play Multi-Follower Bayesian Stackelbe…

200 papers

Two-player mean-payoff Stackelberg games are nonzero-sum infinite duration games played on a bi-weighted graph by Leader (Player 0) and Follower (Player 1). Such games are played sequentially: first, Leader announces her strategy, second,…

Optimization and Control · Mathematics 2021-08-04 Mrudula Balachander , Shibashis Guha , Jean-François Raskin

With the constraint of a no regret follower, will the players in a two-player Stackelberg game still reach Stackelberg equilibrium? We first show when the follower strategy is either reward-average or transform-reward-average, the two…

Computer Science and Game Theory · Computer Science 2024-08-27 Xiangge Huang , Jingyuan Li , Jiaqing Xie

This paper proposes and studies a class of discrete-time finite-time-horizon Stackelberg mean-field games, with one leader and an infinite number of identical and indistinguishable followers. In this game, the objective of the leader is to…

Optimization and Control · Mathematics 2022-10-11 Xin Guo , Anran Hu , Jiacheng Zhang

We consider the problem of asynchronous online combinatorial optimization on a network of communicating agents. At each time step, some of the agents are stochastically activated, requested to make a prediction, and the system pays the…

Machine Learning · Computer Science 2021-02-10 Riccardo Della Vecchia , Tommaso Cesari

We consider the problem of learning in single-player and multiplayer multiarmed bandit models. Bandit problems are classes of online learning problems that capture exploration versus exploitation tradeoffs. In a multiarmed bandit model,…

Machine Learning · Statistics 2016-12-02 Naumaan Nayyar , Dileep Kalathil , Rahul Jain

This article studies linear-quadratic Stackelberg games between two dominating players (or equivalently, leaders) and a large group of followers, each of whom interacts under a mean field game (MFG) framework. Unlike the conventional…

Optimization and Control · Mathematics 2024-07-02 Dantong Chu , Kenneth Tsz Hin Ng , Sheung Chi Phillip Yam , Harry Zheng

We study a bilevel optimization problem which is a zero-sum Stackelberg game. In this problem, there are two players, a leader and a follower, who pick items from a common set. Both the leader and the follower have their own…

Data Structures and Algorithms · Computer Science 2022-04-26 Lin Chen , Xiaoyu Wu , Guochuan Zhang

We consider online no-regret learning in unknown games with bandit feedback, where each player can only observe its reward at each time -- determined by all players' current joint action -- rather than its gradient. We focus on the class of…

Machine Learning · Computer Science 2024-04-01 Wenjia Ba , Tianyi Lin , Jiawei Zhang , Zhengyuan Zhou

To take advantage of strategy commitment, a useful tactic of playing games, a leader must learn enough information about the follower's payoff function. However, this leaves the follower a chance to provide fake information and influence…

Computer Science and Game Theory · Computer Science 2023-06-14 Yurong Chen , Xiaotie Deng , Yuhao Li

This paper is concerned with a Stackelberg stochastic differential game with asymmetric noisy observation, with one follower and one leader. In our model, the follower cannot observe the state process directly, but could observe a noisy…

Optimization and Control · Mathematics 2020-07-14 Yueyang Zheng , Jingtao Shi

Most models of Stackelberg security games assume that the attacker only knows the defender's mixed strategy, but is not able to observe (even partially) the instantiated pure strategy. Such partial observation of the deployed pure strategy…

Computer Science and Game Theory · Computer Science 2015-05-05 Haifeng Xu , Albert X. Jiang , Arunesh Sinha , Zinovi Rabinovich , Shaddin Dughmi , Milind Tambe

This paper is concerned with a linear-quadratic partially observed mean field Stackelberg stochastic differential game, which contains a leader and a large number of followers. Specifically, the followers confront a large-population Nash…

Optimization and Control · Mathematics 2025-12-09 Yu Si , Yueyang Zheng , Jingtao Shi

Partial monitoring is a general model for sequential learning with limited feedback formalized as a game between two players. In this game, the learner chooses an action and at the same time the opponent chooses an outcome, then the learner…

Machine Learning · Statistics 2015-10-01 Junpei Komiyama , Junya Honda , Hiroshi Nakagawa

Resource competition problems are often modeled using Colonel Blotto games, where players take simultaneous actions. However, many real-world scenarios involve sequential decision-making rather than simultaneous moves. To model these…

Computer Science and Game Theory · Computer Science 2025-05-13 Yan Liu , Bonan Ni , Weiran Shen , Zihe Wang , Jie Zhang

Team formation is ubiquitous in many sectors: education, labor markets, sports, etc. A team's success depends on its members' latent types, which are not directly observable but can be (partially) inferred from past performances. From the…

Computer Science and Game Theory · Computer Science 2022-10-18 Matthew Eichhorn , Siddhartha Banerjee , David Kempe

We study a collaborative multi-agent stochastic linear bandit setting, where $N$ agents that form a network communicate locally to minimize their overall regret. In this setting, each agent has its own linear bandit problem (its own reward…

Machine Learning · Computer Science 2022-05-16 Ahmadreza Moradipari , Mohammad Ghavamzadeh , Mahnoosh Alizadeh

We study Stackelberg (leader--follower) tuning of network parameters (tolls, capacities, incentives) in combinatorial congestion games, where selfish users choose discrete routes (or other combinatorial strategies) and settle at a…

Computer Science and Game Theory · Computer Science 2026-02-27 Saeed Masiha , Sepehr Elahi , Negar Kiyavash , Patrick Thiran

We study an online linear classification problem, in which the data is generated by strategic agents who manipulate their features in an effort to change the classification outcome. In rounds, the learner deploys a classifier, and an…

Machine Learning · Computer Science 2017-10-24 Jinshuo Dong , Aaron Roth , Zachary Schutzman , Bo Waggoner , Zhiwei Steven Wu

We study reputation formation where a long-run player repeatedly observes private signals and takes actions. Short-run players observe the long-run player's past actions but not her past signals. The long-run player can thus develop a…

Theoretical Economics · Economics 2025-07-22 Daniel Luo , Alexander Wolitzky

The $1-N$ generalized Stackelberg game (single-leader multi-follower game) is intricately intertwined with the interaction between a leader and followers (hierarchical interaction) and the interaction among followers (simultaneous…

Computer Science and Game Theory · Computer Science 2023-06-12 Jaeyeon Jo , Jihwan Yu , Jinkyoo Park