中文
相关论文

相关论文: Efficient Stackelberg Strategies for Finitely Repe…

200 篇论文

Designing socially optimal policies in multi-agent environments is a fundamental challenge in both economics and artificial intelligence. This paper studies a general framework for learning Stackelberg equilibria in dynamic and uncertain…

系统与控制 · 电气工程与系统科学 2025-09-23 Jun He , Andrew L. Liu , Yihsu Chen

The Stackelberg security game is played between a defender and an attacker, where the defender needs to allocate a limited amount of resources to multiple targets in order to minimize the loss due to adversarial attack by the attacker.…

计算机科学与博弈论 · 计算机科学 2022-04-27 Rufan Bai , Haoxing Lin , Xinyu Yang , Xiaowei Wu , Minming Li , Weijia Jia

Batch reinforcement learning (RL) defines the task of learning from a fixed batch of data lacking exhaustive exploration. Worst-case optimality algorithms, which calibrate a value-function model class from logged experience and perform some…

机器学习 · 统计学 2023-10-03 Wenzhuo Zhou , Annie Qu

Stackelberg games originate where there are market leaders and followers, and the actions of leaders influence the behavior of the followers. Mathematical modelling of such games results in what's called a Bilevel Optimization problem.…

计算机科学与博弈论 · 计算机科学 2023-12-07 Pravesh Koirala , Forrest Laine

The Stackelberg game depicts a leader-follower relationship wherein decisions are made sequentially, and the Stackelberg equilibrium represents an expected optimal solution when the leader can anticipate the rational response of the…

系统与控制 · 电气工程与系统科学 2024-01-17 Yue Chen , Peng Yi

The hierarchical interaction between the actor and critic in actor-critic based reinforcement learning algorithms naturally lends itself to a game-theoretic interpretation. We adopt this viewpoint and model the actor and critic interaction…

机器学习 · 计算机科学 2021-09-28 Liyuan Zheng , Tanner Fiez , Zane Alumbaugh , Benjamin Chasnov , Lillian J. Ratliff

We study a two-player Stackelberg game with incomplete information such that the follower's strategy belongs to a known family of parameterized functions with an unknown parameter vector. We design an adaptive learning approach to…

计算机科学与博弈论 · 计算机科学 2021-01-12 Guosong Yang , Radha Poovendran , João P. Hespanha

Stackelberg equilibrium is a solution concept in two-player games where the leader has commitment rights over the follower. In recent years, it has become a cornerstone of many security applications, including airport patrolling and…

计算机科学与博弈论 · 计算机科学 2021-02-04 Chun Kai Ling , Noam Brown

This paper studies a class of dynamic Stackelberg games under open-loop information structure with constrained linear agent dynamics and quadratic utility functions. We show two important properties for this class of dynamic Stackelberg…

最优化与控制 · 数学 2016-08-09 Sen Li , Wei Zhang , Jianming Lian , Karanjit Kalsi

This paper proposes and studies a class of discrete-time finite-time-horizon Stackelberg mean-field games, with one leader and an infinite number of identical and indistinguishable followers. In this game, the objective of the leader is to…

最优化与控制 · 数学 2022-10-11 Xin Guo , Anran Hu , Jiacheng Zhang

We study the problem of online learning in Stackelberg games with side information between a leader and a sequence of followers. In every round the leader observes contextual information and commits to a mixed strategy, after which the…

In this paper, we are concerned with the stabilizatbility of Stackelberg game-based systems. In particular, two players are involved in the system where one is the follower to minimize the related cost function and the other is the leader…

最优化与控制 · 数学 2021-05-04 Yue Sun , Juanjuan Xu , Huanshui Zhang

This paper addresses a Stackelberg stochastic linear-quadratic (LQ) differential game under closed-loop information, a problem inherently time-inconsistent. Existing approaches rely on solving two coupled Hamilton-Jacobi-Bellman (HJB)…

最优化与控制 · 数学 2026-04-27 Qi Lü , Bowen Ma , Hanxiao Wang

Information uncertainty is one of the major challenges facing applications of game theory. In the context of Stackelberg games, various approaches have been proposed to deal with the leader's incomplete knowledge about the follower's…

计算机科学与博弈论 · 计算机科学 2019-05-21 Jiarui Gan , Haifeng Xu , Qingyu Guo , Long Tran-Thanh , Zinovi Rabinovich , Michael Wooldridge

Infinitely repeated games can support cooperative outcomes that are not equilibria in the one-shot game. The idea is to make sure that any gains from deviating will be offset by retaliation in future rounds. However, this model of…

计算机科学与博弈论 · 计算机科学 2024-06-04 Ratip Emin Berker , Vincent Conitzer

A Stackelberg game is played between a leader and a follower. The leader first chooses an action, then the follower plays his best response. The goal of the leader is to pick the action that will maximize his payoff given the follower's…

数据结构与算法 · 计算机科学 2015-11-19 Aaron Roth , Jonathan Ullman , Zhiwei Steven Wu

This paper is concerned with the closed-loop Stackelberg strategy for linear-quadratic leader-follower game. Completely different from the open-loop and feedback Stackelberg strategy, the solvability of the closed-loop solution even the…

最优化与控制 · 数学 2024-03-19 Hongdan Li , Juanjuan Xu , Hunashui Zhang

We propose a two-layer, semi-decentralized algorithm to compute a local solution to the Stackelberg equilibrium problem in aggregative games with coupling constraints. Specifically, we focus on a single-leader, multiple-follower problem,…

最优化与控制 · 数学 2022-02-17 Filippo Fabiani , Mohammad Amin Tajeddini , Hamed Kebriaei , Sergio Grammatico

In this study, we explore the application of game theory, in particular Stackelberg games, to address the issue of effective coordination strategy generation for heterogeneous robots with one-way communication. To that end, focusing on the…

机器人学 · 计算机科学 2023-08-01 Yuhan Zhao , Baichuan Huang , Jingjin Yu , Quanyan Zhu

We study payoff manipulation in repeated multi-objective Stackelberg games, where a leader may strategically influence a follower's deterministic best response, e.g., by offering a share of their own payoff. We assume that the follower's…

计算机科学与博弈论 · 计算机科学 2025-08-27 Phurinut Srisawad , Juergen Branke , Long Tran-Thanh