中文
相关论文

相关论文: A Linear-quadratic Mean-Field Stochastic Stackelbe…

200 篇论文

The risk-neutral LQR controller is optimal for stochastic linear dynamical systems. However, the classical optimal controller performs inefficiently in the presence of low-probability yet statistically significant (risky) events. The…

系统与控制 · 电气工程与系统科学 2023-07-17 Masoud Roudneshin , Saba Sanami , Amir G. Aghdam

Resource competition problems are often modeled using Colonel Blotto games, where players take simultaneous actions. However, many real-world scenarios involve sequential decision-making rather than simultaneous moves. To model these…

计算机科学与博弈论 · 计算机科学 2025-05-13 Yan Liu , Bonan Ni , Weiran Shen , Zihe Wang , Jie Zhang

Existing methods for learning Stackelberg equilibria typically assume that the followers' (variational, generalized) Nash equilibrium is unique. However, in the presence of multiple equilibria, without a selection convention, the problem…

最优化与控制 · 数学 2026-04-30 Silvia Cianchi , Anibal Sanjab , Sergio Grammatico

We study a two-player Stackelberg game with incomplete information such that the follower's strategy belongs to a known family of parameterized functions with an unknown parameter vector. We design an adaptive learning approach to…

计算机科学与博弈论 · 计算机科学 2021-01-12 Guosong Yang , Radha Poovendran , João P. Hespanha

This paper first presents necessary and sufficient conditions for the solvability of discrete time, mean-field, stochastic linear-quadratic optimal control problems. Then, by introducing several sequences of bounded linear operators, the…

最优化与控制 · 数学 2016-07-25 Robert. J Elliott , Xun Li , Yuan-Hua Ni

In this paper we study a class of matrix-valued linear-quadratic mean-field-type games for both the risk-neutral, risk-sensitive and robust cases. Non-cooperation, full cooperation and adversarial between teams are treated. We provide a…

最优化与控制 · 数学 2019-06-06 Julian Barreiro-Gomez , Tyrone E. Duncan , Hamidou Tembine

Formulating cyber-security problems with attackers and defenders as a partially observable stochastic game has become a trend recently. Among them, the one-sided two-player zero-sum partially observable stochastic game (OTZ-POSG) has…

系统与控制 · 电气工程与系统科学 2021-09-20 Wei Zheng , Taeho Jung , Hai Lin

We address two-player general-sum stochastic Stackelberg games (SSGs), where the leader's policy is optimized considering the best-response follower whose policy is optimal for its reward under the leader. Existing policy gradient and value…

计算机科学与博弈论 · 计算机科学 2026-03-17 Mikoto Kudo , Youhei Akimoto

In this paper, we study large population multi-agent reinforcement learning (RL) in the context of discrete-time linear-quadratic mean-field games (LQ-MFGs). Our setting differs from most existing work on RL for MFGs, in that we consider a…

系统与控制 · 电气工程与系统科学 2020-10-02 Muhammad Aneeq uz Zaman , Kaiqing Zhang , Erik Miehling , Tamer Başar

We study a multi-player one-round game termed Stackelberg Network Pricing Game, in which a leader can set prices for a subset of $m$ priceable edges in a graph. The other edges have a fixed cost. Based on the leader's decision one or more…

数据结构与算法 · 计算机科学 2008-02-21 Patrick Briest , Martin Hoefer , Piotr Krysta

This paper investigates an indefinite linear-quadratic partially observed mean-field game with common noise, incorporating both state-average and control-average effects. In our model, each agent's state is observed through both individual…

最优化与控制 · 数学 2025-08-05 Tian Chen , Tianyang Nie , Zhen Wu

Information uncertainty is one of the major challenges facing applications of game theory. In the context of Stackelberg games, various approaches have been proposed to deal with the leader's incomplete knowledge about the follower's…

计算机科学与博弈论 · 计算机科学 2019-05-21 Jiarui Gan , Haifeng Xu , Qingyu Guo , Long Tran-Thanh , Zinovi Rabinovich , Michael Wooldridge

The paper is concerned with the study of a control system consisting of one major agent and many identical minor agents in the limit case when the number of agents tends to infinity. To study the limiting system we use the mean field…

最优化与控制 · 数学 2022-12-13 Yurii Averboukh

In this paper, we consider a partial observed two-person zero-sum stochastic differential game problem where the system is governed by a stochastic differential equation of mean-field type. Under standard assumptions on the coefficients,…

最优化与控制 · 数学 2016-11-15 Maoning Tang , Qingxin Meng

In this paper, we study the transmission strategy adaptation problem in an RF-powered cognitive radio network, in which hybrid secondary users are able to switch between the harvest-then-transmit mode and the ambient backscatter mode for…

网络与互联网体系结构 · 计算机科学 2018-04-10 Wenbo Wang , Dinh Thai Hoang , Dusit Niyato , Ping Wang , Dong In Kim

We study an online learning problem in general-sum Stackelberg games, where players act in a decentralized and strategic manner. We study two settings depending on the type of information for the follower: (1) the limited information…

机器学习 · 计算机科学 2025-05-06 Yaolong Yu , Haipeng Chen

The paper presents a new method for approximating Strong Stackelberg Equilibrium in general-sum sequential games with imperfect information and perfect recall. The proposed approach is generic as it does not rely on any specific properties…

计算机科学与博弈论 · 计算机科学 2022-08-16 Jan Karwowski , Jacek Mańdziuk

As predictive models are deployed into the real world, they must increasingly contend with strategic behavior. A growing body of work on strategic classification treats this problem as a Stackelberg game: the decision-maker "leads" in the…

机器学习 · 计算机科学 2022-02-01 Tijana Zrnic , Eric Mazumdar , S. Shankar Sastry , Michael I. Jordan

In this study, we investigate $N$-player stochastic differential games with regime switching, where the player dynamics are modulated by a finite-state Markov chain. We analyze the associated Nash system, which consists of a system of…

概率论 · 数学 2025-02-26 Mingrui Wang , Prakash Chakraborty

This paper investigates a conditional mean-field type linear quadratic (LQ) optimal control problem with partial observation and regime switching, where the conditional expectations of the state and control given the history of Markov chain…

最优化与控制 · 数学 2025-12-22 Zhongbin Guo , Guangchen Wang