中文
相关论文

相关论文: Convergence of Learning Dynamics in Stackelberg Ga…

200 篇论文

Starting from a heuristic learning scheme for N-person games, we derive a new class of continuous-time learning dynamics consisting of a replicator-like drift adjusted by a penalty term that renders the boundary of the game's strategy space…

最优化与控制 · 数学 2014-04-08 Pierre Coucheney , Bruno Gaujal , Panayotis Mertikopoulos

A recent body of experimental literature has studied empirical game-theoretical analysis, in which we have partial knowledge of a game, consisting of observations of a subset of the pure-strategy profiles and their associated payoffs to…

计算机科学与博弈论 · 计算机科学 2014-02-13 John Fearnley , Martin Gairing , Paul Goldberg , Rahul Savani

Motivated by Generative Adversarial Networks, we study the computation of Nash equilibrium in concave network zero-sum games (NZSGs), a multiplayer generalization of two-player zero-sum games first proposed with linear payoffs. Extending…

机器学习 · 计算机科学 2020-07-13 Amit Kadan , Hu Fu

In this paper, we examine the Nash equilibrium convergence properties of no-regret learning in general N-player games. For concreteness, we focus on the archetypal follow the regularized leader (FTRL) family of algorithms, and we consider…

计算机科学与博弈论 · 计算机科学 2021-02-05 Angeliki Giannou , Emmanouil-Vasileios Vlatakis-Gkaragkounis , Panayotis Mertikopoulos

In this work, a novel Stackelberg game theoretic framework is proposed for trading energy bidirectionally between the demand-response (DR) aggregator and the prosumers. This formulation allows for flexible energy arbitrage and additional…

机器学习 · 计算机科学 2024-10-28 Styliani I. Kampezidou , Justin Romberg , Kyriakos G. Vamvoudakis , Dimitri N. Mavris

Various social contexts ranging from public goods provision to information collection can be depicted as games of strategic interactions, where a player's well-being depends on her own action as well as on the actions taken by her…

物理与社会 · 物理学 2015-04-01 Giulio Cimini , Claudio Castellano , Angel Sánchez

Inverse game theory is utilized to infer the cost functions of all players based on game outcomes. However, existing inverse game theory methods do not consider the learner as an active participant in the game, which could significantly…

计算机科学与博弈论 · 计算机科学 2025-10-20 Jianguo Chen , Jinlong Lei , Biqiang Mu , Yiguang Hong , Hongsheng Qi

In this paper, we study the transmission strategy adaptation problem in an RF-powered cognitive radio network, in which hybrid secondary users are able to switch between the harvest-then-transmit mode and the ambient backscatter mode for…

网络与互联网体系结构 · 计算机科学 2018-04-10 Wenbo Wang , Dinh Thai Hoang , Dusit Niyato , Ping Wang , Dong In Kim

Stackelberg games are a classic example of bilevel optimization problems, which are often encountered in game theory and economics. These are complex problems with a hierarchical structure, where one optimization task is nested within the…

计算机科学与博弈论 · 计算机科学 2013-07-25 Ankur Sinha , Pekka Malo , Anton Frantsev , Kalyanmoy Deb

We study Stackelberg (leader--follower) tuning of network parameters (tolls, capacities, incentives) in combinatorial congestion games, where selfish users choose discrete routes (or other combinatorial strategies) and settle at a…

计算机科学与博弈论 · 计算机科学 2026-02-27 Saeed Masiha , Sepehr Elahi , Negar Kiyavash , Patrick Thiran

We study Stackelberg games where a principal repeatedly interacts with a non-myopic long-lived agent, without knowing the agent's payoff function. Although learning in Stackelberg games is well-understood when the agent is myopic, dealing…

计算机科学与博弈论 · 计算机科学 2025-05-29 Nika Haghtalab , Thodoris Lykouris , Sloan Nietert , Alexander Wei

Adversarial regularization has been shown to improve the generalization performance of deep learning models in various natural language processing tasks. Existing works usually formulate the method as a zero-sum game, which is solved by…

机器学习 · 计算机科学 2022-04-21 Simiao Zuo , Chen Liang , Haoming Jiang , Xiaodong Liu , Pengcheng He , Jianfeng Gao , Weizhu Chen , Tuo Zhao

Learning in games provides a powerful framework to design control policies for self-interested agents that may be coupled through their dynamics, costs, or constraints. We consider the case where the dynamics of the coupled system can be…

系统与控制 · 电气工程与系统科学 2024-09-18 Mostafa M. Shibl , Vijay Gupta

In this paper, the known deterministic linear-quadratic Stackelberg game is revisited, whose open-loop Stackelberg solution actually possesses the nature of time inconsistency. To handle this time inconsistency, {a two-tier game framework…

最优化与控制 · 数学 2022-03-09 Yuan-Hua Ni , Liping Liu , Xinzhen Zhang

We study the problem of learning in zero-sum matrix games with repeated play and bandit feedback. Specifically, we focus on developing uncoupled algorithms that guarantee, without communication between players, the convergence of the…

机器学习 · 计算机科学 2026-04-20 Côme Fiegel , Pierre Ménard , Tadashi Kozuno , Michal Valko , Vianney Perchet

Zero-sum games are a fundamental setting for adversarial training and decision-making in multi-agent learning (MAL). Existing methods often ensure convergence to (approximate) Nash equilibria by introducing a form of regularization. Yet,…

多智能体系统 · 计算机科学 2026-02-10 Tuo Zhang , Leonardo Stella

We study learnability of mixed-strategy Nash Equilibrium (NE) in general finite games using higher-order replicator dynamics as well as classes of higher-order uncoupled heterogeneous dynamics. In higher-order uncoupled learning dynamics,…

多智能体系统 · 计算机科学 2026-05-06 Sarah A. Toonsi , Jeff S. Shamma

Generative adversarial networks (GANs) represent a zero-sum game between two machine players, a generator and a discriminator, designed to learn the distribution of data. While GANs have achieved state-of-the-art performance in several…

机器学习 · 计算机科学 2020-02-24 Farzan Farnia , Asuman Ozdaglar

Correlated equilibrium generalizes Nash equilibrium by allowing a central coordinator to guide players' actions through shared recommendations, similar to how routing apps guide drivers. We investigate how a coordinator can learn a…

计算机科学与博弈论 · 计算机科学 2025-09-16 Zhenlong Fang , Aryan Deshwal , Yue Yu

Data injection attacks have recently emerged as a significant threat on the smart power grid. By launching data injection attacks, an adversary can manipulate the real-time locational marginal prices to obtain economic benefits. Despite the…

密码学与安全 · 计算机科学 2016-04-04 Anibal Sanjab , Walid Saad