中文
相关论文

相关论文: Convergence of Learning Dynamics in Stackelberg Ga…

200 篇论文

We study repeated games where players use an exponential learning scheme in order to adapt to an ever-changing environment. If the game's payoffs are subject to random perturbations, this scheme leads to a new stochastic version of the…

概率论 · 数学 2010-10-22 Panayotis Mertikopoulos , Aris L. Moustakas

We prove that differential Nash equilibria are generic amongst local Nash equilibria in continuous zero-sum games. That is, there exists an open-dense subset of zero-sum games for which local Nash equilibria are non-degenerate differential…

计算机科学与博弈论 · 计算机科学 2020-02-05 Eric Mazumdar , Lillian Ratliff

We consider a number of questions related to tradeoffs between reward and regret in repeated gameplay between two agents. To facilitate this, we introduce a notion of $\textit{generalized equilibrium}$ which allows for asymmetric regret…

计算机科学与博弈论 · 计算机科学 2023-12-19 William Brown , Jon Schneider , Kiran Vodrahalli

Information uncertainty is one of the major challenges facing applications of game theory. In the context of Stackelberg games, various approaches have been proposed to deal with the leader's incomplete knowledge about the follower's…

计算机科学与博弈论 · 计算机科学 2019-05-21 Jiarui Gan , Haifeng Xu , Qingyu Guo , Long Tran-Thanh , Zinovi Rabinovich , Michael Wooldridge

This paper considers a class of strategic scenarios in which two networks of agents have opposing objectives with regards to the optimization of a common objective function. In the resulting zero-sum game, individual agents collaborate with…

最优化与控制 · 数学 2012-12-24 Bahman Gharesifard , Jorge Cortes

Under what conditions do the behaviors of players, who play a game repeatedly, converge to a Nash equilibrium? If one assumes that the players' behavior is a discrete-time or continuous-time rule whereby the current mixed strategy profile…

计算机科学与博弈论 · 计算机科学 2022-03-29 Jason Milionis , Christos Papadimitriou , Georgios Piliouras , Kelly Spendlove

Zero-sum stochastic games are easy to solve as they can be cast as simple Markov decision processes. This is however not the case with general-sum stochastic games. A fairly general optimization problem formulation is available for…

机器学习 · 计算机科学 2015-07-02 H. L. Prasad , Shalabh Bhatnagar

In stochastic Nash equilibrium problems (SNEPs), it is natural for players to be uncertain about their complex environments and have multi-dimensional unknown parameters in their models. Among various SNEPs, this paper focuses on locally…

最优化与控制 · 数学 2022-04-06 Yuanhanqing Huang , Jianghai Hu

We study online optimization methods for zero-sum games, a fundamental problem in adversarial learning in machine learning, economics, and many other domains. Traditional methods approximate Nash equilibria (NE) using either regret-based…

计算机科学与博弈论 · 计算机科学 2025-07-16 Taemin Kim , James P. Bailey

As predictive models are deployed into the real world, they must increasingly contend with strategic behavior. A growing body of work on strategic classification treats this problem as a Stackelberg game: the decision-maker "leads" in the…

机器学习 · 计算机科学 2022-02-01 Tijana Zrnic , Eric Mazumdar , S. Shankar Sastry , Michael I. Jordan

Distributed Nash equilibrium seeking of aggregative games is investigated and a continuous-time algorithm is proposed. The algorithm is designed by virtue of projected gradient play dynamics and distributed average tracking dynamics, and is…

最优化与控制 · 数学 2021-12-07 Shu Liang , Peng Yi , Yiguang Hong , Kaixiang Peng

Evolutionary anti-coordination games on networks capture real-world strategic situations such as traffic routing and market competition. In such games, agents maximize their utility by choosing actions that differ from their neighbors'…

计算机科学与博弈论 · 计算机科学 2024-04-02 Zirou Qiu , Chen Chen , Madhav V. Marathe , S. S. Ravi , Daniel J. Rosenkrantz , Richard E. Stearns , Anil Vullikanti

We study the existence and computation of Nash equilibria in concave games where the players' admissible strategies are subject to shared coupling constraints. Under playerwise concavity of constraints, we prove existence of Nash…

计算机科学与博弈论 · 计算机科学 2026-02-09 Philip Jordan , Maryam Kamgarpour

The designs of many large-scale systems today, from traffic routing environments to smart grids, rely on game-theoretic equilibrium concepts. However, as the size of an $N$-player game typically grows exponentially with $N$, standard game…

This work explores three-player game training dynamics, under what conditions three-player games converge and the equilibria the converge on. In contrast to prior work, we examine a three-player game architecture in which all players…

机器学习 · 计算机科学 2022-08-16 Kenneth Christofferson , Fernando J. Yanez

This paper is concerned with a linear-quadratic partially observed Stackelberg stochastic differential game with correlated state and observation noises, where the diffusion coefficient does not contain the control variable and the control…

最优化与控制 · 数学 2021-05-25 Yueyang Zheng , Jingtao Shi

This paper is devoted to a high-dimensional mixed leadership stochastic differential game on a finite horizon in feedback information mode, where the control variables enter into the diffusion term of state equation. A verification theorem…

最优化与控制 · 数学 2022-11-28 Qi Huang , Jingtao Shi

Learning in games refers to scenarios where multiple players interact in a shared environment, each aiming to minimize their regret. An equilibrium can be computed at a fast rate of $O(1/T)$ when all players follow the optimistic…

计算机科学与博弈论 · 计算机科学 2025-02-18 Taira Tsuchiya , Shinji Ito , Haipeng Luo

We analyze best response dynamics for finding a Nash equilibrium of an infinite horizon zero-sum stochastic linear quadratic dynamic game (LQDG) with partial and asymmetric information. We derive explicit expressions for each player's best…

系统与控制 · 电气工程与系统科学 2025-09-03 Yuxiang Guan , Iman Shames , Tyler H. Summers

Consider a set of agents who play a network game repeatedly. Agents may not know the network. They may even be unaware that they are interacting with other agents in a network. Possibly, they just understand that their payoffs depend on an…

理论经济学 · 经济学 2022-07-26 Pierpaolo Battigalli , Fabrizio Panebianco , Paolo Pin