中文
相关论文

相关论文: Two-Player Dynamic Potential LQ Games with Sequent…

200 篇论文

We consider multi-agent decision making, where each agent optimizes its cost function subject to constraints. Agents' actions belong to a compact convex Euclidean space and the agents' cost functions are coupled. We propose a distributed…

最优化与控制 · 数学 2016-12-01 Tatiana Tatarenko , Maryam Kamgarpour

This paper investigates the two-person zero-sum stochastic games for piece-wise deterministic Markov decision processes with risk-sensitive finite-horizon cost criterion on a general state space. Here, the transition and cost/reward rates…

最优化与控制 · 数学 2024-05-15 Subrata Golui

In this paper, we consider two-player zero-sum matrix and stochastic games and develop learning dynamics that are payoff-based, convergent, rational, and symmetric between the two players. Specifically, the learning dynamics for matrix…

机器学习 · 计算机科学 2024-09-06 Zaiwei Chen , Kaiqing Zhang , Eric Mazumdar , Asuman Ozdaglar , Adam Wierman

In this paper, we provide exponential rates of convergence to the interior Nash equilibrium for continuous-time dual-space game dynamics such as mirror descent (MD) and actor-critic (AC). We perform our analysis in $N$-player continuous…

最优化与控制 · 数学 2022-02-04 Bolin Gao , Lacra Pavel

This paper studies two-player zero-sum stochastic Bayesian games where each player has its own dynamic state that is unknown to the other player. Using typical techniques, we provide the recursive formulas and sufficient statistics in both…

计算机科学与博弈论 · 计算机科学 2021-05-05 Nabiha Nasir Orpa , Lichun Li

We study linear-quadratic stochastic differential games on directed chains inspired by the directed chain stochastic differential equations introduced by Detering, Fouque, and Ichiba. We solve explicitly for Nash equilibria with a finite…

概率论 · 数学 2020-06-02 Yichen Feng , Jean-Pierre Fouque , Tomoyuki Ichiba

A finite-horizon zero-sum linear-quadratic differential game is considered. Its features are: (i) the control cost of the minimizing player in the game's cost functional is much smaller than the control cost of the maximizing player and the…

最优化与控制 · 数学 2026-04-29 Valery Y. Glizer , Vladimir Turetsky

We analyze independent policy-gradient (PG) learning in $N$-player linear-quadratic (LQ) stochastic differential games. Each player employs a distributed policy that depends only on its own state and updates the policy independently using…

最优化与控制 · 数学 2026-02-19 Philipp Plank , Yufei Zhang

We consider a noncooperative $n$-player principal eigenvalue game which is associated with an infinitesimal generator of a stochastically perturbed multi-channel dynamical system -- where, in the course of such a game, each player attempts…

最优化与控制 · 数学 2018-01-03 Getachew K. Befekadu , Panos J. Antsaklis

Quantum games with incomplete information can be studied within a Bayesian framework. We consider a version of prisoner's dilemma (PD) in this framework with three players and characterize the Nash equilibria. A variation of the standard PD…

量子物理 · 物理学 2017-03-10 Neal Solmeyer , Ricky Dixon , Radhakrishnan Balu

This paper considers discounted infinite horizon mean field games by extending the probabilistic weak formulation of the game as introduced by Carmona and Lacker (2015). Under similar assumptions as in the finite horizon game, we prove…

最优化与控制 · 数学 2024-07-08 René Carmona , Ludovic Tangpi , Kaiwen Zhang

We study a multi-player stochastic differential game, where agents interact through their joint price impact on an asset that they trade to exploit a common trading signal. In this context, we prove that a closed-loop Nash equilibrium…

数理金融 · 定量金融 2023-06-23 Alessandro Micheli , Johannes Muhle-Karbe , Eyal Neuman

Solving feedback Stackelberg games with nonlinear dynamics and coupled constraints, a common scenario in practice, presents significant challenges. This work introduces an efficient method for computing approximate local feedback…

最优化与控制 · 数学 2025-04-03 Jingqi Li , Somayeh Sojoudi , Claire Tomlin , David Fridovich-Keil

The vast majority of products we use daily are supplied to us through complex global supply chains that transform raw materials into finished goods and distribute them to end consumers. This paper proposes a modeling methodology for dynamic…

系统与控制 · 电气工程与系统科学 2024-08-22 Sophie Hall , Laura Guerrini , Florian Dörfler , Dominic Liao-McPherson

Many economic transactions, including those of online markets, have a time lag between the start and end times of transactions. Customers need to wait for completion of their transaction (order fulfillment) and hence are also interested in…

最优化与控制 · 数学 2018-10-19 Manu K. Gupta , N. Hemachandra

Establishing the existence of Nash equilibria for partially observed stochastic dynamic games is known to be quite challenging, with the difficulties stemming from the noisy nature of the measurements available to individual players…

系统与控制 · 计算机科学 2018-06-06 Naci Saldi , Tamer Basar , Maxim Raginsky

We study pure-strategy Nash equilibria in multi-player concurrent deterministic games, for a variety of preference relations. We provide a novel construction, called the suspect game, which transforms a multi-player concurrent game into a…

计算机科学中的逻辑 · 计算机科学 2017-01-11 Patricia Bouyer , Romain Brenguier , Nicolas Markey , Michael Ummels

Finite-horizon probabilistic multiagent concurrent game systems, also known as finite multiplayer stochastic games, are a well-studied model in computer science due to their ability to represent a wide range of real-world scenarios…

计算机科学与博弈论 · 计算机科学 2026-05-27 Senthil Rajasekaran , Moshe Y. Vardi

Dynamic games offer a versatile framework for modeling the evolving interactions of strategic agents, whose steady-state behavior can be captured by the Nash equilibria of the games. Nash equilibria are often computed in feedback, with…

系统与控制 · 电气工程与系统科学 2024-11-22 Chih-Yuan Chiu , Jingqi Li , Maulik Bhatt , Negar Mehr

Game-theoretic MPC (or Receding Horizon Games) is an emerging control methodology for multi-agent systems that generates control actions by solving a dynamic game with coupling constraints in a receding-horizon fashion. This control…

系统与控制 · 电气工程与系统科学 2024-04-19 Sophie Hall , Dominic Liao-McPherson , Giuseppe Belgioioso , Florian Dörfler