中文
相关论文

相关论文: Feedback and Open-Loop Nash Equilibria for LQ Infi…

200 篇论文

We address the problem of finding conditions which guarantee the existence of open-loop Nash equilibria in discrete time dynamic games (DTDGs). The classical approach to DTDGs involves analyzing the problem using optimal control theory…

最优化与控制 · 数学 2015-09-22 Mathew P. Abraham , Ankur A. Kulkarni

In this paper, we consider a linear quadratic stochastic two-person nonzero-sum differential game. Open-loop and closed-loop Nash equilibria are introduced. The existence of the former is characterized by the solvability of a system of…

最优化与控制 · 数学 2016-07-18 Jingrui Sun , Jiongmin Yong

We study reinforcement learning for two-player zero-sum Markov games with simultaneous moves in the finite-horizon setting, where the transition kernel of the underlying Markov games can be parameterized by a linear function over the…

机器学习 · 计算机科学 2022-04-21 Zixiang Chen , Dongruo Zhou , Quanquan Gu

We study the infinite horizon discrete time N-player nonzero-sum Dynkin game ($N \geq 2$) with stopping times as strategies (or pure strategies). We prove existence of an $\varepsilon$-Nash equilibrium point for the game by presenting a…

最优化与控制 · 数学 2022-03-10 Said Hamadène , Mohammed Hassani , Marie-Amélie Morlais

Contemporary applications of machine learning in two-team e-sports and the superior expressivity of multi-agent generative adversarial networks raise important and overlooked theoretical questions regarding optimization in two-team games.…

计算机科学与博弈论 · 计算机科学 2023-04-18 Fivos Kalogiannis , Ioannis Panageas , Emmanouil-Vasileios Vlatakis-Gkaragkounis

We study mean field games and corresponding $N$-player games in continuous time over a finite time horizon where the position of each agent belongs to a finite state space. As opposed to previous works on finite state mean field games, we…

概率论 · 数学 2018-02-01 Alekos Cecchin , Markus Fischer

This paper investigates the convergence time of log-linear learning to an $\epsilon$-efficient Nash equilibrium in potential games, where an efficient Nash equilibrium is defined as the maximizer of the potential function. Previous…

多智能体系统 · 计算机科学 2026-01-13 Anna Maddux , Reda Ouhamma , Maryam Kamgarpour

Consider a strongly monotone game where the players' utility functions include a reward function and a linear term for each dimension, with coefficients that are controlled by the manager. Gradient play converges to a unique Nash…

多智能体系统 · 计算机科学 2026-02-25 Siddharth Chandak , Ilai Bistritz , Nicholas Bambos

This paper combines ideas from Q-learning and fictitious play to define three reinforcement learning procedures which converge to the set of stationary mixed Nash equilibria in identical interest discounted stochastic games. First, we…

计算机科学与博弈论 · 计算机科学 2022-05-17 Lucas Baudin , Rida Laraki

Motivated by Cournot models, this paper proposes novel models of the noncooperative and cooperative differential games with density constraints in infinite dimensions, where markets consist of infinite firms and demand dynamics are governed…

最优化与控制 · 数学 2025-08-20 Zhun Gou , Nan-Jing Huang , Jian-Hao Kang , Jen-Chih Yao

Evolutionary anti-coordination games on networks capture real-world strategic situations such as traffic routing and market competition. In such games, agents maximize their utility by choosing actions that differ from their neighbors'…

计算机科学与博弈论 · 计算机科学 2024-04-02 Zirou Qiu , Chen Chen , Madhav V. Marathe , S. S. Ravi , Daniel J. Rosenkrantz , Richard E. Stearns , Anil Vullikanti

Solving feedback Stackelberg games with nonlinear dynamics and coupled constraints, a common scenario in practice, presents significant challenges. This work introduces an efficient method for computing approximate local feedback…

最优化与控制 · 数学 2025-04-03 Jingqi Li , Somayeh Sojoudi , Claire Tomlin , David Fridovich-Keil

We consider the problem of computing Nash equilibria in potential games where each player's strategy set is subject to private uncoupled constraints. This scenario is frequently encountered in real-world applications like road network…

计算机科学与博弈论 · 计算机科学 2024-02-13 Nikolas Patris , Stelios Stavroulakis , Fivos Kalogiannis , Rose Zhang , Ioannis Panageas

The note considers the problem of computing pure Nash equilibrium (NE) strategies in distributed (i.e., network-based) settings. The paper studies a class of inertial best response dynamics based on the fictitious play (FP) algorithm. It is…

系统与控制 · 计算机科学 2018-04-04 Brian Swenson , Ceyhun Eksin , Soummya Kar , Alejandro Ribeiro

Dynamic games are powerful tools to model multi-agent decision-making, yet computing Nash (generalized Nash) equilibria remains a central challenge in such settings. Complexity arises from tightly coupled optimality conditions, nested…

计算机科学与博弈论 · 计算机科学 2026-02-06 Mahdis Rabbani , Navid Mojahed , Shima Nazari

We analyze novel portfolio liquidation games with self-exciting order flow. Both the N-player game and the mean-field game are considered. We assume that players' trading activities have an impact on the dynamics of future market order…

最优化与控制 · 数学 2020-11-12 Guanxing Fu , Ulrich Horst , Xiaonyu Xia

We develop provably efficient reinforcement learning algorithms for two-player zero-sum finite-horizon Markov games with simultaneous moves. To incorporate function approximation, we consider a family of Markov games where the reward…

机器学习 · 计算机科学 2020-06-25 Qiaomin Xie , Yudong Chen , Zhaoran Wang , Zhuoran Yang

We study the problem of repeated play in a zero-sum game in which the payoff matrix may change, in a possibly adversarial fashion, on each round; we call these Online Matrix Games. Finding the Nash Equilibrium (NE) of a two player zero-sum…

机器学习 · 计算机科学 2020-04-06 Adrian Rivera Cardoso , Jacob Abernethy , He Wang , Huan Xu

We consider stochastic differential games with $N$ nearly identical players, linear-Gaussian dynamics, and infinite horizon discounted quadratic cost. Admissible controls are feedbacks for which the system is ergodic. We first study the…

偏微分方程分析 · 数学 2014-03-18 Fabio S. Priuli

This paper proposes a unifying design framework for dynamic feedback controllers that track solution trajectories of time-varying generalized equations, such as local minimizers of nonlinear programs or competitive equilibria (e.g., Nash)…