English
Related papers

Related papers: Global Convergence of Policy Gradient for Sequenti…

200 papers

We propose locally convergent Nash equilibrium seeking algorithms for $N$-player noncooperative games, which use distributed event-triggered pseudo-gradient estimates. The proposed approach employs sinusoidal perturbations to estimate the…

Optimization and Control · Mathematics 2025-05-13 Victor Hugo Pereira Rodrigues , Tiago Roux Oliveira , Miroslav Krstic , Tamer Basar

We study Nash equilibria learning of a general-sum stochastic game with an unknown transition probability density function. Agents take actions at the current environment state and their joint action influences the transition of the…

Systems and Control · Electrical Eng. & Systems 2022-10-19 Yan Chen , Tao Li

Min-max optimization problems (i.e., min-max games) have attracted a great deal of attention recently as their applicability to a wide range of machine learning problems has become evident. In this paper, we study min-max games with…

Computer Science and Game Theory · Computer Science 2022-08-23 Denizalp Goktas , Amy Greenwald

Reinforcement Learning (RL) has emerged as a powerful framework for sequential decision-making in dynamic environments, particularly when system parameters are unknown. This paper investigates RL-based control for entropy-regularized…

Systems and Control · Electrical Eng. & Systems 2025-12-02 Gabriel Diaz , Lucky Li , Wenhao Zhang

In this paper, we apply the hierarchical strategy to a semilinear weakly degenerate parabolic equation involving a gradient term. We use the Stackelberg-Nash strategy with one leader which tries to drive the solution to zero and two…

Optimization and Control · Mathematics 2022-09-27 Landry Djomegne , Cyrille Kenne , René Dorville , Pascal Zongo

We study incentive designs for a class of stochastic Stackelberg games with one leader and a large number of (finite as well as infinite population of) followers. We investigate whether the leader can craft a strategy under a dynamic…

Computer Science and Game Theory · Computer Science 2024-02-13 Sina Sanjari , Subhonmesh Bose , Tamer Başar

This paper studies an infinite horizon optimal control problem for discrete-time linear system and quadratic criteria, both with random parameters which are independent and identically distributed with respect to time. In this general…

Optimization and Control · Mathematics 2024-03-04 Deyue Li

We present a simple primal-dual algorithm for computing approximate Nash-equilibria in two-person zero-sum sequential games with incomplete information and perfect recall (like Texas Hold'em Poker). Our algorithm is numerically stable,…

Computer Science and Game Theory · Computer Science 2015-12-24 Elvis Dohmatob

This paper presents a novel data-driven approach for approximating the $\varepsilon$-Nash equilibrium in continuous-time linear quadratic Gaussian (LQG) games, where multiple agents interact with each other through their dynamics and…

Systems and Control · Electrical Eng. & Systems 2025-07-22 Zhenhui Xu , Jiayu Chen , Bing-Chang Wang , Tielong Shen

We address two-player general-sum stochastic Stackelberg games (SSGs), where the leader's policy is optimized considering the best-response follower whose policy is optimal for its reward under the leader. Existing policy gradient and value…

Computer Science and Game Theory · Computer Science 2026-03-17 Mikoto Kudo , Youhei Akimoto

We consider the problem of finding a Nash equilibrium (NE) in a general-sum game, where player $i$'s objective is $f_i(x)=f_i(x_1,...,x_n)$, with $x_j\in\mathbb{R}^{d_j}$ denoting the strategy variables of player $j$. Our focus is on…

Computer Science and Game Theory · Computer Science 2026-02-13 Yutong Chao , Jalal Etesami

Despite its popularity in the reinforcement learning community, a provably convergent policy gradient method for continuous space-time control problems with nonlinear state dynamics has been elusive. This paper proposes proximal gradient…

Optimization and Control · Mathematics 2022-12-27 Christoph Reisinger , Wolfgang Stockinger , Yufei Zhang

We study linear quadratic dynamic games where players are uncertain about each other's control policies or goals and consequently seek to be strategically robust. Building on recent work on strategically robust and risk-averse game theory,…

Optimization and Control · Mathematics 2026-04-27 Boris Velasevic , Nicolas Lanzetti , Eric Mazumdar

We address payoff-based decentralized learning in infinite-horizon zero-sum Markov games. In this setting, each player makes decisions based solely on received rewards, without observing the opponent's strategy or actions nor sharing…

Computer Science and Game Theory · Computer Science 2025-02-11 Reda Ouhamma , Maryam Kamgarpour

In this work, we propose a stochastic gradient descent (SGD) framework to design data-driven policy gradient descent algorithms for the linear quadratic regulator problem. Two alternative schemes are considered to estimate the policy…

Systems and Control · Electrical Eng. & Systems 2026-02-24 Bowen Song , Simon Weissmann , Mathias Staudigl , Andrea Iannelli

In this paper, zero-sum mean-field type games (ZSMFTG) with linear dynamics and quadratic cost are studied under infinite-horizon discounted utility function. ZSMFTG are a class of games in which two decision makers whose utilities sum to…

Optimization and Control · Mathematics 2020-09-02 René Carmona , Kenza Hamidouche , Mathieu Laurière , Zongjun Tan

We present a fully-distributed algorithm for Nash equilibrium seeking in aggregative games over networks. The proposed scheme endows each agent with a gradient-based scheme equipped with a tracking mechanism to locally reconstruct the…

Systems and Control · Electrical Eng. & Systems 2025-05-28 Guido Carnevale , Filippo Fabiani , Filiberto Fele , Kostas Margellos , Giuseppe Notarstefano

We develop provably efficient reinforcement learning algorithms for two-player zero-sum finite-horizon Markov games with simultaneous moves. To incorporate function approximation, we consider a family of Markov games where the reward…

Machine Learning · Computer Science 2020-06-25 Qiaomin Xie , Yudong Chen , Zhaoran Wang , Zhuoran Yang

We consider a multi-player stochastic differential game with linear McKean-Vlasov dynamics and quadratic cost functional depending on the variance and mean of the state and control actions of the players in open-loop form. Finite and…

Probability · Mathematics 2018-12-04 Enzo Miller , Huyen Pham

We study decentralized learning in two-player zero-sum discounted Markov games where the goal is to design a policy optimization algorithm for either agent satisfying two properties. First, the player does not need to know the policy of the…

Computer Science and Game Theory · Computer Science 2023-03-07 Zhuoqing Song , Jason D. Lee , Zhuoran Yang