English
Related papers

Related papers: Forward-Backward Dynamic Programming for LQG Dynam…

200 papers

In this article, we study a continuous-time stochastic $H_\infty$ control problem based on reinforcement learning (RL) techniques that can be viewed as solving a stochastic linear-quadratic two-person zero-sum differential game (LQZSG).…

Optimization and Control · Mathematics 2024-10-02 Zhongshi Sun , Guangyan Jia

Limited lookahead has been studied for decades in perfect-information games. We initiate a new direction via two simultaneous deviation points: generalization to imperfect-information games and a game-theoretic approach. We study how one…

Computer Science and Game Theory · Computer Science 2020-03-20 Christian Kroer , Tuomas Sandholm

A fundamental problem in noncooperative dynamic game theory is the computation of Nash equilibria under different information structures, which specify the information available to each agent during decision-making. Prior work has…

Computer Science and Game Theory · Computer Science 2026-03-20 Janani S K , Kushagra Gupta , Ufuk Topcu , David Fridovich-Keil

A multi-agent system operates in an uncertain environment about which agents have different and time varying beliefs that, as time progresses, converge to a common belief. A global utility function that depends on the realized state of the…

Computer Science and Game Theory · Computer Science 2016-02-08 Ceyhun Eksin , Alejandro Ribeiro

We consider distributed learning problem in games with an unknown cost-relevant parameter, and aim to find the Nash equilibrium while learning the true parameter. Inspired by the social learning literature, we propose a distributed…

Optimization and Control · Mathematics 2023-03-14 Shijie Huang , Jinlong Lei , Yiguang Hong

This paper presents a novel approach to numerically solve stochastic differential games for nonlinear systems. The proposed approach relies on the nonlinear Feynman-Kac theorem that establishes a connection between parabolic deterministic…

Optimization and Control · Mathematics 2019-06-13 Ziyi Wang , Keuntaek Lee , Marcus A. Pereira , Ioannis Exarchos , Evangelos A. Theodorou

We investigate a linear quadratic stochastic zero-sum game where two players lobby a political representative to invest in a wind turbine farm. Players are time-inconsistent because they discount performance with a non-constant rate. Our…

General Economics · Economics 2023-09-04 Ali Lazrak , Hanxiao Wang , Jiongmin Yong

An iterative finite difference scheme for mean field games (MFGs) is proposed. The target MFGs are derived from control problems for multidimensional systems with advection terms. For such MFGs, linearization using the Cole-Hopf…

Optimization and Control · Mathematics 2023-04-26 Daisuke Inoue , Yuji Ito , Takahito Kashiwabara , Norikazu Saito , Hiroaki Yoshida

We present a framework that incorporates the idea of bounded rationality into dynamic stochastic pursuit-evasion games. The solution of a stochastic game is characterized, in general, by its (Nash) equilibria in feedback form. However,…

Systems and Control · Electrical Eng. & Systems 2020-03-17 Yue Guan , Dipankar Maity , Christopher M. Kroninger , Panagiotis Tsiotras

This study investigates differential games with motion-payoff uncertainty in continuous-time settings. We propose a framework where players update their beliefs about uncertain parameters using continuous Bayesian updating. Theoretical…

Multiagent Systems · Computer Science 2025-09-16 Jiangjing Zhou , Ovanes Petrosian , Ye Zhang , Hongwei Gao

The mathematical framework of hybrid system is a recent and general tool to treat control systems involving control action of heterogeneous nature. In this paper, we construct and test a semi-Lagrangian numerical scheme for solving the…

Numerical Analysis · Mathematics 2016-08-03 Roberto Ferretti , Achille Sassi

We consider stochastic differential games with $N$ players, linear-Gaussian dynamics in arbitrary state-space dimension, and long-time-average cost with quadratic running cost. Admissible controls are feedbacks for which the system is…

Analysis of PDEs · Mathematics 2014-07-10 Martino Bardi , Fabio S. Priuli

Reinforcement learning is a powerful tool to learn the optimal policy of possibly multiple agents by interacting with the environment. As the number of agents grow to be very large, the system can be approximated by a mean-field problem.…

Optimization and Control · Mathematics 2020-08-18 Weichen Wang , Jiequn Han , Zhuoran Yang , Zhaoran Wang

In this paper, we study large population multi-agent reinforcement learning (RL) in the context of discrete-time linear-quadratic mean-field games (LQ-MFGs). Our setting differs from most existing work on RL for MFGs, in that we consider a…

Systems and Control · Electrical Eng. & Systems 2020-10-02 Muhammad Aneeq uz Zaman , Kaiqing Zhang , Erik Miehling , Tamer Başar

In this paper, zero-sum mean-field type games (ZSMFTG) with linear dynamics and quadratic utility are studied under infinite-horizon discounted utility function. ZSMFTG are a class of games in which two decision makers whose utilities sum…

Optimization and Control · Mathematics 2020-09-07 René Carmona , Kenza Hamidouche , Mathieu Laurière , Zongjun Tan

With the outstanding performance of policy gradient (PG) method in the reinforcement learning field, the convergence theory of it has aroused more and more interest recently. Meanwhile, the significant importance and abundant theoretical…

Optimization and Control · Mathematics 2024-04-19 Xinpei Zhang , Guangyan Jia

We consider the problem of planning under observation and motion uncertainty for nonlinear robotics systems. Determining the optimal solution to this problem, generally formulated as a Partially Observed Markov Decision Process (POMDP), is…

Robotics · Computer Science 2017-07-11 Mohammadhussein Rafieisakhaei , Suman Chakravorty , P. R. Kumar

We develop a hierarchical Bayesian dynamic game for competitive inventory and pricing under incomplete information. Two firms repeatedly choose order quantities and prices while facing two layers of uncertainty: unknown market demand and…

Methodology · Statistics 2026-03-09 Debashis Chatterjee

We consider two-player zero-sum stochastic games and propose a two-timescale $Q$-learning algorithm with function approximation that is payoff-based, convergent, rational, and symmetric between the two players. In two-timescale…

Machine Learning · Computer Science 2023-12-11 Zaiwei Chen , Kaiqing Zhang , Eric Mazumdar , Asuman Ozdaglar , Adam Wierman

We explore reinforcement learning methods for finding the optimal policy in the linear quadratic regulator (LQR) problem. In particular, we consider the convergence of policy gradient methods in the setting of known and unknown parameters.…

Machine Learning · Computer Science 2021-06-25 Ben Hambly , Renyuan Xu , Huining Yang