English
Related papers

Related papers: Mutually Quadratically Invariant Information Struc…

200 papers

We propose the first model-free algorithm that achieves low regret performance for decentralized learning in two-player zero-sum tabular stochastic games with infinite-horizon average-reward objective. In decentralized learning, the…

Machine Learning · Computer Science 2023-01-16 Romain Cravic , Nicolas Gast , Bruno Gaujal

There are only a few learning algorithms applicable to stochastic dynamic teams and games which generalize Markov decision processes to decentralized stochastic control problems involving possibly self-interested decision makers. Learning…

Optimization and Control · Mathematics 2016-05-03 Gürdal Arslan , Serdar Yüksel

Interaction-aware trajectory planning is crucial for closing the gap between autonomous racing cars and human racing drivers. Prior work has applied game theory as it provides equilibrium concepts for non-cooperative dynamic problems. With…

Robotics · Computer Science 2024-02-06 Matthias Rowold , Alexander Langmann , Boris Lohmann , Johannes Betz

In this paper we establish a new connection between a class of 2-player nonzero-sum games of optimal stopping and certain $2$-player nonzero-sum games of singular control. We show that whenever a Nash equilibrium in the game of stopping is…

Optimization and Control · Mathematics 2017-12-29 Tiziano De Angelis , Giorgio Ferrari

We consider a non-cooperative constrained stochastic games with N players with the following special structure. With each player there is an associated controlled Markov chain. The transition probabilities of the i-th Markov chain depend…

Information Theory · Computer Science 2007-07-13 E. Altman , K. Avrachenkov , N. Bonneau , M. Debbah , R. El-Azouzi , D. Sadoc Menasche

We study the global convergence of policy optimization for finding the Nash equilibria (NE) in zero-sum linear quadratic (LQ) games. To this end, we first investigate the landscape of LQ games, viewing it as a nonconvex-nonconcave…

Machine Learning · Computer Science 2021-02-12 Kaiqing Zhang , Zhuoran Yang , Tamer Başar

The focus of this paper is a Bayesian framework for solving a class of problems termed multi-agent inverse reinforcement learning (MIRL). Compared to the well-known inverse reinforcement learning (IRL) problem, MIRL is formalized in the…

Computer Science and Game Theory · Computer Science 2019-07-31 Xiaomin Lin , Peter A. Beling , Randy Cogill

In this paper, we consider a differential stochastic zero-sum game in which two players intervene by adopting impulse controls in a finite time horizon. We provide a numerical solution as an approximation of the value function, which turns…

Optimization and Control · Mathematics 2024-10-14 Antoine Zolome , Brahim El Asri

Designing optimal interdependent networks is important for the robustness and efficiency of national critical infrastructures. Here, we establish a two-person game-theoretic model in which two network designers choose to maximize the global…

Social and Information Networks · Computer Science 2016-02-26 Juntao Chen , Quanyan Zhu

We study multi-agent reinforcement learning (MARL) in infinite-horizon discounted zero-sum Markov games. We focus on the practical but challenging setting of decentralized MARL, where agents make decisions without coordination by a…

Computer Science and Game Theory · Computer Science 2021-12-14 Muhammed O. Sayin , Kaiqing Zhang , David S. Leslie , Tamer Basar , Asuman Ozdaglar

We show that an N-person non-cooperative semi-Markov game under limiting ratio average pay-off has a pure semi-stationary Nash equilibrium. In an earlier paper, the zero-sum two person case has been dealt with. The proof follows by reducing…

Computer Science and Game Theory · Computer Science 2024-02-27 K. G. Bakshi , S. Sinha

In this paper, we consider a learning problem among non-cooperative agents interacting in a time-varying system. Specifically, we focus on repeated linear quadratic network games, in which the network of interactions changes with time and…

Computer Science and Game Theory · Computer Science 2023-10-23 Feras Al Taha , Kiran Rokade , Francesca Parise

In this paper, we study finite-agent linear-quadratic games on graphs. Specifically, we propose a comprehensive framework that extends the existing literature by incorporating heterogeneous and interpretable player interactions. Compared to…

Optimization and Control · Mathematics 2025-11-19 Ruimeng Hu , Jihao Long , Haosheng Zhou

We present an inverse dynamic game-based algorithm to learn parametric constraints from a given dataset of local Nash equilibrium interactions between multiple agents. Specifically, we introduce mixed-integer linear programs (MILP) encoding…

Machine Learning · Computer Science 2026-03-19 Zhouyu Zhang , Chih-Yuan Chiu , Glen Chou

This paper studies robust time-inconsistent (TIC) linear-quadratic stochastic control problems, formulated by stochastic differential games. By a spike variation approach, we derive sufficient conditions for achieving the Nash equilibrium,…

Optimization and Control · Mathematics 2025-04-29 Bingyan Han , Chi Seng Pun , Hoi Ying Wong

In this paper, we present an online learning approach for two-player zero-sum linear quadratic games with unknown dynamics. We develop a framework combining regularized least squares model estimation, high probability confidence sets, and…

Systems and Control · Electrical Eng. & Systems 2026-04-06 Shanting Wang , Weihao Sun , Andreas A. Malikopoulos

Computing Nash equilibrium policies is a central problem in multi-agent reinforcement learning that has received extensive attention both in theory and in practice. However, provable guarantees have been thus far either limited to fully…

In the present work, we consider 2-person zero-sum stochastic differential games with a nonlinear pay-off functional which is defined through a backward stochastic differential equation. Our main objective is to study for such a game the…

Probability · Mathematics 2014-07-29 Rainer Buckdahn , Juan Li , Marc Quincampoix

This paper studies the control problem for safety-critical multi-agent systems based on quadratic programming (QP). Each controlled agent is modeled as a cascade connection of an integrator and an uncertain nonlinear actuation system. In…

Systems and Control · Electrical Eng. & Systems 2022-12-01 Si Wu , Tengfei Liu , Magnus Egerstedt , Zhong-Ping Jiang

The paper investigates the long-time behavior of zero-sum linear-quadratic stochastic differential games, aiming to demonstrate that, under appropriate conditions, both the saddle strategy and the optimal state process exhibit the…

Optimization and Control · Mathematics 2024-06-05 Jingrui Sun , Jiongmin Yong