English
Related papers

Related papers: Policy Optimization for Linear-Quadratic Zero-Sum …

200 papers

We consider zero-sum stochastic games with finite state and action spaces, perfect information, mean payoff criteria, without any irreducibility assumption on the Markov chains associated to strategies (multichain games). The value of such…

Optimization and Control · Mathematics 2012-08-03 Marianne Akian , Jean Cochet-Terrasson , Sylvie Detournay , Stéphane Gaubert

Two-player zero-sum games are a well-established model for synthesising controllers that optimise some performance criterion. In such games one player represents the controller, while the other describes the (adversarial) environment, and…

Computer Science and Game Theory · Computer Science 2010-06-04 Marta Kwiatkowska , Gethin Norman , Ashutosh Trivedi

We study a new class of Markov games, \emph(multi-player) zero-sum Markov Games} with \emph{Networked separable interactions} (zero-sum NMGs), to model the local interaction structure in non-cooperative multi-agent sequential…

Computer Science and Game Theory · Computer Science 2025-07-15 Chanwoo Park , Kaiqing Zhang , Asuman Ozdaglar

In this paper, we investigate the interaction of two populations with a large number of indistinguishable agents. The problem consists in two levels: the interaction between agents of a same population, and the interaction between the two…

Optimization and Control · Mathematics 2018-10-30 Alain Bensoussan , Tao Huang , Mathieu Laurière

We focus on the problem of \emph{Answer-Level Fine-Tuning} (ALFT), where the goal is to optimize a language model based on the correctness or properties of its final answers, rather than the specific reasoning traces used to produce them.…

Machine Learning · Computer Science 2026-05-01 Mehryar Mohri , Jon Schneider , Yifan Wu

Network congestion games are a convenient model for reasoning about routing problems in a network: agents have to move from a source to a target vertex while avoiding congestion, measured as a cost depending on the number of players using…

Computer Science and Game Theory · Computer Science 2022-07-05 Aline Goeminne , Nicolas Markey , Ocan Sankur

This paper establishes a data-driven solution for infinite horizon linear quadratic Gaussian Mean Field Games with network-coupled heterogeneous agent populations where the dynamics of the agents are unknown. The solution technique relies…

Systems and Control · Electrical Eng. & Systems 2026-02-17 Jean Zhu , Shuang Gao

In this paper, we investigate a competitive market involving two agents who consider both their own wealth and the wealth gap with their opponent. Both agents can invest in a financial market consisting of a risk-free asset and a risky…

Optimization and Control · Mathematics 2025-02-10 Junyi Guo , Xia Han , Hao Wang , Kam Chuen Yuen

We present novel techniques for neuro-symbolic concurrent stochastic games, a recently proposed modelling formalism to represent a set of probabilistic agents operating in a continuous-space environment using a combination of neural network…

Computer Science and Game Theory · Computer Science 2022-06-22 Rui Yan , Gabriel Santos , Xiaoming Duan , David Parker , Marta Kwiatkowska

In the present work, we study deterministic mean field games (MFGs) with finite time horizon in which the dynamics of a generic agent is controlled by the acceleration. They are described by a system of PDEs coupling a continuity equation…

Analysis of PDEs · Mathematics 2020-07-29 Yves Achdou , Paola Mannucci , Claudio Marchi , Nicoletta Tchou

We consider the general problem of resource sharing in societal networks, consisting of interconnected communication, transportation, energy and other networks important to the functioning of society. Participants in such network need to…

Computer Science and Game Theory · Computer Science 2018-03-28 Jian Li , Bainan Xia , Xinbo Geng , Hao Ming , Srinivas Shakkottai , Vijay Subramanian , Le Xie

Here, we develop numerical methods for finite-state mean-field games (MFGs) that satisfy a monotonicity condition. MFGs are determined by a system of differential equations with initial and terminal boundary conditions. These non-standard…

Numerical Analysis · Mathematics 2017-05-02 Diogo Gomes , Joao Saude

In this paper, we address linear-quadratic-Gaussian (LQG) risk-sensitive mean field games (MFGs) with common noise. In this framework agents are exposed to a common noise and aim to minimize an exponential cost functional that reflects…

Optimization and Control · Mathematics 2024-03-07 Xin Yue Ren , Dena Firoozi

This paper is concerned with two-person mean-field linear-quadratic non-zero sum stochastic differential games in an infinite horizon. Both open-loop and closed-loop Nash equilibria are introduced. Existence of an open-loop Nash equilibrium…

Optimization and Control · Mathematics 2021-04-09 Xun Li , Jingtao Shi , Jiongmin Yong

Modern reinforcement learning (RL) commonly engages practical problems with large state spaces, where function approximation must be deployed to approximate either the value function or the policy. While recent progresses in RL theory…

Machine Learning · Computer Science 2021-10-14 Chi Jin , Qinghua Liu , Tiancheng Yu

Quantilized mean-field game models involve quantiles of the population's distribution. We study a class of such games with a capacity for ranking games, where the performance of each agent is evaluated based on its terminal state relative…

Optimization and Control · Mathematics 2025-07-02 Rinel Foguen Tchuendom , Dena Firoozi , Michèle Breton

A growing line of work reframes preference-based fine-tuning of large language models game-theoretically: Nash Learning from Human Feedback (NLHF) recasts the problem as a zero-sum game over policies. However, optimization is over expected…

Computer Science and Game Theory · Computer Science 2026-05-14 Max Horwitz , Jake Gonzales , Eric Mazumdar , Lillian J. Ratliff

Subject to reasonable conditions, in large population stochastic dynamics games, where the agents are coupled by the system's mean field (i.e. the state distribution of the generic agent) through their nonlinear dynamics and their nonlinear…

Optimization and Control · Mathematics 2019-05-28 Nevroz Sen , Peter E. Caines

Convergence of the policy iteration method for discrete and continuous optimal control problems holds under general assumptions. Moreover, in some circumstances, it is also possible to show a quadratic rate of convergence for the algorithm.…

Optimization and Control · Mathematics 2022-03-02 Fabio Camilli , Qing Tang

We obtain global, non-asymptotic convergence guarantees for independent learning algorithms in competitive reinforcement learning settings with two agents (i.e., zero-sum stochastic games). We consider an episodic setting where in each…

Machine Learning · Computer Science 2021-01-13 Constantinos Daskalakis , Dylan J. Foster , Noah Golowich