English
Related papers

Related papers: On Bellman's Optimality Principle for zs-POSGs

200 papers

Stochastic games are a well established model for multi-agent sequential decision making under uncertainty. In practical applications, though, agents often have only partial observability of their environment. Furthermore, agents…

Computer Science and Game Theory · Computer Science 2024-07-02 Rui Yan , Gabriel Santos , Gethin Norman , David Parker , Marta Kwiatkowska

A Dynkin game is considered for stochastic differential equations with random coefficients. We first apply Qiu and Tang's maximum principle for backward stochastic partial differential equations to generalize Krylov estimate for the…

Optimization and Control · Mathematics 2011-09-27 Shanjian Tang , Zhou Yang

We study offline multi-agent reinforcement learning (RL) in Markov games, where the goal is to learn an approximate equilibrium -- such as Nash equilibrium and (Coarse) Correlated Equilibrium -- from an offline dataset pre-collected from…

Machine Learning · Computer Science 2023-02-07 Yuheng Zhang , Yu Bai , Nan Jiang

Many processes, such as discrete event systems in engineering or population dynamics in biology, evolve in discrete space and continuous time. We consider the problem of optimal decision making in such discrete state and action space…

Machine Learning · Computer Science 2020-10-27 Bastian Alt , Matthias Schultheis , Heinz Koeppl

This paper considers the problem of two-player zero-sum stochastic differential game with both players adopting impulse controls in finite horizon under rather weak assumptions on the cost functions ($c$ and $\chi$ not decreasing in time).…

Optimization and Control · Mathematics 2018-09-26 Brahim El Asri , Sehail Mazid

We study the computational complexity of solving stochastic games with mean-payoff objectives. Instead of identifying special classes in which simple strategies are sufficient to play $\epsilon$-optimally, or form $\epsilon$-Nash…

Computer Science and Game Theory · Computer Science 2024-05-16 Sougata Bose , Rasmus Ibsen-Jensen , Patrick Totzke

In this paper we study the optimization problem of an economic agent who chooses a job and the time of retirement as well as consumption and portfolio of assets. The agent is constrained in the ability to borrow against future income. We…

Optimization and Control · Mathematics 2021-07-28 Junkee Jeon , Hyeng Keun Koo

Game-theoretic concepts have been extensively studied in economics to provide insight into competitive behaviour and strategic decision making. As computing systems increasingly involve concurrently acting autonomous agents, game-theoretic…

Formal Languages and Automata Theory · Computer Science 2022-07-01 Marta Kwiatkowska , Gethin Norman , David Parker , Gabriel Santos , Rui Yan

The topics treated in this thesis are inherently two-fold. The first part considers the problem of a market maker optimally setting bid/ask quotes over a finite time horizon, to maximize her expected utility. The intensities of the orders…

Optimization and Control · Mathematics 2020-09-15 Diego Zabaljauregui

We describe a nonlinear generalization of dual dynamic programming theory and its application to value function estimation for deterministic control problems over continuous state and action spaces, in a discrete-time infinite horizon…

Optimization and Control · Mathematics 2018-10-05 Joseph Warrington , Paul N. Beuchat , John Lygeros

In this paper, a convex optimization-based method is proposed for numerically solving dynamic programs in continuous state and action spaces. The key idea is to approximate the output of the Bellman operator at a particular state by the…

Optimization and Control · Mathematics 2020-10-23 Insoon Yang

Leader-follower general-sum stochastic games (LF-GSSGs) model sequential decision-making under asymmetric commitment, where a leader commits to a policy and a follower best responds, yielding a strong Stackelberg equilibrium (SSE) with…

Computer Science and Game Theory · Computer Science 2025-12-08 Jilles Steeve Dibangoye , Thibaut Le Marre , Ocan Sankur , François Schwarzentruber

Many real-world decision problems involve the interaction of multiple self-interested agents with limited sensing ability. The partially observable stochastic game (POSG) provides a mathematical framework for modeling these problems,…

Computer Science and Game Theory · Computer Science 2024-10-30 Tyler Becker , Zachary Sunberg

This paper tackles the problem of how two selfish users jointly determine the operating point in the achievable rate region of a two-user Gaussian interference channel through bargaining. In previous work, incentive conditions for two users…

Information Theory · Computer Science 2010-10-05 Xi Liu , Elza Erkip

We study the global convergence of policy optimization for finding the Nash equilibria (NE) in zero-sum linear quadratic (LQ) games. To this end, we first investigate the landscape of LQ games, viewing it as a nonconvex-nonconcave…

Machine Learning · Computer Science 2021-02-12 Kaiqing Zhang , Zhuoran Yang , Tamer Başar

We address the Nash equilibrium problem in a partial-decision information scenario, where each agent can only observe the actions of some neighbors, while its cost possibly depends on the strategies of other agents. Our main contribution is…

Optimization and Control · Mathematics 2021-05-07 Mattia Bianchi , Giuseppe Belgioioso , Sergio Grammatico

We study optimal equilibria in multi-player games. An equilibrium is optimal for a player, if her payoff is maximal. A tempting approach to solving this problem is to seek optimal Nash equilibria, the standard form of equilibria where no…

Computer Science and Game Theory · Computer Science 2013-07-09 Anshul Gupta , Sven Schewe

Finding optimal policies which maximize long term rewards of Markov Decision Processes requires the use of dynamic programming and backward induction to solve the Bellman optimality equation. However, many real-world problems require…

Machine Learning · Computer Science 2023-01-10 Mridul Agarwal , Vaneet Aggarwal

In single-agent Markov decision processes, an agent can optimize its policy based on the interaction with environment. In multi-player Markov games (MGs), however, the interaction is non-stationary due to the behaviors of other players, so…

Computer Science and Game Theory · Computer Science 2021-10-19 Yuanheng Zhu , Dongbin Zhao , Mengchen Zhao , Dong Li

We introduce, to our knowledge, the first direct second-order method for computing Nash equilibria in two-player zero-sum games. To do so, we construct a Douglas-Rachford-style splitting formulation, which we then solve with a semi-smooth…

Computer Science and Game Theory · Computer Science 2025-12-16 David Yang , Yuan Gao , Tianyi Lin , Christian Kroer