English
Related papers

Related papers: Approximate State Abstraction for Markov Games

200 papers

Two-player zero-sum repeated games are well understood. Computing the value of such a game is straightforward. Additionally, if the payoffs are dependent on a random state of the game known to one, both, or neither of the players, the…

Information Theory · Computer Science 2009-11-05 Paul Cuff

A fundamental assumption of reinforcement learning in Markov decision processes (MDPs) is that the relevant decision process is, in fact, Markov. However, when MDPs have rich observations, agents typically learn by way of an abstract state…

Machine Learning · Computer Science 2024-03-18 Cameron Allen , Neev Parikh , Omer Gottesman , George Konidaris

In many multi-player interactions, players incur strictly positive costs each time they execute actions e.g. 'menu costs' or transaction costs in financial systems. Since acting at each available opportunity would accumulate prohibitively…

Multiagent Systems · Computer Science 2024-08-02 David Mguni

A basic question for zero-sum repeated games consists in determining whether the mean payoff per time unit is independent of the initial state. In the special case of "zero-player" games, i.e., of Markov chains equipped with additive…

Optimization and Control · Mathematics 2015-10-20 Marianne Akian , Stéphane Gaubert , Antoine Hochart

In this paper, we investigate a partially observable zero sum games where the state process is a discrete time Markov chain. We consider a general utility function in the optimization criterion. We show the existence of value for both…

Optimization and Control · Mathematics 2022-11-16 Arnab Bhabak , Subhamay saha

The goal of this work is to formally abstract a Markov process evolving in discrete time over a general state space as a finite-state Markov chain, with the objective of precisely approximating its state probability distribution in time,…

Logic in Computer Science · Computer Science 2017-01-11 Sadegh Esmaeil Zadeh Soudjani , Alessandro Abate

We study a model of two-player, zero-sum, stopping games with asymmetric information. We assume that the payoff depends on two continuous-time Markov chains (X, Y), where X is only observed by player 1 and Y only by player 2, implying that…

Optimization and Control · Mathematics 2017-12-06 Fabien Gensbittel , Christine Grün

In this paper, we settle the sampling complexity of solving discounted two-player turn-based zero-sum stochastic games up to polylogarithmic factors. Given a stochastic game with discount factor $\gamma\in(0,1)$ we provide an algorithm that…

Machine Learning · Computer Science 2019-08-30 Aaron Sidford , Mengdi Wang , Lin F. Yang , Yinyu Ye

This paper studies the convergence of mean field games with finite state space to mean field games with a continuous state space. We examine a space discretization of a diffusive dynamics, which is reminiscent of the Markov chain…

Optimization and Control · Mathematics 2024-01-18 Charles Bertucci , Alekos Cecchin

With the increasing ubiquity of safety-critical autonomous systems operating in uncertain environments, there is a need for mathematical methods for formal verification of stochastic models. Towards formally verifying properties of…

Systems and Control · Electrical Eng. & Systems 2026-02-18 Adrien Banse , Giannis Delimpaltadakis , Luca Laurenti , Manuel Mazo , Raphaël M. Jungers

We study Nash equilibrium learning in partially observable Markov games (POMGs), a multi-agent reinforcement learning framework in which agents cannot fully observe the underlying state. Prior work in this setting relies on centralization…

Computer Science and Game Theory · Computer Science 2026-05-08 Philip Jordan , Maryam Kamgarpour

Markov Decision Processes (MDPs) are mathematical models of sequential decision-making under uncertainty that have found applications in healthcare, manufacturing, logistics, and others. In these models, a decision-maker observes the state…

Optimization and Control · Mathematics 2024-05-22 Madeleine Pollack , Lauren N. Steimle

Mean field games formalize dynamic games with a continuum of players and explicit interaction where the players can have heterogeneous states. As they additionally yield approximate equilibria of corresponding $N$-player games, they are of…

Optimization and Control · Mathematics 2020-01-09 Berenice Anne Neumann

This paper considers mean field games in a multi-agent Markov decision process (MDP) framework. Each player has a continuum state and binary action. By active control, a player can bring its state to a resetting point. All players are…

Optimization and Control · Mathematics 2017-01-25 Minyi Huang , Yan Ma

We study a generic family of two-player continuous-time nonzero-sum stopping games modeling a war of attrition with symmetric information and stochastic payoffs that depend on an homogeneous linear diffusion. We first show that any…

Optimization and Control · Mathematics 2022-10-18 Jean-Paul Décamps , Fabien Gensbittel , Thomas Mariotti

Potential games are arguably one of the most important and widely studied classes of normal form games. They define the archetypal setting of multi-agent coordination as all agent utilities are perfectly aligned with each other via a common…

Machine Learning · Computer Science 2025-09-24 Stefanos Leonardos , Will Overman , Ioannis Panageas , Georgios Piliouras

We investigate zero-sum turn-based two-player stochastic games in which the objective of one player is to maximize the amount of rewards obtained during a play, while the other aims at minimizing it. We focus on games in which the minimizer…

Logic in Computer Science · Computer Science 2022-05-20 Pablo F. Castro , Pedro R. D'Argenio , Luciano Putruele , Ramiro Demasi

In the paper we present a model of discrete-time mean-field game with several populations of players. Mean-field games with multiple populations of the players have only been studied in the literature in the continuous-time setting. The…

Optimization and Control · Mathematics 2023-04-07 Piotr Więcek

Planning in adversarial and uncertain environments can be modeled as the problem of devising strategies in stochastic perfect information games. These games are generalizations of Markov decision processes (MDPs): there are two…

Artificial Intelligence · Computer Science 2012-07-09 Krishnendu Chatterjee , Thomas A. Henzinger , Ranjit Jhala , Rupak Majumdar

I give a simple analysis of the game that I previously published in Scientific American which shows the paradoxical behavior whereby two losing games randomly combine to form a winning game. The game, modeled on a random walk, requires only…

Data Analysis, Statistics and Probability · Physics 2007-05-23 R. Dean Astumian