English
Related papers

Related papers: Algorithms for zero-sum stochastic games with the …

200 papers

We study two person nonzero-sum stochastic differential games with risk-sensitive discounted and ergodic cost criteria. Under certain conditions we establish a Nash equilibrium in Markov strategies for the discounted cost criterion and a…

Optimization and Control · Mathematics 2016-04-06 Mrinal K. Ghosh , K. Suresh Kumar , Chandan Pal

Given two bounded convex sets $X\subseteq\RR^m$ and $Y\subseteq\RR^n,$ specified by membership oracles, and a continuous convex-concave function $F:X\times Y\to\RR$, we consider the problem of computing an $\eps$-approximate saddle point,…

Computer Science and Game Theory · Computer Science 2014-05-01 Khaled Elbassioni , Kazuhisa Makino , Kurt Mehlhorn , Fahimeh Ramezani

We consider Gillette's two-person zero-sum stochastic games with perfect information. For each $k \in \ZZ_+$ we introduce an effective reward function, called $k$-total. For $k = 0$ and $1$ this function is known as {\it mean payoff} and…

Discrete Mathematics · Computer Science 2015-08-17 Endre Boros , Khaled Elbassioni , Vladimir Gurvich , Kazuhisa Makino

The notion of approachability was introduced by Blackwell [1] in the context of vector-valued repeated games. The famous Blackwell's approachability theorem prescribes a strategy for approachability, i.e., for `steering' the average cost of…

Machine Learning · Computer Science 2016-06-22 Dileep Kalathil , Vivek Borkar , Rahul Jain

We consider a zero-sum stochastic differential game over elementary mixed feed-back strategies. These are strategies based only on the knowledge of the past state, randomized continuously in time from a sampling distribution which is kept…

Optimization and Control · Mathematics 2014-04-16 Mihai Sîrbu

We study a zero-sum game where the evolution of a spectrally one-sided Levy process is modified by a singular controller and is terminated by the stopper. The singular controller minimizes the expected values of running, controlling and…

Optimization and Control · Mathematics 2014-08-08 Daniel Hernandez-Hernandez , Kazutoshi Yamazaki

This paper investigates value function approximation in the context of zero-sum Markov games, which can be viewed as a generalization of the Markov decision process (MDP) framework to the two-agent case. We generalize error bounds from MDPs…

Artificial Intelligence · Computer Science 2013-01-07 Michail Lagoudakis , Ron Parr

We study a simple adaptive model in the framework of an N -player normal form game. The model consists of a repeated game where the players only know their own action space and their own payoff scored at each stage, not those of the other…

Computer Science and Game Theory · Computer Science 2017-06-12 Mario Bravo

The main purpose of this paper is to approximate several non-local evolution equations by zero-sum repeated games in the spirit of the previous works of Kohn and the second author (2006 and 2009): general fully non-linear parabolic…

Analysis of PDEs · Mathematics 2010-12-07 Cyril Imbert , Sylvia Serfaty

We propose stochastic variance reduced algorithms for solving convex-concave saddle point problems, monotone variational inequalities, and monotone inclusions. Our framework applies to extragradient, forward-backward-forward, and…

Optimization and Control · Mathematics 2022-06-14 Ahmet Alacaoglu , Yura Malitsky

We develop value iteration-based algorithms to solve in a unified manner different classes of combinatorial zero-sum games with mean-payoff type rewards. These algorithms rely on an oracle, evaluating the dynamic programming operator up to…

Computer Science and Game Theory · Computer Science 2024-11-12 Xavier Allamigeon , Stéphane Gaubert , Ricardo D. Katz , Mateusz Skomra

In this paper, we propose a variance-reduced primal-dual algorithm with Bregman distance for solving convex-concave saddle-point problems with finite-sum structure and nonbilinear coupling function. This type of problems typically arises in…

Optimization and Control · Mathematics 2021-06-02 Erfan Yazdandoost Hamedani , Afrooz Jalilzadeh

In this paper, we propose and analyze zeroth-order stochastic approximation algorithms for nonconvex and convex optimization, with a focus on addressing constrained optimization, high-dimensional setting and saddle-point avoiding. To handle…

Optimization and Control · Mathematics 2019-01-16 Krishnakumar Balasubramanian , Saeed Ghadimi

We introduce a new non-zero-sum game of optimal stopping with asymmetric exercise opportunities. Given a stochastic process modelling the value of an asset, one player observes and can act on the process continuously, while the other player…

Probability · Mathematics 2024-05-16 José Luis Pérez , Neofytos Rodosthenous , Kazutoshi Yamazaki

We consider two-player zero-sum concurrent stochastic games (CSGs) played on graphs with reachability and safety objectives. These include degenerate classes such as Markov decision processes or turn-based stochastic games, which can be…

Logic in Computer Science · Computer Science 2025-09-11 Marta Grobelna , Jan Křetínský , Maximilian Weininger

We consider a general nonzero-sum impulse game with two players. The main mathematical contribution of the paper is a verification theorem which provides, under some regularity conditions, a suitable system of quasi-variational inequalities…

Probability · Mathematics 2018-11-09 René Aïd , Matteo Basei , Giorgia Callegaro , Luciano Campi , Tiziano Vargiolu

This paper examines the convergence of no-regret learning in games with continuous action sets. For concreteness, we focus on learning via "dual averaging", a widely used class of no-regret learning schemes where players take small steps…

Optimization and Control · Mathematics 2018-01-17 Panayotis Mertikopoulos , Zhengyuan Zhou

In this paper, we consider a linear quadratic stochastic two-person zero-sum differential game. The controls for both players are allowed to appear in both drift and diffusion of the state equation. The weighting matrices in the performance…

Optimization and Control · Mathematics 2014-01-21 Jingrui Sun , Jiongmin Yong

Many machine learning algorithms minimize a regularized risk, and stochastic optimization is widely used for this task. When working with massive data, it is desirable to perform stochastic optimization in parallel. Unfortunately, many…

Machine Learning · Statistics 2023-11-27 Shin Matsushima , Hyokun Yun , Xinhua Zhang , S. V. N. Vishwanathan

Two-player mean-payoff Stackelberg games are nonzero-sum infinite duration games played on a bi-weighted graph by Leader (Player 0) and Follower (Player 1). Such games are played sequentially: first, Leader announces her strategy, second,…

Optimization and Control · Mathematics 2021-08-04 Mrudula Balachander , Shibashis Guha , Jean-François Raskin