English
Related papers

Related papers: Constant payoff in zero-sum stochastic games

200 papers

In several standard models of dynamic programming (gambling houses, MDPs, POMDPs), we prove the existence of a very robust notion of value for the infinitely repeated problem, namely the pathwise uniform value. This solves two open…

Optimization and Control · Mathematics 2015-09-09 Xavier Venel , Bruno Ziliotto

We prove that for a class of zero-sum differential games with incomplete information on both sides, the value admits a probabilistic representation as the value of a zero-sum stochastic differential game with complete information, where…

Optimization and Control · Mathematics 2017-01-04 Fabien Gensbittel , Catherine Rainer

In this paper the set of value functions of all-possible zero-sum differential games with terminal payoff is characterized. The necessary and sufficient condition for a given function to be a value of some differential game with terminal…

Optimization and Control · Mathematics 2008-11-12 Yurii Averboukh

Mertens [In Proceedings of the International Congress of Mathematicians (Berkeley, Calif., 1986) (1987) 1528-1577 Amer. Math. Soc.] proposed two general conjectures about repeated games: the first one is that, in any two-person zero-sum…

Optimization and Control · Mathematics 2016-03-16 Bruno Ziliotto

This paper introduces alignment games, a new class of zero-sum games modeling strategic interventions where effectiveness depends on alignment with an underlying hidden state. Motivated by operational problems in medical diagnostics,…

Optimization and Control · Mathematics 2025-09-08 Pedro Cesar Lopes Gerum , Thomas Lidbetter

This paper examines the convergence of no-regret learning in games with continuous action sets. For concreteness, we focus on learning via "dual averaging", a widely used class of no-regret learning schemes where players take small steps…

Optimization and Control · Mathematics 2018-01-17 Panayotis Mertikopoulos , Zhengyuan Zhou

In a mean-payoff parity game, one of the two players aims both to achieve a qualitative parity objective and to minimize a quantitative long-term average of payoffs (aka. mean payoff). The game is zero-sum and hence the aim of the other…

Computer Science and Game Theory · Computer Science 2020-01-15 Laure Daviaud , Marcin Jurdzinski , Ranko Lazic

The valuation process that economic agents undergo for investments with uncertain payoff typically depends on their statistical views on possible future outcomes, their attitudes toward risk, and, of course, the payoff structure itself.…

Pricing of Securities · Quantitative Finance 2010-01-11 Constantinos Kardaras

It is well known that the rock-paper-scissors game has no pure saddle point. We show that this holds more generally: A symmetric two-player zero-sum game has a pure saddle point if and only if it is not a generalized rock-paper-scissors…

Computer Science and Game Theory · Computer Science 2013-01-25 Peter Duersch , Joerg Oechssler , Burkhard C. Schipper

In the window mean-payoff objective, given an infinite path, instead of considering a long run average, we consider the minimum payoff that can be ensured at every position of the path over a finite window that slides over the entire path.…

Computer Science and Game Theory · Computer Science 2019-12-09 Benjamin Bordais , Shibashis Guha , Jean-François Raskin

In iterated games, a player can unilaterally exert influence over the outcome through a careful choice of strategy. A powerful class of such "payoff control" strategies was discovered by Press and Dyson (2012). Their so-called…

Computer Science and Game Theory · Computer Science 2022-07-07 Arjun Mirani , Alex McAvoy

We analyze undiscounted continuous-time games of strategic experimentation with two-armed bandits. The risky arm generates payoffs according to a L\'{e}vy process with an unknown average payoff per unit of time which nature draws from an…

Theoretical Economics · Economics 2020-08-26 Godfrey Keller , Sven Rady

The paper investigates the long-time behavior of zero-sum linear-quadratic stochastic differential games, aiming to demonstrate that, under appropriate conditions, both the saddle strategy and the optimal state process exhibit the…

Optimization and Control · Mathematics 2024-06-05 Jingrui Sun , Jiongmin Yong

We present two zero-sum games modeling situations where one player attacks (or hides in) a finite dimensional nonempty compact set, and the other tries to prevent the attack (or find him). The first game, called patrolling game, corresponds…

Optimization and Control · Mathematics 2019-07-03 Tristan Garrec

Two-player zero-sum repeated games are well understood. Computing the value of such a game is straightforward. Additionally, if the payoffs are dependent on a random state of the game known to one, both, or neither of the players, the…

Information Theory · Computer Science 2009-11-05 Paul Cuff

We consider a network of coupled agents playing the Prisoner's Dilemma game, in which players are allowed to pick a strategy in the interval [0,1], with 0 corresponding to defection, 1 to cooperation, and intermediate values representing…

Adaptation and Self-Organizing Systems · Physics 2015-05-28 Francesco Sorrentino , Nicholas Mecholsky

The semigroup game is a two-person zero-sum game defined on a semigroup S as follows: Players 1 and 2 choose elements x and y in S, respectively, and player 1 receives a payoff f(xy) defined by a function f from S to [-1,1]. If the…

Computer Science and Game Theory · Computer Science 2016-07-11 Valerio Capraro , Kent Morrison

We investigate refinements of the mean-payoff criterion in two-player zero-sum perfect-information stochastic games. A strategy is Blackwell optimal if it is optimal in the discounted game for all discount factors sufficiently close to $1$.…

Computer Science and Game Theory · Computer Science 2025-06-24 Stéphane Gaubert , Julien Grand-Clément , Ricardo D. Katz

We consider zero-sum stochastic games for continuous time Markov decision processes with risk-sensitive average cost criterion. Here the transition and cost rates may be unbounded. We prove the existence of the value of the game and a…

Optimization and Control · Mathematics 2021-09-21 Mrinal K. Ghosh , Subrata Golui , Chandan Pal , Somnath Pradhan

We study the interaction between a network designer and an adversary over a dynamical network. The network consists of nodes performing continuous-time distributed averaging. The adversary strategically disconnects a set of links to prevent…

Systems and Control · Computer Science 2015-02-23 Ali Khanafer , Tamer Başar