English
Related papers

Related papers: Expected Window Mean-Payoff

200 papers

We study a two-player, zero-sum, stochastic game with incomplete information on one side in which the players are allowed to play more and more frequently. The informed player observes the realization of a Markov chain on which the payoffs…

Optimization and Control · Mathematics 2013-07-15 Pierre Cardaliaguet , Catherine Rainer , Dinah Rosenberg , Nicolas Vieille

Markovian network equilibrium generalizes the classical Wardrop equilibrium in network games. At a Markovian network equilibrium, each player of the game solves a Markov decision process instead of a shortest path problem. We propose two…

Optimization and Control · Mathematics 2021-10-19 Yue Yu , Dan Calderone , Sarah H. Q. Li , Lillian J. Ratliff , Behçet Açıkmeşe

We study the set of (stationary) feasible payoffs of overlapping generation repeated games that can be achieved by action sequences in which every generation of players plays the same sequence of action profiles. First, we completely…

Theoretical Economics · Economics 2024-12-25 Daehyun Kim , Chihiro Morooka

Markov automata combine non-determinism, probabilistic branching, and exponentially distributed delays. This compositional variant of continuous-time Markov decision processes is used in reliability engineering, performance evaluation and…

Logic in Computer Science · Computer Science 2017-05-11 Tim Quatmann , Sebastian Junges , Joost-Pieter Katoen

In the Time-Windows TSP (TW-TSP) we are given requests at different locations on a network; each request is endowed with a reward and an interval of time; the goal is to find a tour that visits as much reward as possible during the…

Data Structures and Algorithms · Computer Science 2023-04-05 Shuchi Chawla , Dimitris Christou

A version of the secretary problem is considered. The ranks of items, whose values are independent, identically distributed random variables $X_1,X_2,...,X_n$ from a uniform distribution on $[0; 1]$, are observed sequentially by the grader.…

Optimization and Control · Mathematics 2020-11-23 Krzysztof Szajowski

We study a model of two-player, zero-sum, stopping games with asymmetric information. We assume that the payoff depends on two continuous-time Markov chains (X, Y), where X is only observed by player 1 and Y only by player 2, implying that…

Optimization and Control · Mathematics 2017-12-06 Fabien Gensbittel , Christine Grün

The range of a payoff function for an $n$-player finite strategic game is investigated using a novel approach, the notion of extreme points of a non-convex set. The shape of a noncooperative payoff region can be estimated using extreme…

Computer Science and Game Theory · Computer Science 2018-08-07 Yu-Sung Tu , Wei-Torng Juang

We study the problem of evaluating a discrete function by adaptively querying the values of its variables until the values read uniquely determine the value of the function. Reading the value of a variable is done at the expense of some…

Data Structures and Algorithms · Computer Science 2014-06-17 Aline Saettler , Eduardo Laber , Ferdinando Cicalese

We consider some well-known families of two-player, zero-sum, perfect information games that can be viewed as special cases of Shapley's stochastic games. We show that the following tasks are polynomial time equivalent: - Solving simple…

Computer Science and Game Theory · Computer Science 2008-12-03 Vladimir Gurvich , Peter Bro Miltersen

Multi-period mean-variance optimization is a long-standing problem, caused by the failure of dynamic programming principle. This paper studies the mean-variance optimization in a setting of finite-horizon discrete-time Markov decision…

Optimization and Control · Mathematics 2025-07-31 Li Xia , Zhihui Yu

We consider the problem of computing the value and an optimal strategy for minimizing the expected termination time in one-counter Markov decision processes. Since the value may be irrational and an optimal strategy may be rather…

Formal Languages and Automata Theory · Computer Science 2012-05-08 Tomáš Brázdil , Antonín Kučera , Petr Novotný , Dominik Wojtczak

We present the conditional value-at-risk (CVaR) in the context of Markov chains and Markov decision processes with reachability and mean-payoff objectives. CVaR quantifies risk by means of the expectation of the worst p-quantile. As such it…

Logic in Computer Science · Computer Science 2018-05-09 Jan Křetínský , Tobias Meggendorfer

We introduce a mean field game with rank-based reward: competing agents optimize their effort to achieve a goal, are ranked according to their completion time, and paid a reward based on their relative rank. First, we propose a tractable…

Optimization and Control · Mathematics 2017-08-07 Marcel Nutz , Yuchong Zhang

Quantitative games are two-player zero-sum games played on directed weighted graphs. Total-payoff games (that can be seen as a refinement of the well-studied mean-payoff games) are the variant where the payoff of a play is computed as the…

Computer Science and Game Theory · Computer Science 2015-07-15 Thomas Brihaye , Gilles Geeraerts , Axel Haddad , Benjamin Monmege

This paper investigates value function approximation in the context of zero-sum Markov games, which can be viewed as a generalization of the Markov decision process (MDP) framework to the two-agent case. We generalize error bounds from MDPs…

Artificial Intelligence · Computer Science 2013-01-07 Michail Lagoudakis , Ron Parr

A bicriteria approximation algorithm is presented for the unrooted traveling repairman problem, realizing increased profit in return for increased speedup of repairman motion. The algorithm generalizes previous results from the case in…

Data Structures and Algorithms · Computer Science 2011-01-21 Greg N. Frederickson , Barry Wittman

This paper develops a method to upper-bound extreme-values of time-windowed risks for stochastic processes. Examples of such risks include the maximum average or 90% quantile of the current along a transmission line in any 5-minute window.…

Optimization and Control · Mathematics 2024-04-12 Jared Miller , Niklas Schmid , Matteo Tacchi , Didier Henrion , Roy S. Smith

We introduce two-level discounted games played by two players on a perfect-information stochastic game graph. The upper level game is a discounted game and the lower level game is an undiscounted reachability game. Two-level games model…

Logic in Computer Science · Computer Science 2010-06-09 Krishnendu Chatterjee , Rupak Majumdar

We explore a broad class of values for cooperative games in characteristic function form, known as \emph{compromise values\/}. These values efficiently allocate payoffs by linearly combining well-specified upper and lower bounds on payoffs.…

Theoretical Economics · Economics 2025-10-15 Robert P. Gilles , René van den Brink