中文
相关论文

相关论文: Expected Window Mean-Payoff

200 篇论文

A decisionmaker faces $n$ alternatives, each of which represents a potential reward. After investing costly resources into investigating the alternatives, the decisionmaker may select one, or more generally a feasible subset, and obtain the…

计算机科学与博弈论 · 计算机科学 2026-04-02 Robin Bowers , Elias Lindgren , Bo Waggoner

We consider Markov Decision Processes (MDPs) with mean-payoff parity and energy parity objectives. In system design, the parity objective is used to encode \omega-regular specifications, and the mean-payoff and energy objectives can be used…

计算机科学与博弈论 · 计算机科学 2011-04-18 Krishnendu Chatterjee , Laurent Doyen

We study graphs and two-player games in which rewards are assigned to states, and the goal of the players is to satisfy or dissatisfy certain property of the generated outcome, given as a mean payoff property. Since the notion of…

计算机科学中的逻辑 · 计算机科学 2016-04-22 Tomáš Brázdil , Vojtěch Forejt , Antonín Kučera , Petr Novotný

In a zero-sum stochastic game, at each stage, two adversary players take decisions and receive a stage payoff determined by them and by a controlled random variable representing the state of nature. The total payoff is the normalized…

最优化与控制 · 数学 2022-05-06 Olivier Catoni , Miquel Oliu-Barton , Bruno Ziliotto

We consider a dynamic traffic routing game over an urban road network involving a large number of drivers in which each driver selecting a particular route is subject to a penalty that is affine in the logarithm of the number of drivers…

最优化与控制 · 数学 2020-01-22 Takashi Tanaka , Ehsan Nekouei , Ali Reza Pedram , Karl Henrik Johansson

In mean-payoff games, the objective of the protagonist is to ensure that the limit average of an infinite sequence of numeric weights is nonnegative. In energy games, the objective is to ensure that the running sum of weights is always…

计算机科学与博弈论 · 计算机科学 2012-09-17 Yaron Velner , Krishnendu Chatterjee , Laurent Doyen , Thomas A. Henzinger , Alexander Rabinovich , Jean-Francois Raskin

Stochastic games with discounted payoff, introduced by Shapley, model adversarial interactions in stochastic environments where two players try to optimize a discounted sum of rewards. In this model, long-term weights are geometrically…

计算机科学与博弈论 · 计算机科学 2021-10-22 Taylor Dohmen , Ashutosh Trivedi

Two-player games on graphs provide the mathematical foundation for the study of reactive systems. In the quantitative framework, an objective assigns a value to every play, and the goal of player 1 is to minimize the value of the objective.…

计算机科学中的逻辑 · 计算机科学 2014-04-30 Yaron Velner

What payoffs are positionally determined for deterministic two-player antagonistic games on finite directed graphs? In this paper we study this question for payoffs that are continuous. The main reason why continuous positionally determined…

计算机科学与博弈论 · 计算机科学 2024-02-14 Alexander Kozachinskiy

We consider partially observable Markov decision processes (POMDPs) with limit-average payoff, where a reward value in the interval [0,1] is associated to every transition, and the payoff of an infinite path is the long-run average of the…

人工智能 · 计算机科学 2013-08-23 Krishnendu Chatterjee , Martin Chmelík

Markov decision processes (MDPs) with rewards are a widespread and well-studied model for systems that make both probabilistic and nondeterministic choices. A fundamental result about MDPs is that their minimal and maximal expected rewards…

计算机科学中的逻辑 · 计算机科学 2024-11-26 Kevin Batz , Benjamin Lucien Kaminski , Christoph Matheja , Tobias Winkler

This work considers two-player zero-sum semi-Markov games with incomplete information on one side and perfect observation. At the beginning, the system selects a game type according to a given probability distribution and informs to Player…

最优化与控制 · 数学 2021-07-16 Fang Chen , Xianping Guo , Zhong-Wei Liao

A decision maker repeatedly chooses one of a finite set of actions. In each period, the decision maker's payoff depends on fixed basic payoff of the chosen action and the frequency with which the action has been chosen in the past. We…

理论经济学 · 经济学 2024-05-02 Galit Ashkenazi-Golan , Dominik Karos , Ehud Lehrer

Consider a stream of $n$ random points (say, from the unit square) arriving one by one, where a player has to make an irreversible immediate decision for each arriving point whether to pick it. The player has to pick a single point, and the…

计算几何 · 计算机科学 2026-04-28 Sariel Har-Peled

We examine perfect information stochastic mean-payoff games - a class of games containing as special sub-classes the usual mean-payoff games and parity games. We show that deterministic memoryless strategies that are optimal for discounted…

计算机科学与博弈论 · 计算机科学 2010-06-09 Hugo Gimbert , Wiesław Zielonka

In this paper, we provide an effective characterization of all the subgame-perfect equilibria in infinite duration games played on finite graphs with mean-payoff objectives. To this end, we introduce the notion of requirement, and the…

计算机科学与博弈论 · 计算机科学 2022-04-22 Léonard Brice , Jean-François Raskin , Marie Van Den Bogaard

We propose a new deterministic symmetric recursive algorithm for solving mean-payoff games.

计算机科学与博弈论 · 计算机科学 2026-03-10 Pierre Ohlmann

Consider a general path planning problem of a robot on a graph with edge costs, and where each node has a Boolean value of success or failure (with respect to some task) with a given probability. The objective is to plan a path for the…

机器人学 · 计算机科学 2018-08-22 Arjun Muralidharan , Yasamin Mostofi

We study the problem of finding equilibrium strategies in multi-agent games with incomplete payoff information, where the payoff matrices are only known to the players up to some bounded uncertainty sets. In such games, an ex-post…

计算机科学与博弈论 · 计算机科学 2020-07-14 Wenshuo Guo , Mihaela Curmei , Serena Wang , Benjamin Recht , Michael I. Jordan

We consider the problem of maximizing the expected average reward obtained over an infinite time horizon by $n$ weakly coupled Markov decision processes. Our setup is a substantial generalization of the multi-armed restless bandit problem…

最优化与控制 · 数学 2026-04-01 Diego Goldsztajn , Konstantin Avrachenkov