中文
相关论文

相关论文: Expected Window Mean-Payoff

200 篇论文

Advertising options have been recently studied as a special type of guaranteed contracts in online advertising, which are an alternative sales mechanism to real-time auctions. An advertising option is a contract which gives its buyer a…

计算机科学与博弈论 · 计算机科学 2018-08-29 Bowei Chen , Mohan Kankanhalli

We develop value iteration-based algorithms to solve in a unified manner different classes of combinatorial zero-sum games with mean-payoff type rewards. These algorithms rely on an oracle, evaluating the dynamic programming operator up to…

计算机科学与博弈论 · 计算机科学 2024-11-12 Xavier Allamigeon , Stéphane Gaubert , Ricardo D. Katz , Mateusz Skomra

Suppose there are $n$ Markov chains and we need to pay a per-step \emph{price} to advance them. The "destination" states of the Markov chains contain rewards; however, we can only get rewards for a subset of them that satisfy a…

数据结构与算法 · 计算机科学 2019-02-22 Anupam Gupta , Haotian Jiang , Ziv Scully , Sahil Singla

Using methods from the statistical mechanics of disordered systems we analyze the properties of bimatrix games with random payoffs in the limit where the number of pure strategies of each player tends to infinity. We analytically calculate…

无序系统与神经网络 · 物理学 2009-10-31 Johannes Berg

Consider oriented graph nodes requiring periodic visits by a service agent. The agent moves among the nodes and receives a payoff for each completed service task, depending on the time elapsed since the previous visit to a node. We consider…

计算机科学与博弈论 · 计算机科学 2023-05-19 David Klaška , Antonín Kučera , Vít Musil , Vojtěch Řehák

We study stochastic two-player turn-based games in which the objective of one player is to ensure several infinite-horizon total reward objectives, while the other player attempts to spoil at least one of the objectives. The games have…

计算机科学与博弈论 · 计算机科学 2016-05-13 Romain Brenguier , Vojtěch Forejt

In the paper we present a model of discrete-time mean-field game with several populations of players. Mean-field games with multiple populations of the players have only been studied in the literature in the continuous-time setting. The…

最优化与控制 · 数学 2023-04-07 Piotr Więcek

A basic question for zero-sum repeated games consists in determining whether the mean payoff per time unit is independent of the initial state. In the special case of "zero-player" games, i.e., of Markov chains equipped with additive…

最优化与控制 · 数学 2015-10-20 Marianne Akian , Stéphane Gaubert , Antoine Hochart

This paper is concerned with the solution of the optimal stopping problem associated to the valuation of Perpetual American options driven by continuous time Markov chains. We introduce a new dynamic approach for the numerical pricing of…

概率论 · 数学 2019-04-25 Laurent Miclo , Stéphane Villeneuve

In mean-payoff games, the objective of the protagonist is to ensure that the limit average of an infinite sequence of numeric weights is nonnegative. In energy games, the objective is to ensure that the running sum of weights is always…

计算机科学中的逻辑 · 计算机科学 2010-10-05 Krishnendu Chatterjee , Laurent Doyen , Thomas A. Henzinger , Jean-Francois Raskin

We introduce one-way games, a framework motivated by applications in large-scale power restoration, humanitarian logistics, and integrated supply-chains. The distinguishable feature of the games is that the payoff of some player is…

计算机科学与博弈论 · 计算机科学 2015-07-28 Andres Abeliuk , Gerardo Berbeglia , Pascal Van Hentenryck

We study two-player zero-sum repeated games with incomplete information on one side, where the payoff function is tail measurable (and not necessarily the long-run average payoff). We show that the maxmin value equals the concavification of…

最优化与控制 · 数学 2025-12-02 Gil Bar Castellon Koltun , Ehud Lehrer , Eilon Solan

Flip a coin repeatedly, and stop whenever you want. Your payoff is the proportion of heads, and you wish to maximize this payoff in expectation. This so-called Chow-Robbins game is amenable to computer analysis, but while simple-minded…

概率论 · 数学 2012-01-04 Olle Häggström , Johan Wästlund

We consider zero-sum stochastic games with finite state and action spaces, perfect information, mean payoff criteria, without any irreducibility assumption on the Markov chains associated to strategies (multichain games). The value of such…

最优化与控制 · 数学 2012-08-03 Marianne Akian , Jean Cochet-Terrasson , Sylvie Detournay , Stéphane Gaubert

We consider a discrete-time Markov decision process with Borel state and action spaces. The performance criterion is to maximize a total expected {utility determined by unbounded return function. It is shown the existence of optimal…

概率论 · 数学 2018-10-08 François Dufour , Alexandre Genadot

Through a stochastic control theoretic approach, we analyze reputation games where a strategic long-lived player acts in a sequential repeated game against a collection of short-lived players. The key assumption in our model is that the…

最优化与控制 · 数学 2020-01-22 Nuh Aygün Dalkıran , Serdar Yüksel

This paper studies the risk-averse mean-variance optimization in infinite-horizon discounted Markov decision processes (MDPs). The involved variance metric concerns reward variability during the whole process, and future deviations are…

最优化与控制 · 数学 2022-01-19 Shuai Ma , Xiaoteng Ma , Li Xia

This paper develops a path planner that minimizes risk (e.g. motion execution) while maximizing accumulated reward (e.g., quality of sensor viewpoint) motivated by visual assistance or tracking scenarios in unstructured or confined…

机器人学 · 计算机科学 2019-03-11 Xuesu Xiao , Jan Dufek , Robin Murphy

We introduce an algorithm which solves mean payoff games in polynomial time on average, assuming the distribution of the games satisfies a flip invariance property on the set of actions associated with every state. The algorithm is a…

计算机科学与博弈论 · 计算机科学 2014-09-12 Xavier Allamigeon , Pascal Benchimol , Stéphane Gaubert

Many high-stakes decision-making problems, such as those found within cybersecurity and economics, can be modeled as competitive resource allocation games. In these games, multiple players must allocate limited resources to overcome their…

计算机科学与博弈论 · 计算机科学 2024-01-10 N'yoma Diamond , Fabricio Murai