中文
相关论文

相关论文: Expectation in Stochastic Games with Prefix-indepe…

200 篇论文

Delay games are two-player games of infinite duration in which one player may delay her moves to obtain a lookahead on her opponent's moves. Recently, such games with quantitative winning conditions in weak MSO with the unbounding…

计算机科学与博弈论 · 计算机科学 2016-10-11 Felix Klein , Martin Zimmermann

We study the problem of achieving decentralized coordination by a group of strategic decision makers choosing to engage or not in a task in a stochastic setting. First, we define a class of symmetric utility games that encompass a broad…

系统与控制 · 电气工程与系统科学 2023-04-05 Marcos M. Vasconcelos , Behrouz Touri

This paper investigates the two-person zero-sum stochastic games for piece-wise deterministic Markov decision processes with risk-sensitive finite-horizon cost criterion on a general state space. Here, the transition and cost/reward rates…

最优化与控制 · 数学 2024-05-15 Subrata Golui

Recent advancements in algorithms for sequential decision-making under imperfect information have shown remarkable success in large games such as limit- and no-limit poker. These algorithms traditionally formalize the games using the…

计算机科学与博弈论 · 计算机科学 2023-12-07 Vojtěch Kovařík , David Milec , Michal Šustr , Dominik Seitz , Viliam Lisý

We consider a class of two-player zero-sum stochastic games with finite state and compact control spaces, which we call stochastic shortest path (SSP) games. They are undiscounted total cost stochastic dynamic games that have a cost-free…

最优化与控制 · 数学 2014-12-31 Huizhen Yu

We analyse an algorithm solving stochastic mean-payoff games, combining the ideas of relative value iteration and of Krasnoselskii-Mann damping. We derive parameterized complexity bounds for several classes of games satisfying…

最优化与控制 · 数学 2023-05-05 Marianne Akian , Stéphane Gaubert , Ulysse Naepels , Basile Terver

Repeated game has long been the touchstone model for agents' long-run relationships. Previous results suggest that it is particularly difficult for a repeated game player to exert an autocratic control on the payoffs since they are jointly…

计算机科学与博弈论 · 计算机科学 2018-07-19 Dong Hao , Kai Li , Tao Zhou

We formalize the problem of maximizing the mean-payoff value with high probability while satisfying a parity objective in a Markov decision process (MDP) with unknown probabilistic transition function and unknown reward function. Assuming…

人工智能 · 计算机科学 2018-08-24 Jan Křetínský , Guillermo A. Pérez , Jean-François Raskin

We give a converging semidefinite programming hierarchy of outer approximations for the set of quantum correlations of fixed dimension and derive analytical bounds on the convergence speed of the hierarchy. In particular, we give a…

量子物理 · 物理学 2021-07-05 Hyejung H. Jee , Carlo Sparaciari , Omar Fawzi , Mario Berta

The distributed computation of equilibria and optima has seen growing interest in a broad collection of networked problems. We consider the computation of equilibria of convex stochastic Nash games characterized by a possibly nonconvex…

最优化与控制 · 数学 2019-08-05 Jinlong Lei , Uday V. Shanbhag

We study zero-sum differential games with state constraints and one-sided information, where the informed player (Player 1) has a categorical payoff type unknown to the uninformed player (Player 2). The goal of Player 1 is to minimize his…

计算机科学与博弈论 · 计算机科学 2024-06-05 Mukesh Ghimire , Lei Zhang , Zhe Xu , Yi Ren

This paper studies two-player zero-sum games played on graphs and makes contributions toward the following question: given an objective, how much memory is required to play optimally for that objective? We study regular objectives, where…

计算机科学与博弈论 · 计算机科学 2023-09-19 Patricia Bouyer , Nathanaël Fijalkow , Mickael Randour , Pierre Vandenhove

We consider infinite-state turn-based stochastic games of two players, Box and Diamond, who aim at maximizing and minimizing the expected total reward accumulated along a run, respectively. Since the total accumulated reward is unbounded,…

计算机科学与博弈论 · 计算机科学 2012-08-09 Tomáš Brázdil , Antonín Kučera , Petr Novotný

We investigate the problem dependent regime in the stochastic Thresholding Bandit problem (TBP) under several shape constraints. In the TBP, the objective of the learner is to output, at the end of a sequential game, the set of arms whose…

机器学习 · 统计学 2021-06-21 James Cheshire , Pierre Ménard , Alexandra Carpentier

The Shapley value is arguably the most central normative solution concept in cooperative game theory. It specifies a unique way in which the reward from cooperation can be "fairly" divided among players. While it has a wide range of real…

计算机科学与博弈论 · 计算机科学 2014-02-14 Sasan Maleki , Long Tran-Thanh , Greg Hines , Talal Rahwan , Alex Rogers

We show that every two-player stochastic game with finite state and action sets and bounded, Borel-measurable, and shift-invariant payoffs, admits an $\ep$-equilibrium for all $\varepsilon>0$.

最优化与控制 · 数学 2022-03-29 János Flesch , Eilon Solan

This paper considers a class of reinforcement-based learning (namely, perturbed learning automata) and provides a stochastic-stability analysis in repeatedly-played, positive-utility, finite strategic-form games. Prior work in this class of…

计算机科学与博弈论 · 计算机科学 2019-01-29 Georgios C. Chasparis

This paper examines finite zero-sum stochastic games and demonstrates that when the game's duration is sufficiently long, there exists a pair of approximately optimal strategies such that the expected average payoff at any point in the game…

最优化与控制 · 数学 2024-12-02 Thomas Ragel , Bruno Ziliotto

Two-player quantitative zero-sum games provide a natural framework to synthesize controllers with performance guarantees for reactive systems within an uncontrollable environment. Classical settings include mean-payoff games, where the…

计算机科学中的逻辑 · 计算机科学 2015-09-25 Patricia Bouyer , Nicolas Markey , Mickael Randour , Kim G. Larsen , Simon Laursen

We study a two-player discounted zero-sum stochastic game model for dynamic operational planning in military campaigns. At each stage, the players manage multiple commanders who order military actions on objectives that have an open line of…

计算机科学与博弈论 · 计算机科学 2024-03-04 Joseph E. McCarthy , Mathieu Dahan , Chelsea C. White