中文
相关论文

相关论文: A Pseudo-Polynomial Algorithm for Mean Payoff Stoc…

200 篇论文

In two-player games on graphs, the simplest possible strategies are those that can be implemented without any memory. These are called positional strategies. In this paper, we characterize objectives recognizable by deterministic B\"uchi…

计算机科学与博弈论 · 计算机科学 2024-09-04 Patricia Bouyer , Antonio Casares , Mickael Randour , Pierre Vandenhove

We prove that every two-player non-zero-sum Borel game with lower-semi-continuous payoffs admits a subgame-perfect $\ep$-equilibrium. This result complements Example 3 in Solan and Vieille (2003), which shows that a subgame-perfect…

概率论 · 数学 2009-11-18 Ayala Mashiah-Yaakovi , Eilon Solan

We study zero-sum stochastic games for controlled discrete time Markov chains with risk-sensitive average cost criterion with countable state space and Borel action spaces. The payoff function is nonnegative and possibly unbounded. Under a…

最优化与控制 · 数学 2022-01-12 Mrinal K. Ghosh , Subrata Golui , Chandan Pal , Somnath Pradhan

We show that every two-player stochastic game with finite state and action sets and bounded, Borel-measurable, and shift-invariant payoffs, admits an $\ep$-equilibrium for all $\varepsilon>0$.

最优化与控制 · 数学 2022-03-29 János Flesch , Eilon Solan

In the paper we present a model of discrete-time mean-field game with several populations of players. Mean-field games with multiple populations of the players have only been studied in the literature in the continuous-time setting. The…

最优化与控制 · 数学 2023-04-07 Piotr Więcek

Learning from repeated play in a fixed two-player zero-sum game is a classic problem in game theory and online learning. We consider a variant of this problem where the game payoff matrix changes over time, possibly in an adversarial…

机器学习 · 计算机科学 2022-02-01 Mengxiao Zhang , Peng Zhao , Haipeng Luo , Zhi-Hua Zhou

In this paper, we study games with continuous action spaces and non-linear payoff functions. Our key insight is that Lipschitz continuity of the payoff function allows us to provide algorithms for finding approximate equilibria in these…

计算机科学与博弈论 · 计算机科学 2016-03-31 Argyrios Deligkas , John Fearnley , Paul Spirakis

Mean-payoff games on timed automata are played on the infinite weighted graph of configurations of priced timed automata between two players, Player Min and Player Max, by moving a token along the states of the graph to form an infinite…

计算机科学与博弈论 · 计算机科学 2020-01-16 Shibashis Guha , Marcin Jurdzinski , Krishna S. , Ashutosh Trivedi

In this paper we introduce polytopal stochastic games, an extension of two-player, zero-sum, turn-based stochastic games, in which we may have uncertainty over the transition probabilities. In these games the uncertainty over the…

计算机科学中的逻辑 · 计算机科学 2025-02-26 Pablo F. Castro , Pedro D'Argenio

We present a novel framework for {\epsilon}-optimally solving two-player zero-sum partially observable stochastic games (zs-POSGs). These games pose a major challenge due to the absence of a principled connection with dynamic programming…

计算机科学与博弈论 · 计算机科学 2025-11-17 Erwan Christian Escudie , Matthia Sabatelli , Olivier Buffet , Jilles Steeve Dibangoye

We propose a novel algorithm for the solution of mean-payoff games that merges together two seemingly unrelated concepts introduced in the context of parity games, small progress measures and quasi dominions. We show that the integration of…

计算机科学中的逻辑 · 计算机科学 2019-07-16 Massimo Benerecetti , Daniele Dell'Erba , Fabio Mogavero

This paper presents a case study for the application of semiring semantics for fixed-point formulae to the analysis of strategies in B\"uchi games. Semiring semantics generalizes the classical Boolean semantics by permitting multiple truth…

计算机科学中的逻辑 · 计算机科学 2021-09-20 Erich Grädel , Niels Lücking , Matthias Naaf

This paper presents a case study for the application of semiring semantics for fixed-point formulae to the analysis of strategies in B\"uchi games. Semiring semantics generalizes the classical Boolean semantics by permitting multiple truth…

计算机科学中的逻辑 · 计算机科学 2024-08-07 Erich Grädel , Niels Lücking , Matthias Naaf

In \emph{zero-sum two-player hidden stochastic games}, players observe partial information about the state. We address: $(i)$ the existence of the \emph{uniform value}, i.e., a limiting average payoff that both players can guarantee for…

最优化与控制 · 数学 2026-02-09 Krishnendu Chatterjee , David Lurie , Raimundo Saona , Bruno Ziliotto

In statistical decision theory involving a single decision-maker, an information structure is said to be better than another one if for any cost function involving a hidden state variable and an action variable which is restricted to be…

最优化与控制 · 数学 2021-01-07 Ian Hogeboom-Burr , Serdar Yüksel

We consider a two-player zero-sum stochastic differential game in which one of the players has a private information on the game. Both players observe each other, so that the non-informed player can try to guess his missing information. Our…

概率论 · 数学 2011-06-15 Christine Grün

We study best-response type learning dynamics for zero-sum polymatrix games under two information settings. The two settings are distinguished by the type of information that each player has about the game and their opponents' strategy. The…

最优化与控制 · 数学 2025-08-13 Fathima Zarin Faizal , Asuman Ozdaglar , Martin J. Wainwright

The existence of stationary Markov perfect equilibria in stochastic games is shown under a general condition called "(decomposable) coarser transition kernels". This result covers various earlier existence results on correlated equilibria,…

最优化与控制 · 数学 2017-01-24 Wei He , Yeneng Sun

We study a class of two-player zero-sum stochastic games known as \textit{blind stochastic games}, where players neither observe the state nor receive any information about it during the game. A central concept for analyzing long-duration…

最优化与控制 · 数学 2025-11-24 Krishnendu Chatterjee , David Lurie , Raimundo Saona , Bruno Ziliotto

We consider the problem of solving random parity games. We prove that parity games exibit a phase transition threshold above $d_P$, so that when the degree of the graph that defines the game has a degree $d > d_P$ then there exists a…

计算机科学中的逻辑 · 计算机科学 2020-07-17 Richard Combes , Mikael Touati