中文
相关论文

相关论文: HSVI can solve zero-sum Partially Observable Stoch…

200 篇论文

Policy-based methods with function approximation are widely used for solving two-player zero-sum games with large state and/or action spaces. However, it remains elusive how to obtain optimization and statistical guarantees for such…

机器学习 · 计算机科学 2022-03-01 Yulai Zhao , Yuandong Tian , Jason D. Lee , Simon S. Du

Computational equilibrium finding in large zero-sum extensive-form imperfect-information games has led to significant recent AI breakthroughs. The fastest algorithms for the problem are new forms of counterfactual regret minimization [Brown…

计算机科学与博弈论 · 计算机科学 2020-07-01 Brian Hu Zhang , Tuomas Sandholm

Partially observable Markov decision processes (POMDPs) provide an elegant mathematical framework for modeling complex decision and planning problems in stochastic domains in which states of the system are observable only indirectly, via a…

人工智能 · 计算机科学 2011-06-02 M. Hauskrecht

We consider a two-player zero-sum game with integral payoff and with incomplete information on one side, where the payoff is chosen among a continuous set of possible payoffs. We prove that the value function of this game is solution of an…

概率论 · 数学 2012-02-23 Pierre Cardaliaguet , Catherine Rainer

This paper addresses a continuous-time risk-minimizing two-player zero-sum stochastic differential game (SDG), in which each player aims to minimize its probability of failure. Failure occurs in the event when the state of the game enters…

最优化与控制 · 数学 2023-08-23 Apurva Patil , Yujing Zhou , David Fridovich-Keil , Takashi Tanaka

Two-player complete-information game trees are perhaps the simplest possible setting for studying general-sum games and the computational problem of finding equilibria. These games admit a simple bottom-up algorithm for finding subgame…

计算机科学与博弈论 · 计算机科学 2012-07-02 Michael L. Littman , Nishkam Ravi , Arjun Talwar , Martin Zinkevich

In this paper, we study a class of zero-sum two-player stochastic differential games with the controlled stochastic differential equations and the payoff/cost functionals of recursive type. As opposed to the pioneering work by Fleming and…

概率论 · 数学 2021-05-21 Jinniao Qiu , Jing Zhang

Partially observable Markov decision processes (POMDPs) rely on the key assumption that probability distributions are precisely known. Robust POMDPs (RPOMDPs) alleviate this concern by defining imprecise probabilities, referred to as…

人工智能 · 计算机科学 2024-07-30 Eline M. Bovy , Marnix Suilen , Sebastian Junges , Nils Jansen

This work considers two-player zero-sum semi-Markov games with incomplete information on one side and perfect observation. At the beginning, the system selects a game type according to a given probability distribution and informs to Player…

最优化与控制 · 数学 2021-07-16 Fang Chen , Xianping Guo , Zhong-Wei Liao

We prove that for a class of zero-sum differential games with incomplete information on both sides, the value admits a probabilistic representation as the value of a zero-sum stochastic differential game with complete information, where…

最优化与控制 · 数学 2017-01-04 Fabien Gensbittel , Catherine Rainer

We consider discrete time partially observable zero-sum stochastic game with average payoff criterion. We study the game using an equivalent completely observable game. We show that the game has a value and also we come up with a pair of…

最优化与控制 · 数学 2014-09-16 Subhamay Saha

In this paper, we consider a differential stochastic zero-sum game in which two players intervene by adopting impulse controls in a finite time horizon. We provide a numerical solution as an approximation of the value function, which turns…

最优化与控制 · 数学 2024-10-14 Antoine Zolome , Brahim El Asri

We consider a game, in which the dynamics is described by a non-linear Volterra integral equation of Hammerstein type with a weakly-singular kernel and the goals of the first and second players are, respectively, to minimize and maximize a…

最优化与控制 · 数学 2024-04-17 Mikhail I. Gomoyunov

While recent reductions of zero-sum partially observable stochastic games (zs-POSGs) to transition-independent stochastic games (TI-SGs) theoretically admit dynamic programming, practical solutions remain stifled by the inherent…

计算机科学与博弈论 · 计算机科学 2026-05-04 Erwan C. Escudie , Matthia Sabatelli , Jilles S. Dibangoye

Zero-sum stochastic games generalize the notion of Markov Decision Processes (i.e. controlled Markov chains, or stochastic dynamic programming) to the 2-player competitive case : two players jointly control the evolution of a state…

最优化与控制 · 数学 2019-05-17 Jérôme Renault

We consider a class of two-player zero-sum stochastic games with finite state and compact control spaces, which we call stochastic shortest path (SSP) games. They are undiscounted total cost stochastic dynamic games that have a cost-free…

最优化与控制 · 数学 2014-12-31 Huizhen Yu

We examine the problem of the existence of optimal deterministic stationary strategiesintwo-players antagonistic (zero-sum) perfect information stochastic games with finitely many states and actions.We show that the existenceof such…

计算机科学与博弈论 · 计算机科学 2016-11-28 Hugo Gimbert , Wieslaw Zielonka

In this paper we study zero-sum two-player stochastic differential games with the help of theory of Backward Stochastic Differential Equations (BSDEs). At the one hand we generalize the results of the pioneer work of Fleming and Souganidis…

概率论 · 数学 2011-02-19 Rainer Buckdahn , Juan Li

We study a nonzero-sum stochastic differential game with both players adopting impulse controls, on a finite time horizon. The objective of each player is to maximize her total expected discounted profits. The resolution methodology relies…

最优化与控制 · 数学 2021-12-21 René Aïd , Lamia Ben Ajmia , M'hamed Gaïgi , Mohamed Mnif

Simple stochastic games are two-player zero-sum stochastic games with turn-based moves, perfect information, and reachability winning conditions. We present two new algorithms computing the values of simple stochastic games. Both of them…

计算机科学与博弈论 · 计算机科学 2015-07-01 Hugo Gimbert , Florian Horn