English
Related papers

Related papers: {\epsilon}-Optimally Solving Two-Player Zero-Sum P…

200 papers

We study a zero-sum stochastic differential game (SDG) in which one controller plays an impulse control while their opponent plays a stochastic control. We consider an asymmetric setting in which the impulse player commits to, at the start…

Probability · Mathematics 2019-01-31 Parsiad Azimzadeh

We investigate a two-player zero-sum stochastic differential game problem with the state process being constrained in a connected bounded closed domain, and the cost functional described by the solution of a generalized backward stochastic…

Probability · Mathematics 2017-05-12 Lishun Xiao , Dejian Tian

2-TBSG is a two-player game model which aims to find Nash equilibriums and is widely utilized in reinforced learning and AI. Inspired by the fact that the simplex method for solving the deterministic discounted Markov decision processes…

Computer Science and Game Theory · Computer Science 2019-06-11 Zeyu Jia , Zaiwen Wen , Yinyu Ye

We examine the problem of the existence of optimal deterministic stationary strategiesintwo-players antagonistic (zero-sum) perfect information stochastic games with finitely many states and actions.We show that the existenceof such…

Computer Science and Game Theory · Computer Science 2016-11-28 Hugo Gimbert , Wieslaw Zielonka

We consider a stochastic differential game in the context of forward-backward stochastic differential equations, where one player implements an impulse control while the opponent controls the system continuously. Utilizing the notion of…

Optimization and Control · Mathematics 2021-12-20 Magnus Perninge

Simple stochastic games are two-player zero-sum stochastic games with turn-based moves, perfect information, and reachability winning conditions. We present two new algorithms computing the values of simple stochastic games. Both of them…

Computer Science and Game Theory · Computer Science 2015-07-01 Hugo Gimbert , Florian Horn

We consider a two-player zero-sum stochastic differential game in which one of the players has a private information on the game. Both players observe each other, so that the non-informed player can try to guess his missing information. Our…

Probability · Mathematics 2011-06-15 Christine Grün

In this paper we study a zero-sum switching game and its verification theorems expressed in terms of either a system of Reflected Backward Stochastic Differential Equations (RBSDEs in short) with bilateral interconnected obstacles or a…

Probability · Mathematics 2020-06-30 Said Hamadène , Tingshu Mu

We present a new, stochastic variant of the projective splitting (PS) family of algorithms for monotone inclusion problems. It can solve min-max and noncooperative game formulations arising in applications such as robust ML without the…

Optimization and Control · Mathematics 2021-06-25 Patrick R. Johnstone , Jonathan Eckstein , Thomas Flynn , Shinjae Yoo

In this paper, we study a class of zero-sum two-player stochastic differential games with the controlled stochastic differential equations and the payoff/cost functionals of recursive type. As opposed to the pioneering work by Fleming and…

Probability · Mathematics 2021-05-21 Jinniao Qiu , Jing Zhang

Zero-sum Dynkin games under Poisson constraints, where players can only stop at the event times of a Poisson process, have been studied widely in the recent literature. The constraint can be modelled in two ways: either both players share…

Optimization and Control · Mathematics 2025-12-09 David Hobson , Gechun Liang , Edward Wang

Motivated by Generative Adversarial Networks, we study the computation of Nash equilibrium in concave network zero-sum games (NZSGs), a multiplayer generalization of two-player zero-sum games first proposed with linear payoffs. Extending…

Machine Learning · Computer Science 2020-07-13 Amit Kadan , Hu Fu

Policy-based methods with function approximation are widely used for solving two-player zero-sum games with large state and/or action spaces. However, it remains elusive how to obtain optimization and statistical guarantees for such…

Machine Learning · Computer Science 2022-03-01 Yulai Zhao , Yuandong Tian , Jason D. Lee , Simon S. Du

In this paper, we formulate a two-player zero-sum game under dynamic constraints defined by hybrid dynamical equations. The game consists of a min-max problem involving a cost functional that depends on the actions and resulting solutions…

Optimization and Control · Mathematics 2025-05-20 Santiago J. Leudo , Ricardo G. Sanfelice

We present a fast numerical algorithm for large scale zero-sum stochastic games with perfect information, which combines policy iteration and algebraic multigrid methods. This algorithm can be applied either to a true finite state space…

Optimization and Control · Mathematics 2015-03-19 Marianne Akian , Sylvie Detournay

We study a nonzero-sum stochastic differential game with both players adopting impulse controls, on a finite time horizon. The objective of each player is to maximize her total expected discounted profits. The resolution methodology relies…

Optimization and Control · Mathematics 2021-12-21 René Aïd , Lamia Ben Ajmia , M'hamed Gaïgi , Mohamed Mnif

Policy space response oracles (PSRO) is a multi-agent reinforcement learning algorithm that has achieved state-of-the-art performance in very large two-player zero-sum games. PSRO is based on the tabular double oracle (DO) method, an…

Computer Science and Game Theory · Computer Science 2022-02-01 Stephen McAleer , Kevin Wang , John Lanier , Marc Lanctot , Pierre Baldi , Tuomas Sandholm , Roy Fox

This work presents a novel policy iteration algorithm to tackle nonzero-sum stochastic impulse games arising naturally in many applications. Despite the obvious impact of solving such problems, there are no suitable numerical methods…

Optimization and Control · Mathematics 2020-06-29 René Aïd , Francisco Bernal , Mohamed Mnif , Diego Zabaljauregui , Jorge P. Zubelli

We prove that zero-sum Dynkin games in continuous time with partial and asymmetric information admit a value in randomised stopping times when the stopping payoffs of the players are general \cadlag measurable processes. As a by-product of…

Probability · Mathematics 2022-06-08 Tiziano De Angelis , Nikita Merkulov , Jan Palczewski

We show that $\varepsilon$-additive approximations of the optimal value of fixed-size two-player free games with fixed-dimensional entanglement assistance can be computed in time $\mathrm{poly}(1/\varepsilon)$. This stands in contrast to…

Quantum Physics · Physics 2025-07-17 Julius A. Zeiss , Gereon Koßmann , Omar Fawzi , Mario Berta