English
Related papers

Related papers: {\epsilon}-Optimally Solving Two-Player Zero-Sum P…

200 papers

Partially observable stochastic games provide a rich mathematical paradigm for modeling multi-agent dynamic decision making under uncertainty and partial information. However, they generally do not admit closed-form solutions and are…

Optimization and Control · Mathematics 2020-04-15 Yanling Chang , Chelsea C. White

This paper presents a set of continuous-time distributed algorithms that solve unconstrained, separable, convex optimization problems over undirected networks with fixed topologies. The algorithms are developed using a Lyapunov function…

Systems and Control · Computer Science 2011-09-27 Jie Lu , Choon Yik Tang

We study zero-sum stochastic differential games with player dynamics governed by a nondegenerate controlled diffusion process. Under the assumption of uniform stability, we establish the existence of a solution to the Isaac's equation for…

Optimization and Control · Mathematics 2019-03-20 Ari Arapostathis , Vivek S. Borkar , K. Suresh Kumar

The Stackelberg prediction game (SPG) has been extensively used to model the interactions between the learner and data provider in the training process of various machine learning algorithms. Particularly, SPGs played prominent roles in…

Optimization and Control · Mathematics 2021-05-13 Jiali Wang , He Chen , Rujun Jiang , Xudong Li , Zihao Li

The values of two-player general-sum differential games are viscosity solutions to Hamilton-Jacobi-Isaacs (HJI) equations. Value and policy approximations for such games suffer from the curse of dimensionality (CoD). Alleviating CoD through…

Machine Learning · Computer Science 2024-06-04 Lei Zhang , Mukesh Ghimire , Zhe Xu , Wenlong Zhang , Yi Ren

We present version 2.0 of the Partial Exploration Tool (PET), a tool for verification of probabilistic systems. We extend the previous version by adding support for stochastic games, based on a recent unified framework for sound value…

Systems and Control · Electrical Eng. & Systems 2024-05-14 Tobias Meggendorfer , Maximilian Weininger

We prove the existence and uniqueness of viscosity solutions to quasi-variational inequalities (QVIs) with both upper and lower obstacles. In contrast to most previous works, we allow all involved coefficients to depend on the state…

Probability · Mathematics 2024-09-09 Magnus Perninge

We study the alternating gradient descent-ascent (AltGDA) algorithm in two-player zero-sum games. Alternating methods, where players take turns to update their strategies, have long been recognized as simple and practical approaches for…

Computer Science and Game Theory · Computer Science 2026-03-03 Tianlong Nan , Shuvomoy Das Gupta , Garud Iyengar , Christian Kroer

A zero-sum differential game with controlled jump-diffusion driven state is considered, and studied using a combination of dynamic programming and viscosity solution techniques. We prove, under certain conditions, that the value of the game…

Optimization and Control · Mathematics 2010-09-28 Imran H. Biswas

We revisit the problem of solving two-player zero-sum games in the decentralized setting. We propose a simple algorithmic framework that simultaneously achieves the best rates for honest regret as well as adversarial regret, and in addition…

Computer Science and Game Theory · Computer Science 2018-06-07 Ehsan Asadi Kangarshahi , Ya-Ping Hsieh , Mehmet Fatih Sahin , Volkan Cevher

This paper is concerned with a non-zero sum differential game problem of an anticipated forward-backward stochastic differential delayed equation under partial information. We establish a necessary maximum principle and sufficient…

Optimization and Control · Mathematics 2017-02-17 Yi Zhuang

We consider the problem of computing mixed Nash equilibria of two-player zero-sum games with continuous sets of pure strategies and with first-order access to the payoff function. This problem arises for example in game-theory-inspired…

Optimization and Control · Mathematics 2025-09-04 Guillaume Wang , Lénaïc Chizat

This article describes a novel game structure for autonomously optimizing decentralized manufacturing systems with multi-objective optimization challenges, namely Distributed Stackelberg Strategies in State-Based Potential Games (DS2-SbPG).…

Computer Science and Game Theory · Computer Science 2024-08-14 Steve Yuwono , Dorothea Schwung , Andreas Schwung

This paper studies two-player zero-sum repeated Bayesian games in which every player has a private type that is unknown to the other player, and the initial probability of the type of every player is publicly known. The types of players are…

Computer Science and Game Theory · Computer Science 2017-11-08 Lichun Li , Cedric Langbort , Jeff Shamma

We study synthesis problems with constraints in partially observable Markov decision processes (POMDPs), where the objective is to compute a strategy for an agent that is guaranteed to satisfy certain safety and performance specifications.…

In this paper we first investigate zero-sum two-player stochastic differential games with reflection with the help of theory of Reflected Backward Stochastic Differential Equations (RBSDEs). We will establish the dynamic programming…

Probability · Mathematics 2008-09-30 Rainer Buckdahn , Juan Li

In two-player finite-state stochastic games of partial observation on graphs, in every state of the graph, the players simultaneously choose an action, and their joint actions determine a probability distribution over the successor states.…

Computer Science and Game Theory · Computer Science 2011-07-13 Krishnendu Chatterjee , Laurent Doyen

The optimal value computation for turned-based stochastic games with reachability objectives, also known as simple stochastic games, is one of the few problems in $NP \cap coNP$ which are not known to be in $P$. However, there are some…

Computational Complexity · Computer Science 2014-08-10 David Auger , Pierre COUCHENEY , Yann Strozecki

We study the problem of solving discounted, two player, turn based, stochastic games (2TBSGs). Jurdzinski and Savani showed that 2TBSGs with deterministic transitions can be reduced to solving $P$-matrix linear complementarity problems…

Computer Science and Game Theory · Computer Science 2014-12-18 Thomas Dueholm Hansen , Rasmus Ibsen-Jensen

We introduce a contractive abstract dynamic programming framework and related policy iteration algorithms, specifically designed for sequential zero-sum games and minimax problems with a general structure. Aside from greater generality, the…

Computer Science and Game Theory · Computer Science 2021-10-22 Dimitri Bertsekas