English
Related papers

Related papers: {\epsilon}-Optimally Solving Two-Player Zero-Sum P…

200 papers

We study episodic two-player zero-sum Markov games (MGs) in the offline setting, where the goal is to find an approximate Nash equilibrium (NE) policy pair based on a dataset collected a priori. When the dataset does not have uniform…

Machine Learning · Computer Science 2023-01-02 Han Zhong , Wei Xiong , Jiyuan Tan , Liwei Wang , Tong Zhang , Zhaoran Wang , Zhuoran Yang

Zero-sum stochastic games generalize the notion of Markov Decision Processes (i.e. controlled Markov chains, or stochastic dynamic programming) to the 2-player competitive case : two players jointly control the evolution of a state…

Optimization and Control · Mathematics 2019-05-17 Jérôme Renault

This investigation is dedicated to a two-player zero-sum stochastic differential game (SDG), where a cost function is characterized by a backward stochastic differential equation (BSDE) with a continuous and monotonic generator regarding…

Optimization and Control · Mathematics 2024-04-19 Guangchen Wang , Zhuangzhuang Xing

We consider zero-sum stochastic games with perfect information and finitely many states and actions. The payoff is computed by a function which associates to each infinite sequence of states and actions a real number. We prove that if the…

Computer Science and Game Theory · Computer Science 2022-03-29 Hugo Gimbert , Edon Kelmendi

In this paper we use viscosity approach to provide an explicit solution to the problem of a two - player switching game. We characterize the switching regions which reduce the switching problem into one of finding a finite number of…

Optimization and Control · Mathematics 2025-04-22 Brahim El Asri , Magnoudéwa Paka

Leader-follower general-sum stochastic games (LF-GSSGs) model sequential decision-making under asymmetric commitment, where a leader commits to a policy and a follower best responds, yielding a strong Stackelberg equilibrium (SSE) with…

Computer Science and Game Theory · Computer Science 2025-12-08 Jilles Steeve Dibangoye , Thibaut Le Marre , Ocan Sankur , François Schwarzentruber

We generalize the results of Fleming and Souganidis (1989) on zero sum stochastic differential games to the case when the controls are unbounded. We do this by proving a dynamic programming principle using a covering argument instead of…

Optimization and Control · Mathematics 2012-01-17 Erhan Bayraktar , Song Yao

Progressively intricate cyber infiltration mechanisms have made conventional means of defense, such as firewalls and malware detectors, incompetent. These sophisticated infiltration mechanisms can study the defender's behavior, identify…

Artificial Intelligence · Computer Science 2018-10-02 Mohamadreza Ahmadi , Murat Cubuktepe , Nils Jansen , Sebastian Junges , Joost-Pieter Katoen , Ufuk Topcu

We present a new approach to solving games with a countably or uncountably infinite number of players. Such games are often used to model multiagent systems with a large number of agents. The latter are frequently encountered in economics,…

Computer Science and Game Theory · Computer Science 2025-01-17 Carlos Martin , Tuomas Sandholm

In this paper, we propose Posterior Sampling Reinforcement Learning for Zero-sum Stochastic Games (PSRL-ZSG), the first online learning algorithm that achieves Bayesian regret bound of $O(HS\sqrt{AT})$ in the infinite-horizon zero-sum…

Machine Learning · Computer Science 2024-03-12 Mehdi Jafarnia-Jahromi , Rahul Jain , Ashutosh Nayyar

In this paper, we investigate a partially observable zero sum games where the state process is a discrete time Markov chain. We consider a general utility function in the optimization criterion. We show the existence of value for both…

Optimization and Control · Mathematics 2022-11-16 Arnab Bhabak , Subhamay saha

Value methods for solving stochastic games with partial observability model the uncertainty about states of the game as a probability distribution over possible states. The dimension of this belief space is the number of states. For many…

Computer Science and Game Theory · Computer Science 2019-03-14 Karel Horák , Branislav Bošanský , Christopher Kiekintveld , Charles Kamhoua

We consider zero-sum stochastic differential games with possibly path-dependent controlled state. Unlike the previous literature, we allow for weak solutions of the state equation so that the players' controls are automatically of feedback…

Probability · Mathematics 2018-08-14 Dylan Possamaï , Nizar Touzi , Jianfeng Zhang

We derive sublinear-time quantum algorithms for computing the Nash equilibrium of two-player zero-sum games, based on efficient Gibbs sampling methods. We are able to achieve speed-ups for both dense and sparse payoff matrices at the cost…

Quantum Physics · Physics 2019-04-08 Joran van Apeldoorn , András Gilyén

Mean-payoff games (MPGs) are infinite duration two-player zero-sum games played on weighted graphs. Under the hypothesis of perfect information, they admit memoryless optimal strategies for both players and can be solved in…

Logic in Computer Science · Computer Science 2015-04-14 Paul Hunter , Guillermo A. Pérez , Jean-François Raskin

Saddle point optimization is a critical problem employed in numerous real-world applications, including portfolio optimization, generative adversarial networks, and robotics. It has been extensively studied in cases where the objective…

Machine Learning · Computer Science 2025-03-25 Shubhankar Agarwal , Hamzah I. Khan , Sandeep P. Chinchali , David Fridovich-Keil

We study Nash equilibrium learning in partially observable Markov games (POMGs), a multi-agent reinforcement learning framework in which agents cannot fully observe the underlying state. Prior work in this setting relies on centralization…

Computer Science and Game Theory · Computer Science 2026-05-08 Philip Jordan , Maryam Kamgarpour

The Policy-Space Response Oracles (PSRO) framework scales equilibrium computation to large zero-sum games by iteratively expanding a restricted strategy set using deep reinforcement learning (DRL). A central challenge is to construct, under…

Artificial Intelligence · Computer Science 2026-05-28 Junyu Zhang , Feihong Yang , Jian Wang , Chao Wang , Xudong Zhang

We address two-player general-sum stochastic Stackelberg games (SSGs), where the leader's policy is optimized considering the best-response follower whose policy is optimal for its reward under the leader. Existing policy gradient and value…

Computer Science and Game Theory · Computer Science 2026-03-17 Mikoto Kudo , Youhei Akimoto

The problem of two-player zero-sum Markov games has recently attracted increasing interests in theoretical studies of multi-agent reinforcement learning (RL). In particular, for finite-horizon episodic Markov decision processes (MDPs), it…

Machine Learning · Computer Science 2024-06-07 Songtao Feng , Ming Yin , Yu-Xiang Wang , Jing Yang , Yingbin Liang
‹ Prev 1 3 4 5 6 7 10 Next ›