English
Related papers

Related papers: Algorithms for zero-sum stochastic games with the …

200 papers

In this paper, we investigate a partially observable zero sum games where the state process is a discrete time Markov chain. We consider a general utility function in the optimization criterion. We show the existence of value for both…

Optimization and Control · Mathematics 2022-11-16 Arnab Bhabak , Subhamay saha

In this paper, we study the problem of escaping from saddle points in smooth nonconvex optimization problems subject to a convex set $\mathcal{C}$. We propose a generic framework that yields convergence to a second-order stationary point of…

Machine Learning · Computer Science 2018-10-10 Aryan Mokhtari , Asuman Ozdaglar , Ali Jadbabaie

We consider strongly-convex-strongly-concave saddle-point problems with general non-bilinear objective and different condition numbers with respect to the primal and the dual variables. First, we consider such problems with smooth composite…

Optimization and Control · Mathematics 2021-06-15 Vladislav Tominin , Yaroslav Tominin , Ekaterina Borodich , Dmitry Kovalev , Alexander Gasnikov , Pavel Dvurechensky

Stochastic games are an important class of problems that generalize Markov decision processes to game theoretic scenarios. We consider finite state two-player zero-sum stochastic games over an infinite time horizon with discounted rewards.…

Optimization and Control · Mathematics 2008-06-17 Parikshit Shah , Pablo A. Parrilo

We study a zero-sum stochastic differential switching game in infinite horizon. We prove the existence of the value of the game and characterize it as the unique viscosity solution of the associated system of quasi-variational inequalities…

Optimization and Control · Mathematics 2018-05-04 Brahim El Asri , Sehail Mazid

This paper deals with N-person nonzero-sum discrete-time Markov games under a probability criterion, in which the transition probabilities and reward functions are allowed to vary with time. Differing from the existing works on the expected…

Probability · Mathematics 2025-05-16 Xin Guo , Xin Wen

We consider a general class of nonzero-sum $N$-player stochastic games with impulse controls, where players control the underlying dynamics with discrete interventions. We adopt a verification approach and provide sufficient conditions for…

Optimization and Control · Mathematics 2020-10-06 Matteo Basei , Haoyang Cao , Xin Guo

This work develops an approximation procedure for a class of non-zero-sum stochastic differential investment and reinsurance games between two insurance companies. Both proportional reinsurance and excess-of loss reinsurance policies are…

Optimization and Control · Mathematics 2018-09-17 Trang Bui , Xiang Cheng , Zhuo Jin , George Yin

In stochastic games with stage duration h, players act at times 0, h, 2h, and so on. The payoff and leaving probabilities are proportional to h. As h approaches 0, such discrete-time games approximate games played in continuous time. The…

Optimization and Control · Mathematics 2024-09-25 Ivan Novikov

We study Bayesian learning in episodic, finite-horizon zero-sum Markov games with unknown transition and reward models. We investigate a posterior algorithm in which each player maintains a Bayesian posterior over the game model,…

Machine Learning · Computer Science 2026-03-24 Chang-Wei Yueh , Andy Zhao , Ashutosh Nayyar , Rahul Jain

We study policy optimization algorithms for computing correlated equilibria in multi-player general-sum Markov Games. Previous results achieve $O(T^{-1/2})$ convergence rate to a correlated equilibrium and an accelerated $O(T^{-3/4})$…

Machine Learning · Computer Science 2024-05-03 Yang Cai , Haipeng Luo , Chen-Yu Wei , Weiqiang Zheng

In this paper, we investigate the robustness of stationary mean-field equilibria in the presence of model uncertainties, specifically focusing on infinite-horizon discounted cost functions. To achieve this, we initially establish…

Systems and Control · Electrical Eng. & Systems 2026-04-10 Uğur Aydın , Naci Saldi

An algorithm of searching a zero of an unknown undimensional function is considered, measured at a point x with some error. The step sizes are random positive values and are calculated according to the rule: if two consecutive iterations…

Statistics Theory · Mathematics 2007-06-13 Alexander Plakhov , Pedro Cruz

We study Stackelberg equilibria in finitely repeated games, where the leader commits to a strategy that picks actions in each round and can be adaptive to the history of play (i.e. they commit to an algorithm). In particular, we study…

Computer Science and Game Theory · Computer Science 2024-03-08 Natalie Collina , Eshwar Ram Arunachaleswaran , Michael Kearns

This paper investigates group distributionally robust optimization (GDRO) with the goal of learning a model that performs well over $m$ different distributions. First, we formulate GDRO as a stochastic convex-concave saddle-point problem,…

Machine Learning · Computer Science 2024-11-21 Lijun Zhang , Haomin Bai , Peng Zhao , Tianbao Yang , Zhi-Hua Zhou

Quantitative games are two-player zero-sum games played on directed weighted graphs. Total-payoff games (that can be seen as a refinement of the well-studied mean-payoff games) are the variant where the payoff of a play is computed as the…

Computer Science and Game Theory · Computer Science 2015-07-15 Thomas Brihaye , Gilles Geeraerts , Axel Haddad , Benjamin Monmege

Recent applications that arise in machine learning have surged significant interest in solving min-max saddle point games. This problem has been extensively studied in the convex-concave regime for which a global equilibrium solution can be…

Optimization and Control · Mathematics 2019-11-01 Maher Nouiehed , Maziar Sanjabi , Tianjian Huang , Jason D. Lee , Meisam Razaviyayn

This paper is concerned with a non-zero sum differential game problem of an anticipated forward-backward stochastic differential delayed equation under partial information. We establish a necessary maximum principle and sufficient…

Optimization and Control · Mathematics 2017-02-17 Yi Zhuang

We explore the use of policy approximations to reduce the computational cost of learning Nash equilibria in zero-sum stochastic games. We propose a new Q-learning type algorithm that uses a sequence of entropy-regularized soft policies to…

Machine Learning · Computer Science 2021-06-29 Yue Guan , Qifan Zhang , Panagiotis Tsiotras

This paper focuses on zero-sum stochastic differential games in the framework of forward-backward stochastic differential equations on a finite time horizon with both players adopting impulse controls. By means of BSDE methods, in…

Optimization and Control · Mathematics 2021-04-08 Liangquan Zhang
‹ Prev 1 8 9 10 Next ›