English
Related papers

Related papers: Policy iteration algorithm for zero-sum multichain…

200 papers

In the literature, existence of mean-field equilibria has been established for discrete-time mean field games under both the discounted cost and the average cost optimality criteria. In this paper, we provide a value iteration algorithm to…

Systems and Control · Electrical Eng. & Systems 2020-04-02 Berkay Anahtarci , Can Deha Kariksiz , Naci Saldi

We study the computational complexity of solving stochastic games with mean-payoff objectives. Instead of identifying special classes in which simple strategies are sufficient to play $\epsilon$-optimally, or form $\epsilon$-Nash…

Computer Science and Game Theory · Computer Science 2024-05-16 Sougata Bose , Rasmus Ibsen-Jensen , Patrick Totzke

Convergence of the policy iteration method for discrete and continuous optimal control problems holds under general assumptions. Moreover, in some circumstances, it is also possible to show a quadratic rate of convergence for the algorithm.…

Optimization and Control · Mathematics 2022-03-02 Fabio Camilli , Qing Tang

Zero-sum Markov Games (MGs) has been an efficient framework for multi-agent systems and robust control, wherein a minimax problem is constructed to solve the equilibrium policies. At present, this formulation is well studied under tabular…

Machine Learning · Computer Science 2022-12-06 Yangang Ren , Yao Lyu , Wenxuan Wang , Shengbo Eben Li , Zeyang Li , Jingliang Duan

In this paper we formulate and solve a mean-field game described by a linear stochastic dynamics and a quadratic or exponential-quadratic cost functional for each generic player. The optimal strategies for the players are given explicitly…

Optimization and Control · Mathematics 2014-12-02 Djehiche Boualem , Tembine Hamidou

Optimization under uncertainty is a fundamental problem in learning and decision-making, particularly in multi-agent systems. Previously, Feldman, Kalai, and Tennenholtz [2010] demonstrated the ability to efficiently compete in repeated…

Computer Science and Game Theory · Computer Science 2026-01-29 Daniel Ablin , Alon Cohen

A zero-sum two-person Perfect Information Semi-Markov game (PISMG) under limiting ratio average payoff has a value and both the maximiser and the minimiser have optimal pure semi-stationary strategies. We arrive at the result by first…

Computer Science and Game Theory · Computer Science 2023-02-15 S. Sinha , K. G. Bakshi

We propose projection-free sequential algorithms for linear-quadratic dynamics games. These policy gradient based algorithms are akin to Stackelberg leadership model and can be extended to model-free settings. We show that if the leader…

Systems and Control · Electrical Eng. & Systems 2019-11-13 Jingjing Bu , Lillian J. Ratliff , Mehran Mesbahi

Model-based algorithms -- algorithms that explore the environment through building and utilizing an estimated model -- are widely used in reinforcement learning practice and theoretically shown to achieve optimal sample efficiency for…

Machine Learning · Computer Science 2021-02-09 Qinghua Liu , Tiancheng Yu , Yu Bai , Chi Jin

In this paper, we consider the problem of optimization and learning for constrained and multi-objective Markov decision processes, for both discounted rewards and expected average rewards. We formulate the problems as zero-sum games where…

Optimization and Control · Mathematics 2021-03-05 Ather Gattami , Qinbo Bai , Vaneet Agarwal

We study stochastic zero-sum games on graphs, which are prevalent tools to model decision-making in presence of an antagonistic opponent in a random environment. In this setting, an important question is the one of strategy complexity: what…

Computer Science and Game Theory · Computer Science 2024-02-14 Patricia Bouyer , Youssouf Oualhadj , Mickael Randour , Pierre Vandenhove

A growing line of work reframes preference-based fine-tuning of large language models game-theoretically: Nash Learning from Human Feedback (NLHF) recasts the problem as a zero-sum game over policies. However, optimization is over expected…

Computer Science and Game Theory · Computer Science 2026-05-14 Max Horwitz , Jake Gonzales , Eric Mazumdar , Lillian J. Ratliff

In this paper, we consider a zero-sum undiscounted stochastic game which has finite state space and finitely many pure actions. Also, we assume the transition probability of the undiscounted stochastic game is controlled by one player and…

Optimization and Control · Mathematics 2022-09-23 Purba Das , T. Parthasarathy , G Ravindran

Recently, in [K.R. Apt and S. Simon: Well-founded extensive games with perfect information, TARK21], we studied well-founded games, a natural extension of finite extensive games with perfect information in which all plays are finite. We…

Computer Science and Game Theory · Computer Science 2023-07-18 Krzysztof R. Apt , Sunil Simon

Matrix games constitute a fundamental problem of game theory and describe a situation of two players with completely conflicting interests. We show how methods from statistical mechanics can be used to investigate the statistical properties…

Disordered Systems and Neural Networks · Physics 2009-10-31 J. Berg , A. Engel

This paper studies a multi-player, general-sum stochastic game characterized by a dual-stage temporal structure per period. The agents face uncertainty regarding the time-evolving state that is realized at the beginning of each period.…

Computer Science and Game Theory · Computer Science 2023-10-09 Tao Zhang , Quanyan Zhu

In two-player zero-sum stochastic games, where two competing players make decisions under uncertainty, a pair of optimal strategies is traditionally described by Nash equilibrium and computed under the assumption that the players have…

Optimization and Control · Mathematics 2019-07-30 Yagiz Savas , Mohamadreza Ahmadi , Takashi Tanaka , Ufuk Topcu

In the present work, we consider 2-person zero-sum stochastic differential games with a nonlinear pay-off functional which is defined through a backward stochastic differential equation. Our main objective is to study for such a game the…

Probability · Mathematics 2014-07-29 Rainer Buckdahn , Juan Li , Marc Quincampoix

The classical, complete-information two-player games assume that the problem data (in particular the payoff matrix) is known exactly by both players. In a now famous result, Nash has shown that any such game has an equilibrium in mixed…

Computer Science and Game Theory · Computer Science 2015-12-11 Nicolas Loizou

Mean-payoff games (MPGs) are infinite duration two-player zero-sum games played on weighted graphs. Under the hypothesis of perfect information, they admit memoryless optimal strategies for both players and can be solved in…

Logic in Computer Science · Computer Science 2015-04-14 Paul Hunter , Guillermo A. Pérez , Jean-François Raskin
‹ Prev 1 4 5 6 7 8 10 Next ›