English
Related papers

Related papers: Zero-sum turn games using Q-learning: finite compu…

200 papers

We propose the first model-free algorithm that achieves low regret performance for decentralized learning in two-player zero-sum tabular stochastic games with infinite-horizon average-reward objective. In decentralized learning, the…

Machine Learning · Computer Science 2023-01-16 Romain Cravic , Nicolas Gast , Bruno Gaujal

For two-person dynamic zero-sum games (both discrete and continuous settings), we investigate the limit of value functions of finite horizon games with long run average cost as the time horizon tends to infinity and the limit of value…

Optimization and Control · Mathematics 2017-09-26 Dmitry Khlopin

Making decisions in the presence of a strategic opponent requires one to take into account the opponent's ability to actively mask its intended objective. To describe such strategic situations, we introduce the non-cooperative inverse…

Computer Science and Game Theory · Computer Science 2020-01-07 Xiangyuan Zhang , Kaiqing Zhang , Erik Miehling , Tamer Başar

Simple stochastic games are turn-based 2.5-player zero-sum graph games with a reachability objective. The problem is to compute the winning probability as well as the optimal strategies of both players. In this paper, we compare the three…

Computer Science and Game Theory · Computer Science 2022-07-21 Jan Kretinsky , Emanuel Ramneantu , Alexander Slivinskiy , Maximilian Weininger

We consider a stochastic differential game in the context of forward-backward stochastic differential equations, where one player implements an impulse control while the opponent controls the system continuously. Utilizing the notion of…

Optimization and Control · Mathematics 2021-12-20 Magnus Perninge

Although recent work in AI has made great progress in solving large, zero-sum, extensive-form games, the underlying assumption in most past work is that the parameters of the game itself are known to the agents. This paper deals with the…

Machine Learning · Computer Science 2018-06-29 Chun Kai Ling , Fei Fang , J. Zico Kolter

We consider the problem of a particular kind of quantum correlation that arises in some two-party games. In these games, one player is presented with a question they must answer, yielding an outcome of either 'win' or 'lose'. Molina and…

Quantum Physics · Physics 2017-03-14 Srinivasan Arunachalam , Abel Molina , Vincent Russo

The theory of quantum games permits players to choose strategies that prepare and measure quantum states. Whereas conventional game theory provides guarantees for fixed-point stability in non-cooperative games, so-called Nash equilibria, we…

Quantum Physics · Physics 2017-12-13 Faisal Shah Khan , Travis S. Humble

We revisit the problem of learning in two-player zero-sum Markov games, focusing on developing an algorithm that is uncoupled, convergent, and rational, with non-asymptotic convergence rates. We start from the case of stateless matrix game…

Computer Science and Game Theory · Computer Science 2023-11-10 Yang Cai , Haipeng Luo , Chen-Yu Wei , Weiqiang Zheng

Reinforcement learning from self-play has recently reported many successes. Self-play, where the agents compete with themselves, is often used to generate training data for iterative policy improvement. In previous work, heuristic rules are…

Machine Learning · Computer Science 2020-09-15 Yuanyi Zhong , Yuan Zhou , Jian Peng

Finite turn-based safety games have been used for very different problems such as the synthesis of linear temporal logic (LTL), the synthesis of schedulers for computer systems running on multiprocessor platforms, and also for the…

Logic in Computer Science · Computer Science 2014-05-08 Gilles Geeraerts , Joël Goossens , Amélie Stainer

This paper proposes a new method for finding closed-loop saddle points in zero-sum linear-quadratic stochastic differential games by decoupling their inherent structure. Specifically, we develop a nested iterative scheme that constructs a…

Optimization and Control · Mathematics 2025-12-10 Yiyuan Wang

We consider an infinite collection of agents who make decisions, sequentially, about an unknown underlying binary state of the world. Each agent, prior to making a decision, receives an independent private signal whose distribution depends…

Computer Science and Game Theory · Computer Science 2012-09-07 Kimon Drakopoulos , Asuman Ozdaglar , John Tsitsiklis

Value-based algorithms are a cornerstone of off-policy reinforcement learning due to their simplicity and training stability. However, their use has traditionally been restricted to discrete action spaces, as they rely on estimating…

Machine Learning · Computer Science 2025-10-23 Yigit Korkmaz , Urvi Bhuwania , Ayush Jain , Erdem Bıyık

Robust Reinforcement Learning (RRL) is a promising Reinforcement Learning (RL) paradigm aimed at training robust to uncertainty or disturbances models, making them more efficient for real-world applications. Following this paradigm,…

Machine Learning · Computer Science 2024-05-06 Anton Plaksin , Vitaly Kalev

Although significant progress has been made in decision-making for automated driving, challenges remain for deployment in the real world. One challenge lies in addressing interaction-awareness. Most existing approaches oversimplify…

Robotics · Computer Science 2026-02-10 Karim Essalmi , Fernando Garrido , Fawzi Nashashibi

We introduce FlipDyn with control, a finite-horizon zero-sum resource takeover game, where a defender and an adversary decide when to takeover and how to control a common resource. At each discrete-time step, the players can take over or…

Systems and Control · Electrical Eng. & Systems 2025-10-21 Sandeep Banik , Shaunak D. Bopardikar

The paper is concerned with a variant of the continuous-time finite state Markov game of control and stopping where both players can affect transition rates, while only one player can choose a stopping time. We use the dynamic programming…

Optimization and Control · Mathematics 2022-08-09 Yurii Averboukh

A significant roadblock to the development of principled multi-agent reinforcement learning is the fact that desired solution concepts like Nash equilibria may be intractable to compute. To overcome this obstacle, we take inspiration from…

Computer Science and Game Theory · Computer Science 2024-08-28 Eric Mazumdar , Kishan Panaganti , Laixi Shi

We study a zero-sum game where the evolution of a spectrally one-sided Levy process is modified by a singular controller and is terminated by the stopper. The singular controller minimizes the expected values of running, controlling and…

Optimization and Control · Mathematics 2014-08-08 Daniel Hernandez-Hernandez , Kazutoshi Yamazaki