English
Related papers

Related papers: Near-Optimal Reinforcement Learning with Self-Play

200 papers

This paper investigates posterior sampling algorithms for competitive reinforcement learning (RL) in the context of general function approximations. Focusing on zero-sum Markov games (MGs) under two critical settings, namely self-play and…

Machine Learning · Computer Science 2023-11-01 Shuang Qiu , Ziyu Dai , Han Zhong , Zhaoran Wang , Zhuoran Yang , Tong Zhang

This work studies Nash equilibrium seeking for a class of stochastic aggregative games, where each player has an expectation-valued objective function depending on its local strategy and the aggregate of all players' strategies. We propose…

Optimization and Control · Mathematics 2022-05-17 Tongyu Wang , Peng Yi , Jie Chen

Many important real-world settings contain multiple players interacting over an unknown duration with probabilistic state transitions, and are naturally modeled as stochastic games. Prior research on algorithms for stochastic games has…

Computer Science and Game Theory · Computer Science 2021-02-19 Sam Ganzfried

Creating strong agents for games with more than two players is a major open problem in AI. Common approaches are based on approximating game-theoretic solution concepts such as Nash equilibrium, which have strong theoretical guarantees in…

Computer Science and Game Theory · Computer Science 2018-11-07 Sam Ganzfried , Austin Nowak , Joannier Pinales

We investigate a class of reinforcement learning dynamics where players adjust their strategies based on their actions' cumulative payoffs over time - specifically, by playing mixed strategies that maximize their expected cumulative payoff…

Optimization and Control · Mathematics 2016-02-10 Panayotis Mertikopoulos , William H. Sandholm

Multi-agent reinforcement learning (MARL), as a thriving field, explores how multiple agents independently make decisions in a shared dynamic environment. Due to environmental uncertainties, policies in MARL must remain robust to tackle the…

Machine Learning · Computer Science 2025-12-02 Na Li , Zewu Zheng , Wei Ni , Hangguan Shan , Wenjie Zhang , Xinyu Li

In this paper, we consider reinforcement learning of Markov Decision Processes (MDP) with peak constraints, where an agent chooses a policy to optimize an objective and at the same time satisfy additional constraints. The agent has to take…

Optimization and Control · Mathematics 2019-12-09 Ather Gattami

In this paper we consider the problem of computing an $\epsilon$-approximate Nash Equilibrium of a zero-sum game in a payoff matrix $A \in \mathbb{R}^{m \times n}$ with $O(1)$-bounded entries given access to a matrix-vector product oracle…

Optimization and Control · Mathematics 2025-09-05 Ishani Karmarkar , Liam O'Carroll , Aaron Sidford

We introduce, to our knowledge, the first direct second-order method for computing Nash equilibria in two-player zero-sum games. To do so, we construct a Douglas-Rachford-style splitting formulation, which we then solve with a semi-smooth…

Computer Science and Game Theory · Computer Science 2025-12-16 David Yang , Yuan Gao , Tianyi Lin , Christian Kroer

Existing methods for learning Stackelberg equilibria typically assume that the followers' (variational, generalized) Nash equilibrium is unique. However, in the presence of multiple equilibria, without a selection convention, the problem…

Optimization and Control · Mathematics 2026-04-30 Silvia Cianchi , Anibal Sanjab , Sergio Grammatico

Nash equilibrium has long been a desired solution concept in multi-player games, especially for those on continuous strategy spaces, which have attracted a rapidly growing amount of interests due to advances in research applications such as…

Computer Science and Game Theory · Computer Science 2019-10-29 Zehao Dou , Xiang Yan , Dongge Wang , Xiaotie Deng

We consider the problem of two-player zero-sum games. This problem is formulated as a min-max Markov game in the literature. The solution of this game, which is the min-max payoff, starting from a given state is called the min-max value of…

Machine Learning · Computer Science 2022-03-21 Raghuram Bharadwaj Diddigi , Chandramouli Kamanchi , Shalabh Bhatnagar

Repeated games consider a situation where multiple agents are motivated by their independent rewards throughout learning. In general, the dynamics of their learning become complex. Especially when their rewards compete with each other like…

Computer Science and Game Theory · Computer Science 2023-05-23 Yuma Fujimoto , Kaito Ariu , Kenshi Abe

We study techniques to incentivize self-interested agents to form socially desirable solutions in scenarios where they benefit from mutual coordination. Towards this end, we consider coordination games where agents have different intrinsic…

Computer Science and Game Theory · Computer Science 2014-04-21 Elliot Anshelevich , Shreyas Sekar

We motivate and propose a new model for non-cooperative Markov game which considers the interactions of risk-aware players. This model characterizes the time-consistent dynamic "risk" from both stochastic state transitions (inherent to the…

Computer Science and Game Theory · Computer Science 2019-11-22 Wenjie Huang , Pham Viet Hai , William B. Haskell

Game-theoretic approaches and Nash equilibrium have been widely applied across various engineering domains. However, practical challenges such as disturbances, delays, and actuator limitations can hinder the precise execution of Nash…

Computer Science and Game Theory · Computer Science 2026-03-17 Mahdis Rabbani , Navid Mojahed , Shima Nazari

Equilibria of realistic multiplayer games constitute a key solution concept both in practical applications, such as online advertising auctions and electricity markets, and in analytical frameworks used to study strategic voting in…

Computer Science and Game Theory · Computer Science 2025-11-18 Jakub Černý , Shuvomoy Das Gupta , Christian Kroer

Motivated by the scarcity of accurate payoff feedback in practical applications of game theory, we examine a class of learning dynamics where players adjust their choices based on past payoff observations that are subject to noise and…

Optimization and Control · Mathematics 2016-06-03 Mario Bravo , Panayotis Mertikopoulos

We consider multi-agent decision making, where each agent optimizes its cost function subject to constraints. Agents' actions belong to a compact convex Euclidean space and the agents' cost functions are coupled. We propose a distributed…

Optimization and Control · Mathematics 2016-12-01 Tatiana Tatarenko , Maryam Kamgarpour

Nash equilibrium is a central concept in game theory. Several Nash solvers exist, yet none scale to normal-form games with many actions and many players, especially those with payoff tensors too big to be stored in memory. In this work, we…

Computer Science and Game Theory · Computer Science 2022-02-07 Ian Gemp , Rahul Savani , Marc Lanctot , Yoram Bachrach , Thomas Anthony , Richard Everett , Andrea Tacchetti , Tom Eccles , János Kramár
‹ Prev 1 8 9 10 Next ›