English
Related papers

Related papers: Empirical Policy Optimization for $n$-Player Marko…

200 papers

Behavioral diversity, expert imitation, fairness, safety goals and others give rise to preferences in sequential decision making domains that do not decompose additively across time. We introduce the class of convex Markov games that allow…

Computer Science and Game Theory · Computer Science 2025-06-17 Ian Gemp , Andreas Haupt , Luke Marris , Siqi Liu , Georgios Piliouras

Feedback Nash equilibrium strategies in multi-agent dynamic games require availability of all players' state information to compute control actions. However, in real-world scenarios, sensing and communication limitations between agents make…

Computer Science and Game Theory · Computer Science 2025-04-10 Xinjie Liu , Jingqi Li , Filippos Fotiadis , Mustafa O. Karabag , Jesse Milzman , David Fridovich-Keil , Ufuk Topcu

We study a new class of Markov games, \emph(multi-player) zero-sum Markov Games} with \emph{Networked separable interactions} (zero-sum NMGs), to model the local interaction structure in non-cooperative multi-agent sequential…

Computer Science and Game Theory · Computer Science 2025-07-15 Chanwoo Park , Kaiqing Zhang , Asuman Ozdaglar

Softmax policy gradient is a popular algorithm for policy optimization in single-agent reinforcement learning, particularly since projection is not needed for each gradient update. However, in multi-agent systems, the lack of central…

Optimization and Control · Mathematics 2022-11-01 Runyu Zhang , Jincheng Mei , Bo Dai , Dale Schuurmans , Na Li

Multi-agent reinforcement learning has been successfully applied to fully-cooperative and fully-competitive environments, but little is currently known about mixed cooperative/competitive environments. In this paper, we focus on a…

Machine Learning · Computer Science 2021-10-22 Roy Fox , Stephen McAleer , Will Overman , Ioannis Panageas

In this paper, we investigate the seeking of Nash equilibrium (NE) in a non-cooperative quadratic game where all agents exchange their delayed strategy information with their neighbors. To extend best-response algorithms to the delayed…

Systems and Control · Electrical Eng. & Systems 2026-02-24 Kaichen Jiang , Yuyue Yan , Mingda Yue , Yuhu Wu

We present a framework for computing approximate mixed-strategy Nash equilibria of continuous-action games. It is a modification of the traditional double oracle algorithm, extended to multiple players and continuous action spaces. Unlike…

Computer Science and Game Theory · Computer Science 2024-06-14 Carlos Martin , Tuomas Sandholm

We discuss similarities and differences between systems of interacting players maximizing their individual payoffs and particles minimizing their interaction energy. Long-run behavior of stochastic dynamics of spatial games with multiple…

Statistical Mechanics · Physics 2009-11-10 Jacek Miekisz

Game theory studies situations in which strategic players can modify the state of a given system, due to the absence of a central authority. Solution concepts, such as Nash equilibrium, are defined to predict the outcome of such situations.…

Computer Science and Game Theory · Computer Science 2013-11-08 Diodato Ferraioli , Paul W. Goldberg , Carmine Ventre

We explore the use of policy approximations to reduce the computational cost of learning Nash equilibria in zero-sum stochastic games. We propose a new Q-learning type algorithm that uses a sequence of entropy-regularized soft policies to…

Machine Learning · Computer Science 2021-06-29 Yue Guan , Qifan Zhang , Panagiotis Tsiotras

In this work, we consider dynamic influence maximization games over social networks with multiple players (influencers). The goal of each influencer is to maximize their own reward subject to their limited total budget rate constraints.…

Computer Science and Game Theory · Computer Science 2023-09-26 Melih Bastopcu , S. Rasoul Etesami , Tamer Başar

We investigate mean field games for players, who are weakly coupled via their empirical measure. To this end we investigate time-dependent pure jump type propagators over a finite space in the framework of non-linear Markov processes. We…

Optimization and Control · Mathematics 2015-03-25 Rani Basna , Astrid Hilbert , Vassili N. Kolokoltsov

Reinforcement learning is a powerful tool to learn the optimal policy of possibly multiple agents by interacting with the environment. As the number of agents grow to be very large, the system can be approximated by a mean-field problem.…

Optimization and Control · Mathematics 2020-08-18 Weichen Wang , Jiequn Han , Zhuoran Yang , Zhaoran Wang

Due to the lack of coordination, it is unlikely that the selfish players of a strategic game reach a socially good state. A possible way to cope with selfishness is to compute a desired outcome (if it is tractable) and impose it. However…

Computer Science and Game Theory · Computer Science 2010-12-20 Bruno Escoffier , Laurent Gourvès , Jérôme Monnot

Partially Observable Markov Games (POMGs) provide a general framework for modeling multi-agent sequential decision-making under asymmetric information. A common approach is to reformulate a POMG as a fully observable Markov game over belief…

Multiagent Systems · Computer Science 2026-04-08 Lan Sang , Chinmay Maheshwari

We consider graphical $n$-person games with perfect information that have no Nash equilibria in pure stationary strategies. Solving these games in mixed strategies, we introduce probabilistic distributions in all non-terminal positions. The…

Combinatorics · Mathematics 2023-08-21 Vladimir Gurvich , Mariya Naumova

We study Markov potential games under the infinite horizon average reward criterion. Most previous studies have been for discounted rewards. We prove that both algorithms based on independent policy gradient and independent natural policy…

Machine Learning · Computer Science 2024-03-12 Min Cheng , Ruida Zhou , P. R. Kumar , Chao Tian

Motivated by the scarcity of accurate payoff feedback in practical applications of game theory, we examine a class of learning dynamics where players adjust their choices based on past payoff observations that are subject to noise and…

Optimization and Control · Mathematics 2016-06-03 Mario Bravo , Panayotis Mertikopoulos

Multiagent learning settings are inherently more difficult than single-agent learning because each agent interacts with other simultaneously learning agents in a shared environment. An effective approach in multiagent reinforcement learning…

Computer Science and Game Theory · Computer Science 2022-10-31 Dong-Ki Kim , Matthew Riemer , Miao Liu , Jakob N. Foerster , Gerald Tesauro , Jonathan P. How

In a social network, individuals express their opinions on several interdependent topics, and therefore the evolution of their opinions on these topics is also mutually dependent. In this work, we propose a differential game model for the…

Social and Information Networks · Computer Science 2024-02-23 Hossein B. Jond , Aykut Yıldız