English
Related papers

Related papers: A non-zero-sum game with reinforcement learning un…

200 papers

The paper studies an oligopolistic equilibrium model of financial agents who aim to share their random endowments. The risk-sharing securities and their prices are endogenously determined as the outcome of a strategic game played among all…

General Finance · Quantitative Finance 2016-05-18 Michail Anthropelos

Zero-sum games such as chess and poker are, abstractly, functions that evaluate pairs of agents, for example labeling them `winner' and `loser'. If the game is approximately transitive, then self-play generates sequences of agents of…

Machine Learning · Computer Science 2019-05-14 David Balduzzi , Marta Garnelo , Yoram Bachrach , Wojciech M. Czarnecki , Julien Perolat , Max Jaderberg , Thore Graepel

Contemporary applications of machine learning in two-team e-sports and the superior expressivity of multi-agent generative adversarial networks raise important and overlooked theoretical questions regarding optimization in two-team games.…

Computer Science and Game Theory · Computer Science 2023-04-18 Fivos Kalogiannis , Ioannis Panageas , Emmanouil-Vasileios Vlatakis-Gkaragkounis

Graphon games have been introduced to study games with many players who interact through a weighted graph of interaction. By passing to the limit, a game with a continuum of players is obtained, in which the interactions are through a…

Optimization and Control · Mathematics 2024-04-02 Mathieu Laurière , Ludovic Tangpi , Xuchen Zhou

Successful algorithms have been developed for computing Nash equilibrium in a variety of finite game classes. However, solving continuous games -- in which the pure strategy space is (potentially uncountably) infinite -- is far more…

Computer Science and Game Theory · Computer Science 2021-06-02 Sam Ganzfried

We study a nonzero-sum stochastic differential game with both players adopting impulse controls, on a finite time horizon. The objective of each player is to maximize her total expected discounted profits. The resolution methodology relies…

Optimization and Control · Mathematics 2021-12-21 René Aïd , Lamia Ben Ajmia , M'hamed Gaïgi , Mohamed Mnif

We discuss a natural game of competition and solve the corresponding mean field game with \emph{common noise} when agents' rewards are \emph{rank dependent}. We use this solution to provide an approximate Nash equilibrium for the finite…

Probability · Mathematics 2016-10-18 Erhan Bayraktar , Yuchong Zhang

Frequent violations of fair principles in real-life settings raise the fundamental question of whether such principles can guarantee the existence of a self-enforcing equilibrium in a free economy. We show that elementary principles of…

Theoretical Economics · Economics 2021-07-28 Ghislain H. Demeze-Jouatsa , Roland Pongou , Jean-Baptiste Tondji

We investigate stochastic utility maximization games under relative performance concerns in both finite-agent and infinite-agent (graphon) settings. An incomplete market model is considered where agents with power (CRRA) utility functions…

Optimization and Control · Mathematics 2024-12-05 Zongxia Liang , Keyu Zhang , Yaqi Zhuang

When a vehicle drives on the road, its behaviors will be affected by surrounding vehicles. Prediction and decision should not be considered as two separate stages because all vehicles make decisions interactively. This paper constructs the…

Artificial Intelligence · Computer Science 2023-02-09 Xujie Song , Zexi Lin

This paper studies the connection between a class of mean-field games and a social welfare optimization problem. We consider a mean-field game in function spaces with a large population of agents, and each agent seeks to minimize an…

Optimization and Control · Mathematics 2018-02-15 Sen Li , Wei Zhang , Lin Zhao

We study a heterogeneous agent macroeconomic model with an infinite number of households and firms competing in a labor market. Each household earns income and engages in consumption at each time step while aiming to maximize a concave…

General Economics · Economics 2023-03-10 Ruitu Xu , Yifei Min , Tianhao Wang , Zhaoran Wang , Michael I. Jordan , Zhuoran Yang

The standard risk minimization paradigm of machine learning is brittle when operating in environments whose test distributions are different from the training distribution due to spurious correlations. Training on data from many…

Machine Learning · Computer Science 2020-03-20 Kartik Ahuja , Karthikeyan Shanmugam , Kush R. Varshney , Amit Dhurandhar

Examining the behavior of multi-agent systems is vitally important to many emerging distributed applications - game theory has emerged as a powerful tool set in which to do so. The main approach of game-theoretic techniques is to model…

Computer Science and Game Theory · Computer Science 2024-06-03 Rohit Konda , Rahul Chandan , Jason Marden

A growing line of work reframes preference-based fine-tuning of large language models game-theoretically: Nash Learning from Human Feedback (NLHF) recasts the problem as a zero-sum game over policies. However, optimization is over expected…

Computer Science and Game Theory · Computer Science 2026-05-14 Max Horwitz , Jake Gonzales , Eric Mazumdar , Lillian J. Ratliff

This paper considers the problem of designing optimal algorithms for reinforcement learning in two-player zero-sum games. We focus on self-play algorithms which learn the optimal policy by playing against itself without any direct…

Machine Learning · Computer Science 2020-07-15 Yu Bai , Chi Jin , Tiancheng Yu

Reinforcement learning is a powerful tool to learn the optimal policy of possibly multiple agents by interacting with the environment. As the number of agents grow to be very large, the system can be approximated by a mean-field problem.…

Optimization and Control · Mathematics 2020-08-18 Weichen Wang , Jiequn Han , Zhuoran Yang , Zhaoran Wang

We approach the continuous-time mean-variance (MV) portfolio selection with reinforcement learning (RL). The problem is to achieve the best tradeoff between exploration and exploitation, and is formulated as an entropy-regularized, relaxed…

Portfolio Management · Quantitative Finance 2019-05-07 Haoran Wang , Xun Yu Zhou

Offline Reinforcement Learning (RL) enables policy improvement from fixed datasets without online interactions, making it highly suitable for real-world applications lacking efficient simulators. Despite its success in the single-agent…

Multiagent Systems · Computer Science 2025-10-15 Jingxiao Chen , Weiji Xie , Weinan Zhang , Yong yu , Ying Wen

We consider a number of questions related to tradeoffs between reward and regret in repeated gameplay between two agents. To facilitate this, we introduce a notion of $\textit{generalized equilibrium}$ which allows for asymmetric regret…

Computer Science and Game Theory · Computer Science 2023-12-19 William Brown , Jon Schneider , Kiran Vodrahalli