English
Related papers

Related papers: Generalized conditional gradient and learning in p…

200 papers

Deep reinforcement learning provides a promising approach for text-based games in studying natural language communication between humans and artificial agents. However, the generalization still remains a big challenge as the agents depend…

Computation and Language · Computer Science 2021-09-22 Yunqiu Xu , Meng Fang , Ling Chen , Yali Du , Chengqi Zhang

We introduce a generalization of zero-sum network multiagent matrix games and prove that alternating gradient descent converges to the set of Nash equilibria at rate $O(1/T)$ for this set of games. Alternating gradient descent obtains this…

Computer Science and Game Theory · Computer Science 2021-10-07 James P. Bailey

We introduce a mean field game with rank-based reward: competing agents optimize their effort to achieve a goal, are ranked according to their completion time, and paid a reward based on their relative rank. First, we propose a tractable…

Optimization and Control · Mathematics 2017-08-07 Marcel Nutz , Yuchong Zhang

Gradient-based Meta-RL (GMRL) refers to methods that maintain two-level optimisation procedures wherein the outer-loop meta-learner guides the inner-loop gradient-based reinforcement learner to achieve fast adaptations. In this paper, we…

Machine Learning · Computer Science 2024-03-26 Xidong Feng , Bo Liu , Jie Ren , Luo Mai , Rui Zhu , Haifeng Zhang , Jun Wang , Yaodong Yang

We study a class of stochastic dynamic games that exhibit strategic complementarities between players; formally, in the games we consider, the payoff of a player has increasing differences between her own state and the empirical…

Computer Science and Game Theory · Computer Science 2010-12-13 Sachin Adlakha , Ramesh Johari

We derive a policy gradient theorem for Cumulative Prospect Theory (CPT) objectives in finite-horizon Reinforcement Learning (RL), generalizing the standard policy gradient theorem and encompassing distortion-based risk objectives as…

Machine Learning · Computer Science 2026-02-18 Olivier Lepel , Anas Barakat

Mean field games are studied by means of the weak formulation of stochastic optimal control. This approach allows the mean field interactions to enter through both state and control processes and take a form which is general enough to…

Probability · Mathematics 2015-04-09 Rene Carmona , Daniel Lacker

We consider a general-sum N-player linear-quadratic game with stochastic dynamics over a finite horizon and prove the global convergence of the natural policy gradient method to the Nash equilibrium. In order to prove the convergence of the…

Optimization and Control · Mathematics 2022-08-16 Ben Hambly , Renyuan Xu , Huining Yang

Multi-agent interactions are increasingly important in the context of reinforcement learning, and the theoretical foundations of policy gradient methods have attracted surging research interest. We investigate the global convergence of…

Optimization and Control · Mathematics 2023-03-21 Sarath Pattathil , Kaiqing Zhang , Asuman Ozdaglar

We analyze a system of partial differential equations that model a potential mean field game of controls, briefly MFGC. Such a game describes the interaction of infinitely many negligible players competing to optimize a personal value…

Analysis of PDEs · Mathematics 2020-10-27 Jameson Graber , Alan Mullenix , Laurent Pfeiffer

Fictitious play (FP) is a well-studied algorithm that enables agents to learn Nash equilibrium in games with certain reward structures. However, when agents have no prior knowledge of the reward functions, FP faces a major challenge: the…

Computer Science and Game Theory · Computer Science 2025-08-28 Semih Kara , Tamer Başar

This paper studies the n-player game and the mean field game under the CRRA relative performance on terminal wealth, in which the interaction occurs by peer competition. In the model with n agents, the price dynamics of underlying risky…

Mathematical Finance · Quantitative Finance 2023-02-10 Lijun Bo , Shihua Wang , Xiang Yu

We introduce a novel framework to model and solve mean-field game systems with nonlocal interactions. Our approach relies on kernel-based representations of mean-field interactions and feature-space expansions in the spirit of kernel…

Optimization and Control · Mathematics 2020-04-29 Siting Liu , Matthew Jacobs , Wuchen Li , Levon Nurbekyan , Stanley J. Osher

We explore a class of stochastic multiplayer games where each player in the game aims to optimize its objective under uncertainty and adheres to some expectation constraints. The study employs an offline learning paradigm, leveraging a…

Optimization and Control · Mathematics 2025-09-09 Yuanhanqing Huang , Jianghai Hu

We develop the linear programming approach to mean-field games in a general setting. This relaxed control approach allows to prove existence results under weak assumptions, and lends itself well to numerical implementation. We consider…

Optimization and Control · Mathematics 2020-11-24 Roxana Dumitrescu , Marcos Leutscher , Peter Tankov

Federated Reinforcement Learning (FRL) allows multiple agents to collaboratively build a decision making policy without sharing raw trajectories. However, if a small fraction of these agents are adversarial, it can lead to catastrophic…

Machine Learning · Computer Science 2024-11-06 Swetha Ganesh , Jiayu Chen , Gugan Thoppe , Vaneet Aggarwal

We consider a mean-field game model where the cost functions depend on a fixed parameter, called \textit{state}, which is unknown to players. Players learn about the state from a a stream of private signals they receive throughout the game.…

Optimization and Control · Mathematics 2024-02-01 Eran Shmaya , Bruno Ziliotto

In this paper, we consider a class of mean field games in which the optimal strategy of a representative agent depends on the statistical distribution of the states and controls. We prove some existence results for the forward-backward…

Analysis of PDEs · Mathematics 2020-07-13 Z Kobeissi

The goal of reinforcement learning algorithms is to estimate and/or optimise the value function. However, unlike supervised learning, no teacher or oracle is available to provide the true value function. Instead, the majority of…

Machine Learning · Computer Science 2018-05-25 Zhongwen Xu , Hado van Hasselt , David Silver

Many real-world problems modeled by stochastic games have huge state and/or action spaces, leading to the well-known curse of dimensionality. The complexity of the analysis of large-scale systems is dramatically reduced by exploiting mean…

Systems and Control · Computer Science 2015-03-19 H. Tembine