中文
相关论文

相关论文: Reinforcement Learning for Mean Field Games with S…

200 篇论文

In this paper we study iterative procedures for stationary equilibria in games with large number of players. Most of learning algorithms for games with continuous action spaces are limited to strict contraction best reply maps in which the…

机器学习 · 计算机科学 2012-10-18 Hamidou Tembine , Raul Tempone , Pedro Vilanova

Multi-team games, prevalent in robotics and resource management, involve team members striving for a joint best response against other teams. Team-Nash equilibrium (TNE) predicts the outcomes of such coordinated interactions. However, can…

计算机科学与博弈论 · 计算机科学 2024-11-01 Ahmed Said Donmez , Yuksel Arslantas , Muhammed O. Sayin

Mean-field games have been studied under the assumption of very large number of players. For such large systems, the basic idea consists to approximate large games by a stylized game model with a continuum of players. The approach has been…

计算机科学与博弈论 · 计算机科学 2014-04-08 Hamidou Tembine

We study the problem of learning a Nash equilibrium (NE) in Markov games which is a cornerstone in multi-agent reinforcement learning (MARL). In particular, we focus on infinite-horizon adversarial team Markov games (ATMGs) in which agents…

计算机科学与博弈论 · 计算机科学 2024-10-10 Fivos Kalogiannis , Jingming Yan , Ioannis Panageas

Games are natural models for multi-agent machine learning settings, such as generative adversarial networks (GANs). The desirable outcomes from algorithmic interactions in these games are encoded as game theoretic equilibrium concepts, e.g.…

计算机科学与博弈论 · 计算机科学 2022-02-25 Gabriel P. Andrade , Rafael Frongillo , Georgios Piliouras

Mean-Field Games are games with a continuum of players that incorporate the time-dimension through a control-theoretic approach. Recently, simpler approaches relying on the Best Reply Strategy have been proposed. They assume that the agents…

最优化与控制 · 数学 2014-12-24 Pierre Degond , Michael Herty , Jian-Guo Liu

In a regular mean field game (MFG), the agents are assumed to be insignificant, they do not realize their effect on the population level and this may result in a phenomenon coined as the Tragedy of the Commons by the economists. However, in…

最优化与控制 · 数学 2024-09-13 Gokce Dayanikli , Mathieu Lauriere

Federated learning is a setting where agents, each with access to their own data source, combine models from local data to create a global model. If agents are drawing their data from different distributions, though, federated learning…

计算机科学与博弈论 · 计算机科学 2020-12-18 Kate Donahue , Jon Kleinberg

This paper presents a novel model-free method to solve linear quadratic Gaussian mean field social control problems in the presence of multiplicative noise. The objective is to achieve a social optimum by solving two algebraic Riccati…

最优化与控制 · 数学 2025-11-11 Zhenhui Xu , Bing-Chang Wang , Tielong Shen

This chapter examines monotonicity techniques in the theory of mean-field games(MFGs). Originally, monotonicity ideas were used to establish the uniqueness of solutions for MFGs. Later, monotonicity methods and monotone operators were…

偏微分方程分析 · 数学 2025-02-28 Rita Ferreira , Diogo Gomes , Teruo Tada

Mean field games (MFGs) describe the collective behavior of large populations of interacting agents. In this work, we tackle ill-posed inverse problems in potential MFGs, aiming to recover the agents' population, momentum, and environmental…

机器学习 · 计算机科学 2025-02-18 Jingguo Zhang , Xianjin Yang , Chenchen Mou , Chao Zhou

The mean-field framework has been used to find approximate solutions to problems involving very large populations of symmetric, anonymous agents, which may be intractable by other methods. The cooperative mean-field control (MFC) problem…

多智能体系统 · 计算机科学 2025-12-23 Patrick Benjamin , Alessandro Abate

We establish a connection between federated learning, a concept from machine learning, and mean-field games, a concept from game theory and control theory. In this analogy, the local federated learners are considered as the players and the…

机器学习 · 统计学 2021-07-09 Arash Mehrjou

In the present work, we study deterministic mean field games (MFGs) with finite time horizon in which the dynamics of a generic agent is controlled by the acceleration. They are described by a system of PDEs coupling a continuity equation…

偏微分方程分析 · 数学 2020-07-29 Yves Achdou , Paola Mannucci , Claudio Marchi , Nicoletta Tchou

We propose a discrete time graphon game formulation on continuous state and action spaces using a representative player to study stochastic games with heterogeneous interaction among agents. This formulation admits both philosophical and…

最优化与控制 · 数学 2024-06-07 Fuzhong Zhou , Chenyu Zhang , Xu Chen , Xuan Di

Estimating the unknown reward functions driving agents' behaviors is of central interest in inverse reinforcement learning and game theory. To tackle this problem, we develop a unified framework for reward function recovery in two-player…

机器学习 · 计算机科学 2026-05-20 Junyi Liao , Zihan Zhu , Ethan Fang , Zhuoran Yang , Vahid Tarokh

Mean Field Games (MFG) provide a theoretical frame to model socio-economic systems. In this letter, we study a particular class of MFG which shows strong analogies with the {\em non-linear Schr\"odinger and Gross-Pitaevski equations}…

物理与社会 · 物理学 2016-03-30 Igor Swiecicki , Thierry Gobron , Denis Ullmo

Concave Utility Reinforcement Learning (CURL) extends RL from linear to concave utilities in the occupancy measure induced by the agent's policy. This encompasses not only RL but also imitation learning and exploration, among others. Yet,…

Max-min fairness (MMF) is a widely known approach to a fair allocation of bandwidth to each of the users in a network. This allocation can be computed by uniformly raising the bandwidths of all users without violating capacity constraints.…

网络与互联网体系结构 · 计算机科学 2014-01-15 Tobias Harks , Martin Hoefer , Kevin Schewior , Alexander Skopalik

In this paper, we introduce a regularized mean-field game and study learning of this game under an infinite-horizon discounted reward function. Regularization is introduced by adding a strongly concave regularization function to the…

最优化与控制 · 数学 2022-11-11 Berkay Anahtarci , Can Deha Kariksiz , Naci Saldi
‹ 上一页 1 8 9 10 下一页 ›