中文
相关论文

相关论文: On Imitation in Mean-field Games

200 篇论文

The designs of many large-scale systems today, from traffic routing environments to smart grids, rely on game-theoretic equilibrium concepts. However, as the size of an $N$-player game typically grows exponentially with $N$, standard game…

In this tutorial, we provide an introduction to machine learning methods for finding Nash equilibria in games with large number of agents. These types of problems are important for the operations research community because of their…

最优化与控制 · 数学 2024-06-18 Gokce Dayanikli , Mathieu Lauriere

Many large-scale platforms and networked control systems have a centralized decision maker interacting with a massive population of agents under strict observability constraints. Motivated by such applications, we study a cooperative Markov…

多智能体系统 · 计算机科学 2026-05-12 Emile Anand , Ishani Karmarkar

Adversarial imitation learning (AIL) has become a popular alternative to supervised imitation learning that reduces the distribution shift suffered by the latter. However, AIL requires effective exploration during an online reinforcement…

机器学习 · 计算机科学 2023-10-16 Trevor Ablett , Bryan Chan , Jonathan Kelly

We study Nash equilibria for a sequence of symmetric $N$-player stochastic games of finite-fuel capacity expansion with singular controls and their mean-field game (MFG) counterpart. We construct a solution of the MFG via a simple iterative…

概率论 · 数学 2022-01-19 Luciano Campi , Tiziano De Angelis , Maddalena Ghio , Giulia Livieri

Multi-agent reinforcement learning (MARL), despite its popularity and empirical success, suffers from the curse of dimensionality. This paper builds the mathematical framework to approximate cooperative MARL by a mean-field control (MFC)…

机器学习 · 计算机科学 2021-10-04 Haotian Gu , Xin Guo , Xiaoli Wei , Renyuan Xu

Establishing the existence of Nash equilibria for partially observed stochastic dynamic games is known to be quite challenging, with the difficulties stemming from the noisy nature of the measurements available to individual players…

系统与控制 · 计算机科学 2018-06-06 Naci Saldi , Tamer Basar , Maxim Raginsky

In this paper, we consider a mean field game (MFG) model perturbed by small common noise. Our goal is to give an approximation of the Nash equilibrium strategy of this game using a solution from the original no common noise MFG whose…

概率论 · 数学 2017-07-31 Saran Ahuja , Weiluo Ren , Tzu-Wei Yang

In this paper, we consider a mean field game model inspired by crowd motion in which several interacting populations evolving in $\mathbb R^d$ aim at reaching given target sets in minimal time. The movement of each agent is described by a…

最优化与控制 · 数学 2022-08-16 Saeed Sadeghi Arjmand , Guilherme Mazanti

This paper studies a large population dynamic game involving nonlinear stochastic dynamical systems with agents of the following mixed types: (i) a major agent, and (ii) a population of $N$ minor agents where $N$ is very large. The major…

最优化与控制 · 数学 2013-06-07 Mojtaba Nourian , Peter E. Caines

In this paper, we introduce discrete-time linear mean-field games subject to an infinite-horizon discounted-cost optimality criterion. The state space of a generic agent is a compact Borel space. At every time, each agent is randomly…

系统与控制 · 电气工程与系统科学 2023-01-18 Naci Saldi

We present a simulation-based approach for solution of mean field games (MFGs), using the framework of empirical game-theoretical analysis (EGTA). Our primary method employs a version of the double oracle, iteratively adding strategies…

多智能体系统 · 计算机科学 2023-02-14 Yongzhao Wang , Michael P. Wellman

We study discrete-time mean-field Markov games with infinite numbers of agents where each agent aims to minimize its ergodic cost. We consider the setting where the agents have identical linear state transitions and quadratic cost…

最优化与控制 · 数学 2019-10-17 Zuyue Fu , Zhuoran Yang , Yongxin Chen , Zhaoran Wang

Generative Adversarial Imitation Learning (GAIL) is a powerful and practical approach for learning sequential decision-making policies. Different from Reinforcement Learning (RL), GAIL takes advantage of demonstration data by experts (e.g.,…

机器学习 · 计算机科学 2020-01-14 Minshuo Chen , Yizhou Wang , Tianyi Liu , Zhuoran Yang , Xingguo Li , Zhaoran Wang , Tuo Zhao

In recent years, the development of robotics and artificial intelligence (AI) systems has been nothing short of remarkable. As these systems continue to evolve, they are being utilized in increasingly complex and unstructured environments,…

机器学习 · 计算机科学 2024-10-28 Maryam Zare , Parham M. Kebria , Abbas Khosravi , Saeid Nahavandi

Learning the behavior of large agent populations is an important task for numerous research areas. Although the field of multi-agent reinforcement learning (MARL) has made significant progress towards solving these systems, solutions for…

多智能体系统 · 计算机科学 2024-02-26 Christian Fabian , Kai Cui , Heinz Koeppl

Learning to imitate expert behavior from demonstrations can be challenging, especially in environments with high-dimensional, continuous observations and unknown dynamics. Supervised learning methods based on behavioral cloning (BC) suffer…

机器学习 · 计算机科学 2019-09-27 Siddharth Reddy , Anca D. Dragan , Sergey Levine

In this book, we present a curated collection of existing results on inverse problems for Mean Field Games (MFGs), a cutting-edge and rapidly evolving field of research. Our aim is to provide fresh insights, novel perspectives, and a…

偏微分方程分析 · 数学 2025-03-20 Hongyu Liu , Catharine W. K. Lo , Shen Zhang

Mean field games (MFGs) tractably model behavior in large agent populations. The literature on learning MFG equilibria typically focuses on finding Nash equilibria (NE), which assume perfectly rational agents and are hence implausible in…

计算机科学与博弈论 · 计算机科学 2025-01-31 Yannick Eich , Christian Fabian , Kai Cui , Heinz Koeppl

Imitation Learning (IL) is one of the most widely used methods in machine learning. Yet, many works find it is often unable to fully recover the underlying expert behavior, even in constrained environments like single-agent games. However,…

机器学习 · 计算机科学 2024-12-20 Jens Tuyls , Dhruv Madeka , Kari Torkkola , Dean Foster , Karthik Narasimhan , Sham Kakade