中文
相关论文

相关论文: Deep Learning for Principal-Agent Mean Field Games

200 篇论文

Model-based algorithms -- algorithms that explore the environment through building and utilizing an estimated model -- are widely used in reinforcement learning practice and theoretically shown to achieve optimal sample efficiency for…

机器学习 · 计算机科学 2021-02-09 Qinghua Liu , Tiancheng Yu , Yu Bai , Chi Jin

In this paper, we consider a mean field game (MFG) with a major and $N$ minor agents. We first consider the limiting problem and allow the coefficients to vary with the conditional distribution in a nonlinear way. We use the stochastic…

最优化与控制 · 数学 2024-11-05 Ziyu Huang , Shanjian Tang

We investigate stochastic utility maximization games under relative performance concerns in both finite-agent and infinite-agent (graphon) settings. An incomplete market model is considered where agents with power (CRRA) utility functions…

最优化与控制 · 数学 2024-12-05 Zongxia Liang , Keyu Zhang , Yaqi Zhuang

We test the performance of deep deterministic policy gradient (DDPG), a deep reinforcement learning algorithm, able to handle continuous state and action spaces, to learn Nash equilibria in a setting where firms compete in prices. These…

计算机科学与博弈论 · 计算机科学 2025-09-30 Christoph Graf , Viktor Zobernig , Johannes Schmidt , Claude Klöckl

Prediction is a well-studied machine learning task, and prediction algorithms are core ingredients in online products and services. Despite their centrality in the competition between online companies who offer prediction-based products,…

计算机科学与博弈论 · 计算机科学 2019-05-09 Omer Ben-Porat , Moshe Tennenholtz

We initiate the study of how to perturb the reward in a zero-sum Markov game with two players to induce a desirable Nash equilibrium, namely arbitrating. Such a problem admits a bi-level optimization formulation. The lower level requires…

多智能体系统 · 计算机科学 2023-02-21 Jing Wang , Meichen Song , Feng Gao , Boyi Liu , Zhaoran Wang , Yi Wu

We study a class of dynamic decision problems of mean field type with time inconsistent cost functionals, and derive a stochastic maximum principle to characterize subgame perfect Nash equilibrium points. Subsequently, this approach is…

最优化与控制 · 数学 2014-03-26 Boualem Djehiche , Minyi Huang

Nash equilibrium is a key concept in game theory fundamental for elucidating the equilibrium state of strategic interactions, finding applications in diverse fields such as economics, political science, and biology. However, the Nash…

计算机科学与博弈论 · 计算机科学 2024-04-02 Elie Eshoa , Ali R. Zomorrodi

Self-trained autonomous agents developed using machine learning are showing great promise in a variety of control settings, perhaps most remarkably in applications involving autonomous vehicles. The main challenge associated with…

机器学习 · 计算机科学 2022-11-11 Patrik Hammersborg , Inga Strümke

This paper addresses the problem of distributed online generalized Nash equilibrium (GNE) learning for multi-cluster games with delayed feedback information. Specifically, each agent in the game is assumed to be informed a sequence of local…

最优化与控制 · 数学 2024-07-08 Bingqian Liu , Guanghui Wen , Xiao Fang , Tingwen Huang , Guanrong Chen

We study the problem of computing an approximate Nash equilibrium of continuous-action game without access to gradients. Such game access is common in reinforcement learning settings, where the environment is typically treated as a black…

计算机科学与博弈论 · 计算机科学 2023-08-30 Carlos Martin , Tuomas Sandholm

This paper studies the problem of Nash equilibrium approximation in large-scale heterogeneous mean-field games under communication and computation constraints. A deterministic mean-field game is considered in which the non-linear utility…

最优化与控制 · 数学 2017-09-20 Ehsan Nekouei , Tansu Alpcan , Girish Nair

Mean-field games arise in various fields including economics, engineering, and machine learning. They study strategic decision making in large populations where the individuals interact via certain mean-field quantities. The ground metrics…

最优化与控制 · 数学 2020-07-23 Lisang Ding , Wuchen Li , Stanley Osher , Wotao Yin

This paper presents a novel data-driven approach for approximating the $\varepsilon$-Nash equilibrium in continuous-time linear quadratic Gaussian (LQG) games, where multiple agents interact with each other through their dynamics and…

系统与控制 · 电气工程与系统科学 2025-07-22 Zhenhui Xu , Jiayu Chen , Bing-Chang Wang , Tielong Shen

We study $n$-agent Bayesian Games with $m$-dimensional vector types and linear payoffs, also called Linear Multidimensional Bayesian Games. This class of games is equivalent with $n$-agent, $m$-game Uniform Multigames. We distinguish…

计算机科学与博弈论 · 计算机科学 2023-10-24 Sébastien Huot , Abbas Edalat

Financial markets are often driven by latent factors which traders cannot observe. Here, we address an algorithmic trading problem with collections of heterogeneous agents who aim to perform optimal execution or statistical arbitrage, where…

数理金融 · 定量金融 2019-04-02 Philippe Casgrain , Sebastian Jaimungal

We study Nash equilibria learning of a general-sum stochastic game with an unknown transition probability density function. Agents take actions at the current environment state and their joint action influences the transition of the…

系统与控制 · 电气工程与系统科学 2022-10-19 Yan Chen , Tao Li

In this paper, we study a class of discrete-time mean-field games under the infinite-horizon risk-sensitive discounted-cost optimality criterion. Risk-sensitivity is introduced for each agent (player) via an exponential utility function. In…

最优化与控制 · 数学 2018-10-08 Naci Saldi , Tamer Basar , Maxim Raginsky

Generalized Nash Equilibrium Problems (GNEPs) arise in many applications, including non-cooperative multi-agent control problems. Although many methods exist for finding generalized Nash equilibria, most of them rely on assuming knowledge…

计算机科学与博弈论 · 计算机科学 2026-03-19 Pablo Krupa , Alberto Bemporad

Learning by experience in Multi-Agent Systems (MAS) is a difficult and exciting task, due to the lack of stationarity of the environment, whose dynamics evolves as the population learns. In order to design scalable algorithms for systems…

最优化与控制 · 数学 2020-02-24 Romuald Elie , Julien Pérolat , Mathieu Laurière , Matthieu Geist , Olivier Pietquin