中文
相关论文

相关论文: Reinforcement Learning Algorithm for Mixed Mean Fi…

200 篇论文

Mean-field reinforcement learning has become a popular theoretical framework for efficiently approximating large-scale multi-agent reinforcement learning (MARL) problems exhibiting symmetry. However, questions remain regarding the…

计算机科学与博弈论 · 计算机科学 2024-02-09 Batuhan Yardim , Artur Goldman , Niao He

In this tutorial, we provide an introduction to machine learning methods for finding Nash equilibria in games with large number of agents. These types of problems are important for the operations research community because of their…

最优化与控制 · 数学 2024-06-18 Gokce Dayanikli , Mathieu Lauriere

In this paper, we investigate a competitive market involving two agents who consider both their own wealth and the wealth gap with their opponent. Both agents can invest in a financial market consisting of a risk-free asset and a risky…

最优化与控制 · 数学 2025-02-10 Junyi Guo , Xia Han , Hao Wang , Kam Chuen Yuen

The Mean-Field approximation is a tractable approach for studying large population dynamics. However, its assumption on homogeneity and universal connections among all agents limits its applicability in many real-world scenarios.…

计算机科学与博弈论 · 计算机科学 2023-10-26 Peihan Huo , Oscar Peralta , Junyu Guo , Qiaomin Xie , Andreea Minca

Reinforcement learning for multi-agent games has attracted lots of attention recently. However, given the challenge of solving Nash equilibria for large population games, existing works with guaranteed polynomial complexities either focus…

最优化与控制 · 数学 2025-09-04 Anran Hu , Junzi Zhang

Mean Field Games (MFG) are the class of games with a very large number of agents and the standard equilibrium concept is a Mean Field Equilibrium (MFE). Algorithms for learning MFE in dynamic MFGs are unknown in general. Our focus is on an…

最优化与控制 · 数学 2021-02-02 Kiyeob Lee , Desik Rengarajan , Dileep Kalathil , Srinivas Shakkottai

We characterize zonal ancillary market coupling relying on noncooperative game theory. To that purpose, we formulate the ancillary market as a multi-leader single follower bilevel problem, that we subsequently cast as a generalized Nash…

多智能体系统 · 计算机科学 2025-11-13 Francesco Morri , Hélène Le Cadre , Pierre Gruet , Luce Brotcorne

In Multi-task learning (MTL), a joint model is trained to simultaneously make predictions for several tasks. Joint training reduces computation costs and improves data efficiency; however, since the gradients of these different tasks may…

机器学习 · 计算机科学 2022-07-11 Aviv Navon , Aviv Shamsian , Idan Achituve , Haggai Maron , Kenji Kawaguchi , Gal Chechik , Ethan Fetaya

Multi-agent reinforcement learning is a challenging and active field of research due to the inherent nonstationary property and coupling between agents. A popular approach to modeling the multi-agent interactions underlying the multi-agent…

多智能体系统 · 计算机科学 2025-10-07 Jushan Chen , Santiago Paternain

This thesis is going to give a gentle introduction to Mean Field Games. It aims to produce a coherent text beginning for simple notions of deterministic control theory progressively to current Mean Field Games theory. The framework…

最优化与控制 · 数学 2019-07-03 Athanasios Vasiliadis

We establish the convergence of the unified two-timescale Reinforcement Learning (RL) algorithm presented in a previous work by Angiuli et al. This algorithm provides solutions to Mean Field Game (MFG) or Mean Field Control (MFC) problems…

最优化与控制 · 数学 2024-05-02 Andrea Angiuli , Jean-Pierre Fouque , Mathieu Laurière , Mengrui Zhang

The designs of many large-scale systems today, from traffic routing environments to smart grids, rely on game-theoretic equilibrium concepts. However, as the size of an $N$-player game typically grows exponentially with $N$, standard game…

We study discrete-time, finite-state mean-field games (MFGs) under model uncertainty, where agents face ambiguity about the state transition probabilities. Each agent maximizes its expected payoff against the worst-case transitions within…

最优化与控制 · 数学 2026-01-21 Zongxia Liang , Zhou Zhou , Yaqi Zhuang , Bin Zou

We propose a policy iteration method to solve an inverse problem for a mean-field game (MFG) model, specifically to reconstruct the obstacle function in the game from the partial observation data of value functions, which represent the…

最优化与控制 · 数学 2026-02-12 Kui Ren , Nathan Soedjak , Shanyin Tong

In this work, we study an equilibrium-based continuous asset pricing problem which seeks to form a price process endogenously by requiring it to balance the flow of sales-and-purchase orders in the exchange market, where a large number of…

数理金融 · 定量金融 2021-09-28 Masaaki Fujii , Akihiko Takahashi

We investigate reinforcement learning in the setting of Markov decision processes for a large number of exchangeable agents interacting in a mean field manner. Applications include, for example, the control of a large number of robots…

最优化与控制 · 数学 2025-04-30 René Carmona , Mathieu Laurière , Zongjun Tan

The use of reinforcement learning algorithms in financial trading is becoming increasingly prevalent. However, the autonomous nature of these algorithms can lead to unexpected outcomes that deviate from traditional game-theoretical…

交易与市场微观结构 · 定量金融 2026-02-16 Fabrizio Lillo , Andrea Macrì

We present the development and analysis of a reinforcement learning (RL) algorithm designed to solve continuous-space mean field game (MFG) and mean field control (MFC) problems in a unified manner. The proposed approach pairs the…

最优化与控制 · 数学 2025-03-07 Andrea Angiuli , Jean-Pierre Fouque , Ruimeng Hu , Alan Raydan

Reinforcement learning algorithms, just like any other Machine learning algorithm pose a serious threat from adversaries. The adversaries can manipulate the learning algorithm resulting in non-optimal policies. In this paper, we analyze the…

机器学习 · 计算机科学 2021-03-12 Aqeel Anwar , Arijit Raychowdhury

A mean-field game (MFG) seeks the Nash Equilibrium of a game involving a continuum of players, where the Nash Equilibrium corresponds to a fixed point of the best-response mapping. However, simple fixed-point iterations do not always…

最优化与控制 · 数学 2025-07-15 Jiajia Yu , Xiuyuan Cheng , Jian-Guo Liu , Hongkai Zhao