中文
相关论文

相关论文: Mean Field LQG Social Optimization: A Reinforcemen…

200 篇论文

State-space models provide an important body of techniques for analyzing time-series, but their use requires estimating unobserved states. The optimal estimate of the state is its conditional expectation given the observation histories, and…

We study the linear-quadratic control problem for a class of non-exchangeable mean-field systems, which model large populations of heterogeneous interacting agents. We explicitly characterize the optimal control in terms of a new…

最优化与控制 · 数学 2025-12-30 Anna de Crescenzo , Filippo de Feo , Huyên Pham

This paper investigates the so-called reward-balancing methods, a novel class of algorithms for solving discounted-return reinforcement learning (RL) problems. These methods consist of iteratively adjusting the reward function to transform…

最优化与控制 · 数学 2026-04-23 Simone Baroncini , Bahman Gharesifard , Giuseppe Notarstefano

In this paper we discuss a class of mean field linear-quadratic-Gaussian (LQG) games for large population system which has never been addressed by existing literature. The features of our works are sketched as follows. First of all, our…

概率论 · 数学 2013-08-09 Jianhui Huang , Xun Li , Tianxiao Wang

This paper studies open-loop and feedback solutions to leader-follower mean field linear-quadratic-Gaussian games with multiplicative noise by the direct approach. The leader-follower game involves a leader and many followers, where the…

最优化与控制 · 数学 2025-12-04 Bing-Chang Wang , Huanshui Zhang , Ji-Feng Zhang

We present a simulation-based approach for solution of mean field games (MFGs), using the framework of empirical game-theoretical analysis (EGTA). Our primary method employs a version of the double oracle, iteratively adding strategies…

多智能体系统 · 计算机科学 2023-02-14 Yongzhao Wang , Michael P. Wellman

The recent mean field game (MFG) formalism facilitates otherwise intractable computation of approximate Nash equilibria in many-agent settings. In this paper, we consider discrete-time finite MFGs subject to finite-horizon objectives. We…

多智能体系统 · 计算机科学 2022-07-11 Kai Cui , Heinz Koeppl

Mean field control (MFC) problems have been introduced to study social optima in very large populations of strategic agents. The main idea is to consider an infinite population and to simplify the analysis by using a mean field…

最优化与控制 · 数学 2023-03-01 Sebastian Baudelet , Brieuc Frénais , Mathieu Laurière , Amal Machtalay , Yuchen Zhu

Variational inference with a factorized Gaussian posterior estimate is a widely used approach for learning parameters and hidden variables. Empirically, a regularizing effect can be observed that is poorly understood. In this work, we show…

机器学习 · 计算机科学 2019-09-04 Julius Kunze , Louis Kirsch , Hippolyt Ritter , David Barber

In a regular mean field game (MFG), the agents are assumed to be insignificant, they do not realize their effect on the population level and this may result in a phenomenon coined as the Tragedy of the Commons by the economists. However, in…

最优化与控制 · 数学 2024-09-13 Gokce Dayanikli , Mathieu Lauriere

This paper studies a class of partial information linear-quadratic mean-field game problems. A general stochastic large-population system is considered, where the diffusion term of the dynamic of each agent can depend on the state and…

最优化与控制 · 数学 2022-03-22 Min Li , Tianyang Nie , Zhen Wu

While the Matrix Generalized Inverse Gaussian ($\mathcal{MGIG}$) distribution arises naturally in some settings as a distribution over symmetric positive semi-definite matrices, certain key properties of the distribution and effective ways…

机器学习 · 统计学 2016-08-23 Farideh Fazayeli , Arindam Banerjee

Many machine learning problems optimize an objective that must be measured with noise. The primary method is a first order stochastic gradient descent using one or more Monte Carlo (MC) samples at each step. There are settings where…

机器学习 · 计算机科学 2021-04-22 Sifan Liu , Art B. Owen

This paper considers a class of reinforcement learning problems, which involve systems with two types of states: stochastic and pseudo-stochastic. In such systems, stochastic states follow a stochastic transition kernel while the…

机器学习 · 计算机科学 2023-11-09 Honghao Wei , Xin Liu , Weina Wang , Lei Ying

Traditional control theory-based methods require tailored engineering for each system and constant fine-tuning. In power plant control, one often needs to obtain a precise representation of the system dynamics and carefully design the…

系统与控制 · 电气工程与系统科学 2024-09-21 Yixuan Sun , Sami Khairy , Richard B. Vilim , Rui Hu , Akshay J. Dave

We consider the reinforcement learning (RL) problem with general utilities which consists in maximizing a function of the state-action occupancy measure. Beyond the standard cumulative reward RL setting, this problem includes as particular…

机器学习 · 计算机科学 2023-06-06 Anas Barakat , Ilyas Fatkhullin , Niao He

We develop a convex analysis approach for solving LQG optimal control problems and apply it to major-minor (MM) LQG mean-field game (MFG) systems. The approach retrieves the best response strategies for the major agent and all minor agents…

系统与控制 · 计算机科学 2020-06-15 Dena Firoozi , Sebastian Jaimungal , Peter E. Caines

Reinforcement learning (RL) is a classical tool to solve network control or policy optimization problems in unknown environments. The original Q-learning suffers from performance and complexity challenges across very large networks. Herein,…

机器学习 · 计算机科学 2024-09-02 Talha Bozkus , Urbashi Mitra

Many large-scale platforms and networked control systems have a centralized decision maker interacting with a massive population of agents under strict observability constraints. Motivated by such applications, we study a cooperative Markov…

多智能体系统 · 计算机科学 2026-05-12 Emile Anand , Ishani Karmarkar

This paper introduces a novel data-driven approach to design a linear quadratic regulator (LQR) using a reinforcement learning (RL) algorithm that does not require a system model. The key contribution is to perform policy iteration (PI) by…

系统与控制 · 电气工程与系统科学 2023-11-20 Soroush Asri , Luis Rodrigues