中文
相关论文

相关论文: LQG Risk-Sensitive Single-Agent and Major-Minor Me…

200 篇论文

Many large-scale platforms and networked control systems have a centralized decision maker interacting with a massive population of agents under strict observability constraints. Motivated by such applications, we study a cooperative Markov…

多智能体系统 · 计算机科学 2026-05-12 Emile Anand , Ishani Karmarkar

In single-agent Markov decision processes, an agent can optimize its policy based on the interaction with environment. In multi-player Markov games (MGs), however, the interaction is non-stationary due to the behaviors of other players, so…

计算机科学与博弈论 · 计算机科学 2021-10-19 Yuanheng Zhu , Dongbin Zhao , Mengchen Zhao , Dong Li

This thesis is going to give a gentle introduction to Mean Field Games. It aims to produce a coherent text beginning for simple notions of deterministic control theory progressively to current Mean Field Games theory. The framework…

最优化与控制 · 数学 2019-07-03 Athanasios Vasiliadis

In this work, we revisit the Linear Quadratic Gaussian (LQG) optimal control problem from a behavioral perspective. Motivated by the suitability of behavioral models for data-driven control, we begin with a reformulation of the LQG problem…

系统与控制 · 电气工程与系统科学 2022-09-20 Abed AlRahman Al Makdah , Vishaal Krishnan , Vaibhav Katewa , Fabio Pasqualetti

Control of multi-agent systems via game theory is investigated. Assume a system level object is given, the utility functions for individual agents are designed to convert a multi-agent system into a potential game. First, for fixed…

最优化与控制 · 数学 2016-08-02 Ting Liu , Jinhuan Wang , Daizhan Cheng

In this paper we study reduction by symmetry for optimality conditions in optimal control problems of left-invariant affine multi-agent control systems, with partial symmetry breaking cost functions. Our approach emphasizes the role of…

最优化与控制 · 数学 2022-04-14 Efstratios Stratoglou , Leonardo Colombo , Tomoki Ohsawa

Motivated by the self-pursuit of controlled objects, we consider the exact controllability of a linear mean-field type game-based control system (MF-GBCS, for short) generated by a linear-quadratic (LQ, for short) Nash game. A Gram-type…

最优化与控制 · 数学 2023-03-01 Cui Chen , Zhiyong Yu

Nash equilibria provide a principled framework for modeling interactions in multi-agent decision-making and control. However, many equilibrium-seeking methods implicitly assume that each agent has access to the other agents' objectives and…

计算机科学与博弈论 · 计算机科学 2026-03-19 Mahdis Rabbani , Navid Mojahed , Shima Nazari

Mean field games (MFGs) tractably model behavior in large agent populations. The literature on learning MFG equilibria typically focuses on finding Nash equilibria (NE), which assume perfectly rational agents and are hence implausible in…

计算机科学与博弈论 · 计算机科学 2025-01-31 Yannick Eich , Christian Fabian , Kai Cui , Heinz Koeppl

We propose and investigate a discrete-time mean field game model involving risk-averse agents. The model under study is a coupled system of dynamic programming equations with a Kolmogorov equation. The agents' risk aversion is modeled by…

最优化与控制 · 数学 2020-12-29 J. Frédéric Bonnans , Pierre Lavigne , Laurent Pfeiffer

Mean field games are studied by means of the weak formulation of stochastic optimal control. This approach allows the mean field interactions to enter through both state and control processes and take a form which is general enough to…

概率论 · 数学 2015-04-09 Rene Carmona , Daniel Lacker

Large language models (LLMs) demonstrate strong reasoning abilities across mathematical, strategic, and linguistic tasks, yet little is known about how well they reason in dynamic, real-time, multi-agent scenarios, such as collaborative…

多智能体系统 · 计算机科学 2026-01-01 Shaurya Mallampati , Rashed Shelim , Walid Saad , Naren Ramakrishnan

Existing multi-agent reinforcement learning methods are limited typically to a small number of agents. When the agent number increases largely, the learning becomes intractable due to the curse of the dimensionality and the exponential…

多智能体系统 · 计算机科学 2020-12-16 Yaodong Yang , Rui Luo , Minne Li , Ming Zhou , Weinan Zhang , Jun Wang

We examine global non-asymptotic convergence properties of policy gradient methods for multi-agent reinforcement learning (RL) problems in Markov potential games (MPG). To learn a Nash equilibrium of an MPG in which the size of state space…

机器学习 · 计算机科学 2022-08-08 Dongsheng Ding , Chen-Yu Wei , Kaiqing Zhang , Mihailo R. Jovanović

We introduce a generic solver for dynamic portfolio allocation problems when the market exhibits return predictability, price impact and partial observability. We assume that the price modeling can be encoded into a linear state-space and…

投资组合管理 · 定量金融 2016-11-07 M. Abeille , E. Serie , A. Lazaric , X. Brokmann

Trust region methods are widely applied in single-agent reinforcement learning problems due to their monotonic performance-improvement guarantee at every iteration. Nonetheless, when applied in multi-agent settings, the guarantee of trust…

多智能体系统 · 计算机科学 2021-06-15 Ying Wen , Hui Chen , Yaodong Yang , Zheng Tian , Minne Li , Xu Chen , Jun Wang

We study the policy gradient method (PGM) for the linear quadratic Gaussian (LQG) dynamic output-feedback control problem using an input-output-history (IOH) representation of the closed-loop system. First, we show that any dynamic…

最优化与控制 · 数学 2025-10-23 Tomonori Sadamoto , Takashi Tanaka

Multiagent systems where agents interact among themselves and with a stochastic environment can be formalized as stochastic games. We study a subclass named Markov potential games (MPGs) that appear often in economic and engineering…

多智能体系统 · 计算机科学 2018-05-23 Sergio Valcarcel Macua , Javier Zazo , Santiago Zazo

In this paper, a robust nonlinear control scheme is proposed for a nonlinear multi-input multi-output (MIMO) system subject to bounded time varying uncertainty which satisfies a certain integral quadratic constraint condition. The scheme…

系统与控制 · 计算机科学 2016-08-14 Obaid Ur Rehman , Ian R. Petersen , Barış Fidan

In this tutorial, we provide an introduction to machine learning methods for finding Nash equilibria in games with large number of agents. These types of problems are important for the operations research community because of their…

最优化与控制 · 数学 2024-06-18 Gokce Dayanikli , Mathieu Lauriere