中文
相关论文

相关论文: Two system transformation data-driven algorithms f…

200 篇论文

In this manuscript, we study a class of linear-quadratic (LQ) mean field control problems with a common noise and their corresponding $N$-particle systems. The mean field control problems considered are not standard LQ mean field control…

最优化与控制 · 数学 2024-12-02 Mengzhen Li , Chenchen Mou , Zhen Wu , Chao Zhou

This paper investigates an infinite-horizon linear quadratic stochastic (LQS) optimal control problem for a class of continuous-time stochastic systems. By employing the technique of adaptive dynamic programming (ADP), we propose a novel…

最优化与控制 · 数学 2022-10-11 Heng Zhang

This paper studies a class of linear quadratic mean field games where the coefficients of quadratic cost functions depend on both the mean and the variance of the population's state distribution through its quantile function. Such a…

最优化与控制 · 数学 2024-11-05 Shuang Gao , Roland P. Malhamé

We explore the use of transformers for solving quadratic programs and how this capability benefits decision-making problems that involve covariance matrices. We first show that the linear attention mechanism can provably solve unconstrained…

机器学习 · 计算机科学 2026-02-17 Kutay Tire , Yufan Zhang , Ege Onur Taga , Samet Oymak

A new algorithm for solving the solution of the linear-quadratic optimization problem (LQP) with unseparated boundary conditions in the continuous case is given. Using the properties of symmetry of the corresponding Hamiltonian matrix, the…

最优化与控制 · 数学 2019-04-16 Fikret Aliev , M. Mutallimov

This paper is concerned with a general linear quadratic (LQ) control problem of mean-field backward stochastic differential equation (BSDE). Here, the weighting matrices in the cost functional are allowed to be indefinite. Necessary and…

最优化与控制 · 数学 2024-12-31 Wencan Wang , Huanjun Zhang

A linear quadratic (LQ) stochastic optimization system involving large population, which is driven by forward-backward stochastic differential equation (FBSDE), is investigated in this paper. Agents cooperate with each other to minimize the…

最优化与控制 · 数学 2024-04-30 Guangchen Wang , Shujun Wang , Jie Xiong

We formulate a new class of two-person zero-sum differential games, in a stochastic setting, where a specification on a target terminal state distribution is imposed on the players. We address such added specification by introducing…

系统与控制 · 电气工程与系统科学 2019-09-13 Yongxin Chen , Tryphon T. Georgiou , Michele Pavon

We propose a new viewpoint on variational mean-field games with diffusion and quadratic Hamiltonian. We show the equivalence of such mean-field games with a relative entropy minimization at the level of probabilities on curves. We also…

最优化与控制 · 数学 2019-04-01 Jean-David Benamou , Guillaume Carlier , Simone Di Marino , Luca Nenna

Integrating data-driven techniques with mechanism-driven insights has recently gained popularity as a powerful learning approach to solving traditional LQR problems for designing intelligent controllers in complex dynamic systems. However,…

最优化与控制 · 数学 2025-12-10 Xiushan Jiang , Dong Wang , Weihai Zhang , Daniel W. C. Ho , Yuanqing Wu

We create classical (non-quantum) dynamic data structures supporting queries for recommender systems and least-squares regression that are comparable to their quantum analogues. De-quantizing such algorithms has received a flurry of…

数据结构与算法 · 计算机科学 2022-06-30 Nadiia Chepurko , Kenneth L. Clarkson , Lior Horesh , Honghao Lin , David P. Woodruff

In this work, we propose, for the first time, a reinforcement learning framework specifically designed for zero-sum linear-quadratic stochastic differential games. This approach offers a generalized solution for scenarios in which accurate…

最优化与控制 · 数学 2026-02-10 Yiyuan Wang

This paper is concerned with an overlapping information linear-quadratic (LQ) Stackelberg stochastic differential game with two leaders and two followers, where the diffusion terms of the state equation contain both the control and state…

最优化与控制 · 数学 2024-01-17 Yu Si , Jingtao Shi

This paper presents a sample-efficient, data-driven control framework for finite-horizon linear quadratic (LQ) control of linear time-varying (LTV) systems. In contrast to the time-invariant case, the time-varying LQ problem involves a…

系统与控制 · 电气工程与系统科学 2025-09-30 Sahel Vahedi Noori , Maryam Babazadeh

We consider a Mean Field Games model where the dynamics of the agents is given by a controlled Langevin equation and the cost is quadratic. A change of variables, introduced in [9], transforms the Mean Field Games system into a system of…

偏微分方程分析 · 数学 2021-01-12 Fabio Camilli

We introduce two algorithms based on a policy iteration method to numerically solve time-dependent Mean Field Game systems of partial differential equations with non-separable Hamiltonians. We prove the convergence of such algorithms in…

最优化与控制 · 数学 2022-10-03 Mathieu Laurière , Jiahao Song , Qing Tang

In this paper, an off-policy reinforcement learning algorithm is designed to solve the continuous-time LQR problem using only input-state data measured from the system. Different from other algorithms in the literature, we propose the use…

系统与控制 · 电气工程与系统科学 2023-04-03 Victor G. Lopez , Matthias A. Müller

We provide a thorough study of a general class of linear-quadratic extended mean field games and control problems in any dimensions where the mean field terms are allowed to be unbounded and there are also presence of cross terms in the…

最优化与控制 · 数学 2023-11-10 Alain Bensoussan , Bohan Li , Sheung Chi Phillip Yam

$ $This paper addresses the inverse problem for Linear-Quadratic (LQ) nonzero-sum $N$-player differential games, where the goal is to learn parameters of an unknown cost function for the game, called observed, given the demonstrated…

最优化与控制 · 数学 2024-10-28 Emin Martirosyan , Ming Cao

In this second part of our two-part paper, we invoke the stochastic maximum principle, conditional Hamiltonian and the coupled backward-forward stochastic differential equations of the first part [1] to derive team optimal decentralized…

最优化与控制 · 数学 2013-02-15 Charalambos D. Charalambous , Nasir U. Ahmed