中文
相关论文

相关论文: Linear-Quadratic Delayed Mean-Field Social Optimiz…

200 篇论文

This paper is concerned with a general non-homogeneous stochastic linear quadratic (LQ) control problem with regime switching and random coefficients. We obtain the explicit optimal state feedback control and optimal value for this problem…

最优化与控制 · 数学 2023-07-17 Ying Hu , Xiaomin Shi , Zuo Quan Xu

In this paper, we continue our study on a general time-inconsistent stochastic linear--quadratic (LQ) control problem originally formulated in [6]. We derive a necessary and sufficient condition for equilibrium controls via a flow of…

投资组合管理 · 定量金融 2015-05-27 Ying Hu , Hanqing Jin , Xun Yu Zhou

This paper studies a discrete-time stochastic control problem with linear quadratic criteria over an infinite-time horizon. We focus on a class of control systems whose system matrices are associated with random parameters involving unknown…

最优化与控制 · 数学 2022-01-17 Zhaorong Zhang , Juanjuan Xu , Xun Li

We consider optimal control of an unknown multi-agent linear quadratic (LQ) system where the dynamics and the cost are coupled across the agents through the mean-field (i.e., empirical mean) of the states and controls. Directly using…

系统与控制 · 电气工程与系统科学 2020-11-11 Mukul Gagrani , Sagar Sudhakara , Aditya Mahajan , Ashutosh Nayyar , Yi Ouyang

A network of noisy bistable elements with global time-delayed couplings is considered. A dichotomous mean field model has recently been developed describing the collective dynamics in such systems with uniform time delays near the…

统计力学 · 物理学 2007-05-23 Daniel Huber , Lev Tsimring

This paper considers linear-quadratic (LQ) stochastic leader-follower Stackelberg differential games for jump-diffusion systems with random coefficients. We first solve the LQ problem of the follower using the stochastic maximum principle…

最优化与控制 · 数学 2020-10-07 Jun Moon

This paper presents the numerical discretization methods of the continuous-time linear-quadratic optimal control problems (LQ-OCPs) with time delays. We describe the weight matrices of the LQ-OCPs as differential equations systems, allowing…

系统与控制 · 电气工程与系统科学 2024-04-15 Zhanhao Zhang , Steen Hørsholt , John Bagterp Jørgensen

Reinforcement learning (RL) has seen significant research and application results but often requires large amounts of training data. This paper proposes two data-efficient off-policy RL methods that use parametrized Q-learning. In these…

系统与控制 · 电气工程与系统科学 2025-04-09 J. S. van Hulst , W. P. M. H. Heemels , D. J. Antunes

We consider distributed optimization under communication constraints for training deep learning models. We propose a new algorithm, whose parameter updates rely on two forces: a regular gradient step, and a corrective direction dictated by…

机器学习 · 计算机科学 2022-04-29 Yunfei Teng , Wenbo Gao , Francois Chalus , Anna Choromanska , Donald Goldfarb , Adrian Weller

This paper focuses on the discrete-time backward stochastic linear quadratic (BSLQ) optimal control problem with nonhomogeneous system terms and cost function cross terms. The terminal constraint of such systems distinguishes it from…

最优化与控制 · 数学 2026-04-14 Hu Ligui , Meng Qingxin , Tang Maoning

We consider a class of stochastic control problems with a delayed control, both in drift and diffusion, of the type dX t = $\alpha$ t--d (bdt + $\sigma$dW t). We provide a new characterization of the solution in terms of a set of Riccati…

最优化与控制 · 数学 2021-02-25 William Lefebvre , Enzo Miller

We consider a class of dynamic collective choice models with social interactions, whereby a large number of non-uniform agents have to individually settle on one of multiple discrete alternative choices, with the relevance of their would-be…

系统与控制 · 计算机科学 2017-08-21 Rabih Salhab , Roland P. Malhamé , Jerome Le Ny

In this paper we are introducing a new reinforcement learning method for control problems in environments with delayed feedback. Specifically, our method employs stochastic planning, versus previous methods that used deterministic planning.…

机器学习 · 计算机科学 2024-02-02 Zhiyuan Yao , Ionut Florescu , Chihoon Lee

This paper studies the formation stabilization problem of asynchronous nonlinear multi-agent systems (MAS) subject to parametric uncertainties, external disturbances and bounded time-varying communication delays. A self-triggered min-max…

系统与控制 · 电气工程与系统科学 2022-02-22 Henglai Wei , Kunwu Zhang , Yang Shi

In this paper, we study the linear quadratic (LQ) optimal control problem of linear systems with private input and measurement information. The main challenging lies in the unavailability of other regulators' historical input information.…

最优化与控制 · 数学 2023-05-29 Juanjuan Xu , Huanshui Zhang

One of the most widely used methods for solving large-scale stochastic optimization problems is distributed asynchronous stochastic gradient descent (DASGD), a family of algorithms that result from parallelizing stochastic gradient descent…

最优化与控制 · 数学 2021-07-08 Zhengyuan Zhou , Panayotis Mertikopoulos , Nicholas Bambos , Peter W. Glynn , Yinyu Ye

In this paper, the problem of non-fragile finite-time stabilization for linear discrete mean-field stochastic systems is studied. The uncertain characteristics in control parameters are assumed to be random satisfying the Bernoulli…

最优化与控制 · 数学 2023-01-24 Tianliang Zhang , Feiqi Deng , Peng Shi

Distributed optimal control is known to be challenging and can become intractable even for linear-quadratic regulator problems. In this work, we study a special class of such problems where distributed state feedback controllers can give…

系统与控制 · 电气工程与系统科学 2024-03-14 Johan Olsson , Runyu Zhang , Emma Tegling , Na Li

This papers studies multi-agent (convex and \emph{nonconvex}) optimization over static digraphs. We propose a general distributed \emph{asynchronous} algorithmic framework whereby i) agents can update their local variables as well as…

最优化与控制 · 数学 2019-09-12 Ye Tian , Ying Sun , Gesualdo Scutari

Non-stationarity is a fundamental challenge in multi-agent reinforcement learning (MARL), where agents update their behaviour as they learn. Many theoretical advances in MARL avoid the challenge of non-stationarity by coordinating the…

计算机科学与博弈论 · 计算机科学 2025-03-19 Bora Yongacoglu , Gürdal Arslan , Serdar Yüksel
‹ 上一页 1 8 9 10 下一页 ›