中文
相关论文

相关论文: Learning the Linear Quadratic Regulator from Nonli…

200 篇论文

We study reinforcement learning (RL) for a class of continuous-time linear-quadratic (LQ) control problems for diffusions, where states are scalar-valued and running control rewards are absent but volatilities of the state processes depend…

机器学习 · 计算机科学 2025-07-25 Yilie Huang , Yanwei Jia , Xun Yu Zhou

``Sim2real gap", in which the system learned in simulations is not the exact representation of the real system, can lead to loss of stability and performance when controllers learned using data from the simulated system are used on the real…

系统与控制 · 电气工程与系统科学 2025-05-15 Shivam Bajaj , Prateek Jaiswal , Vijay Gupta

An optimal control law for networked control systems with a discrete-time linear time-invariant (LTI) system as plant and networks between sensor and controller as well as between controller and actuator is proposed. This controller is…

系统与控制 · 电气工程与系统科学 2021-07-09 Marijan Palmisano , Martin Steinberger , Martin Horn

We study the Linear-Quadratic optimal control problem for a general class of infinite-dimensional passive systems, allowing for unbounded input and output operators. We show that under mild assumptions, the finite cost condition is always…

最优化与控制 · 数学 2025-06-05 Anthony Hastir , Birgit Jacob

Policy gradient algorithms are widely used in reinforcement learning and belong to the class of approximate dynamic programming methods. This paper studies two key policy gradient algorithms, the Natural Policy Gradient and the Gauss-Newton…

系统与控制 · 电气工程与系统科学 2026-05-11 Bowen Song , Sebastien Gros , Andrea Iannelli

Unlike traditional model-based reinforcement learning approaches that estimate system parameters from data, non-model-based data-driven control learns the optimal policy directly from input-state data without any intermediate model…

最优化与控制 · 数学 2026-05-05 Leilei Cui , Zhong-Ping Jiang , Petter N. Kolm , Grégoire G. Macqueron

Current research suggests the use of a liner quadratic performance index for optimal control of regulators in various applications. Some examples include correcting the trajectory of rocket and air vehicles, vibration suppression of…

综合数学 · 数学 2007-05-23 Alexander Bolonkin , Robert Sierakowski

In this paper we provide direct data-driven expressions for the Linear Quadratic Regulator (LQR), the Kalman filter, and the Linear Quadratic Gaussian (LQG) controller using a finite dataset of noisy input, state, and output trajectories.…

最优化与控制 · 数学 2023-09-21 Abed AlRahman Al Makdah , Fabio Pasqualetti

Direct policy search has achieved great empirical success in reinforcement learning. Many recent studies have revisited its theoretical foundation for continuous control, which reveals elegant nonconvex geometry in various benchmark…

最优化与控制 · 数学 2023-12-27 Yang Zheng , Chih-fan Pai , Yujie Tang

This paper studies distributed Q-learning for Linear Quadratic Regulator (LQR) in a multi-agent network. The existing results often assume that agents can observe the global system state, which may be infeasible in large-scale systems due…

多智能体系统 · 计算机科学 2020-12-24 Hang Wang , Sen Lin , Hamid Jafarkhani , Junshan Zhang

This paper addresses the end-to-end sample complexity bound for learning the H2 optimal controller (the Linear Quadratic Gaussian (LQG) problem) with unknown dynamics, for potentially unstable Linear Time Invariant (LTI) systems. The robust…

系统与控制 · 电气工程与系统科学 2022-07-05 Yifei Zhang , Sourav Kumar Ukil , Ephraim Neimand , Serban Sabau , Myron E. Hohil

Individual agents in a multi-agent system (MAS) may have decoupled open-loop dynamics, but a cooperative control objective usually results in coupled closed-loop dynamics thereby making the control design computationally expensive. The…

系统与控制 · 电气工程与系统科学 2021-03-09 Gangshan Jing , He Bai , Jemin George , Aranya Chakrabortty

We address the problem of model-free distributed stabilization of heterogeneous multi-agent systems using reinforcement learning (RL). Two algorithms are developed. The first algorithm solves a centralized linear quadratic regulator (LQR)…

系统与控制 · 电气工程与系统科学 2021-03-09 Gangshan Jing , He Bai , Jemin George , Aranya Chakrabortty , Piyush K. Sharma

This paper investigates the performance of Newton's method, iterative Linear Quadratic Regulator (iLQR), and Differential Dynamic Programming (DDP) in solving discrete-time optimal control problems. We offer a unified perspective on these…

最优化与控制 · 数学 2026-05-26 Abhijeet , Suman Chakravorty

This paper delves into designing stabilizing feedback control gains for continuous linear systems with unknown state matrix, in which the control is subject to a general structural constraint. We bring forth the ideas from reinforcement…

系统与控制 · 电气工程与系统科学 2025-11-11 Sayak Mukherjee , Thanh Long Vu

We consider the problem of controlling a linear dynamical system from bilinear observations with minimal quadratic cost. Despite the similarity of this problem to standard linear quadratic Gaussian (LQG) control, we show that when the…

最优化与控制 · 数学 2025-10-23 Yahya Sattar , Sunmook Choi , Yassir Jedra , Maryam Fazel , Sarah Dean

It is well-known that linear quadratic regulators (LQR) enjoy guaranteed stability margins, whereas linear quadratic Gaussian regulators (LQG) do not. In this letter, we consider systems and compensators defined over directed acyclic…

系统与控制 · 电气工程与系统科学 2023-05-29 Mruganka Kashyap , Laurent Lessard

This paper focuses on the linear quadratic control (LQC) design of systems corrupted by both stochastic noise and bounded noise simultaneously. When only of these noises are considered, the LQC strategy leads to stochastic or robust…

最优化与控制 · 数学 2025-12-15 Xuehui Ma , Shiliang Zhang , Xiaohui Zhang , Jing Xin , Hector Garcia de Marina

This paper is concerned with the linear quadratic (LQ) optimal control of continuous-time system with terminal state constraint. In particular, multiple agents exist in the system which can only access partial information of the matrix…

最优化与控制 · 数学 2025-10-21 Wenjing Yang , Zhaorong Zhang , Juanjuan Xu

We study the exploration problem in episodic MDPs with rich observations generated from a small number of latent states. Under certain identifiability assumptions, we demonstrate how to estimate a mapping from the observations to latent…

机器学习 · 计算机科学 2021-09-10 Simon S. Du , Akshay Krishnamurthy , Nan Jiang , Alekh Agarwal , Miroslav Dudík , John Langford