English
Related papers

Related papers: Data-Driven LQR with Finite-Time Experiments via E…

200 papers

This work theoretically studies a ubiquitous reinforcement learning policy for controlling the canonical model of continuous-time stochastic linear-quadratic systems. We show that randomized certainty equivalent policy addresses the…

Machine Learning · Computer Science 2022-08-23 Mohamad Kazem Shirani Faradonbeh

The output regulation problem for unknown linear systems has been studied using state-based and output-based internal model approaches in the special case with no disturbances. This paper further investigates the output regulation problem…

Optimization and Control · Mathematics 2026-01-07 Haoyan Lin , Jie Huang

We consider the linear quadratic regulator (LQR) for one-dimensional linear evolution partial differential equations (PDEs) on a finite interval in space. The control is applied as an additive forcing term to PDEs. Existing methods for…

Systems and Control · Electrical Eng. & Systems 2025-05-26 Zhexian Li , Athanassios S. Fokas , Ketan Savla

This paper proposes a novel lifting method which converts the standard discrete-time linear periodic system to an augmented linear time-invariant system. The linear quadratic optimal control is then based on the solution of the…

Optimization and Control · Mathematics 2018-06-21 Yaguang Yang

This work uses the entropy-regularised relaxed stochastic control perspective as a principled framework for designing reinforcement learning (RL) algorithms. Herein agent interacts with the environment by generating noisy controls…

Machine Learning · Computer Science 2023-09-18 Lukasz Szpruch , Tanut Treetanthiploet , Yufei Zhang

Linear temporal logic (LTL) offers a simplified way of specifying tasks for policy optimization that may otherwise be difficult to describe with scalar reward functions. However, the standard RL framework can be too myopic to find maximally…

Machine Learning · Computer Science 2023-03-06 Cameron Voloshin , Abhinav Verma , Yisong Yue

We consider the linear quadratic Gaussian control problem with a discounted cost functional for descriptor systems on the infinite time horizon. Based on recent results from the deterministic framework, we characterize the feasibility of…

Optimization and Control · Mathematics 2020-04-21 Hermann Mena , Lena-Maria Pfurtscheller , Matthias Voigt

With the outstanding performance of policy gradient (PG) method in the reinforcement learning field, the convergence theory of it has aroused more and more interest recently. Meanwhile, the significant importance and abundant theoretical…

Optimization and Control · Mathematics 2024-04-19 Xinpei Zhang , Guangyan Jia

Policy gradient methods are a powerful family of reinforcement learning algorithms for continuous control that optimize a policy directly. However, standard first-order methods often converge slowly. Second-order methods can accelerate…

Systems and Control · Electrical Eng. & Systems 2025-11-05 Amirreza Valaei , Arash Bahari Kordabad , Sadegh Soudjani

In this technical note we analyse the performance improvement and optimality properties of the Learning Model Predictive Control (LMPC) strategy for linear deterministic systems. The LMPC framework is a policy iteration scheme where…

Optimization and Control · Mathematics 2022-02-02 Ugo Rosolia , Yingzhao Lian , Emilio T. Maddalena , Giancarlo Ferrari-Trecate , Colin N. Jones

This paper is concerned with a backward stochastic linear-quadratic (LQ, for short) optimal control problem with deterministic coefficients. The weighting matrices are allowed to be indefinite, and cross-product terms in the control and…

Optimization and Control · Mathematics 2021-04-13 Jingrui Sun , Zhen Wu , Jie Xiong

This manuscript surveys reinforcement learning from the perspective of optimization and control with a focus on continuous control applications. It surveys the general formulation, terminology, and typical experimental implementations of…

Optimization and Control · Mathematics 2018-11-13 Benjamin Recht

In this paper we provide direct data-driven expressions for the Linear Quadratic Regulator (LQR), the Kalman filter, and the Linear Quadratic Gaussian (LQG) controller using a finite dataset of noisy input, state, and output trajectories.…

Optimization and Control · Mathematics 2023-09-21 Abed AlRahman Al Makdah , Fabio Pasqualetti

This article explores the discrete-time stochastic optimal LQR control with delay and quadratic constraints. The inclusion of delay, compared to delay-free optimal LQR control with quadratic constraints, significantly increases the…

Optimization and Control · Mathematics 2024-11-19 Dawei Liu , Juanjuan Xu , huanshui Zhang

Many applications -- including power systems, robotics, and economics -- involve a dynamical system interacting with a stochastic and hard-to-model environment. We adopt a reinforcement learning approach to control such systems.…

Optimization and Control · Mathematics 2025-08-26 Abed AlRahman Al Makdah , Oliver Kosut , Lalitha Sankar , Shaofeng Zou

The data-driven linear quadratic regulator (ddLQR) is a widely studied control method for unknown dynamical systems with disturbance. Existing approaches, both indirect, i.e., those that identify a model followed by model-based design, and…

Optimization and Control · Mathematics 2026-04-13 Thierry Schwaller , Feiran Zhao , Florian Dörfler

An online policy learning problem of linear control systems is studied. In this problem, the control system is known and linear, and a sequence of quadratic cost functions is revealed to the controller in hindsight, and the controller…

Optimization and Control · Mathematics 2021-01-27 Mohammad Akbari , Bahman Gharesifard , Tamas Linder

Linear-Quadratic optimal controls are computed for a class of boundary controlled, boundary observed hyperbolic infinite-dimensional systems, which may be viewed as networks of waves. The main results of this manuscript consist in…

Optimization and Control · Mathematics 2025-02-06 Anthony Hastir , Birgit Jacob , Hans Zwart

This letter presents a robust data-driven receding-horizon control framework for the discrete time linear quadratic regulator (LQR) with input constraints. Unlike existing data-driven approaches that design a controller from initial data…

Optimization and Control · Mathematics 2025-10-08 Jian Zheng , Mario Sznaier

We propose controller synthesis for state regulation problems in which a human operator shares control with an autonomy system, running in parallel. The autonomy system continuously improves over human action, with minimal intervention, and…

Systems and Control · Computer Science 2019-09-23 Murad Abu-Khalaf , Sertac Karaman , Daniela Rus