English
Related papers

Related papers: Single Time-scale Actor-critic Method to Solve the…

200 papers

An optimal control law for networked control systems with a discrete-time linear time-invariant (LTI) system as plant and networks between sensor and controller as well as between controller and actuator is proposed. This controller is…

Systems and Control · Electrical Eng. & Systems 2021-07-09 Marijan Palmisano , Martin Steinberger , Martin Horn

We present a continuous-time equivalent to the well-known iterative linear-quadratic algorithm including an implementation of a backtracking line-search policy and a novel regularization approach based on the necessary conditions in the…

Systems and Control · Electrical Eng. & Systems 2025-05-22 Juraj Lieskovský , Jaroslav Bušek , Tomáš Vyhlídal

``Sim2real gap", in which the system learned in simulations is not the exact representation of the real system, can lead to loss of stability and performance when controllers learned using data from the simulated system are used on the real…

Systems and Control · Electrical Eng. & Systems 2025-05-15 Shivam Bajaj , Prateek Jaiswal , Vijay Gupta

In this report, linear quadratic regulator is used to design adaptive cruise control system. In the regulator, Q and R parameters vary with time according to current traffic situations. Phase-plant method is used to give constraints on Q…

Systems and Control · Electrical Eng. & Systems 2020-08-06 Yuncheng Jiang

This study focuses on the numerical discretization methods for the continuous-time discounted linear-quadratic optimal control problem (LQ-OCP) with time delays. By assuming piecewise constant inputs, we formulate the discrete system…

Optimization and Control · Mathematics 2024-07-29 Zhanhao Zhang , Steen Hørsholt , John Bagterp Jørgensen

For various typical cases and situations where the formulation results in an optimal control problem, the Linear Quadratic Regulator (LQR) approach and its variants continue to be highly attractive. In certain scenarios, it can happen that…

Optimization and Control · Mathematics 2023-02-14 Jun Ma , Zilong Cheng , Xiaocong Li , Wenxin Wang , Masayoshi Tomizuka , Tong Heng Lee

As an important type of reinforcement learning algorithms, actor-critic (AC) and natural actor-critic (NAC) algorithms are often executed in two ways for finding optimal policies. In the first nested-loop design, actor's one update of…

Machine Learning · Computer Science 2020-05-11 Tengyu Xu , Zhe Wang , Yingbin Liang

We consider the Linear-Quadratic-Regulator (LQR) problem in terms of optimizing a real-valued matrix function over the set of feedback gains. Such a setup facilitates examining the implications of a natural initial-state independent…

Systems and Control · Electrical Eng. & Systems 2019-07-31 Jingjing Bu , Afshin Mesbahi , Maryam Fazel , Mehran Mesbahi

For linear time-invariant (LTI) systems, the design of an optimal controller is a commonly encountered problem in many applications. Among all the optimization approaches available, the linear quadratic regulator (LQR) methodology certainly…

Optimization and Control · Mathematics 2022-03-29 Zilong Cheng , Jun Ma , Xiaocong Li , Masayoshi Tomizuka , Tong Heng Lee

Policy optimization has drawn increasing attention in reinforcement learning, particularly in the context of derivative-free methods for linear quadratic regulator (LQR) problems with unknown dynamics. This paper focuses on characterizing…

Optimization and Control · Mathematics 2025-06-17 Weijian Li , Panagiotis Kounatidis , Zhong-Ping Jiang , Andreas A. Malikopoulos

We revisit in this paper the discrete-time linear quadratic regulator (LQR) problem from the perspective of receding-horizon policy gradient (RHPG), a newly developed model-free learning framework for control applications. We provide a…

Optimization and Control · Mathematics 2024-02-02 Xiangyuan Zhang , Tamer Başar

We study a structured bi-level optimization problem where the upper-level objective is a smooth function and the lower-level problem is policy optimization in a Markov decision process (MDP). The upper-level decision variable parameterizes…

Machine Learning · Computer Science 2026-04-23 Sihan Zeng , Sujay Bhatt , Sumitra Ganesh , Alec Koppel

In this paper, the solvability of discrete-time stochastic linear-quadratic (LQ) optimal control problem in finite horizon is considered. Firstly, it shows that the closed-loop solvability for the LQ control problem is optimal if and only…

Optimization and Control · Mathematics 2025-02-25 Yue Sun , Xianping Wu , Xun Li

This work presents an algorithmic scheme for solving the infinite-time constrained linear quadratic regulation problem. We employ an accelerated version of a popular proximal gradient scheme, commonly known as the Forward-Backward Splitting…

Optimization and Control · Mathematics 2015-01-20 Giorgos Stathopoulos , Milan Korda , Colin N. Jones

While there has been substantial success for solving continuous control with actor-critic methods, simpler critic-only methods such as Q-learning find limited application in the associated high-dimensional action spaces. However, most…

Machine Learning · Computer Science 2023-09-27 Tim Seyde , Peter Werner , Wilko Schwarting , Igor Gilitschenski , Martin Riedmiller , Daniela Rus , Markus Wulfmeier

The purpose of this paper is to close the remaining gaps in the understanding of the role that the constrained generalized continuous algebraic Riccati equation plays in singular linear-quadratic (LQ) optimal control. Indeed, in spite of…

Optimization and Control · Mathematics 2014-04-08 Augusto Ferrante , Lorenzo Ntogramatzidis

We address the issue of estimation bias in deep reinforcement learning (DRL) by introducing solution mechanisms that include a new, twin TD-regularized actor-critic (TDR) method. It aims at reducing both over and under-estimation errors.…

Machine Learning · Computer Science 2023-11-08 Junmin Zhong , Ruofan Wu , Jennie Si

This paper is concerned with a linear-quadratic (LQ, for short) optimal control problem for backward stochastic differential equations (BSDEs, for short), where the coefficients of the backward control system and the weighting matrices in…

Optimization and Control · Mathematics 2021-05-14 Jingrui Sun , Hanxiao Wang

In this work, we propose a stochastic gradient descent (SGD) framework to design data-driven policy gradient descent algorithms for the linear quadratic regulator problem. Two alternative schemes are considered to estimate the policy…

Systems and Control · Electrical Eng. & Systems 2026-02-24 Bowen Song , Simon Weissmann , Mathias Staudigl , Andrea Iannelli

This paper is concerned with the linear quadratic optimal control of discrete-time time-varying system with terminal state constraint. The main contribution is to propose a Q-learning algorithm for the optimal controller when the…

Optimization and Control · Mathematics 2023-07-20 Juanjuan Xu , Jingmei Liu , Zhaorong Zhang , Wei Wang