中文
相关论文

相关论文: Single Time-scale Actor-critic Method to Solve the…

200 篇论文

An optimal control law for networked control systems with a discrete-time linear time-invariant (LTI) system as plant and networks between sensor and controller as well as between controller and actuator is proposed. This controller is…

系统与控制 · 电气工程与系统科学 2021-07-09 Marijan Palmisano , Martin Steinberger , Martin Horn

We present a continuous-time equivalent to the well-known iterative linear-quadratic algorithm including an implementation of a backtracking line-search policy and a novel regularization approach based on the necessary conditions in the…

系统与控制 · 电气工程与系统科学 2025-05-22 Juraj Lieskovský , Jaroslav Bušek , Tomáš Vyhlídal

``Sim2real gap", in which the system learned in simulations is not the exact representation of the real system, can lead to loss of stability and performance when controllers learned using data from the simulated system are used on the real…

系统与控制 · 电气工程与系统科学 2025-05-15 Shivam Bajaj , Prateek Jaiswal , Vijay Gupta

In this report, linear quadratic regulator is used to design adaptive cruise control system. In the regulator, Q and R parameters vary with time according to current traffic situations. Phase-plant method is used to give constraints on Q…

系统与控制 · 电气工程与系统科学 2020-08-06 Yuncheng Jiang

This study focuses on the numerical discretization methods for the continuous-time discounted linear-quadratic optimal control problem (LQ-OCP) with time delays. By assuming piecewise constant inputs, we formulate the discrete system…

最优化与控制 · 数学 2024-07-29 Zhanhao Zhang , Steen Hørsholt , John Bagterp Jørgensen

For various typical cases and situations where the formulation results in an optimal control problem, the Linear Quadratic Regulator (LQR) approach and its variants continue to be highly attractive. In certain scenarios, it can happen that…

最优化与控制 · 数学 2023-02-14 Jun Ma , Zilong Cheng , Xiaocong Li , Wenxin Wang , Masayoshi Tomizuka , Tong Heng Lee

As an important type of reinforcement learning algorithms, actor-critic (AC) and natural actor-critic (NAC) algorithms are often executed in two ways for finding optimal policies. In the first nested-loop design, actor's one update of…

机器学习 · 计算机科学 2020-05-11 Tengyu Xu , Zhe Wang , Yingbin Liang

We consider the Linear-Quadratic-Regulator (LQR) problem in terms of optimizing a real-valued matrix function over the set of feedback gains. Such a setup facilitates examining the implications of a natural initial-state independent…

系统与控制 · 电气工程与系统科学 2019-07-31 Jingjing Bu , Afshin Mesbahi , Maryam Fazel , Mehran Mesbahi

For linear time-invariant (LTI) systems, the design of an optimal controller is a commonly encountered problem in many applications. Among all the optimization approaches available, the linear quadratic regulator (LQR) methodology certainly…

最优化与控制 · 数学 2022-03-29 Zilong Cheng , Jun Ma , Xiaocong Li , Masayoshi Tomizuka , Tong Heng Lee

Policy optimization has drawn increasing attention in reinforcement learning, particularly in the context of derivative-free methods for linear quadratic regulator (LQR) problems with unknown dynamics. This paper focuses on characterizing…

最优化与控制 · 数学 2025-06-17 Weijian Li , Panagiotis Kounatidis , Zhong-Ping Jiang , Andreas A. Malikopoulos

We revisit in this paper the discrete-time linear quadratic regulator (LQR) problem from the perspective of receding-horizon policy gradient (RHPG), a newly developed model-free learning framework for control applications. We provide a…

最优化与控制 · 数学 2024-02-02 Xiangyuan Zhang , Tamer Başar

We study a structured bi-level optimization problem where the upper-level objective is a smooth function and the lower-level problem is policy optimization in a Markov decision process (MDP). The upper-level decision variable parameterizes…

机器学习 · 计算机科学 2026-04-23 Sihan Zeng , Sujay Bhatt , Sumitra Ganesh , Alec Koppel

In this paper, the solvability of discrete-time stochastic linear-quadratic (LQ) optimal control problem in finite horizon is considered. Firstly, it shows that the closed-loop solvability for the LQ control problem is optimal if and only…

最优化与控制 · 数学 2025-02-25 Yue Sun , Xianping Wu , Xun Li

This work presents an algorithmic scheme for solving the infinite-time constrained linear quadratic regulation problem. We employ an accelerated version of a popular proximal gradient scheme, commonly known as the Forward-Backward Splitting…

最优化与控制 · 数学 2015-01-20 Giorgos Stathopoulos , Milan Korda , Colin N. Jones

While there has been substantial success for solving continuous control with actor-critic methods, simpler critic-only methods such as Q-learning find limited application in the associated high-dimensional action spaces. However, most…

The purpose of this paper is to close the remaining gaps in the understanding of the role that the constrained generalized continuous algebraic Riccati equation plays in singular linear-quadratic (LQ) optimal control. Indeed, in spite of…

最优化与控制 · 数学 2014-04-08 Augusto Ferrante , Lorenzo Ntogramatzidis

We address the issue of estimation bias in deep reinforcement learning (DRL) by introducing solution mechanisms that include a new, twin TD-regularized actor-critic (TDR) method. It aims at reducing both over and under-estimation errors.…

机器学习 · 计算机科学 2023-11-08 Junmin Zhong , Ruofan Wu , Jennie Si

This paper is concerned with a linear-quadratic (LQ, for short) optimal control problem for backward stochastic differential equations (BSDEs, for short), where the coefficients of the backward control system and the weighting matrices in…

最优化与控制 · 数学 2021-05-14 Jingrui Sun , Hanxiao Wang

In this work, we propose a stochastic gradient descent (SGD) framework to design data-driven policy gradient descent algorithms for the linear quadratic regulator problem. Two alternative schemes are considered to estimate the policy…

系统与控制 · 电气工程与系统科学 2026-02-24 Bowen Song , Simon Weissmann , Mathias Staudigl , Andrea Iannelli

This paper is concerned with the linear quadratic optimal control of discrete-time time-varying system with terminal state constraint. The main contribution is to propose a Q-learning algorithm for the optimal controller when the…

最优化与控制 · 数学 2023-07-20 Juanjuan Xu , Jingmei Liu , Zhaorong Zhang , Wei Wang