English
Related papers

Related papers: Policy Gradient-based Algorithms for Continuous-ti…

200 papers

Current research suggests the use of a liner quadratic performance index for optimal control of regulators in various applications. Some examples include correcting the trajectory of rocket and air vehicles, vibration suppression of…

General Mathematics · Mathematics 2007-05-23 Alexander Bolonkin , Robert Sierakowski

The framework of Integral Quadratic Constraints (IQCs) is used to perform an analysis of gradient descent with varying step sizes. Two performance metrics are considered: convergence rate and noise amplification. We assume that the step…

Optimization and Control · Mathematics 2025-05-13 Ram Padmanabhan , Peter Seiler

Linear time-invariant control systems can be considered as finitely generated modules over the commutative principal ideal ring $\mathbb{R}[\frac{d}{dt}]$ of linear differential operators with respect to the time derivative. The Kalman…

Optimization and Control · Mathematics 2025-12-15 Cédric Join , Emmanuel Delaleau , Michel Fliess

Multi-agent reinforcement learning has been successfully applied to a number of challenging problems. Despite these empirical successes, theoretical understanding of different algorithms is lacking, primarily due to the curse of…

Machine Learning · Computer Science 2021-12-28 Yuwei Luo , Zhuoran Yang , Zhaoran Wang , Mladen Kolar

This paper presents a novel Lyapunov-Based Quantum Reinforcement Learning (LQRL) framework that integrates quantum policy optimization with Lyapunov stability analysis for continuous-time vehicle control. The proposed approach combines the…

Feedback control problems involving autonomous polynomial systems are prevalent, yet there are limited algorithms and software for approximating their solution. This paper represents a step forward by considering the special case of the…

Optimization and Control · Mathematics 2020-09-24 Jeff Borggaard , Lizette Zietsman

From the perspective of control theory, the gradient descent optimization methods can be regarded as a dynamic system where various control techniques can be designed to enhance the performance of the optimization method. In this paper, we…

Optimization and Control · Mathematics 2025-03-19 Osama F. Abdel Aal , Necdet Sinan Ozbek , Jairo Viola , YangQuan Chen

We study deterministic, discrete linear time-invariant systems with infinite-horizon discounted quadratic cost. It is well-known that standard stabilizability and detectability properties are not enough in general to conclude stability…

Optimization and Control · Mathematics 2025-09-04 Jonathan de Brusse , Jamal Daafouz , Mathieu Granzotto , Romain Postoyan , Dragan Nesic

Reinforcement learning (RL) has achieved significant success across a wide range of domains, however, most existing methods are formulated in discrete time. In this work, we introduce a novel RL method for continuous-time control, where…

Machine Learning · Computer Science 2025-10-21 Chengxiu Hua , Jiawen Gu , Yushun Tang

We consider the optimal control problem for a linear conditional McKean-Vlasov equation with quadratic cost functional. The coefficients of the system and the weigh-ting matrices in the cost functional are allowed to be adapted processes…

Probability · Mathematics 2017-03-09 Huyên Pham

Inspired by REINFORCE, we introduce a novel receding-horizon algorithm for the Linear Quadratic Regulator (LQR) problem with unknown dynamics. Unlike prior methods, our algorithm avoids reliance on two-point gradient estimates while…

Optimization and Control · Mathematics 2025-10-07 Amirreza Neshaei Moghaddam , Alex Olshevsky , Bahman Gharesifard

This technical report is concerned with the convergence properties of what we call the split optimal policy iteration for coupled LQR problems; see section 3.1 in the manuscript. Interestingly, the iteration shows different convergence…

Optimization and Control · Mathematics 2014-04-22 Péter Koltai

Feedback-based methods have gained significant attention as an alternative training paradigm for the Quantum Approximate Optimization Algorithm (QAOA) in solving combinatorial optimization problems such as MAX-CUT. In particular, Quantum…

Quantum Physics · Physics 2026-02-16 Masih Mozakka , Mohsen Heidari

This study presents the design, discretization and implementation of the continuous-time linear-quadratic model predictive control (CT-LMPC). The control model of the CT-LMPC is parameterized as transfer functions with time delays, and they…

Optimization and Control · Mathematics 2025-03-18 Zhanhao Zhang , Anders Hilmar Damm Christensen , Steen Hørsholt , John Bagterp Jørgensen

In this paper, we investigate the infinite-horizon risk-constrained linear quadratic regulator problem (RC-QR), which augments the classical LQR formulation with a statistical constraint on the variability of the system state to incorporate…

Optimization and Control · Mathematics 2025-10-28 Weijian Li , Andreas A. Malikopoulos

This paper considers a risk-sensitive optimal control problem for a field-mediated interconnection of a quantum plant with a coherent (measurement-free) quantum controller. The plant and the controller are multimode open quantum harmonic…

Optimization and Control · Mathematics 2023-08-09 Igor G. Vladimirov , Ian R. Petersen

In recent times, significant advancements have been made in delving into the optimization landscape of policy gradient methods for achieving optimal control in linear time-invariant (LTI) systems. Compared with state-feedback control,…

Optimization and Control · Mathematics 2023-10-31 Jingliang Duan , Jie Li , Xuyang Chen , Kai Zhao , Shengbo Eben Li , Lin Zhao

This paper revisits the classical Linear Quadratic Gaussian (LQG) control from a modern optimization perspective. We analyze two aspects of the optimization landscape of the LQG problem: 1) connectivity of the set of stabilizing controllers…

Optimization and Control · Mathematics 2021-02-09 Yang Zheng , Yujie Tang , Na Li

We consider a discrete-time Linear-Quadratic-Gaussian (LQG) control problem in which Massey's directed information from the observed output of the plant to the control input is minimized while required control performance is attainable.…

Optimization and Control · Mathematics 2017-06-13 Takashi Tanaka , Peyman Mohajerin Esfahani , Sanjoy K. Mitter

This paper studies the Linear Quadratic Regulator (LQR) problem for continuous-time Markov Jump Linear Systems (MJLS) governed by general finite-state Markov chains that may include transient, absorbing, or non-communicating states. The…

Optimization and Control · Mathematics 2025-11-20 Alfredo R. R. Narváez , Jeinny Peralta , M. A. C. Candezano