中文
相关论文

相关论文: Policy Gradient Method for LQG Control via Input-O…

200 篇论文

We present a model-based globally convergent policy gradient method (PGM) for linear quadratic Gaussian (LQG) control. Firstly, we establish equivalence between optimizing dynamic output feedback controllers and designing a static feedback…

最优化与控制 · 数学 2024-02-27 Tomonori Sadamoto , Fumiya Nakamata

This paper proposes model-based and model-free policy gradient methods (PGMs) for designing dynamic output feedback controllers for discrete-time partially observable systems. To fulfill this objective, we first show that any dynamic output…

系统与控制 · 电气工程与系统科学 2024-07-26 Tomonori Sadamoto , Takumi Hirai

We consider solutions to the linear quadratic Gaussian (LQG) regulator problem via policy gradient (PG) methods. Although PG methods have demonstrated strong theoretical guarantees in solving the linear quadratic regulator (LQR) problem,…

最优化与控制 · 数学 2025-07-15 Kasra Fallah , Leonardo F. Toso , James Anderson

In this work, we revisit the Linear Quadratic Gaussian (LQG) optimal control problem from a behavioral perspective. Motivated by the suitability of behavioral models for data-driven control, we begin with a reformulation of the LQG problem…

系统与控制 · 电气工程与系统科学 2022-09-20 Abed AlRahman Al Makdah , Vishaal Krishnan , Vaibhav Katewa , Fabio Pasqualetti

Motivated by recent advances of reinforcement learning and direct data-driven control, we propose policy gradient adaptive control (PGAC) for the linear quadratic regulator (LQR), which uses online closed-loop data to improve the control…

最优化与控制 · 数学 2025-06-16 Feiran Zhao , Alessandro Chiuso , Florian Dörfler

While the optimization landscape of policy gradient methods has been recently investigated for partially observed linear systems in terms of both static output feedback and dynamical controllers, they only provide convergence guarantees to…

最优化与控制 · 数学 2023-04-25 Feiran Zhao , Xingyun Fu , Keyou You

In recent times, significant advancements have been made in delving into the optimization landscape of policy gradient methods for achieving optimal control in linear time-invariant (LTI) systems. Compared with state-feedback control,…

最优化与控制 · 数学 2023-10-31 Jingliang Duan , Jie Li , Xuyang Chen , Kai Zhao , Shengbo Eben Li , Lin Zhao

The convergence of policy gradient algorithms hinges on the optimization landscape of the underlying optimal control problem. Theoretical insights into these algorithms can often be acquired from analyzing those of linear quadratic control.…

最优化与控制 · 数学 2023-11-02 Jingliang Duan , Wenhan Cao , Yang Zheng , Lin Zhao

This paper studies the linear quadratic regulation (LQR) problem of unknown discrete-time systems via dynamic output feedback learning control. In contrast to the state feedback, the optimality of the dynamic output feedback control for…

系统与控制 · 电气工程与系统科学 2025-05-29 Kedi Xie , Martin Guay , Shimin Wang , Fang Deng , Maobin Lu

This paper is concerned with the Coherent Quantum Linear Quadratic Gaussian (CQLQG) control problem of finding a stabilizing measurement-free quantum controller for a quantum plant so as to minimize an infinite-horizon mean square…

量子物理 · 物理学 2015-02-03 Arash Kh. Sichani , Igor G. Vladimirov , Ian R. Petersen

We consider direct policy optimization for the linear-quadratic Gaussian (LQG) setting. Over the past few years, it has been recognized that the landscape of dynamic output-feedback controllers of relevance to LQG has an intricate geometry,…

最优化与控制 · 数学 2024-08-19 Spencer Kraisler , Mehran Mesbahi

We consider the static output feedback control for Linear Quadratic Regulator problems with structured constraints under the assumption that system parameters are unknown. To solve the problem in the model free setting, we propose the…

最优化与控制 · 数学 2023-03-21 Shokichi Takakura , Kazuhiro Sato

The linear quadratic regulator is the fundamental problem of optimal control. Its state feedback version was set and solved in the early 1960s. However the static output feedback problem has no explicit-form solution. It is suggested to…

最优化与控制 · 数学 2020-11-03 Ilyas Fatkhullin , Boris Polyak

In this paper, we investigate a model-free optimal control design that minimizes an infinite horizon average expected quadratic cost of states and control actions subject to a probabilistic risk or chance constraint using input-output data.…

系统与控制 · 电气工程与系统科学 2024-11-11 Arunava Naha , Subhrakanti Dey

This paper is concerned with coherent quantum linear quadratic Gaussian (CQLQG) control. The problem is to find a stabilizing measurement-free quantum controller for a quantum plant so as to minimize a mean square cost for the fully quantum…

量子物理 · 物理学 2016-09-27 Arash Kh. Sichani , Igor G. Vladimirov , Ian R. Petersen

We consider the continuous-time Linear-Quadratic-Regulator (LQR) problem in terms of optimizing a real-valued matrix function over the set of feedback gains. The results developed are in parallel to those in Bu et al. [1] for discrete-time…

系统与控制 · 电气工程与系统科学 2020-06-17 Jingjing Bu , Afshin Mesbahi , Mehran Mesbahi

We present a formulation of feedback in quantum systems in which the best estimates of the dynamical variables are obtained continuously from the measurement record, and fed back to control the system. We apply this method to the problem of…

量子物理 · 物理学 2009-10-31 A. C. Doherty , K. Jacobs

The purpose of this paper is to study the mixed linear quadratic Gaussian (LQG) and $H_\infty$ optimal control problem for linear quantum stochastic systems, where the controller itself is also a quantum system, often referred to as…

量子物理 · 物理学 2016-11-15 Lei Cui , Zhiyuan Dong , Guofeng Zhang , Heung Wing Joseph Lee

With the outstanding performance of policy gradient (PG) method in the reinforcement learning field, the convergence theory of it has aroused more and more interest recently. Meanwhile, the significant importance and abundant theoretical…

最优化与控制 · 数学 2024-04-19 Xinpei Zhang , Guangyan Jia

The paper is concerned with the coherent quantum Linear Quadratic Gaussian (CQLQG) control problem for time-varying quantum plants governed by linear quantum stochastic differential equations over a bounded time interval. A controller is…

量子物理 · 物理学 2012-05-21 Igor G. Vladimirov , Ian R. Petersen
‹ 上一页 1 2 3 10 下一页 ›