中文
相关论文

相关论文: On the Sample Complexity of the Linear Quadratic G…

200 篇论文

Inspired by REINFORCE, we introduce a novel receding-horizon algorithm for the Linear Quadratic Regulator (LQR) problem with unknown dynamics. Unlike prior methods, our algorithm avoids reliance on two-point gradient estimates while…

最优化与控制 · 数学 2025-10-07 Amirreza Neshaei Moghaddam , Alex Olshevsky , Bahman Gharesifard

Consider a discrete-time Linear Quadratic Regulator (LQR) problem solved using policy gradient descent when the system matrices are unknown. The gradient is transmitted across a noisy channel over a finite time horizon using analog…

最优化与控制 · 数学 2025-07-22 Ashwin Verma , Aritra Mitra , Lintao Ye , Vijay Gupta

We consider transport processes that are modeled by first order hyperbolic partial differential equations. Our goal is to find a full state feedback that makes a given reference profile locally asymptotically stable. To accomplish this we…

最优化与控制 · 数学 2025-08-22 Arthur J. Krener

This paper synthesizes a gain-scheduled controller to stabilize all possible Linear Parameter-Varying (LPV) plants that are consistent with measured input/state data records. Inspired by prior work in data informativity and LTI…

最优化与控制 · 数学 2022-12-08 Jared Miller , Mario Sznaier

The convergence of policy gradient algorithms in reinforcement learning hinges on the optimization landscape of the underlying optimal control problem. Theoretical insights into these algorithms can often be acquired from analyzing those of…

机器学习 · 计算机科学 2023-11-01 Jingliang Duan , Wenhan Cao , Yang Zheng , Lin Zhao

The purpose of this paper is to present a theoretic and numerical study of utilizing squeezing and phase shift in coherent feedback control of linear quantum optical systems. A quadrature representation with built-in phase shifters is…

量子物理 · 物理学 2012-06-19 Guofeng Zhang , Heung Wing Joseph Lee , Bo Huang , Hu Zhang

This paper is concerned with the linear quadratic (LQ) optimal control of continuous-time system with terminal state constraint. In particular, multiple agents exist in the system which can only access partial information of the matrix…

最优化与控制 · 数学 2025-10-21 Wenjing Yang , Zhaorong Zhang , Juanjuan Xu

With the continuous development of large-scale complex hybrid AC-DC grids, the fast adjustability of HVDC systems is required by the grid to provide frequency regulation services. This paper develops a fully data-driven linear quadratic…

系统与控制 · 电气工程与系统科学 2022-12-05 Qianni Cao , Ye Liu , Chen Shen

A code for communication over the k-receiver additive white Gaussian noise broadcast channel with feedback is presented and analyzed using tools from the theory of linear quadratic Gaussian optimal control. It is shown that the performance…

信息论 · 计算机科学 2011-02-17 Ehsan Ardestanizadeh , Paolo Minero , Massimo Franceschetti

Reinforcement learning (RL) has been successfully used to solve many continuous control tasks. Despite its impressive results however, fundamental questions regarding the sample complexity of RL on continuous problems remain open. We study…

机器学习 · 计算机科学 2017-12-27 Stephen Tu , Benjamin Recht

We develop the framework for a non-intrusive, quadrature-based method for approximate balanced truncation (QuadBT) of linear systems with quadratic outputs, thus extending the applicability of QuadBT, which was originally designed for…

数值分析 · 数学 2025-09-17 Reetish Padhi , Ion Victor Gosea , Igor Pontes Duff , Serkan Gugercin

In this paper we present a linear quadratic Gaussian (LQG) feedback control strategy for a class of linear non-Markovian quantum systems. The feedback control law is designed based on the estimated states of a whitening quantum filter for…

量子物理 · 物理学 2017-02-20 Shibei Xue , Matthew R. James , Valery Ugrinovskii , Ian R. Petersen

Stochastic Optimal Control models represent the state-of-the-art in modeling goal-directed human movements. The linear-quadratic sensorimotor (LQS) model based on signal-dependent noise processes in state and output equation is the current…

最优化与控制 · 数学 2023-03-28 Philipp Karg , Simon Stoll , Simon Rothfuß , Sören Hohmann

The linear quadratic regulator (LQR) problem has reemerged as an important theoretical benchmark for reinforcement learning-based control of complex dynamical systems with continuous state and action spaces. In contrast with nearly all…

机器学习 · 计算机科学 2020-05-04 Benjamin Gravell , Peyman Mohajerin Esfahani , Tyler Summers

We describe a compact and reliable method to calculate the Fisher information for the estimation of a dynamical parameter in a continuously measured linear Gaussian quantum system. Unlike previous methods in the literature, which involve…

量子物理 · 物理学 2017-06-08 Marco G. Genoni

We investigate the benefits of combining regular and impulsive inputs for the control of sampled-data linear time-invariant systems. We first observe that adding an impulsive term to a regular, zero-order-hold controller may help enlarging…

系统与控制 · 电气工程与系统科学 2024-09-04 Jamal Daafouz , Jérôme Lohéac , Romain Postoyan

In this paper we design suboptimal control laws for an unknown linear system on the basis of measured data. We focus on the suboptimal linear quadratic regulator problem and the suboptimal H2 control problem. For both problems, we establish…

最优化与控制 · 数学 2020-05-08 Henk J. van Waarde , Mehran Mesbahi

This paper examines learning the optimal filtering policy, known as the Kalman gain, for a linear system with unknown noise covariance matrices using noisy output data. The learning problem is formulated as a stochastic policy optimization…

系统与控制 · 电气工程与系统科学 2023-10-27 Shahriar Talebi , Amirhossein Taghvaei , Mehran Mesbahi

This paper develops a robust extended Kalman filter to estimate the rotor angles and the rotor speeds of synchronous generators of a multimachine power system. Using a batch-mode regression form, the filter processes together predicted…

系统与控制 · 电气工程与系统科学 2021-04-06 Marcos Netto , Junbo Zhao , Lamine Mili

In this paper we explore the Linear-Quadratic Regulator (LQR) to model movement of the mouse pointer. We propose a model in which users are assumed to behave optimally with respect to a certain cost function. Users try to minimize the…

人机交互 · 计算机科学 2020-02-27 Florian Fischer , Arthur Fleig , Markus Klar , Lars Gruene , Joerg Mueller
‹ 上一页 1 8 9 10 下一页 ›