中文
相关论文

相关论文: Imitation and Transfer Learning for LQG Control

200 篇论文

This paper considers the Linear Quadratic Regulator problem for linear systems with unknown dynamics, a central problem in data-driven control and reinforcement learning. We propose a method that uses data to directly return a controller…

系统与控制 · 电气工程与系统科学 2020-05-05 Claudio De Persis , Pietro Tesi

The purpose of this paper is to study the mixed linear quadratic Gaussian (LQG) and $H_\infty$ optimal control problem for linear quantum stochastic systems, where the controller itself is also a quantum system, often referred to as…

量子物理 · 物理学 2016-11-15 Lei Cui , Zhiyuan Dong , Guofeng Zhang , Heung Wing Joseph Lee

This paper addresses the problem of distributed coordination control for multi-robot systems (MRSs) in the presence of localization uncertainty using a Linear Quadratic Gaussian (LQG) approach. We introduce a stochastic LQG control strategy…

系统与控制 · 电气工程与系统科学 2025-04-07 Tohid Kargar Tasooji , Sakineh Khodadadi

This paper investigates a conditional mean-field type linear quadratic (LQ) optimal control problem with partial observation and regime switching, where the conditional expectations of the state and control given the history of Markov chain…

最优化与控制 · 数学 2025-12-22 Zhongbin Guo , Guangchen Wang

In this letter, we consider a Linear Quadratic Gaussian (LQG) control system where feedback occurs over a noiseless binary channel and derive lower bounds on the minimum communication cost (quantified via the channel bitrate) required to…

信息论 · 计算机科学 2022-06-06 Travis C. Cuvelier , Takashi Tanaka , Robert W. Heath

Imitation learning methods have demonstrated considerable success in teaching autonomous systems complex tasks through expert demonstrations. However, a limitation of these methods is their lack of interpretability, particularly in…

机器学习 · 计算机科学 2025-07-21 Wenliang Liu , Danyang Li , Erfan Aasi , Daniela Rus , Roberto Tron , Calin Belta

In decentralized control systems with linear dynamics, quadratic cost, and Gaussian disturbance (also called decentralized LQG systems) linear control strategies are not always optimal. Nonetheless, linear control strategies are appealing…

最优化与控制 · 数学 2014-03-13 Aditya Mahajan , Ashutosh Nayyar

An important issue in quadcopter control is that an accurate dynamic model of the system is nonlinear, complex, and costly to obtain. This limits achievable control performance in practice. Gaussian process (GP) based estimation is an…

系统与控制 · 电气工程与系统科学 2021-12-23 Yuhan Liu , Roland Tóth

In this work, we propose an architecture of LLM Modules that enables the transfer of knowledge from a large pre-trained model to a smaller model using an Enhanced Cross-Attention mechanism. In the proposed scheme, the Qwen2-1.5B model is…

计算与语言 · 计算机科学 2025-02-13 Konstantin Kolomeitsev

The goal of this work is to enable a team of quadrotors to learn how to accurately track a desired trajectory while holding a given formation. We solve this problem in a distributed manner, where each vehicle has only access to the…

机器人学 · 计算机科学 2016-09-27 Andreas Hock , Angela P. Schoellig

In this paper, we investigate a data-driven framework to solve Linear Quadratic Regulator (LQR) problems when the dynamics is unknown, with the additional challenge of providing stability certificates for the overall learning and control…

系统与控制 · 电气工程与系统科学 2026-04-13 Lorenzo Sforni , Guido Carnevale , Ivano Notarnicola , Giuseppe Notarstefano

Learning-based control methods utilize run-time data from the underlying process to improve the controller performance under model mismatch and unmodeled disturbances. This is beneficial for optimizing industrial processes, where the…

系统与控制 · 电气工程与系统科学 2021-11-22 Efe C. Balta , Kira Barton , Dawn M. Tilbury , Alisa Rupenyan , John Lygeros

This paper studies the linear quadratic regulator (LQR) problem over an unknown Bernoulli packet loss channel. The unknown loss rate is estimated using finite channel samples and a certainty-equivalence (CE) optimal controller is then…

系统与控制 · 电气工程与系统科学 2025-06-17 Zhenning Zhang , Liang Xu , Yilin Mo , Xiaofan Wang

Model-free approaches for reinforcement learning (RL) and continuous control find policies based only on past states and rewards, without fitting a model of the system dynamics. They are appealing as they are general purpose and easy to…

机器学习 · 计算机科学 2018-10-09 Yasin Abbasi-Yadkori , Nevena Lazic , Csaba Szepesvari

This paper develops a controller synthesis algorithm for distributed LQG control problems under output feedback. We consider a system consisting of three interconnected linear subsystems with a delayed information sharing structure. While…

系统与控制 · 计算机科学 2016-11-15 Hamid Reza Feyzmahdavian , Assad Alam , Ather Gattami

In this paper, we study the irregular output feedback linear quadratic (LQ) control problem, which is a continuous work of previous works for irregular LQ control [33] where the state is assumed to be exactly known priori. Different from…

最优化与控制 · 数学 2019-05-17 Juanjuan Xu , Huanshui Zhang

While differentiable control has emerged as a powerful paradigm combining model-free flexibility with model-based efficiency, the iterative Linear Quadratic Regulator (iLQR) remains underexplored as a differentiable component. The…

机器人学 · 计算机科学 2025-06-24 Shuyuan Wang , Philip D. Loewen , Michael Forbes , Bhushan Gopaluni , Wei Pan

Complex planning and scheduling problems have long been solved using various optimization or heuristic approaches. In recent years, imitation learning that aims to learn from expert demonstrations has been proposed as a viable alternative…

机器学习 · 计算机科学 2024-05-24 Qian Shao , Pradeep Varakantham , Shih-Fen Cheng

Quadruped locomotion provides a natural setting for understanding when model-free learning can outperform model-based control design, by exploiting data patterns to bypass the difficulty of optimizing over discrete contacts and the…

机器学习 · 计算机科学 2026-03-10 Ruipeng Zhang , Hongzhan Yu , Ya-Chien Chang , Chenghao Li , Henrik I. Christensen , Sicun Gao

Consider a discrete-time Linear Quadratic Regulator (LQR) problem solved using policy gradient descent when the system matrices are unknown. The gradient is transmitted across a noisy channel over a finite time horizon using analog…

最优化与控制 · 数学 2025-07-22 Ashwin Verma , Aritra Mitra , Lintao Ye , Vijay Gupta