中文
相关论文

相关论文: Imitation and Transfer Learning for LQG Control

200 篇论文

Many control tasks can be formulated as a tracking problem of a known or unknown reference signal. Examples are movement compensation in collaborative robotics, the synchronisation of oscillations for power systems or reference tracking of…

最优化与控制 · 数学 2019-11-26 Janine Matschek , Andreas Himmel , Kai Sundmacher , Rolf Findeisen

One of the key benefits of model predictive control is the capability of controlling a system proactively in the sense of taking the future system evolution into account. However, often external disturbances or references are not a priori…

最优化与控制 · 数学 2019-12-03 Janine Matschek , Tim Gonschorek , Magnus Hanses , Norbert Elkmann , Frank Ortmeier , Rolf Findeisen

Iterative linear quadratic regulator (iLQR) has gained wide popularity in addressing trajectory optimization problems with nonlinear system models. However, as a model-based shooting method, it relies heavily on an accurate system model to…

机器学习 · 计算机科学 2022-09-16 Zilong Cheng , Yulin Li , Kai Chen , Jun Ma , Tong Heng Lee

We consider the linear quadratic (LQ) optimal control problem for a class of evolution equations in infinite dimensions, in the presence of distributed and nonlocal inputs. Following the perspective taken in our previous research work on…

最优化与控制 · 数学 2024-07-23 Paolo Acquistapace , Francesca Bucci

We present an algorithm that learns to imitate expert behavior and can transfer to previously unseen domains without retraining. Such an algorithm is extremely relevant in real-world applications such as robotic learning because 1) reward…

机器学习 · 计算机科学 2023-10-11 Alvaro Cauderan , Gauthier Boeshertz , Florian Schwarb , Calvin Zhang

This paper is concerned with the linear quadratic (LQ) optimal control of continuous-time system with terminal state constraint. In particular, multiple agents exist in the system which can only access partial information of the matrix…

最优化与控制 · 数学 2025-10-21 Wenjing Yang , Zhaorong Zhang , Juanjuan Xu

Integer-order calculus often falls short in capturing the long-range dependencies and memory effects found in many real-world processes. Fractional calculus addresses these gaps via fractional-order integrals and derivatives, but…

系统与控制 · 电气工程与系统科学 2025-10-20 Xiaole Zhang , Peiyu Zhang , Xiongye Xiao , Shixuan Li , Vasileios Tzoumas , Vijay Gupta , Paul Bogdan

We introduce a new problem setting for continuous control called the LQR with Rich Observations, or RichLQR. In our setting, the environment is summarized by a low-dimensional continuous latent state with linear dynamics and quadratic…

Robust imitation learning seeks to mimic expert controller behavior while ensuring stability, but current methods require accurate plant models. Here, robust imitation learning is addressed for stabilizing poorly modeled plants with linear…

系统与控制 · 电气工程与系统科学 2022-10-04 Amy K. Strong , Ethan J. LoCicero , Leila Bridgeman

Regularization is a well recognized powerful strategy to improve the performance of a learning machine and $l^q$ regularization schemes with $0<q<\infty$ are central in use. It is known that different $q$ leads to different properties of…

机器学习 · 计算机科学 2014-09-26 Shaobo Lin , Jinshan Zeng , Jian Fang , Zongben Xu

Mitigating measurement errors in quantum systems without relying on quantum error correction is of critical importance for the practical development of quantum technology. Deep learning-based quantum measurement error mitigation has…

量子物理 · 物理学 2024-08-12 ChangWon Lee , Daniel K. Park

This paper proposes a differentiable robust LQR layer for reinforcement learning and imitation learning under model uncertainty and stochastic dynamics. The robust LQR layer can exploit the advantages of robust optimal control and…

机器人学 · 计算机科学 2021-06-11 Ngo Anh Vien , Gerhard Neumann

In modern machine learning, models can often fit training data in numerous ways, some of which perform well on unseen (test) data, while others do not. Remarkably, in such cases gradient descent frequently exhibits an implicit bias that…

机器学习 · 计算机科学 2024-06-04 Noam Razin , Yotam Alexander , Edo Cohen-Karlik , Raja Giryes , Amir Globerson , Nadav Cohen

In this paper, we study the use of state-of-the-art nonlinear system identification techniques for the optimal control of nonlinear systems. We show that the nonlinear systems identification problem is equivalent to estimating the…

最优化与控制 · 数学 2023-10-23 Aayushman Sharma , Suman Chakravorty

We analyze offline designs of linear quadratic regulator (LQR) strategies with uncertain disturbances. First, we consider the scenario where the exogenous variable can be estimated in a controlled environment, and subsequently, consider a…

系统与控制 · 电气工程与系统科学 2025-09-26 Sayak Mukherjee , Ramij R. Hossain , Mahantesh Halappanavar

This paper investigates the design of self-triggered control for networked control systems (NCS), where the dynamics of the plant is unknown apriori. To deal with the nature of the self-triggered control, in which state measurements are…

系统与控制 · 电气工程与系统科学 2022-02-22 Wang Zhijun , Kazumune Hashimoto , Wataru Hashimoto , Shigemasa Takai

Recent literature has made much progress in understanding \emph{online LQR}: a modern learning-theoretic take on the classical control problem in which a learner attempts to optimally control an unknown linear dynamical system with fully…

机器学习 · 计算机科学 2020-10-06 Max Simchowitz

In this paper we explore the Linear-Quadratic Regulator (LQR) to model movement of the mouse pointer. We propose a model in which users are assumed to behave optimally with respect to a certain cost function. Users try to minimize the…

人机交互 · 计算机科学 2020-02-27 Florian Fischer , Arthur Fleig , Markus Klar , Lars Gruene , Joerg Mueller

Off-policy, value-based reinforcement learning methods such as Q-learning are appealing because they can learn from arbitrary experience, including data collected by older policies or other agents. In practice, however, bootstrapping makes…

人工智能 · 计算机科学 2026-05-12 Armaan A. Abraham , Lucy Xiaoyang Shi , Chelsea Finn

Recently, motion generation by machine learning has been actively researched to automate various tasks. Imitation learning is one such method that learns motions from data collected in advance. However, executing long-term tasks remains…

机器人学 · 计算机科学 2022-03-17 Kazuki Hayashi , Sho Sakaino , Toshiaki Tsuji
‹ 上一页 1 8 9 10 下一页 ›