中文
相关论文

相关论文: Learning to control from expert demonstrations

200 篇论文

For a parameter-unknown linear descriptor system, this paper proposes data-driven methods to testify the system's type and controllability and then to stabilize it. First, a data-based condition is developed to identify whether this unknown…

系统与控制 · 电气工程与系统科学 2022-01-03 Jiabao He , Xuan Zhang , Feng Xu , Junbo Tan , Xueqian Wang

We study the problem of imitating an expert demonstrator in a discrete-time, continuous state-and-action control system. We show that, even if the dynamics satisfy a control-theoretic property called exponential stability (i.e. the effects…

机器学习 · 计算机科学 2025-07-29 Max Simchowitz , Daniel Pfrommer , Ali Jadbabaie

In this work, we propose a framework to learn feedback control policies with guarantees on closed-loop generalization and adversarial robustness. These policies are learned directly from expert demonstrations, contained in a dataset of…

机器学习 · 计算机科学 2022-11-03 Abed AlRahman Al Makdah , Vishaal Krishnan , Fabio Pasqualetti

We consider a nonlinear control system modeled as an ordinary differential equation subject to disturbance, with a state feedback controller parameterized as a feedforward neural network. We propose a framework for training controllers with…

机器学习 · 计算机科学 2025-11-12 Akash Harapanahalli , Samuel Coogan

A high-gain observer is used for a class of feedback linearisable nonlinear systems to synthesize safety-preserving controllers over the observer output. A bound on the distance between trajectories under state and output feedback is…

系统与控制 · 计算机科学 2016-03-23 Kendra Lesser , Alessandro Abate

This paper investigates how to utilize different forms of human interaction to safely train autonomous systems in real-time by learning from both human demonstrations and interventions. We implement two components of the Cycle-of-Learning…

Semi-supervised learning improves the performance of supervised machine learning by leveraging methods from unsupervised learning to extract information not explicitly available in the labels. Through the design of a system that enables a…

机器人学 · 计算机科学 2020-07-27 Simón C. Smith , Subramanian Ramamoorthy

The study of controlled hybrid systems requires practical tools for approximation and comparison of system behaviors. Existing approaches to these problems impose undue restrictions on the system's continuous and discrete dynamics.…

最优化与控制 · 数学 2015-04-15 Samuel Burden , Humberto Gonzalez , Ramanarayan Vasudevan , Ruzena Bajcsy , S. Shankar Sastry

Imitation learning algorithms learn a policy from demonstrations of expert behavior. We show that, for deterministic experts, imitation learning can be done by reduction to reinforcement learning with a stationary reward. Our theoretical…

机器学习 · 统计学 2022-03-16 Kamil Ciosek

This paper addresses the optimal control problem known as the Linear Quadratic Regulator in the case when the dynamics are unknown. We propose a multi-stage procedure, called Coarse-ID control, that estimates a model from a few experimental…

最优化与控制 · 数学 2018-12-17 Sarah Dean , Horia Mania , Nikolai Matni , Benjamin Recht , Stephen Tu

A method is presented to learn neural network (NN) controllers with stability and safety guarantees through imitation learning (IL). Convex stability and safety conditions are derived for linear time-invariant plant dynamics with NN…

系统与控制 · 电气工程与系统科学 2021-04-08 He Yin , Peter Seiler , Ming Jin , Murat Arcak

We propose a two-component data-driven controller to safely perform docking maneuvers for satellites. Reinforcement Learning is used to deduce an optimal control policy based on measurement data. To safeguard the learning phase, an…

最优化与控制 · 数学 2024-07-30 Simon Gottschalk , Lukas Lanza , Karl Worthmann , Kerstin Lux-Gottschalk

We present a novel method for imitation learning for control requirements expressed using Signal Temporal Logic (STL). More concretely we focus on the problem of training a neural network to imitate a complex controller. The learning…

机器人学 · 计算机科学 2024-03-26 Thao Dang , Alexandre Donzé , Inzemamul Haque , Nikolaos Kekatos , Indranil Saha

Model predictive control allows solving complex control tasks with control and state constraints. However, an optimal control problem must be solved in real-time to predict the future system behavior, which is hardly possible on embedded…

系统与控制 · 电气工程与系统科学 2023-04-13 Jan Olucak , Walter Fichter , Torbjørn Cunis

An algorithm for constructing a control function that transfers a wide class of stationary nonlinear systems of ordinary differential equations from an initial state to a final state under certain control restrictions is proposed. The…

最优化与控制 · 数学 2017-03-01 Alexander N. Kvitko , Oksana S. Firyulina , Alexey S. Eremin

Achieving precise control of colloidal self-assembly into specific patterns remains a longstanding challenge due to the complex process dynamics. Recently, machine learning-based state representation and reinforcement learning-based control…

软凝聚态物质 · 物理学 2025-12-19 Andres Lizano-Villalobos , Fangyuan Ma , Wentao Tang , Wei Sun , Xun Tang

Two different aspects of formation control of multiple agents subjected to linear transformation have been addressed in this paper. We consider a set of complex single integrator systems so that the dimension of the system reduces to half…

系统与控制 · 计算机科学 2015-06-03 Soumic Sarkar , Indra Narayan Kar

Imitation learning has achieved great success in many sequential decision-making tasks, in which a neural agent is learned by imitating collected human demonstrations. However, existing algorithms typically require a large number of…

机器学习 · 计算机科学 2023-06-14 Tianxiang Zhao , Wenchao Yu , Suhang Wang , Lu Wang , Xiang Zhang , Yuncong Chen , Yanchi Liu , Wei Cheng , Haifeng Chen

This paper discusses the systematic design of an adaptive feedback linearizing neurocontroller for a high-order model of the synchronous machine/infinite bus power system. The power system is first modelled as an input-output nonlinear…

最优化与控制 · 数学 2007-05-23 Kingsley Fregene , Diane Kennedy

Recently, there has been a surge in interest in safe and robust techniques within reinforcement learning (RL). Current notions of risk in RL fail to capture the potential for systemic failures such as abrupt stoppages from system failures…

系统与控制 · 计算机科学 2019-10-09 David Mguni