中文
相关论文

相关论文: Stabilizing reinforcement learning control: A modu…

200 篇论文

This work introduces a data-driven control approach for stabilizing high-dimensional dynamical systems from scarce data. The proposed context-aware controller inference approach is based on the observation that controllers need to act…

最优化与控制 · 数学 2023-02-23 Steffen W. R. Werner , Benjamin Peherstorfer

For an unknown linear system, starting from noisy open-loop input-state data collected during a finite-length experiment, we directly design a linear feedback controller that guarantees robust invariance of a given polyhedral set of the…

系统与控制 · 电气工程与系统科学 2022-05-25 Andrea Bisoffi , Claudio De Persis , Pietro Tesi

Through the method of Learning Feedback Linearization, we seek to learn a linearizing controller to simplify the process of controlling a car to race autonomously. A soft actor-critic approach is used to learn a decoupling matrix and drift…

最优化与控制 · 数学 2021-10-22 Michael Estrada , Sida Li , Xiangyu Cai

This paper presents a novel linear robust Youla controller output observation system for tracking vehicle motion trajectories using a simple nonlinear kinematic vehicle model, supplemented with positional data from a radar sensor. The…

系统与控制 · 电气工程与系统科学 2025-06-16 Rongfei Li , Francis Assadian. Iman Soltani

The application of reinforcement learning to safety-critical systems is limited by the lack of formal methods for verifying the robustness and safety of learned policies. This paper introduces a novel framework that addresses this gap by…

人工智能 · 计算机科学 2025-08-22 Ahmed Nasir , Abdelhafid Zenati

Model-based Reinforcement Learning (MBRL) has shown many desirable properties for intelligent control tasks. However, satisfying safety and stability constraints during training and rollout remains an open question. We propose a new…

系统与控制 · 电气工程与系统科学 2024-05-28 Harry Zhang

Vision-Language-Action (VLA) models have recently emerged as powerful general-purpose policies for robotic manipulation, benefiting from large-scale multi-modal pre-training. However, they often fail to generalize reliably in…

机器人学 · 计算机科学 2025-12-02 Hongyin Zhang , Shuo Zhang , Junxi Jin , Qixin Zeng , Runze Li , Donglin Wang

Deep learning methods have demonstrated significant potential for addressing complex nonlinear control problems. For real-world safety-critical tasks, however, it is crucial to provide formal stability guarantees for the designed…

系统与控制 · 电气工程与系统科学 2025-06-10 Han Wang , Keyan Miao , Diego Madeira , Antonis Papachristodoulou

We present a novel approach to control design for nonlinear systems which leverages model-free policy optimization techniques to learn a linearizing controller for a physical plant with unknown dynamics. Feedback linearization is a…

In this paper, a novel robust output regulation control framework is proposed for the system subject to noise, modeled disturbance and unmodeled disturbance to seek tracking performance and robustness simultaneously. The output regulation…

系统与控制 · 电气工程与系统科学 2022-10-28 Zhicheng Zhang , Zhiqiang Zuo , Xiang Chen , Ying Tan , Yijing Wang

In this paper, the stability and stabilization problem of positive nonlinear systems, described by the Takagi-Sugeno discrete-time fuzzy model, is studied. The proposed approach is based on the linear co-positive Lyapunov function and…

系统与控制 · 电气工程与系统科学 2019-12-17 Elham Ahmadi , Jafar Zarei

Despite their effectiveness and popularity in offline or model-based reinforcement learning (RL), transformers remain underexplored in online model-free RL due to their sensitivity to training setups and model design decisions such as how…

机器学习 · 计算机科学 2025-10-16 Nikita Kachaev , Daniil Zelezetsky , Egor Cherepanov , Alexey K. Kovelev , Aleksandr I. Panov

This paper presents a robust data-driven controller design based on the noisy input-output data without assumptions on the statistical properties of the noises. We start with the direct data-representation of system models that take…

最优化与控制 · 数学 2023-02-24 Chin-Yao Chang , Andrey Bernstein

Linear quadratic regulator with unmeasurable states and unknown system matrix parameters better aligns with practical scenarios. However, for this problem, balancing the optimality of the resulting controller and the leniency of the…

最优化与控制 · 数学 2025-09-04 Jun Xie , Yuan-Hua Ni , Yiqin Yang , Bo Xu

This paper addresses the end-to-end sample complexity bound for learning in closed loop the state estimator-based robust H2 controller for an unknown (possibly unstable) Linear Time Invariant (LTI) system, when given a fixed state-feedback…

系统与控制 · 电气工程与系统科学 2022-12-23 Yifei Zhang , Sourav Kumar Ukil , Andrei Sperila , Serban Sabau

Stability is a basic requirement when studying the behavior of dynamical systems. However, stabilizing dynamical systems via reinforcement learning is challenging because only little data can be collected over short time horizons before…

最优化与控制 · 数学 2024-11-01 Steffen W. R. Werner , Benjamin Peherstorfer

Reinforcement learning has been established over the past decade as an effective tool to find optimal control policies for dynamical systems, with recent focus on approaches that guarantee safety during the learning and/or execution phases.…

系统与控制 · 电气工程与系统科学 2021-10-06 S M Nahid Mahmud , Scott A Nivison , Zachary I. Bell , Rushikesh Kamalapurkar

This paper proposes an optimization with penalty-based feedback design framework for safe stabilization of control affine systems. Our starting point is the availability of a control Lyapunov function (CLF) and a control barrier function…

最优化与控制 · 数学 2022-07-26 Pol Mestres , Jorge Cortés

Understanding simplicity biases in deep learning offers a promising path toward developing reliable AI. A common metric for this, inspired by Boolean function analysis, is average sensitivity, which captures a model's robustness to…

机器学习 · 计算机科学 2026-02-10 Themistoklis Haris , Zihan Zhang , Yuichi Yoshida

We consider the problem of designing robust state-feedback controllers for discrete-time linear time-invariant systems, based directly on measured data. The proposed design procedures require no model knowledge, but only a single open-loop…

系统与控制 · 电气工程与系统科学 2020-10-27 Julian Berberich , Anne Romer , Carsten W. Scherer , Frank Allgöwer