中文
相关论文

相关论文: Stabilizing reinforcement learning control: A modu…

200 篇论文

Neural network controllers have shown potential in achieving superior performance in feedback control systems. Although a neural network can be trained efficiently using deep and reinforcement learning methods, providing formal guarantees…

最优化与控制 · 数学 2024-01-10 Han Wang , Zuxun Xiong , Liqun Zhao , Antonis Papachristodoulou

A hierarchical 2DOF (2-degree-of-freedom) structure combining Youla-Kucera (YK) parameterization and model predictive control (MPC) is presented in this paper. The YK parameterization employs the coprime factorization of the nominal system…

系统与控制 · 电气工程与系统科学 2026-05-12 Zhiheng Zhao , Hans Henrik Niemann , John Bagterp Jørgensen

When neural networks are used to model dynamics, properties such as stability of the dynamics are generally not guaranteed. In contrast, there is a recent method for learning the dynamics of autonomous systems that guarantees global…

机器学习 · 计算机科学 2022-03-21 Kenji Kashima , Ryota Yoshiuchi , Yu Kawano

We introduce a novel distributed control architecture for heterogeneous platoons of linear time--invariant autonomous vehicles. Our approach is based on a generalization of the concept of {\em leader--follower} controllers for which we…

系统与控制 · 计算机科学 2015-11-02 Serban Sabau , Cristian Oara , Sean Warnick , Ali Jadbabaie

Model-free RL-based recommender systems have recently received increasing research attention due to their capability to handle partial feedback and long-term rewards. However, most existing research has ignored a critical feature in…

机器学习 · 计算机科学 2023-08-28 Tianchi Cai , Shenliao Bao , Jiyan Jiang , Shiji Zhou , Wenpeng Zhang , Lihong Gu , Jinjie Gu , Guannan Zhang

This work provides a framework for data-driven control of discrete time systems with unknown input-output dynamics and outputs controllable by the inputs. This framework leads to stable and robust real-time control of the system such that a…

系统与控制 · 电气工程与系统科学 2021-04-02 Amit K. Sanyal

Urban autonomous driving decision making is challenging due to complex road geometry and multi-agent interactions. Current decision making methods are mostly manually designing the driving policy, which might result in sub-optimal solutions…

机器学习 · 计算机科学 2019-10-23 Jianyu Chen , Bodi Yuan , Masayoshi Tomizuka

For a parameter-unknown linear descriptor system, this paper proposes data-driven methods to testify the system's type and controllability and then to stabilize it. First, a data-based condition is developed to identify whether this unknown…

系统与控制 · 电气工程与系统科学 2022-01-03 Jiabao He , Xuan Zhang , Feng Xu , Junbo Tan , Xueqian Wang

In this paper, we propose a deep unfolding-based framework for the output feedback control of systems with input saturation. Although saturation commonly arises in several practical control systems, there is still a scarce of effective…

系统与控制 · 电气工程与系统科学 2021-01-28 Koki Kobayashi , Masaki Ogura , Taisuke Kobayashi , Kenji Sugimoto

Robust controllers ensure stability in feedback loops designed under uncertainty but at the cost of performance. Model uncertainty in time-invariant systems can be reduced by recently proposed learning-based methods, which improve the…

系统与控制 · 电气工程与系统科学 2023-01-18 Alexander von Rohr , Friedrich Solowjow , Sebastian Trimpe

We present a framework for systematically combining data of an unknown linear time-invariant system with prior knowledge on the system matrices or on the uncertainty for robust controller design. Our approach leads to linear matrix…

系统与控制 · 电气工程与系统科学 2024-12-04 Julian Berberich , Carsten W. Scherer , Frank Allgöwer

We consider the design of state feedback control laws for both the switching signal and the continuous input of an unknown switched linear system, given past noisy input-state trajectories measurements. Based on Lyapunov-Metzler…

最优化与控制 · 数学 2025-06-05 Mattia Bianchi , Sergio Grammatico , Jorge Cortés

This paper studies the robustness of reinforcement learning algorithms to errors in the learning process. Specifically, we revisit the benchmark problem of discrete-time linear quadratic regulation (LQR) and study the long-standing open…

最优化与控制 · 数学 2021-03-16 Bo Pang , Zhong-Ping Jiang

Dynamic metabolic control allows key metabolic fluxes to be modulated in real time, enhancing bioprocess flexibility and expanding available optimization degrees of freedom. This is achieved, e.g., via targeted modulation of metabolic…

系统与控制 · 电气工程与系统科学 2025-10-03 Sebastián Espinel-Ríos , River Walser , Dongda Zhang

Learning solution operators for differential equations with neural networks has shown great potential in scientific computing, but ensuring their stability under input perturbations remains a critical challenge. This paper presents a robust…

机器学习 · 计算机科学 2026-01-13 Chutian Huang , Chang Ma , Kaibo Wang , Yang Xiang

Existing reinforcement learning (RL)-based post-training methods for large language models have advanced rapidly, yet their design has largely been guided by heuristics rather than systematic theoretical principles. This gap limits our…

机器学习 · 统计学 2026-01-16 Zixun Huang , Jiayi Sheng , Zeyu Zheng

This note studies the robust output feedback stabilization problem of a class of multi-input multi-output invertible nonlinear systems, for which an "ideal" state feedback based on feedback linearization can be designed under certain mild…

系统与控制 · 电气工程与系统科学 2021-01-07 Lei Wang , Christopher M. Kellett

This paper considers a class of bilinear systems with a neural network in the loop. These arise naturally when employing machine learning techniques to approximate general, non-affine in the input, control systems. We propose a controller…

系统与控制 · 电气工程与系统科学 2025-06-02 Dhruv Shah , Jorge Cortés

Stabilizing a dynamical system is a fundamental problem that serves as a cornerstone for many complex tasks in the field of control systems. The problem becomes challenging when the system model is unknown. Among the Reinforcement Learning…

系统与控制 · 电气工程与系统科学 2026-01-30 Ankang Zhang , Ming Chi , Xiaoling Wang , Lintao Ye

Reinforcement learning is a model-free optimal control method that optimizes a control policy through direct interaction with the environment. For reaching tasks that end in regulation, popular discrete-action methods are not well suited…

机器人学 · 计算机科学 2021-06-23 Wouter Caarls