中文
相关论文

相关论文: Stabilizing reinforcement learning control: A modu…

200 篇论文

Deep model-based Reinforcement Learning (RL) has the potential to substantially improve the sample-efficiency of deep RL. While various challenges have long held it back, a number of papers have recently come out reporting success with deep…

机器学习 · 计算机科学 2020-12-04 Harm van Seijen , Hadi Nekoei , Evan Racah , Sarath Chandar

We have witnessed the emergence of several controller parameterizations and the corresponding synthesis methods, including Youla, system level, input-output, and many other new proposals. Meanwhile, under the same synthesis method, there…

最优化与控制 · 数学 2022-02-11 Shih-Hao Tseng

Complex mechanical systems such as vehicle powertrains are inherently subject to multiple nonlinearities and uncertainties arising from parametric variations. Modeling errors are therefore unavoidable, making the transfer of control systems…

系统与控制 · 电气工程与系统科学 2026-02-13 Heisei Yonezawa , Ansei Yonezawa , Itsuro Kajiwara

Online reinforcement learning is concerned with training an agent on-the-fly via dynamic interaction with the environment. Here, due to the specifics of the application, it is not generally possible to perform long pre-training, as it is…

系统与控制 · 电气工程与系统科学 2022-11-17 Grigory Yaremenko , Georgiy Malaniya , Pavel Osinenko

Reinforcement Learning (RL) has been shown to be effective in many scenarios. However, it typically requires the exploration of a sufficiently large number of state-action pairs, some of which may be unsafe. Consequently, its application to…

系统与控制 · 电气工程与系统科学 2022-06-24 Yousef Emam , Gennaro Notomista , Paul Glotfelter , Zsolt Kira , Magnus Egerstedt

We introduce the framework of performative reinforcement learning where the policy chosen by the learner affects the underlying reward and transition dynamics of the environment. Following the recent literature on performative…

机器学习 · 计算机科学 2023-06-08 Debmalya Mandal , Stelios Triantafyllou , Goran Radanovic

This paper revisits a classical challenge in the design of stabilizing controllers for nonlinear systems with a norm-bounded input constraint. By extending Lin-Sontag's universal formula and introducing a generic (state-dependent) scaling…

系统与控制 · 电气工程与系统科学 2026-04-22 Ming Li , Zhiyong Sun , Siep Weiland

Control design for general nonlinear robotic systems with guaranteed stability and/or safety in the presence of model uncertainties is a challenging problem. Recent efforts attempt to learn a controller and a certificate (e.g., a Lyapunov…

系统与控制 · 电气工程与系统科学 2025-06-05 Vivek Sharma , Pan Zhao , Naira Hovakimyan

Dynamical models identified from data are frequently employed in control system design. However, decoupling system identification from controller synthesis can result in situations where no suitable controller exists after a model has been…

系统与控制 · 电气工程与系统科学 2025-12-30 Sampath Kumar Mulagaleti , Alberto Bemporad

The convergence of policy gradient algorithms in reinforcement learning hinges on the optimization landscape of the underlying optimal control problem. Theoretical insights into these algorithms can often be acquired from analyzing those of…

机器学习 · 计算机科学 2023-11-01 Jingliang Duan , Wenhan Cao , Yang Zheng , Lin Zhao

In large-scale networks of uncertain dynamical systems, where communication is limited and there is a strong interaction among subsystems, learning local models and control policies offers great potential for designing high-performance…

系统与控制 · 电气工程与系统科学 2021-11-08 Andrea Carron , Jerome Sieber , Melanie N. Zeilinger

We consider the problem of stabilization of a linear system, under state and control constraints, and subject to bounded disturbances and unknown parameters in the state matrix. First, using a simple least square solution and available…

系统与控制 · 电气工程与系统科学 2020-07-22 Edouard Leurent , Denis Efimov , Odalric-Ambrym Maillard

Purpose: This study aims to address the challenges of controlling unstable and nonlinear systems by proposing an adaptive PID controller based on predictive reinforcement learning (PRL-PID), where the PRL-PID combines the advantages of both…

系统与控制 · 电气工程与系统科学 2025-06-11 Chaoqun Ma , Zhiyong Zhang

Robust stability and stochastic stability have separately seen intense study in control theory for many decades. In this work we establish relations between these properties for discrete-time systems and employ them for robust control…

动力系统 · 数学 2020-04-20 Benjamin Gravell , Peyman Mohajerin Esfahani , Tyler Summers

This work primarily focuses on an operator inference methodology aimed at constructing low-dimensional dynamical models based on a priori hypotheses about their structure, often informed by established physics or expert insights. Stability…

机器学习 · 计算机科学 2024-03-04 Igor Pontes Duff , Pawan Goyal , Peter Benner

Transient stability of power systems is becoming increasingly important because of the growing integration of renewable resources. These resources lead to a reduction in mechanical inertia but also provide increased flexibility in frequency…

系统与控制 · 电气工程与系统科学 2021-05-07 Wenqi Cui , Baosen Zhang

The ability to learn and execute optimal control policies safely is critical to realization of complex autonomy, especially where task restarts are not available and/or the systems are safety-critical. Safety requirements are often…

系统与控制 · 电气工程与系统科学 2021-10-06 S M Nahid Mahmud , Moad Abudia , Scott A Nivison , Zachary I. Bell , Rushikesh Kamalapurkar

Controllable multimodal generation is commonly formulated as an inference-time conditioning problem using prompts, guidance, or auxiliary modules. While effective, such approaches do not explicitly structure how semantic attributes evolve,…

计算机视觉与模式识别 · 计算机科学 2026-05-19 Jamuna S. Murthy , Amin Karimi Monsefi , Rajiv Ramnath

To meet the demands of instantaneous control of instabilities over long time horizons in plasma fusion, we design a dynamic feedback control strategy for the Vlasov-Poisson system by constructing an operator that maps state perturbations to…

数值分析 · 数学 2026-01-01 Jingcheng Lu , Li Wang , Jeff Calder

We introduce the notion of descriptor embedding for nonlinear systems and use it for the data-driven design of stabilizing controllers. Specifically, we provide sufficient data-dependent LMI conditions which, if feasible, return a…

最优化与控制 · 数学 2025-11-04 Mohammad Alsalti , Claudio De Persis , Victor G. Lopez , Matthias A. Müller