中文
相关论文

相关论文: Certifying Stability of Reinforcement Learning Pol…

200 篇论文

Reinforcement learning-based controller design methods often require substantial data in the initial training phase. Moreover, the training process tends to exhibit strong randomness and slow convergence. It often requires considerable time…

系统与控制 · 电气工程与系统科学 2025-09-24 Chenxu Ke , Congling Tian , Kaichen Xu , Ye Li , Lingcong Bao

We study gradient descent for rank-1 matrix factorization through a certificate-based viewpoint. The central object is a parameterized quadratic certificate $I(\delta;\,\cdot)$ whose level sets shrink along the dynamics, thereby inducing a…

数值分析 · 数学 2026-05-01 Jaehong Moon

This paper develops an approach to learn a policy of a dynamical system that is guaranteed to be both provably safe and goal-reaching. Here, the safety means that a policy must not drive the state of the system to any unsafe region, while…

系统与控制 · 电气工程与系统科学 2020-06-16 Wanxin Jin , Zhaoran Wang , Zhuoran Yang , Shaoshuai Mou

This paper addresses the stability problem for discrete-time switched systems under autonomous switching. Each mode of the switched system is modeled as a Linear Parameter Varying (LPV) system, the time-varying parameters can vary…

系统与控制 · 电气工程与系统科学 2020-05-13 Márcio J. Lacerda , Cristiano M. Agulhari

Model-free Deep Reinforcement Learning (DRL) controllers have demonstrated promising results on various challenging non-linear control tasks. While a model-free DRL algorithm can solve unknown dynamics and high-dimensional problems, it…

机器人学 · 计算机科学 2022-03-03 Zikang Xiong , Joe Eappen , Ahmed H. Qureshi , Suresh Jagannathan

Reinforcement learning (RL) is a powerful data-driven control method that has been largely explored in autonomous driving tasks. However, conventional RL approaches learn control policies through trial-and-error interactions with the…

机器人学 · 计算机科学 2021-11-03 Tianyu Shi , Dong Chen , Kaian Chen , Zhaojian Li

The problem of safely learning and controlling a dynamical system - i.e., of stabilizing an originally (partially) unknown system while ensuring that it does not leave a prescribed 'safe set' - has recently received tremendous attention in…

系统与控制 · 电气工程与系统科学 2023-10-10 Jafar Abbaszadeh Chekan , Cedric Langbort

Ensuring the safety of reinforcement learning (RL) algorithms is crucial to unlock their potential for many real-world tasks. However, vanilla RL and most safe RL approaches do not guarantee safety. In recent years, several methods have…

机器学习 · 计算机科学 2023-11-21 Hanna Krasowski , Jakob Thumm , Marlon Müller , Lukas Schäfer , Xiao Wang , Matthias Althoff

This paper considers the problem of solving constrained reinforcement learning (RL) problems with anytime guarantees, meaning that the algorithmic solution must yield a constraint-satisfying policy at every iteration of its evolution. Our…

系统与控制 · 电气工程与系统科学 2025-10-03 Pol Mestres , Arnau Marzabal , Jorge Cortés

Deploying Reinforcement Learning (RL) agents in the real-world require that the agents satisfy safety constraints. Current RL agents explore the environment without considering these constraints, which can lead to damage to the hardware or…

机器学习 · 计算机科学 2021-03-17 Harshit Sikchi , Wenxuan Zhou , David Held

We propose a sampling-based approach to learn Lyapunov functions for a class of discrete-time autonomous hybrid systems that admit a mixed-integer representation. Such systems include autonomous piecewise affine systems, closed-loop…

最优化与控制 · 数学 2020-12-23 Shaoru Chen , Mahyar Fazlyab , Manfred Morari , George J. Pappas , Victor M. Preciado

Robust stabilization conditions for uncertain switched affine systems subject to a unitary input delay are presented. They are obtained through the Lyapunov framework and a min-switching state-feedback predictive control law. The result…

系统与控制 · 电气工程与系统科学 2026-04-20 Gerson Portilla , Carolina Albea , Alexandre Seuret

PID control has been the dominant control strategy in the process industry due to its simplicity in design and effectiveness in controlling a wide range of processes. However, traditional methods on PID tuning often require extensive domain…

系统与控制 · 电气工程与系统科学 2022-02-14 Ayub I. Lakhani , Myisha A. Chowdhury , Qiugang Lu

Most of nonlinear robust control methods just consider the affine nonlinear nominal model. When the nominal model is assumed to be affine nonlinear, available information about existing non-affine nonlinearities is ignored. For non-affine…

系统与控制 · 电气工程与系统科学 2019-12-30 Chaolun Lu , Yongqiang Li , Zhongsheng Hou , Yuanjing Feng , Yu Feng , Ronghu Chi , Xuhui Bu

In safe Reinforcement Learning (RL), safety cost is typically defined as a function dependent on the immediate state and actions. In practice, safety constraints can often be non-Markovian due to the insufficient fidelity of state…

机器学习 · 计算机科学 2024-05-07 Siow Meng Low , Akshat Kumar

Reinforcement learning (RL) methods have demonstrated their efficiency in simulation environments. However, many applications for which RL offers great potential, such as autonomous driving, are also safety critical and require a certified…

系统与控制 · 电气工程与系统科学 2021-01-19 Kim P. Wabersich , Lukas Hewing , Andrea Carron , Melanie N. Zeilinger

Many modern nonlinear control methods aim to endow systems with guaranteed properties, such as stability or safety, and have been successfully applied to the domain of robotics. However, model uncertainty remains a persistent challenge,…

机器人学 · 计算机科学 2020-11-20 Andrew J. Taylor , Victor D. Dorobantu , Hoang M. Le , Yisong Yue , Aaron D. Ames

Reinforcement learning (RL) is recognized as lacking generalization and robustness under environmental perturbations, which excessively restricts its application for real-world robotics. Prior work claimed that adding regularization to the…

机器学习 · 计算机科学 2023-12-06 Yuan Zhang , Jianhong Wang , Joschka Boedecker

In systems where the ability to actuate is a scarce resource, e.g., spacecrafts, it is desirable to only apply a given controller in an intermittent manner--with periods where the controller is on and periods where it is off. Motivated by…

系统与控制 · 电气工程与系统科学 2022-04-08 Pio Ong , Gilbert Bahati , Aaron D. Ames

Reinforcement Learning (RL) has been shown to be effective and convenient for a number of tasks in robotics. However, it requires the exploration of a sufficiently large number of state-action pairs, many of which may be unsafe or…

‹ 上一页 1 8 9 10 下一页 ›