中文
相关论文

相关论文: HJB-based online safety-embedded critic learning f…

200 篇论文

For continuous systems modeled by dynamical equations such as ODEs and SDEs, Bellman's Principle of Optimality takes the form of the Hamilton-Jacobi-Bellman (HJB) equation, which provides the theoretical target of reinforcement learning…

机器学习 · 计算机科学 2025-10-28 Haruki Settai , Naoya Takeishi , Takehisa Yairi

In this paper time-driven learning refers to the machine learning method that updates parameters in a prediction model continuously as new data arrives. Among existing approximate dynamic programming (ADP) and reinforcement learning (RL)…

系统与控制 · 电气工程与系统科学 2020-06-17 Qingtao Zhao , Jennie Si , Jian Sun

Safety assurance is critical in the planning and control of robotic systems. For robots operating in the real world, the safety-critical design often needs to explicitly address uncertainties and the pre-computed guarantees often rely on…

机器人学 · 计算机科学 2024-07-09 Hao Zhou , Yanze Zhang , Wenhao Luo

We introduce a novel extension to robust control theory that explicitly addresses uncertainty in the value function's gradient, a form of uncertainty endemic to applications like reinforcement learning where value functions are…

机器学习 · 计算机科学 2025-07-22 Qian Qi

Safety is of paramount importance in control systems to avoid costly risks and catastrophic damages. The control barrier function (CBF) method, a promising solution for safety-critical control, poses a new challenge of enhancing control…

系统与控制 · 电气工程与系统科学 2025-03-26 Shengbo Wang , Ke Li , Zheng Yan , Zhenyuan Guo , Song Zhu , Guanghui Wen , Shiping Wen

A learning technique for finite horizon optimal control problems and its approximation based on polynomials is analyzed. It allows to circumvent, in part, the curse dimensionality which is involved when the feedback law is constructed by…

最优化与控制 · 数学 2023-02-21 Karl Kunisch , Donato Vásquez-Varas

In this paper, we explore a new class of stochastic control problems characterized by specific control constraints. Specifically, the admissible controls are subject to the ratcheting constraint, meaning they must be non-decreasing over…

最优化与控制 · 数学 2024-12-17 Mingxin Guo , Zuo Quan Xu

Adaptive control has focused on online control of dynamic systems in the presence of parametric uncertainties, with solutions guaranteeing stability and control performance. Safety, a related property to stability, is becoming increasingly…

系统与控制 · 电气工程与系统科学 2023-09-12 Johannes Autenrieb , Anuradha M. Annaswamy

This paper addresses learning safe output feedback control laws from partial observations of expert demonstrations. We assume that a model of the system dynamics and a state estimator are available along with corresponding error bounds,…

系统与控制 · 电气工程与系统科学 2024-04-04 Lars Lindemann , Alexander Robey , Lejun Jiang , Satyajeet Das , Stephen Tu , Nikolai Matni

Learning-based control has recently shown great efficacy in performing complex tasks for various applications. However, to deploy it in real systems, it is of vital importance to guarantee the system will stay safe. Control Barrier…

系统与控制 · 电气工程与系统科学 2024-09-05 Fernando Castañeda , Jason J. Choi , Wonsuhk Jung , Bike Zhang , Claire J. Tomlin , Koushil Sreenath

An optimal control problem is considered for a stochastic differential equation containing a state-dependent regime switching, with a recursive cost functional. Due to the non-exponential discounting in the cost functional, the problem is…

最优化与控制 · 数学 2017-12-29 Hongwei Mei , Jiongmin Yong

Racing demands each vehicle to drive at its physical limits, when any safety infraction could lead to catastrophic failure. In this work, we study the problem of safe reinforcement learning (RL) for autonomous racing, using the vehicle's…

机器人学 · 计算机科学 2021-12-02 Bingqing Chen , Jonathan Francis , Jean Oh , Eric Nyberg , Sylvia L. Herbert

This study introduces a mathematical framework to investigate the viability and reachability of production systems under constraints. We develop a model that incorporates key decision variables, such as pricing policy, quality investment,…

最优化与控制 · 数学 2025-09-16 Achraf Bouhmady , Mustapha Serhani , Nadia Raissi

We address the problem of safe policy learning in multi-agent safety-critical autonomous systems. In such systems, it is necessary for each agent to meet the safety requirements at all times while also cooperating with other agents to…

Ensuring safety in the sense of constraint satisfaction for learning-based control is a critical challenge, especially in the model-free case. While safety filters address this challenge in the model-based setting by modifying unsafe…

系统与控制 · 电气工程与系统科学 2026-01-15 Kanghui He , Shengling Shi , Ton van den Boom , Bart De Schutter

In advanced manufacturing, strict safety guarantees are required to allow humans and robots to work together in a shared workspace. One of the challenges in this application field is the variety and unpredictability of human behavior,…

机器人学 · 计算机科学 2023-08-22 Dianhao Zhang , Mien Van , Stephen Mcllvanna , Yuzhu Sun , Seán McLoone

The need for robust control laws is especially important in safety-critical applications. We propose robust hybrid control barrier functions as a means to synthesize control laws that ensure robust safety. Based on this notion, we formulate…

系统与控制 · 电气工程与系统科学 2021-05-14 Alexander Robey , Lars Lindemann , Stephen Tu , Nikolai Matni

This paper addresses the problem of safety-critical control for non-affine control systems. It has been shown that optimizing quadratic costs subject to state and control constraints can be sub-optimally reduced to a sequence of quadratic…

系统与控制 · 电气工程与系统科学 2024-02-15 Wei Xiao , Ross Allen , Daniela Rus

Hybrid dynamical systems are ubiquitous as practical robotic applications often involve both continuous states and discrete switchings. Safety is a primary concern for hybrid robotic systems. Existing safety-critical control approaches for…

机器人学 · 计算机科学 2024-12-02 Shuo Yang , Yu Chen , Xiang Yin , George J. Pappas , Rahul Mangharam

Reinforcement Learning (RL) algorithms have found limited success beyond simulated applications, and one main reason is the absence of safety guarantees during the learning process. Real world systems would realistically fail or break…

机器学习 · 计算机科学 2019-03-22 Richard Cheng , Gabor Orosz , Richard M. Murray , Joel W. Burdick