中文
相关论文

相关论文: HJB-based online safety-embedded critic learning f…

200 篇论文

This paper develops a model-based reinforcement learning (MBRL) framework for learning online the value function of an infinite-horizon optimal control problem while obeying safety constraints expressed as control barrier functions (CBFs).…

机器学习 · 计算机科学 2022-11-10 Max H. Cohen , Calin Belta

Many optimal control problems are formulated as two point boundary value problems (TPBVPs) with conditions of optimality derived from the Hamilton-Jacobi-Bellman (HJB) equations. In most cases, it is challenging to solve HJBs due to the…

最优化与控制 · 数学 2019-07-25 Sixiong You , Ran Dai , Ping Lu

In recent times, a variety of Reinforcement Learning (RL) algorithms have been proposed for optimal tracking problem of continuous time nonlinear systems with input constraints. Most of these algorithms are based on the notion of uniform…

系统与控制 · 电气工程与系统科学 2020-06-16 Amardeep Mishra , Satadal Ghosh

Autonomous systems have witnessed a rapid increase in their capabilities, but it remains a challenge for them to perform tasks both effectively and safely. The fact that performance and safety can sometimes be competing objectives renders…

系统与控制 · 电气工程与系统科学 2024-12-04 Hao Wang , Adityaya Dhande , Somil Bansal

We present a framework to \emph{certify} Hamilton--Jacobi (HJ) reachability learned by reinforcement learning (RL). Building on a discounted initial time \emph{travel-cost} formulation that makes small-step RL value iteration provably…

系统与控制 · 电气工程与系统科学 2026-02-19 Prashant Solanki , Isabelle El-Hajj , Jasper J. van Beers , Erik-Jan van Kampen , Coen C. de Visser

Optimal control strategies are often combined with safety certificates to ensure both performance and safety in safety-critical systems. A prominent example is combining Model Predictive Control (MPC) with Control Barrier Functions (CBF).…

系统与控制 · 电气工程与系统科学 2025-12-05 Kerim Dzhumageldyev , Filippo Airaldi , Azita Dabiri

In many safety-critical control systems, possibly opposing safety restrictions and control performance objectives arise. To confront such a conflict, this letter proposes a novel methodology that embeds safety into stability of control…

系统与控制 · 电气工程与系统科学 2021-08-23 Hassan Almubarak , Nader Sadegh , Evangelos A. Theodorou

Optimal feedback control with implicit Hamiltonians poses a fundamental challenge for learning-based value function methods due to the absence of closed-form optimal control laws. Recent work~\cite{gelphman2025end} introduced an implicit…

最优化与控制 · 数学 2026-04-28 Eric Gelphman , Deepanshu Verma , Nicole Tianjiao Yang , Stanley Osher , Samy Wu Fung

Feedback controllers for port-Hamiltonian systems reveal an intrinsic inverse optimality property since each passivating state feedback controller is optimal with respect to some specific performance index. Due to the nonlinear…

最优化与控制 · 数学 2020-07-20 Lukas Kölsch , Pol Jané Soneira , Felix Strehle , Sören Hohmann

This paper, which is the natural continuation of a previous paper by the same authors, studies a class of optimal control problems with state constraints where the state equation is a differential equation with delays. This class includes…

最优化与控制 · 数学 2009-07-10 Salvatore Federico , Ben Goldys , Fausto Gozzi

Microgrids have more operational flexibilities as well as uncertainties than conventional power grids, especially when renewable energy resources are utilized. An energy storage based feedback controller can compensate undesired dynamics of…

系统与控制 · 电气工程与系统科学 2022-03-10 Tianwei Xia , Kai Sun , Wei Kang

Designing optimal controllers for nonlinear dynamical systems often relies on reinforcement learning and adaptive dynamic programming (ADP) to approximate solutions of the Hamilton Jacobi Bellman (HJB) equation. However, these methods…

最优化与控制 · 数学 2025-11-27 Akash Vyas , Shreyas Kumar , Jayant Kumar Mohanta , Ravi Prakash

Recent developments in autonomous driving and robotics underscore the necessity of safety-critical controllers. Control barrier functions (CBFs) are a popular method for appending safety guarantees to a general control framework, but they…

机器人学 · 计算机科学 2025-05-21 Matthew Kim , William Sharpless , Hyun Joe Jeong , Sander Tonkens , Somil Bansal , Sylvia Herbert

We study a class of optimal control problems with state constraints where the state equation is a differential equation with delays. This class includes some problems arising in economics, in particular the so-called models with time to…

最优化与控制 · 数学 2009-07-09 Salvatore Federico , Ben Goldys , Fausto Gozzi

Safety assurance is a fundamental requirement for deploying learning-enabled autonomous systems. Hamilton-Jacobi (HJ) reachability analysis is a fundamental method for formally verifying safety and generating safe controllers. However,…

机器学习 · 计算机科学 2025-11-21 Ihab Tabbara , Yuxuan Yang , Hussein Sibai

In this paper, we propose Q-learning algorithms for continuous-time deterministic optimal control problems with Lipschitz continuous controls. Our method is based on a new class of Hamilton-Jacobi-Bellman (HJB) equations derived from…

机器学习 · 计算机科学 2020-10-28 Jeongho Kim , Jaeuk Shin , Insoon Yang

In this paper, we present a novel algorithm named synchronous integral Q-learning, which is based on synchronous policy iteration, to solve the continuous-time infinite horizon optimal control problems of input-affine system dynamics. The…

系统与控制 · 电气工程与系统科学 2021-05-20 Lei Guo , Han Zhao

We study the problem of generating control laws for systems with unknown dynamics. Our approach is to represent the controller and the value function with neural networks, and to train them using loss functions adapted from the…

机器人学 · 计算机科学 2023-02-21 Selim Engin , Volkan Isler

Learning-based control with safety guarantees usually requires real-time safety certification and modifications of possibly unsafe learning-based policies. The control barrier function (CBF) method uses a safety filter containing a…

系统与控制 · 电气工程与系统科学 2024-10-25 Kanghui He , Shengling Shi , Ton van den Boom , Bart De Schutter

Deep Reinforcement Learning (RL) has shown remarkable success in robotics with complex and heterogeneous dynamics. However, its vulnerability to unknown disturbances and adversarial attacks remains a significant challenge. In this paper, we…

机器人学 · 计算机科学 2024-10-01 Hanyang Hu , Xilun Zhang , Xubo Lyu , Mo Chen