中文
相关论文

相关论文: HJB-based online safety-embedded critic learning f…

200 篇论文

In this paper, we propose a novel image restoration framework that integrates optimal control techniques with the Hamilton-Jacobi-Bellman (HJB) equation. Motivated by models from production planning, our method restores degraded images by…

偏微分方程分析 · 数学 2025-05-13 Dragos-Patru Covei

This paper presents the optimal control and synchronization problem of a multilevel network of R\"ossler chaotic oscillators. Using the Hamilton-Jacobi-Bellman (HJB) technique, the optimal control law with three-state variables feedback is…

As autonomous robots move into complex, dynamic real-world environments, they must learn to navigate safely in real time, yet anticipating all possible behaviors is infeasible. We propose a composable, model-free reinforcement learning…

机器人学 · 计算机科学 2026-02-16 Xinhuan Sang , Abdelrahman Abdelgawad , Roberto Tron

Providing safety guarantees for learning-based controllers is important for real-world applications. One approach to realizing safety for arbitrary control policies is safety filtering. If necessary, the filter modifies control inputs to…

系统与控制 · 电气工程与系统科学 2023-12-18 Lukas Brunke , Siqi Zhou , Mingxuan Che , Angela P. Schoellig

Hamilton-Jacobi (HJ) reachability analysis is a widely adopted verification tool to provide safety and performance guarantees for autonomous systems. However, it involves solving a partial differential equation (PDE) to compute a safety…

机器人学 · 计算机科学 2025-05-12 Aditya Singh , Zeyuan Feng , Somil Bansal

Control barrier functions have shown great success in addressing control problems with safety guarantees. These methods usually find the next safe control input by solving an online quadratic programming problem. However, model uncertainty…

系统与控制 · 电气工程与系统科学 2020-11-24 Chuanzheng Wang , Yinan Li , Yiming Meng , Stephen L. Smith , Jun Liu

Learning-based approaches for controlling safety-critical systems are rapidly growing in popularity; thus, it is important to assure their performance and safety. Hamilton-Jacobi (HJ) reachability analysis is a popular formal verification…

机器人学 · 计算机科学 2024-04-11 Albert Lin , Somil Bansal

This paper proposes a safety-critical control design approach for nonlinear control affine systems in the presence of matched and unmatched uncertainties. Our constructive framework couples control barrier function (CBF) theory with a new…

系统与控制 · 电气工程与系统科学 2025-02-03 Ersin Das , Joel W. Burdick

In this paper, we present a framework for enabling autonomous vehicles to interact with cyclists in a manner that balances safety and optimality. The approach integrates Hamilton-Jacobi reachability analysis with deep Q-learning to jointly…

机器人学 · 计算机科学 2026-02-23 Aarati Andrea Noronha , Jean Oh

In this paper, a novel online, output-feedback, critic-only, model-based reinforcement learning framework is developed for safety-critical control systems operating in complex environments. The developed framework ensures system stability…

系统与控制 · 电气工程与系统科学 2024-06-28 Tochukwu Elijah Ogri , Muzaffar Qureshi , Zachary I. Bell , Rushikesh Kamalapurkar

In this paper, we study a time-inconsistent stochastic optimal control problem with a recursive cost functional by a multi-person hierarchical differential game approach. An equilibrium strategy of this problem is constructed and a…

最优化与控制 · 数学 2016-06-13 Qingmeng Wei , Jiongmin Yong , Zhiyong Yu

We study the problem of optimal portfolio selection under stochastic volatility within a continuous time reinforcement learning framework with portfolio constraints. Exploration is modeled through entropy-regularized relaxed controls, where…

数理金融 · 定量金融 2026-04-27 Thai Nguyen , Pertiny Nkuize

Recently, safe reinforcement learning (RL) with the actor-critic structure for continuous control tasks has received increasing attention. It is still challenging to learn a near-optimal control policy with safety and convergence…

机器学习 · 计算机科学 2024-02-06 Xinglong Zhang , Yaoqian Peng , Biao Luo , Wei Pan , Xin Xu , Haibin Xie

We propose and analyze a randomization scheme for a general class of impulse control problems. The solution to this randomized problem is characterized as the fixed point of a compound operator which consists of a regularized nonlocal…

最优化与控制 · 数学 2026-05-26 Haoyang Cao , Yuchao Dong , Zhouhao Yang

Autonomous vehicles face tremendous challenges while interacting with human drivers in different kinds of scenarios. Developing control methods with safety guarantees while performing interactions with uncertainty is an ongoing research…

机器人学 · 计算机科学 2021-04-30 Yiwei Lyu , Wenhao Luo , John M. Dolan

This note lays part of the theoretical ground for a definition of differential systems modeling reinforcement learning in continuous time non-Markovian rough environments. Specifically we focus on optimal relaxed control of rough equations…

最优化与控制 · 数学 2024-02-29 Prakash Chakraborty , Harsha Honnappa , Samy Tindel

Recent learning-based safety filters have outperformed conventional methods, such as hand-crafted Control Barrier Functions (CBFs), by effectively adapting to complex constraints. However, these learning-based approaches lack formal safety…

机器学习 · 计算机科学 2025-06-23 Jiaxing Li , Hanjiang Hu , Yujie Yang , Changliu Liu

Ensuring safe exploration in high-dimensional systems with unknown dynamics remains a significant challenge. Existing safe reinforcement learning methods often provide safety guarantees only in expectation, which can still lead to safety…

机器学习 · 计算机科学 2026-04-28 Rahul Narava , Siddharth Verma , Ojas Jain , Shashi Shekhar Jha , Mayank Shekhar Jha

Recent literature has proposed approaches that learn control policies with high performance while maintaining safety guarantees. Synthesizing Hamilton-Jacobi (HJ) reachable sets has become an effective tool for verifying safety and…

系统与控制 · 电气工程与系统科学 2024-08-23 Milan Ganai , Sicun Gao , Sylvia Herbert

A control theoretic approach is presented in this paper for both batch and instantaneous updates of weights in feed-forward neural networks. The popular Hamilton-Jacobi-Bellman (HJB) equation has been used to generate an optimal weight…

神经与进化计算 · 计算机科学 2015-04-29 Vipul Arora , Laxmidhar Behera , Ajay Pratap Yadav