English
Related papers

Related papers: Unifying Hamilton-Jacobi Reachability and Reinforc…

200 papers

Optimal control and the associated second-order path-dependent Hamilton-Jacobi-Bellman (PHJB) equation are studied for unbounded functional stochastic evolution systems in Hilbert spaces. The notion of viscosity solution without…

Optimization and Control · Mathematics 2024-02-27 Shanjian Tang , Jianjun Zhou

We study a stochastic control problem on a bounded domain, which arises from a continuous-time optimal management model. Via the corresponding Hamilton-Jacobi-Bellman equation the value function is shown to be jointly continuous and to…

Probability · Mathematics 2017-10-24 Ruoting Gong , Christian Houdré

We introduce a novel extension to robust control theory that explicitly addresses uncertainty in the value function's gradient, a form of uncertainty endemic to applications like reinforcement learning where value functions are…

Machine Learning · Computer Science 2025-07-22 Qian Qi

We present a simple and easy to implement method for the numerical solution of a rather general class of Hamilton-Jacobi-Bellman (HJB) equations. In many cases, the considered problems have only a viscosity solution, to which, fortunately,…

Computational Finance · Quantitative Finance 2011-02-17 Jan Hendrik Witte , Christoph Reisinger

Safe reinforcement learning (RL) that solves constraint-satisfactory policies provides a promising way to the broader safety-critical applications of RL in real-world problems such as robotics. Among all safe RL approaches, model-based…

Robotics · Computer Science 2022-10-17 Dongjie Yu , Wenjun Zou , Yujie Yang , Haitong Ma , Shengbo Eben Li , Jingliang Duan , Jianyu Chen

Offline goal-conditioned reinforcement learning (GCRL) learns goal-conditioned policies from static pre-collected datasets. However, accurate value estimation remains a challenge due to the limited coverage of the state-action space. Recent…

Machine Learning · Computer Science 2026-02-27 Hrishikesh Viswanath , Juanwu Lu , S. Talha Bukhari , Damon Conover , Ziran Wang , Aniket Bera

We study policy iteration (PI) for deterministic infinite-horizon discounted optimal control problems, whose value function is characterized by a stationary Hamilton--Jacobi--Bellman (HJB) equation. At the PDE level, PI is fundamentally…

Optimization and Control · Mathematics 2026-04-14 Namkyeong Cho , Yeoneung Kim

Recent results in the study of the Hamilton Jacobi Bellman (HJB) equation have led to the discovery of a formulation of the value function as a linear Partial Differential Equation (PDE) for stochastic nonlinear systems with a mild…

Optimization and Control · Mathematics 2014-02-13 Matanya B. Horowitz , Joel W. Burdick

In this paper, we present a novel method for computing the optimal feedback gain of the infinite-horizon Linear Quadratic Regulator (LQR) problem via an ordinary differential equation. We introduce a novel continuous-time Bellman error,…

Systems and Control · Electrical Eng. & Systems 2026-04-17 Armin Gießler , Albertus Johannes Malan , Sören Hohmann

Autonomous systems like aircraft and assistive robots often operate in scenarios where guaranteeing safety is critical. Methods like Hamilton-Jacobi reachability can provide guaranteed safe sets and controllers for such systems. However,…

We study a random process on R n moving in straight lines and changing randomly its velocity at random exponential times. We focus more precisely on the Kolmogorov equation in the hyperbolic scale (t, x, v) $\to$ t $\epsilon$, x $\epsilon$,…

Analysis of PDEs · Mathematics 2016-08-08 Nils Caillerie

Constrained reinforcement learning (CRL) has gained significant interest recently, since safety constraints satisfaction is critical for real-world problems. However, existing CRL methods constraining discounted cumulative costs generally…

Machine Learning · Computer Science 2022-06-08 Dongjie Yu , Haitong Ma , Shengbo Eben Li , Jianyu Chen

Merton portfolio management problem is studied in this paper within a stochastic volatility, non constant time discount rate, and power utility framework. This problem is time inconsistent and the way out of this predicament is to consider…

Portfolio Management · Quantitative Finance 2024-02-09 Oumar Mbodji , Traian A. Pirvu

Safe value functions, such as control barrier functions, characterize a safe set and synthesize a safety filter, overriding unsafe actions, for a dynamic system. While function approximators like neural networks can synthesize approximately…

Robotics · Computer Science 2024-09-10 Sander Tonkens , Alex Toofanian , Zhizhen Qin , Sicun Gao , Sylvia Herbert

The goal of this paper is to prove a comparison principle for viscosity solutions of semilinear Hamilton-Jacobi equations in the space of probability measures. The method involves leveraging differentiability properties of the…

Analysis of PDEs · Mathematics 2023-08-30 Samuel Daudin , Benjamin Seeger

Following the recent resurgence in establishing linear control theoretic benchmarks for reinforcement leaning (RL)-based policy optimization (PO) for complex dynamical systems with continuous state and action spaces, an optimal control…

Systems and Control · Electrical Eng. & Systems 2023-06-30 Leilei Cui , Lekan Molu

Autonomous systems have witnessed a rapid increase in their capabilities, but it remains a challenge for them to perform tasks both effectively and safely. The fact that performance and safety can sometimes be competing objectives renders…

Systems and Control · Electrical Eng. & Systems 2024-12-04 Hao Wang , Adityaya Dhande , Somil Bansal

This paper investigates the optimal control problems for the finite-horizon continuous-time Markov decision processes with delay-dependent control policies. We develop compactification methods in decision processes, and show that the…

Probability · Mathematics 2023-07-06 Zhong-Wei Liao , Jinghai Shao

The ergodic control problem for a non-degenerate controlled diffusion controlled through its drift is considered under a uniform stability condition that ensures the well-posedness of the associated Hamilton-Jacobi-Bellman (HJB) equation. A…

Optimization and Control · Mathematics 2019-03-20 Ari Arapostathis , Vivek S. Borkar

We consider stochastic impulse control problems when the impulses cost functions are arbitrary. We use the dynamic programming principle and viscosity solutions approach to show that the value function is a unique viscosity solution for the…

Optimization and Control · Mathematics 2019-01-17 Brahim El Asri , Sehail Mazid