English
Related papers

Related papers: Hamilton-Jacobi-Bellman Equations for Q-Learning i…

200 papers

This paper studies the q-learning, recently coined as the continuous time counterpart of Q-learning by Jia and Zhou (2023), for continuous time Mckean-Vlasov control problems in the setting of entropy-regularized reinforcement learning. In…

Machine Learning · Computer Science 2024-11-04 Xiaoli Wei , Xiang Yu

An optimal control problem is considered for a stochastic differential equation containing a state-dependent regime switching, with a recursive cost functional. Due to the non-exponential discounting in the cost functional, the problem is…

Optimization and Control · Mathematics 2017-12-29 Hongwei Mei , Jiongmin Yong

We develop a continuous-time reinforcement learning framework for a class of singular stochastic control problems without entropy regularization. The optimal singular control is characterized as the optimal singular control law, which is a…

Optimization and Control · Mathematics 2026-05-14 Zongxia Liang , Xiaodong Luo , Xiang Yu

We investigate in this work a fully-discrete semi-Lagrangian approximation of second order possibly degenerate Hamilton-Jacobi-Bellman (HJB) equations on a bounded domain with oblique boundary conditions. These equations appear naturally in…

Numerical Analysis · Mathematics 2021-09-22 Elisa Calzola , Elisabetta Carlini , Xavier Dupuis , Francisco J. Silva

This paper studies the continuous-time reinforcement learning (RL) for optimal switching problems across multiple regimes. We consider a type of exploratory formulation under entropy regularization where the agent randomizes both the timing…

Optimization and Control · Mathematics 2025-12-23 Yijie Huang , Mengge Li , Xiang Yu , Zhou Zhou

In this work we investigate regularity properties of a large class of Hamilton-Jacobi-Bellman (HJB) equations with or without obstacles, which can be stochastically interpreted in form of a stochastic control system which nonlinear cost…

Probability · Mathematics 2012-02-08 Rainer Buckdahn , Jianhui Huang , Juan Li

Deterministic optimal impulse control problem with terminal state constraint is considered. Due to the appearance of the terminal state constraint, the value function might be discontinuous in general. The main contribution of this paper is…

Optimization and Control · Mathematics 2020-11-10 Yue Zhou , Xinwei Feng , Jiongmin Yong

In this paper we study stochastic optimal control problems of fully coupled forward-backward stochastic differential equations (FBSDEs). The recursive cost functionals are defined by controlled fully coupled FBSDEs. We study two cases of…

Optimization and Control · Mathematics 2013-02-06 Juan Li , Qingmeng Wei

For pricing American options, %after suitable discretization in space and time, a sequence of discrete linear complementarity problems (LCPs) or equivalently Hamilton-Jacobi-Bellman (HJB) equations need to be solved in a sequential…

Numerical Analysis · Mathematics 2024-05-15 Xian-Ming Gu , Jun Liu , Cornelis W. Oosterlee

In this paper we propose a new computational method for designing optimal regulators for high-dimensional nonlinear systems. The proposed approach leverages physics-informed machine learning to solve high-dimensional Hamilton-Jacobi-Bellman…

Optimization and Control · Mathematics 2021-04-09 Tenavi Nakamura-Zimmerer , Qi Gong , Wei Kang

We introduce a stochastic version of the optimal transport problem. We provide an analysis by means of the study of the associated Hamilton-Jacobi-Bellman equation, which is set on the set of probability measures. We introduce a new…

Analysis of PDEs · Mathematics 2024-05-22 Charles Bertucci

The approximation of solutions to second order Hamilton--Jacobi--Bellman (HJB) equations by deep neural networks is investigated. It is shown that for HJB equations that arise in the context of the optimal control of certain Markov…

Numerical Analysis · Mathematics 2021-03-11 Philipp Grohs , Lukas Herrmann

We study a family of stationary Hamilton-Jacobi-Bellman (HJB) equations in Hilbert spaces arising from stochastic optimal control problems. The main difficulties to treat such problems are: the lack of smoothing properties of the linear…

Optimization and Control · Mathematics 2025-10-31 Gabriele Bolli , Fausto Gozzi

We propose \emph{Choquet regularizers} to measure and manage the level of exploration for reinforcement learning (RL), and reformulate the continuous-time entropy-regularized RL problem of Wang et al. (2020, JMLR, 21(198)) in which we…

Machine Learning · Statistics 2022-08-19 Xia Han , Ruodu Wang , Xun Yu Zhou

This work concerns the optimal control problem for McKean-Vlasov SDEs. In order to characterize the value function, we develop the viscosity solution theory for Hamilton-Jacobi-Bellman (HJB) equations on the Wasserstein space using…

Probability · Mathematics 2023-10-19 Jinghai Shao

In this paper, we show that the value functions of mean field control problems with common noise are the unique viscosity solutions to fully second-order Hamilton-Jacobi-Bellman equations, in a Crandall-Lions-like framework. We allow the…

Optimization and Control · Mathematics 2025-01-06 Erhan Bayraktar , Hang Cheung , Ibrahim Ekren , Jinniao Qiu , Ho Man Tai , Xin Zhang

In this paper, we study the optimal stopping problem in the so-called exploratory framework, in which the agent takes actions randomly conditioning on current state and an entropy-regularized term is added to the reward functional. Such a…

Optimization and Control · Mathematics 2023-09-04 Yuchao Dong

Solving the Hamilton-Jacobi-Bellman equation is important in many domains including control, robotics and economics. Especially for continuous control, solving this differential equation and its extension the Hamilton-Jacobi-Isaacs…

Robotics · Computer Science 2021-10-06 Michael Lutter , Boris Belousov , Shie Mannor , Dieter Fox , Animesh Garg , Jan Peters

We study optimal control problems for interacting branching diffusion processes, a class of measure-valued dynamics capturing both spatial motion and branching mechanisms. From the perspective of the dynamic programming principle, we…

Optimization and Control · Mathematics 2026-01-19 Antonio Ocello

This paper presents an implicit solution formula for the Hamilton-Jacobi partial differential equation (HJ PDE). The formula is derived using the method of characteristics and is shown to coincide with the Hopf and Lax formulas in the case…

Machine Learning · Computer Science 2025-02-03 Yesom Park , Stanley Osher