English
Related papers

Related papers: A Temporal Difference Method for Stochastic Contin…

200 papers

This paper aims to explore the relationship between maximum principle and dynamic programming principle for stochastic recursive control problem with random coefficients. Under certain regular conditions for the coefficients, the…

Optimization and Control · Mathematics 2020-12-10 Yuchao Dong , Qingxin Meng , Qi Zhang

Designing optimal controllers for nonlinear dynamical systems often relies on reinforcement learning and adaptive dynamic programming (ADP) to approximate solutions of the Hamilton Jacobi Bellman (HJB) equation. However, these methods…

Optimization and Control · Mathematics 2025-11-27 Akash Vyas , Shreyas Kumar , Jayant Kumar Mohanta , Ravi Prakash

This paper is concerned with stochastic impulse control problems in which the running cost changes depending on the impulse control. Because of such a dependence, it brings several difficulties when the usual dynamic programming principle…

Optimization and Control · Mathematics 2025-11-11 Yuchen Cao , Jiongmin Yong

We unify Hamilton-Jacobi (HJ) reachability and Reinforcement Learning (RL) through a proposed running cost formulation. We prove that the resultant travel-cost value function is the unique bounded viscosity solution of a time-dependent…

Systems and Control · Electrical Eng. & Systems 2026-05-12 Prashant Solanki , Isabelle El-Hajj , Jasper van Beers , Erik-Jan van Kampen , Coen de Visser

In this paper we present a novel sampling-based numerical scheme designed to solve a certain class of stochastic optimal control problems, utilizing forward and backward stochastic differential equations (FBSDEs). By means of a nonlinear…

Systems and Control · Computer Science 2020-06-18 Ioannis Exarchos , Evangelos A. Theodorou

This paper is concerned with an optimal control problem for a forward-backward stochastic differential equation (FBSDE, for short) with a recursive cost functional determined by a backward stochastic Volterra integral equation (BSVIE, for…

Optimization and Control · Mathematics 2022-09-20 Hanxiao Wang , Jiongmin Yong , Chao Zhou

Continuous-time stochastic processes underlie many natural and engineered systems. In healthcare, autonomous driving, and industrial control, direct interaction with the environment is often unsafe or impractical, motivating offline…

Machine Learning · Statistics 2025-11-14 Nicolas Hoischen , Petar Bevanda , Max Beier , Stefan Sosnowski , Boris Houska , Sandra Hirche

Despite impressive results, reinforcement learning (RL) suffers from slow convergence and requires a large variety of tuning strategies. In this paper, we investigate the ability of RL algorithms on simple continuous control tasks. We show…

Robotics · Computer Science 2024-02-16 Daniel Layeghi , Steve Tonneau , Michael Mistry

We propose a new probabilistic numerical scheme for fully nonlinear equation of Hamilton-Jacobi-Bellman (HJB) type associated to stochastic control problem, which is based on the Feynman-Kac representation in [12] by means of control…

Probability · Mathematics 2019-06-28 Idris Kharroubi , Nicolas Langrené , Huyên Pham

Recent studies have extended the use of the stochastic Hamilton-Jacobi-Bellman (HJB) equation to include complex variables for deriving quantum mechanical equations. However, these studies often assume that it is valid to apply the HJB…

Quantum Physics · Physics 2024-10-14 Vasil Yordanov

We develop a discrete analogue of Hamilton-Jacobi theory in the framework of discrete Hamiltonian mechanics. The resulting discrete Hamilton-Jacobi equation is discrete only in time. We describe a discrete analogue of Jacobi's solution and…

Optimization and Control · Mathematics 2011-08-15 Tomoki Ohsawa , Anthony M. Bloch , Melvin Leok

This paper studies the optimal dividend problem with a bounded payout rate in a partially observed regime-switching diffusion model, where, in practice, the market regime is unobserved and key model parameters are unknown. To address this…

Optimization and Control · Mathematics 2026-01-29 Zhongqin Gao , Yan Lv , Jingmin He

We propose a novel data-driven neural network (NN) optimization framework for solving an optimal stochastic control problem under stochastic constraints. Customized activation functions for the output layers of the NN are applied, which…

Optimization and Control · Mathematics 2023-06-21 Marc Chen , Mohammad Shirazi , Peter A. Forsyth , Yuying Li

In this paper, we study a stochastic recursive optimal control problem in which the value functional is defined by the solution of a backward stochastic differential equation (BSDE) under $\tilde{G}$-expectation. Under standard assumptions,…

Optimization and Control · Mathematics 2021-06-08 Mingshang Hu , Shaolin Ji , Xiaojuan Li

We study the structure of a simple dynamic optimization problem consisting of one state and one control variable, from a physicist's point of view. By using an analogy to a physical model, we study this system in the classical and quantum…

Mathematical Finance · Quantitative Finance 2017-04-05 Mauricio Contreras , Rely Pellicer , Marcelo Villena

Non-Markovian dynamics are commonly found in real-world environments due to long-range dependencies, partial observability, and memory effects. The Bellman equation that is the central pillar of Reinforcement learning (RL) becomes only…

Machine Learning · Computer Science 2026-02-09 Zuyuan Zhang , Sizhe Tang , Tian Lan

In this paper, we present a novel method for computing the optimal feedback gain of the infinite-horizon Linear Quadratic Regulator (LQR) problem via an ordinary differential equation. We introduce a novel continuous-time Bellman error,…

Systems and Control · Electrical Eng. & Systems 2026-04-17 Armin Gießler , Albertus Johannes Malan , Sören Hohmann

Learning optimal feedback control laws capable of executing optimal trajectories is essential for many robotic applications. Such policies can be learned using reinforcement learning or planned using optimal control. While reinforcement…

Machine Learning · Computer Science 2019-10-14 Michael Lutter , Boris Belousov , Kim Listmann , Debora Clever , Jan Peters

We address two major challenges in scientific machine learning (SciML): interpretability and computational efficiency. We increase the interpretability of certain learning processes by establishing a new theoretical connection between…

Machine Learning · Computer Science 2024-05-08 Paula Chen , Tingwei Meng , Zongren Zou , Jérôme Darbon , George Em Karniadakis

Stochastic optimal principle leads to the resolution of a partial differential equation (PDE), namely the Hamilton-Jacobi-Bellman (HJB) equation. In general, this equation cannot be solved analytically, thus numerical algorithms are the…

Numerical Analysis · Mathematics 2021-09-14 Christelle Dleuna Nyoumbi , Antoine Tambue