English
Related papers

Related papers: Hamilton-Jacobi Deep Q-Learning for Deterministic …

200 papers

Neural network approaches that parameterize value functions have succeeded in approximating high-dimensional optimal feedback controllers when the Hamiltonian admits explicit formulas. However, many practical problems, such as the space…

Optimization and Control · Mathematics 2025-10-08 Eric Gelphman , Deepanshu Verma , Nicole Tianjiao Yang , Stanley Osher , Samy Wu Fung

The optimal \(H_{\infty}\) control problem over an infinite time horizon, which incorporates a performance function with a discount factor \(e^{-\alpha t}\) (\(\alpha > 0\)), is important in various fields. Solving this optimal…

Optimization and Control · Mathematics 2024-10-04 Guoyuan Chen , Yi Wang , Qinglong Zhou

Commonly in reinforcement learning (RL), rewards are discounted over time using an exponential function to model time preference, thereby bounding the expected long-term reward. In contrast, in economics and psychology, it has been shown…

Machine Learning · Computer Science 2022-12-08 Matthias Schultheis , Constantin A. Rothkopf , Heinz Koeppl

We consider the portfolio optimisation problem where the terminal function is an S-shaped utility applied at the difference between the wealth and a random benchmark process. We develop several numerical methods for solving the problem…

Computational Finance · Quantitative Finance 2024-10-10 Ashley Davey , Harry Zheng

In an episodic Markov Decision Process (MDP) problem, an online algorithm chooses from a set of actions in a sequence of $H$ trials, where $H$ is the episode length, in order to maximize the total payoff of the chosen actions. Q-learning,…

Machine Learning · Computer Science 2019-07-11 Xu Zhu

We consider the problem of time-optimal path planning for simple nonholonomic vehicles. In previous similar work, the vehicle has been simplified to a point mass and the obstacles have been stationary. Our formulation accounts for a…

Optimization and Control · Mathematics 2021-11-22 Christian Parkinson , Madeline Ceccia

We study policy iteration (PI) for deterministic infinite-horizon discounted optimal control problems, whose value function is characterized by a stationary Hamilton--Jacobi--Bellman (HJB) equation. At the PDE level, PI is fundamentally…

Optimization and Control · Mathematics 2026-04-14 Namkyeong Cho , Yeoneung Kim

Optimal control and the associated second-order Hamilton-Jacobi-Bellman (HJB) equation are studied for unbounded stochastic evolution systems in Hilbert spaces. A new notion of viscosity solution, featured by absence of B-continuity, is…

Optimization and Control · Mathematics 2026-02-10 Shanjian Tang , Jianjun Zhou

We study the properties of the value function associated with an optimal control problem with uncertainties, known as average or Riemann-Stieltjes problem. Uncertainties are assumed to belong to a compact metric probability space, and…

Optimization and Control · Mathematics 2024-07-19 M. Soledad Aronna , Michele Palladino , Oscar Sierra

In this work, we propose a class of numerical schemes for solving semilinear Hamilton-Jacobi-Bellman-Isaacs (HJBI) boundary value problems which arise naturally from exit time problems of diffusion processes with controlled drift. We…

Numerical Analysis · Mathematics 2020-02-14 Kazufumi Ito , Christoph Reisinger , Yufei Zhang

We revisit the linear programming approach to deterministic, continuous time, infinite horizon discounted optimal control problems. In the first part, we relax the original problem to an infinite-dimensional linear program over a measure…

Optimization and Control · Mathematics 2017-06-08 Angeliki Kamoutsi , Tobias Sutter , Peyman Mohajerin Esfahani , John Lygeros

We study a stochastic optimal control problem for a partially observed diffusion. By using the control randomization method in [4], we prove a corresponding randomized dynamic programming principle (DPP) for the value function, which is…

Probability · Mathematics 2016-09-12 Elena Bandini , Andrea Cosso , Marco Fuhrman , Huyên Pham

We consider Hamilton Jacobi Bellman equations in an inifinite dimensional Hilbert space, with quadratic (respectively superquadratic) hamiltonian and with continuous (respectively lipschitz continuous) final conditions. This allows to study…

Probability · Mathematics 2013-04-10 Federica Masiero

This paper studies the stochastic optimal control of jump-diffusion processes and the associated fully nonlinear backward stochastic Hamilton--Jacobi--Bellman (BSHJB) equations. We establish the dynamic programming principle (DPP) via…

Optimization and Control · Mathematics 2026-05-21 Dunxiang Liang , Qingxin Meng

The Bellman equation and its continuous-time counterpart, the Hamilton-Jacobi-Bellman (HJB) equation, serve as necessary conditions for optimality in reinforcement learning and optimal control. While the value function is known to be the…

Machine Learning · Computer Science 2025-03-07 Haoxiang You , Lekan Molu , Ian Abraham

In this paper we study stochastic optimal control problems of fully coupled forward-backward stochastic differential equations (FBSDEs). The recursive cost functionals are defined by controlled fully coupled FBSDEs. We study two cases of…

Optimization and Control · Mathematics 2013-02-06 Juan Li , Qingmeng Wei

We present a partial-differential-equation-based optimal path-planning framework for curvature constrained motion, with application to vehicles in 2- and 3-spatial-dimensions. This formulation relies on optimal control theory, dynamic…

Numerical Analysis · Mathematics 2024-04-17 Christian Parkinson , Isabelle Boyle

We present a simple and easy to implement method for the numerical solution of a rather general class of Hamilton-Jacobi-Bellman (HJB) equations. In many cases, the considered problems have only a viscosity solution, to which, fortunately,…

Computational Finance · Quantitative Finance 2011-02-17 Jan Hendrik Witte , Christoph Reisinger

We present a kernel-based linear matrix inequality (LMI) approach for the approximate solution of Hamilton--Jacobi--Bellman (HJB) equations arising in nonlinear optimal control. The method represents the gradient of the value function in a…

Dynamical Systems · Mathematics 2026-05-19 Boumediene Hamzi , Umesh Vaidya

We present methods for locally solving the Dynamic Programming Equations (DPE) and the Hamilton Jacobi Bellman (HJB) PDE that arise in the infinite horizon optimal control problem. The method for solving the DPE is the discrete time version…

Optimization and Control · Mathematics 2007-05-23 Carmeliza Luna Navasca