English
Related papers

Related papers: Hamilton-Jacobi-Bellman Equations for Q-Learning i…

200 papers

We consider an initial value problem for a Hamilton--Jacobi equation with a quadratic and degenerate Hamiltonian. Our Hamiltonian comes from the dynamics of $N$-peakon in the Camassa--Holm equation. It is given by a quadratic form with a…

Analysis of PDEs · Mathematics 2020-07-06 Tomasz Cieślak , Jakub Siemianowski , Andrzej Święch

This paper presents a new methodology to craft navigation functions for nonlinear systems with stochastic uncertainty. The method relies on the transformation of the Hamilton-Jacobi-Bellman (HJB) equation into a linear partial differential…

Robotics · Computer Science 2014-09-23 Matanya B. Horowitz , Joel W. Burdick

In this manuscript, we study optimal control problems for stochastic delay differential equations using the dynamic programming approach in Hilbert spaces via viscosity solutions of the associated Hamilton-Jacobi-Bellman equations. We show…

Optimization and Control · Mathematics 2024-12-24 Filippo de Feo , Andrzej Święch

Recent studies have extended the use of the stochastic Hamilton-Jacobi-Bellman (HJB) equation to include complex variables for deriving quantum mechanical equations. However, these studies often assume that it is valid to apply the HJB…

Quantum Physics · Physics 2024-10-14 Vasil Yordanov

We discuss a class of time-dependent Hamilton-Jacobi equations, where an unknown function of time is intended to keep the maximum of the solution to the constant value 0. Our main result is that the full problem has a unique viscosity…

Analysis of PDEs · Mathematics 2015-05-25 Sepideh Mirrahimi , Jean-Michel Roquejoffre

Feedback controllers for port-Hamiltonian systems reveal an intrinsic inverse optimality property since each passivating state feedback controller is optimal with respect to some specific performance index. Due to the nonlinear…

Optimization and Control · Mathematics 2020-07-20 Lukas Kölsch , Pol Jané Soneira , Felix Strehle , Sören Hohmann

A new algorithm for time dependent Hamilton Jacobi equations on networks, based on semi Lagrangian scheme, is proposed. It is based on the definition of viscosity solution for this kind of problems recently given in. A thorough convergence…

Numerical Analysis · Mathematics 2023-10-11 Elisabetta Carlini , Antonio Siconolfi

We consider a stochastic optimal control problem where the controller can anticipate the evolution of the driving noise over some dynamically changing time window. The controlled state dynamics are understood as a rough differential…

Optimization and Control · Mathematics 2025-10-07 Peter Bank , Franziska Bielert

We propose a supervised learning scheme for the first order Hamilton--Jacobi PDEs in high dimensions. The scheme is designed by using the geometric structure of Wasserstein Hamiltonian flows via a density coupling strategy. It is…

Numerical Analysis · Mathematics 2025-11-05 Jianbo Cui , Shu Liu , Haomin Zhou

This study introduces a mathematical framework to investigate the viability and reachability of production systems under constraints. We develop a model that incorporates key decision variables, such as pricing policy, quality investment,…

Optimization and Control · Mathematics 2025-09-16 Achraf Bouhmady , Mustapha Serhani , Nadia Raissi

In this manuscript we consider a class optimal control problem for stochastic differential delay equations. First, we rewrite the problem in a suitable infinite-dimensional Hilbert space. Then, using the dynamic programming approach, we…

Optimization and Control · Mathematics 2023-02-20 Filippo de Feo , Salvatore Federico , Andrzej Święch

We investigate the large-time behavior of the value functions of the optimal control problems on the $n$-dimensional torus which appear in the dynamic programming for the system whose states are governed by random changes. From the point of…

Analysis of PDEs · Mathematics 2013-03-13 Hiroyoshi Mitake , Hung V. Tran

For Hamilton-Jacobi-Bellman (HJB) equations, with the standard definitions of viscosity super-solution and sub-solution, it is known that there is a comparison between any (viscosity) super-solutions and sub-solutions. This should be the…

Analysis of PDEs · Mathematics 2021-02-08 Yue Zhou , Xinwei Feng , Jiongmin Yong

Continuous-time stochastic processes underlie many natural and engineered systems. In healthcare, autonomous driving, and industrial control, direct interaction with the environment is often unsafe or impractical, motivating offline…

Machine Learning · Statistics 2025-11-14 Nicolas Hoischen , Petar Bevanda , Max Beier , Stefan Sosnowski , Boris Houska , Sandra Hirche

This is the first in a series of papers in which we study an efficient approximation scheme for solving the Hamilton-Jacobi-Bellman equation for multi-dimensional problems in stochastic control theory. The method is a combination of a WKB…

Computational Finance · Quantitative Finance 2014-06-26 Sakda Chaiworawitkul , Patrick S. Hagan , Andrew Lesniewski

In this paper infinite horizon optimal control problems for nonlinear high-dimensional dynamical systems are studied. Nonlinear feedback laws can be computed via the value function characterized as the unique viscosity solution to the…

Optimization and Control · Mathematics 2016-02-22 Alessandro Alla , Maurizio Falcone , Stefan Volkwein

In this paper we establish H\"older continuity estimates for viscosity solutions to first order Hamilton-Jacobi equations linked to linear control systems satisfying the Kalman rank condition. Our model Hamiltonians are non-convex in the…

Analysis of PDEs · Mathematics 2026-05-08 Megan Griffin-Pickering , Alpár R. Mészáros

In this paper we propose and analyze a method based on the Riccati transformation for solving the evolutionary Hamilton-Jacobi-Bellman equation arising from the stochastic dynamic optimal allocation problem. We show how the fully nonlinear…

Portfolio Management · Quantitative Finance 2013-07-25 Sona Kilianova , Daniel Sevcovic

We consider continuous-state and continuous-time control problems where the admissible trajectories of the system are constrained to remain on a union of half-planes which share a common straight line. This set will be named a junction. We…

Optimization and Control · Mathematics 2014-12-10 Salomé Oudet

Q-learning is a promising method for solving optimal control problems for uncertain systems without the explicit need for system identification. However, approaches for continuous-time Q-learning have limited provable safety guarantees,…

Systems and Control · Electrical Eng. & Systems 2024-01-30 Soutrik Bandyopadhyay , Shubhendu Bhasin