English
Related papers

Related papers: Hamilton-Jacobi-Bellman Equations for Q-Learning i…

200 papers

In this paper, we propose Q-learning algorithms for continuous-time deterministic optimal control problems with Lipschitz continuous controls. Our method is based on a new class of Hamilton-Jacobi-Bellman (HJB) equations derived from…

Machine Learning · Computer Science 2020-10-28 Jeongho Kim , Jaeuk Shin , Insoon Yang

We address the crucial yet underexplored stability properties of the Hamilton--Jacobi--Bellman (HJB) equation in model-free reinforcement learning contexts, specifically for Lipschitz continuous optimal control problems. We bridge the gap…

Optimization and Control · Mathematics 2024-04-23 Namkyeong Cho , Yeoneung Kim

Maximum entropy reinforcement learning (RL) methods have been successfully applied to a range of challenging sequential decision-making and control tasks. However, most of existing techniques are designed for discrete-time systems. As a…

Optimization and Control · Mathematics 2020-09-29 Jeongho Kim , Insoon Yang

This paper proposes a new framework to model control systems in which a dynamic friction occurs. The model consists in a controlled differential inclusion with a discontinuous right hand side, which still preserves existence and uniqueness…

Optimization and Control · Mathematics 2020-12-02 Fabio Tedone , Michele Palladino

The Bellman equation and its continuous-time counterpart, the Hamilton-Jacobi-Bellman (HJB) equation, serve as necessary conditions for optimality in reinforcement learning and optimal control. While the value function is known to be the…

Machine Learning · Computer Science 2025-03-07 Haoxiang You , Lekan Molu , Ian Abraham

Continuous-time reinforcement learning offers an appealing formalism for describing control problems in which the passage of time is not naturally divided into discrete increments. Here we consider the problem of predicting the distribution…

Machine Learning · Computer Science 2022-06-20 Harley Wiltzer , David Meger , Marc G. Bellemare

In this paper we study the fully nonlinear stochastic Hamilton-Jacobi-Bellman (HJB) equation for the optimal stochastic control problem of stochastic differential equations with random coefficients. The notion of viscosity solution is…

Optimization and Control · Mathematics 2018-07-16 Jinniao Qiu

In this paper, we present a novel algorithm named synchronous integral Q-learning, which is based on synchronous policy iteration, to solve the continuous-time infinite horizon optimal control problems of input-affine system dynamics. The…

Systems and Control · Electrical Eng. & Systems 2021-05-20 Lei Guo , Han Zhao

Many optimal control problems are formulated as two point boundary value problems (TPBVPs) with conditions of optimality derived from the Hamilton-Jacobi-Bellman (HJB) equations. In most cases, it is challenging to solve HJBs due to the…

Optimization and Control · Mathematics 2019-07-25 Sixiong You , Ran Dai , Ping Lu

We present a new formulation for the computation of solutions of a class of Hamilton Jacobi Bellman (HJB) equations on closed smooth surfaces of co-dimension one. For the class of equations considered in this paper, the viscosity solution…

Numerical Analysis · Mathematics 2020-08-06 Lindsay Martin , Richard Tsai

In this article, a class of optimal control problems of differential equations with delays are investigated for which the associated Hamilton-Jacobi-Bellman (HJB) equations are nonlinear partial differential equations with delays. This type…

Optimization and Control · Mathematics 2015-07-16 Jianjun Zhou

We study a class of optimal control problems with state constraints where the state equation is a differential equation with delays. This class includes some problems arising in economics, in particular the so-called models with time to…

Optimization and Control · Mathematics 2009-07-09 Salvatore Federico , Ben Goldys , Fausto Gozzi

We study the exploratory Hamilton--Jacobi--Bellman (HJB) equation arising from the entropy-regularized exploratory control problem, which was formulated by Wang, Zariphopoulou and Zhou (J. Mach. Learn. Res., 21, 2020) in the context of…

Optimization and Control · Mathematics 2021-09-22 Wenpin Tang , Paul Yuming Zhang , Xun Yu Zhou

Convex Q-learning is a recent approach to reinforcement learning, motivated by the possibility of a firmer theory for convergence, and the possibility of making use of greater a priori knowledge regarding policy or value function structure.…

Optimization and Control · Mathematics 2022-10-18 Fan Lu , Joel Mathias , Sean Meyn , Karanjit Kalsi

The purpose of this paper is to describe the numerical solution of the Hamilton-Jacobi-Bellman (HJB) for an optimal control problem for quantum spin systems. This HJB equation is a first order nonlinear partial differential equation defined…

Quantum Physics · Physics 2011-10-05 Srinivas Sridharan , Matthew R. James

Optimal control and the associated second-order Hamilton-Jacobi-Bellman (HJB) equation are studied for unbounded stochastic evolution systems in Hilbert spaces. A new notion of viscosity solution, featured by absence of B-continuity, is…

Optimization and Control · Mathematics 2026-02-10 Shanjian Tang , Jianjun Zhou

Computing optimal feedback controls for nonlinear systems generally requires solving Hamilton-Jacobi-Bellman (HJB) equations, which are notoriously difficult when the state dimension is large. Existing strategies for high-dimensional…

Optimization and Control · Mathematics 2021-04-09 Tenavi Nakamura-Zimmerer , Qi Gong , Wei Kang

In this article, a notion of viscosity solutions is introduced for second order path-dependent Hamilton-Jacobi-Bellman (PHJB) equations associated with optimal control problems for path-dependent stochastic differential equations. We…

Optimization and Control · Mathematics 2022-12-26 Jianjun Zhou

In this article, a notion of viscosity solutions is introduced for second order path-dependent Hamilton-Jacobi-Bellman (PHJB) equations associated with optimal control problems for path-dependent stochastic evolution equations in Hilbert…

Probability · Mathematics 2020-09-14 Jianjun Zhou

In this article, a notion of viscosity solutions is introduced for first order path-dependent Hamilton-Jacobi-Bellman (PHJB) equations associated with optimal control problems for path-dependent evolution equations in Hilbert space. We…

Probability · Mathematics 2020-07-09 Jianjun Zhou
‹ Prev 1 2 3 10 Next ›