English
Related papers

Related papers: Distributional Hamilton-Jacobi-Bellman Equations f…

200 papers

We provide a unified approach to find equilibrium solutions for time-inconsistent problems with distribution dependent rewards, which are important to the study of behavioral finance and economics. Our approach is based on {\it equilibrium…

Mathematical Finance · Quantitative Finance 2022-04-11 Zongxia Liang , Fengyi Yuan

In this paper, we study the optimal dividend problem under the continuous time diffusion model with the bounded dividend rate from the Reinforcement Learning (RL) perspective. Unlike the standard literature, our main focus will be on…

Optimization and Control · Mathematics 2026-03-30 Lihua Bai , Thejani Gamage , Jin Ma , Gaozhan Wang

This work reformulates language generation as a stochastic optimal control problem, providing a unified theoretical perspective to analyze autoregressive and diffusion models and explain their limitations (Efficiency-Fidelity Paradox,…

Computation and Language · Computer Science 2026-05-18 ZiYi Dong , Yuliang Huang , Weijian Deng , Xiangyang Ji , Liang Lin , Pengxu Wei

In this paper, training a neural network is identified, exactly, as a search through Hamilton--Jacobi initial-value problems: each gradient step selects the initial data of a viscous Hamilton--Jacobi equation whose Hopf--Cole propagator…

Machine Learning · Computer Science 2026-05-29 Jose Marie Antonio Miñoza , Erika Fille T. Legara , Christopher P. Monterola

In this paper, a stochastic optimal control problem is investigated in which the system is governed by a stochastic functional differential equation. In the framework of functional It\^o calculus, we build the dynamic programming principle…

Optimization and Control · Mathematics 2013-01-03 Shaolin Ji , Shuzhen Yang

This work addresses stochastic optimal control problems where the unknown state evolves in continuous time while partial, noisy, and possibly controllable measurements are only available in discrete time. We develop a framework for…

Optimization and Control · Mathematics 2025-08-19 Christian Bayer , Boualem Djehiche , Eliza Rezvanova , Raul Fidel Tempone

This paper introduces a new type of second order stochastic backward Hamilton-Jacobi-Bellman (HJB) equations for optimal stochastic control problems with a currently observable but non-predicable parameter process, in addition to the…

Optimization and Control · Mathematics 2020-03-04 Nikolai Dokuchaev

We consider fully nonlinear Hamilton-Jacobi-Bellman equations associated to diffusion control problems involving a finite set-valued (or switching) control and possibly a continuum-valued control. In previous works (Akian, Fodjo, 2016 and…

Optimization and Control · Mathematics 2018-01-08 Marianne Akian , Eric Fodjo

An advantageous feature of piecewise constant policy timestepping for Hamilton-Jacobi-Bellman (HJB) equations is that different linear approximation schemes, and indeed different meshes, can be used for the resulting linear equations for…

Numerical Analysis · Mathematics 2016-01-21 Christoph Reisinger , Peter Forsyth

This paper derives recursion equations for a robust smoothing problem for a class of nonlinear systems with uncertainties in modeling and exogenous noise sources. The systems considered operate in discrete-time and the uncertainties are…

Optimization and Control · Mathematics 2013-03-27 Abhijit G. Kallapur , Ian R. Petersen

This paper aims to explore the relationship between maximum principle and dynamic programming principle for stochastic recursive control problem with random coefficients. Under certain regular conditions for the coefficients, the…

Optimization and Control · Mathematics 2020-12-10 Yuchao Dong , Qingxin Meng , Qi Zhang

The Bellman equation and its continuous form, the Hamilton-Jacobi-Bellman equation, are ubiquitous in reinforcement learning and control theory. However, these equations become intractable for high-dimensional or nonlinear systems. This…

Artificial Intelligence · Computer Science 2026-05-04 Preston Rozwood , Edward Mehrez , Ludger Paehler , Wen Sun , Steven L. Brunton

From the Hamilton-Jacobi-Bellman equation for the value function we derive a non-linear partial differential equation for the optimal portfolio strategy (the dynamic control). The equation is general in the sense that it does not depend on…

Portfolio Management · Quantitative Finance 2013-11-20 Mads Nielsen

We consider a dynamic portfolio optimization problem that incorporates predictable returns, instantaneous transaction costs, price impact, and stochastic volatility, extending the classical results of Garleanu and Pedersen (2013), which…

Computational Finance · Quantitative Finance 2025-07-24 Patrick Chan , Ronnie Sircar , Iosif Zimbidis

In distributional reinforcement learning not only expected returns but the complete return distributions of a policy are taken into account. The return distribution for a fixed policy is given as the solution of an associated distributional…

Machine Learning · Statistics 2023-05-29 Julian Gerstenberg , Ralph Neininger , Denis Spiegel

In this paper, we study the optimal stopping problem in the so-called exploratory framework, in which the agent takes actions randomly conditioning on current state and an entropy-regularized term is added to the reward functional. Such a…

Optimization and Control · Mathematics 2023-09-04 Yuchao Dong

Stochastic optimal control problems governed by delay equations with delay in the control are usually more difficult to study than the the ones when the delay appears only in the state. This is particularly true when we look at the…

Probability · Mathematics 2021-03-22 F. Gozzi , F. Masiero

This paper proposes a fully data-driven approach for optimal control of nonlinear control-affine systems represented by a stochastic diffusion. The focus is on the scenario where both the nonlinear dynamics and stage cost functions are…

Optimization and Control · Mathematics 2025-11-03 Nicolas Hoischen , Petar Bevanda , Stefan Sosnowski , Sandra Hirche , Boris Houska

The Bellman equation and its continuous-time counterpart, the Hamilton-Jacobi-Bellman (HJB) equation, serve as necessary conditions for optimality in reinforcement learning and optimal control. While the value function is known to be the…

Machine Learning · Computer Science 2025-03-07 Haoxiang You , Lekan Molu , Ian Abraham

We consider a class of exit time stochastic control problems for diffusion processes with discounted criterion, where the controller can utilize a given amount of resource, called "fuel". In contrast to the vast majority of existing…

Optimization and Control · Mathematics 2015-01-30 Dmitry B. Rokhlin , Georgii Mironenko
‹ Prev 1 4 5 6 7 8 10 Next ›