English
Related papers

Related papers: A neural network based policy iteration algorithm …

200 papers

We consider a kind of stochastic exit time optimal control problems, in which the cost function is defined through a nonlinear backward stochastic differential equation. We study the regularity of the value function for such a control…

Probability · Mathematics 2016-03-15 Rainer Buckdahn , Tianyang Nie

This paper proposes two algorithms for solving stochastic control problems with deep learning, with a focus on the utility maximisation problem. The first algorithm solves Markovian problems via the Hamilton Jacobi Bellman (HJB) equation.…

Computational Finance · Quantitative Finance 2024-10-15 Ashley Davey , Harry Zheng

We study the problem of learning the optimal control policy for fine-tuning a given diffusion process, using general value function approximation. We develop a new class of algorithms by solving a variational inequality problem based on the…

Machine Learning · Computer Science 2025-09-03 Wenlong Mou

We introduce two algorithms based on a policy iteration method to numerically solve time-dependent Mean Field Game systems of partial differential equations with non-separable Hamiltonians. We prove the convergence of such algorithms in…

Optimization and Control · Mathematics 2022-10-03 Mathieu Laurière , Jiahao Song , Qing Tang

This paper develops an algorithm for upper- and lower-bounding the value function for a class of linear time-varying games subject to convex control sets. In particular, a two-player zero-sum differential game is considered where the…

Optimization and Control · Mathematics 2025-03-12 Vincent Liu , Chris Manzie , Peter M. Dower

This paper introduces Deep Policy Iteration (DPI), a novel approach that integrates the strengths of Neural Networks with the stability and convergence advantages of Policy Iteration (PI) to address high-dimensional stochastic Mean Field…

Optimization and Control · Mathematics 2024-07-15 Mouhcine Assouli , Badr Missaoui

This paper investigates a class of multiscale stochastic control problems driven by $\alpha$-stable L\'evy noises, where the controlled dynamics evolve across separate slow and fast time scales. The associated value functions are governed…

Optimization and Control · Mathematics 2025-11-11 Qi Zhang , Yanjie Zhang , Ao Zhang

We propose a new numerical method for solving the Hamilton-Jacobi-Bellman quasi-variational inequality associated with the combined impulse and stochastic optimal control problem over a finite time horizon. Our method corresponds to an…

Numerical Analysis · Mathematics 2015-02-05 Masashi Ieda

In this paper we deal with the problem of existence of a smooth solution of the Hamilton-Jacobi-Bellman-Isaacs (HJBI for short) system of equations associated with nonzero-sum stochastic differential games. We consider the problem in…

Analysis of PDEs · Mathematics 2018-10-24 Said Hamadene , Paola Mannucci

We study a class of optimal control problems with state constraints where the state equation is a differential equation with delays. This class includes some problems arising in economics, in particular the so-called models with time to…

Optimization and Control · Mathematics 2009-07-09 Salvatore Federico , Ben Goldys , Fausto Gozzi

Many optimal control problems are formulated as two point boundary value problems (TPBVPs) with conditions of optimality derived from the Hamilton-Jacobi-Bellman (HJB) equations. In most cases, it is challenging to solve HJBs due to the…

Optimization and Control · Mathematics 2019-07-25 Sixiong You , Ran Dai , Ping Lu

A control theoretic approach is presented in this paper for both batch and instantaneous updates of weights in feed-forward neural networks. The popular Hamilton-Jacobi-Bellman (HJB) equation has been used to generate an optimal weight…

Neural and Evolutionary Computing · Computer Science 2015-04-29 Vipul Arora , Laxmidhar Behera , Ajay Pratap Yadav

We study the convergence rates of policy iteration (PI) for nonconvex viscous Hamilton--Jacobi equations using a discrete space-time scheme, where both space and time variables are discretized. We analyze the case with an uncontrolled…

Numerical Analysis · Mathematics 2025-03-05 Xiaoqin Guo , Hung Vinh Tran , Yuming Paul Zhang

In this note we study the convergence of monotone P1 finite element methods on unstructured meshes for fully non-linear Hamilton-Jacobi-Bellman equations arising from stochastic optimal control problems with possibly degenerate, isotropic…

Numerical Analysis · Mathematics 2013-02-25 Max Jensen , Iain Smears

In this paper, we are concerned with the classical solvability of a class of second-order Hamilton-Jacobi-Bellman equations (HJB equations) arising from stochastic optimal control problems with linear dynamics and uniformly convex cost…

Optimization and Control · Mathematics 2025-12-19 Jinghua Li , Zhiyong Yu

In this paper, we investigate a fully nonlinear evolutionary Hamilton-Jacobi-Bellman (HJB) parabolic equation utilizing the monotone operator technique. We consider the HJB equation arising from portfolio optimization selection, where the…

Mathematical Finance · Quantitative Finance 2021-04-14 Daniel Sevcovic , Cyril Izuchukwu Udeani

We consider a class of exit time stochastic control problems for diffusion processes with discounted criterion, where the controller can utilize a given amount of resource, called "fuel". In contrast to the vast majority of existing…

Optimization and Control · Mathematics 2015-01-30 Dmitry B. Rokhlin , Georgii Mironenko

This paper, which is the natural continuation of a previous paper by the same authors, studies a class of optimal control problems with state constraints where the state equation is a differential equation with delays. This class includes…

Optimization and Control · Mathematics 2009-07-10 Salvatore Federico , Ben Goldys , Fausto Gozzi

This paper establishes the existence, uniqueness, and global $C^{1,\beta}$ regularity of positive classical solutions to a class of quasilinear Hamilton--Jacobi--Bellman (HJB) equations with Dirichlet boundary conditions on bounded convex…

Analysis of PDEs · Mathematics 2026-03-10 Dragos-Patru Covei

We consider the portfolio optimisation problem where the terminal function is an S-shaped utility applied at the difference between the wealth and a random benchmark process. We develop several numerical methods for solving the problem…

Computational Finance · Quantitative Finance 2024-10-10 Ashley Davey , Harry Zheng
‹ Prev 1 4 5 6 7 8 10 Next ›