中文
相关论文

相关论文: Neural Actor-Critic Methods for Hamilton-Jacobi-Be…

200 篇论文

Stochastic optimal control problems governed by delay equations with delay in the control are usually more difficult to study than the the ones when the delay appears only in the state. This is particularly true when we look at the…

概率论 · 数学 2015-06-22 Fausto Gozzi , Federica Masiero

We study policy iteration (PI) for deterministic infinite-horizon discounted optimal control problems, whose value function is characterized by a stationary Hamilton--Jacobi--Bellman (HJB) equation. At the PDE level, PI is fundamentally…

最优化与控制 · 数学 2026-04-14 Namkyeong Cho , Yeoneung Kim

This paper presents a physics-informed machine learning approach for synthesizing optimal feedback control policy for infinite-horizon optimal control problems by solving the Hamilton-Jacobi-Bellman (HJB) partial differential equation(PDE).…

系统与控制 · 电气工程与系统科学 2025-11-24 Tanay Raghunandan Srinivasa , Suraj Kumar

We study the problem of optimal portfolio selection under stochastic volatility within a continuous time reinforcement learning framework with portfolio constraints. Exploration is modeled through entropy-regularized relaxed controls, where…

数理金融 · 定量金融 2026-04-27 Thai Nguyen , Pertiny Nkuize

We propose a neural network approach that yields approximate solutions for high-dimensional optimal control problems and demonstrate its effectiveness using examples from multi-agent path finding. Our approach yields controls in a feedback…

最优化与控制 · 数学 2022-06-29 Derek Onken , Levon Nurbekyan , Xingjian Li , Samy Wu Fung , Stanley Osher , Lars Ruthotto

Optimal control problems are crucial in various domains, including path planning, robotics, and humanoid control, demonstrating their broad applicability. The connection between optimal control and Hamilton-Jacobi (HJ) partial differential…

最优化与控制 · 数学 2024-03-06 Tingwei Meng , Siting Liu , Wuchen Li , Stanley Osher

We address finding the semi-global solutions to optimal feedback control and the Hamilton--Jacobi--Bellman (HJB) equation. Using the solution of an HJB equation, a feedback optimal control law can be implemented in real-time with minimum…

最优化与控制 · 数学 2016-06-17 Wei Kang , Lucas C. Wilcox

We propose new and original mathematical connections between Hamilton-Jacobi (HJ) partial differential equations (PDEs) with initial data and neural network architectures. Specifically, we prove that some classes of neural networks…

最优化与控制 · 数学 2020-03-10 Jerome Darbon , Gabriel P. Langlois , Tingwei Meng

Stochastic optimal control problems governed by delay equations with delay in the control are usually more difficult to study than the the ones when the delay appears only in the state. This is particularly true when we look at the…

概率论 · 数学 2021-03-22 F. Gozzi , F. Masiero

In this paper, we are concerned with the classical solvability of a class of second-order Hamilton-Jacobi-Bellman equations (HJB equations) arising from stochastic optimal control problems with linear dynamics and uniformly convex cost…

最优化与控制 · 数学 2025-12-19 Jinghua Li , Zhiyong Yu

This paper is concerned with a stochastic recursive optimal control problem with time delay, where the controlled system is described by a stochastic differential delayed equation (SDDE) and the cost functional is formulated as the solution…

最优化与控制 · 数学 2014-08-26 Jingtao Shi , Huanshui Zhang

Actor-critic algorithms are widely used in reinforcement learning, but are challenging to mathematically analyse due to the online arrival of non-i.i.d. data samples. The distribution of the data samples dynamically changes as the model is…

机器学习 · 计算机科学 2023-09-20 Ziheng Wang , Justin Sirignano

Actor-critic algorithms have become a cornerstone in reinforcement learning (RL), leveraging the strengths of both policy-based and value-based methods. Despite recent progress in understanding their statistical efficiency, no existing work…

机器学习 · 统计学 2025-05-07 Kevin Tan , Wei Fan , Yuting Wei

Controlling systems of ordinary differential equations (ODEs) is ubiquitous in science and engineering. For finding an optimal feedback controller, the value function and associated fundamental equations such as the Bellman equation and the…

最优化与控制 · 数学 2021-04-14 Mathias Oster , Leon Sallandt , Reinhold Schneider

We propose an actor-critic framework to solve the time-continuous stochastic optimal control problem. A least square temporal difference method is applied to compute the value function for the critic. The policy gradient method is…

最优化与控制 · 数学 2025-01-27 Mo Zhou , Jianfeng Lu

Physics-informed neural solvers offer a promising route to model-based reinforcement learning in continuous time, where optimal feedback synthesis is governed by Hamilton--Jacobi--Bellman (HJB) equations. Practical implementations often…

机器学习 · 计算机科学 2026-05-11 Minseok Kim , Yeongjong Kim , Namkyeong Cho , Yeoneung Kim

For an infinite-horizon control problem, the optimal control can be represented by the stable manifold of the characteristic Hamiltonian system of Hamilton-Jacobi-Bellman (HJB) equation in a semiglobal domain. In this paper, we first…

最优化与控制 · 数学 2024-05-14 Guoyuan Chen

Optimal control problem is typically solved by first finding the value function through Hamilton-Jacobi equation (HJE) and then taking the minimizer of the Hamiltonian to obtain the control. In this work, instead of focusing on the value…

最优化与控制 · 数学 2021-09-10 Alain Bensoussan , Jiayue Han , Sheung Chi Phillip Yam , Xiang Zhou

In this note, we study a class of indefinite stochastic McKean-Vlasov linear-quadratic (LQ in short) control problem under the control taking nonnegative values. In contrast to the conventional issue, both the classical dynamic programming…

最优化与控制 · 数学 2023-10-05 Xun Li , Liangquan Zhang

We study a family of stationary Hamilton-Jacobi-Bellman (HJB) equations in Hilbert spaces arising from stochastic optimal control problems. The main difficulties to treat such problems are: the lack of smoothing properties of the linear…

最优化与控制 · 数学 2025-10-31 Gabriele Bolli , Fausto Gozzi