中文
相关论文

相关论文: Simpler near-optimal controllers through direct su…

200 篇论文

To sidestep the curse of dimensionality when computing solutions to Hamilton-Jacobi-Bellman partial differential equations (HJB PDE), we propose an algorithm that leverages a neural network to approximate the value function. We show that…

机器学习 · 计算机科学 2017-03-28 Frank Jiang , Glen Chou , Mo Chen , Claire J. Tomlin

In optimal control problems of control-affine systems, whose solutions are bang-bang or singular type, verification of optimality using the Hamilton-Jacobi-Bellman (HJB) equation involves the computation of partial derivatives of switching…

最优化与控制 · 数学 2020-09-15 Victor Riquelme

The Hamilton Jacobi Bellman Equation (HJB) provides the globally optimal solution to large classes of control problems. Unfortunately, this generality comes at a price, the calculation of such solutions is typically intractible for systems…

最优化与控制 · 数学 2014-09-23 Matanya B. Horowitz , Anil Damle , Joel W. Burdick

This paper introduces a reinforcement learning-based tracking control approach for a class of nonlinear systems using neural networks. In this approach, adversarial attacks were considered both in the actuator and on the outputs. This…

系统与控制 · 电气工程与系统科学 2022-09-20 Farshad Rahimi , Sepideh Ziaei

Optimal control problems are crucial in various domains, including path planning, robotics, and humanoid control, demonstrating their broad applicability. The connection between optimal control and Hamilton-Jacobi (HJ) partial differential…

最优化与控制 · 数学 2024-03-06 Tingwei Meng , Siting Liu , Wuchen Li , Stanley Osher

We establish the convergence of the deep Galerkin method (DGM), a deep learning-based scheme for solving high-dimensional nonlinear PDEs, for Hamilton-Jacobi-Bellman (HJB) equations that arise from the study of mean field control problems…

最优化与控制 · 数学 2024-05-24 William Hofgard , Jingruo Sun , Asaf Cohen

We show that necessary and sufficient conditions of optimality in periodic optimization problems can be stated in terms of a solution of the corresponding HJB inequality, the latter being equivalent to a max-min type variational problem…

最优化与控制 · 数学 2013-09-10 Vladimir Gaitsgory , Ludmila Manic

Controlling systems of ordinary differential equations (ODEs) is ubiquitous in science and engineering. For finding an optimal feedback controller, the value function and associated fundamental equations such as the Bellman equation and the…

最优化与控制 · 数学 2021-04-14 Mathias Oster , Leon Sallandt , Reinhold Schneider

In this paper, a highly parallel and derivative-free martingale neural network learning method is proposed to solve Hamilton-Jacobi-Bellman (HJB) equations arising from stochastic optimal control problems (SOCPs), as well as general…

最优化与控制 · 数学 2024-12-23 Wei Cai , Shuixin Fang , Wenzhong Zhang , Tao Zhou

An optimal control problem is considered for a stochastic differential equation containing a state-dependent regime switching, with a recursive cost functional. Due to the non-exponential discounting in the cost functional, the problem is…

最优化与控制 · 数学 2017-12-29 Hongwei Mei , Jiongmin Yong

We investigate feedback control for infinite horizon optimal control problems for partial differential equations. The method is based on the coupling between Hamilton-Jacobi-Bellman (HJB) equations and model reduction techniques. It is…

最优化与控制 · 数学 2016-07-11 Alessandro Alla , Andreas Schmidt , Bernard Haasdonk

Designing optimal controllers for nonlinear dynamical systems often relies on reinforcement learning and adaptive dynamic programming (ADP) to approximate solutions of the Hamilton Jacobi Bellman (HJB) equation. However, these methods…

最优化与控制 · 数学 2025-11-27 Akash Vyas , Shreyas Kumar , Jayant Kumar Mohanta , Ravi Prakash

For continuous systems modeled by dynamical equations such as ODEs and SDEs, Bellman's Principle of Optimality takes the form of the Hamilton-Jacobi-Bellman (HJB) equation, which provides the theoretical target of reinforcement learning…

机器学习 · 计算机科学 2025-10-28 Haruki Settai , Naoya Takeishi , Takehisa Yairi

This paper introduces the Hamilton-Jacobi-Bellman Proximal Policy Optimization (HJBPPO) algorithm into reinforcement learning. The Hamilton-Jacobi-Bellman (HJB) equation is used in control theory to evaluate the optimality of the value…

机器学习 · 计算机科学 2023-02-02 Amartya Mukherjee , Jun Liu

Environmental management optimizing a long-run objective is an ergodic control problem whose resolution can be achieved by solving an associated non-local Hamilton-Jacobi-Bellman (HJB) equation having an effective Hamiltonian. Focusing on…

最优化与控制 · 数学 2022-05-11 Hidekazu Yoshioka , Motoh Tsujimura , Yuta Yaegashi

We propose a neural network approach that yields approximate solutions for high-dimensional optimal control problems and demonstrate its effectiveness using examples from multi-agent path finding. Our approach yields controls in a feedback…

最优化与控制 · 数学 2022-06-29 Derek Onken , Levon Nurbekyan , Xingjian Li , Samy Wu Fung , Stanley Osher , Lars Ruthotto

In this note, we demonstrate that a locally semiconvex viscosity supersolution to a possibly degenerate fully nonlinear elliptic Hamilton-Jacobi-Bellman (HJB) equation is differentiable along the directions spanned by the range of the…

最优化与控制 · 数学 2025-01-28 Salvatore Federico , Giorgio Ferrari , Mauro Rosestolato

The purpose of this paper is to describe the numerical solution of the Hamilton-Jacobi-Bellman (HJB) for an optimal control problem for quantum spin systems. This HJB equation is a first order nonlinear partial differential equation defined…

量子物理 · 物理学 2011-10-05 Srinivas Sridharan , Matthew R. James

In this paper we study a first extension of the theory of mild solutions for HJB equations in Hilbert spaces to the case when the domain is not the whole space. More precisely, we consider a half-space as domain, and a semilinear…

最优化与控制 · 数学 2022-09-30 Alessandro Calvia , Gianluca Cappa , Fausto Gozzi , Enrico Priola

We treat infinite horizon optimal control problems by solving the associated stationary Hamilton-Jacobi-Bellman (HJB) equation numerically to compute the value function and an optimal feedback law. The dynamical systems under consideration…

最优化与控制 · 数学 2021-05-19 Mathias Oster , Leon Sallandt , Reinhold Schneider