中文
相关论文

相关论文: Simpler near-optimal controllers through direct su…

200 篇论文

The approximation of solutions to second order Hamilton--Jacobi--Bellman (HJB) equations by deep neural networks is investigated. It is shown that for HJB equations that arise in the context of the optimal control of certain Markov…

数值分析 · 数学 2021-03-11 Philipp Grohs , Lukas Herrmann

We study a class of optimal control problems with state constraints where the state equation is a differential equation with delays. This class includes some problems arising in economics, in particular the so-called models with time to…

最优化与控制 · 数学 2009-07-09 Salvatore Federico , Ben Goldys , Fausto Gozzi

In this paper, we are concerned with the classical solvability of a class of second-order Hamilton-Jacobi-Bellman equations (HJB equations) arising from stochastic optimal control problems with linear dynamics and uniformly convex cost…

最优化与控制 · 数学 2025-12-19 Jinghua Li , Zhiyong Yu

Continuous-time reinforcement learning offers an appealing formalism for describing control problems in which the passage of time is not naturally divided into discrete increments. Here we consider the problem of predicting the distribution…

机器学习 · 计算机科学 2022-06-20 Harley Wiltzer , David Meger , Marc G. Bellemare

Presented is a method for efficient computation of the Hamilton-Jacobi (HJ) equation for time-optimal control problems using the generalized Hopf formula. Typically, numerical methods to solve the HJ equation rely on a discrete grid of the…

系统与控制 · 计算机科学 2019-10-22 Matthew R. Kirchner , Gary Hewer , Jerome Darbon , Stanley Osher

Recent research reveals that deep learning is an effective way of solving high dimensional Hamilton-Jacobi-Bellman equations. The resulting feedback control law in the form of a neural network is computationally efficient for real-time…

动力系统 · 数学 2022-10-10 Wei Kang , Qi Gong , Tenavi Nakamura-Zimmerer

We propose a variant of consensus-based optimization (CBO) algorithms, controlled-CBO, which introduces a feedback control term to improve convergence towards global minimizers of non-convex functions in multiple dimensions. The feedback…

最优化与控制 · 数学 2025-07-29 Yuyang Huang , Michael Herty , Dante Kalise , Nikolas Kantas

Optimal feedback controllers for nonlinear systems can be derived by solving the Hamilton-Jacobi-Bellman (HJB) equation. However, because the HJB is a nonlinear partial differential equation, numerical methods typically provide only…

最优化与控制 · 数学 2026-03-25 Morgan Jones , Matthew Peet

Maximum entropy reinforcement learning (RL) methods have been successfully applied to a range of challenging sequential decision-making and control tasks. However, most of existing techniques are designed for discrete-time systems. As a…

最优化与控制 · 数学 2020-09-29 Jeongho Kim , Insoon Yang

This work proposes an optimal safe controller minimizing an infinite horizon cost functional subject to control barrier functions (CBFs) safety conditions. The constrained optimal control problem is reformulated as a minimization problem of…

系统与控制 · 电气工程与系统科学 2022-02-03 Hassan Almubarak , Evangelos A. Theodorou , Nader Sadegh

A deep learning approach for the approximation of the Hamilton-Jacobi-Bellman partial differential equation (HJB PDE) associated to the Nonlinear Quadratic Regulator (NLQR) problem. A state-dependent Riccati equation control law is first…

最优化与控制 · 数学 2022-07-20 Anastasia Borovykh , Dante Kalise , Alexis Laignelet , Panos Parpas

This is the first in a series of papers in which we study an efficient approximation scheme for solving the Hamilton-Jacobi-Bellman equation for multi-dimensional problems in stochastic control theory. The method is a combination of a WKB…

计算金融 · 定量金融 2014-06-26 Sakda Chaiworawitkul , Patrick S. Hagan , Andrew Lesniewski

We address the crucial yet underexplored stability properties of the Hamilton--Jacobi--Bellman (HJB) equation in model-free reinforcement learning contexts, specifically for Lipschitz continuous optimal control problems. We bridge the gap…

最优化与控制 · 数学 2024-04-23 Namkyeong Cho , Yeoneung Kim

Sampling from probability densities is a common challenge in fields such as Uncertainty Quantification (UQ) and Generative Modelling (GM). In GM in particular, the use of reverse-time diffusion processes depending on the log-densities of…

机器学习 · 统计学 2024-02-26 David Sommer , Robert Gruhlke , Max Kirstein , Martin Eigel , Claudia Schillings

We consider an extension of the well-known Hamilton-Jacobi-Bellman (HJB) equation for fractional order dynamical systems in which a generalized performance index is considered for the related optimal control problem. Owing to the…

最优化与控制 · 数学 2018-11-29 Abolhassan Razminia , Mehdi AsadiZadehShiraz , Delfim F. M. Torres

Neural network approaches that parameterize value functions have succeeded in approximating high-dimensional optimal feedback controllers when the Hamiltonian admits explicit formulas. However, many practical problems, such as the space…

最优化与控制 · 数学 2025-10-08 Eric Gelphman , Deepanshu Verma , Nicole Tianjiao Yang , Stanley Osher , Samy Wu Fung

This paper presents a two-stage framework for constrained near-optimal feedback control of input-affine nonlinear systems. An approximate value function for the unconstrained control problem is computed offline by solving the…

系统与控制 · 电气工程与系统科学 2026-03-18 Milad Alipour Shahraki , Laurent Lessard

The objective of designing a control system is to steer a dynamical system with a control signal, guiding it to exhibit the desired behavior. The Hamilton-Jacobi-Bellman (HJB) partial differential equation offers a framework for optimal…

机器学习 · 计算机科学 2025-10-22 Jostein Barry-Straume , Adwait D. Verulkar , Arash Sarshar , Andrey A. Popov , Adrian Sandu

We introduce a new numerical method to approximate the solution of a finite horizon deterministic optimal control problem. We exploit two Hamilton-Jacobi-Bellman PDE, arising by considering the dynamics in forward and backward time. This…

最优化与控制 · 数学 2023-04-21 Marianne Akian , Stéphane Gaubert , Shanqing Liu

The main goal of this paper is to establish existence, regularity and uniqueness results for the solution of a Hamilton-Jacobi-Bellman (HJB) equation, whose operator is an elliptic integro-differential operator. The HJB equation studied in…

最优化与控制 · 数学 2016-12-01 Harold A. Moreno-Franco