中文
相关论文

相关论文: Actor-Critic Methods using Physics-Informed Neural…

200 篇论文

We propose a novel numerical method for high dimensional Hamilton--Jacobi--Bellman (HJB) type elliptic partial differential equations (PDEs). The HJB PDEs, reformulated as optimal control problems, are tackled by the actor-critic framework…

最优化与控制 · 数学 2022-01-07 Mo Zhou , Jiequn Han , Jianfeng Lu

The objective of designing a control system is to steer a dynamical system with a control signal, guiding it to exhibit the desired behavior. The Hamilton-Jacobi-Bellman (HJB) partial differential equation offers a framework for optimal…

机器学习 · 计算机科学 2025-10-22 Jostein Barry-Straume , Adwait D. Verulkar , Arash Sarshar , Andrey A. Popov , Adrian Sandu

This paper introduces the Hamilton-Jacobi-Bellman Proximal Policy Optimization (HJBPPO) algorithm into reinforcement learning. The Hamilton-Jacobi-Bellman (HJB) equation is used in control theory to evaluate the optimality of the value…

机器学习 · 计算机科学 2023-02-02 Amartya Mukherjee , Jun Liu

We mathematically analyze and numerically study an actor-critic machine learning algorithm for solving high-dimensional Hamilton-Jacobi-Bellman (HJB) partial differential equations from stochastic control theory. The architecture of the…

最优化与控制 · 数学 2026-05-20 Samuel N. Cohen , Jackson Hebner , Deqing Jiang , Justin Sirignano

We propose a physics-informed neural network policy iteration (PINN-PI) framework for solving stochastic optimal control problems governed by second-order Hamilton--Jacobi--Bellman (HJB) equations. At each iteration, a neural network is…

机器学习 · 计算机科学 2025-08-05 Yeongjong Kim , Yeoneung Kim , Minseok Kim , Namkyeong Cho

The aim of this work is to develop a deep learning method for solving high-dimensional stochastic control problems based on the Hamilton--Jacobi--Bellman (HJB) equation and physics-informed learning. Our approach is to parameterize the…

最优化与控制 · 数学 2025-06-23 Zhe Jiao , Wantao Jia , Weiqiu Zhu

We propose a physics-informed neural networks (PINNs) framework to solve the infinite-horizon optimal control problem of nonlinear systems. In particular, since PINNs are generally able to solve a class of partial differential equations…

系统与控制 · 电气工程与系统科学 2025-05-29 Filippos Fotiadis , Kyriakos G. Vamvoudakis

This paper presents a physics-informed machine learning approach for synthesizing optimal feedback control policy for infinite-horizon optimal control problems by solving the Hamilton-Jacobi-Bellman (HJB) partial differential equation(PDE).…

系统与控制 · 电气工程与系统科学 2025-11-24 Tanay Raghunandan Srinivasa , Suraj Kumar

This paper addresses the model-free nonlinear optimal problem with generalized cost functional, and a data-based reinforcement learning technique is developed. It is known that the nonlinear optimal control problem relies on the solution of…

系统与控制 · 计算机科学 2013-11-20 Biao Luo , Huai-Ning Wu , Tingwen Huang , Derong Liu

We treat infinite horizon optimal control problems by solving the associated stationary Hamilton-Jacobi-Bellman (HJB) equation numerically to compute the value function and an optimal feedback law. The dynamical systems under consideration…

最优化与控制 · 数学 2021-05-19 Mathias Oster , Leon Sallandt , Reinhold Schneider

Stochastic optimal principle leads to the resolution of a partial differential equation (PDE), namely the Hamilton-Jacobi-Bellman (HJB) equation. In general, this equation cannot be solved analytically, thus numerical algorithms are the…

数值分析 · 数学 2021-09-14 Christelle Dleuna Nyoumbi , Antoine Tambue

We introduce a new numerical method to approximate the solution of a finite horizon deterministic optimal control problem. We exploit two Hamilton-Jacobi-Bellman PDE, arising by considering the dynamics in forward and backward time. This…

最优化与控制 · 数学 2023-04-21 Marianne Akian , Stéphane Gaubert , Shanqing Liu

We study the problem of optimal control of dissipative quantum dynamics. Although under most circumstances dissipation leads to an increase in entropy (or a decrease in purity) of the system, there is an important class of problems for…

量子物理 · 物理学 2009-11-10 Shlomo E. Sklarz , David J. Tannor , Navin Khaneja

The control of a battery thermal management system (BTMS) is essential for the thermal safety, energy efficiency, and durability of electric vehicles (EVs) in hot weather. To address the battery cooling optimization problem, this paper…

系统与控制 · 电气工程与系统科学 2023-08-08 Yue Wu , Zhiwu Huang , Dongjun Li , Heng Li , Jun Peng , Daniel Stroe , Ziyou Song

We introduce a new and efficient numerical method for multicriterion optimal control and single criterion optimal control under integral constraints. The approach is based on extending the state space to include information on a "budget"…

最优化与控制 · 数学 2016-01-06 Ajeet Kumar , Alexander Vladimirsky

We propose a novel data-driven neural network (NN) optimization framework for solving an optimal stochastic control problem under stochastic constraints. Customized activation functions for the output layers of the NN are applied, which…

最优化与控制 · 数学 2023-06-21 Marc Chen , Mohammad Shirazi , Peter A. Forsyth , Yuying Li

Passivity-based control (PBC) for port-Hamiltonian systems provides an intuitive way of achieving stabilization by rendering a system passive with respect to a desired storage function. However, in most instances the control law is obtained…

系统与控制 · 计算机科学 2019-03-29 Olivier Sprangers , Gabriel A. D. Lopes , Robert Babuska

In this paper, we propose Q-learning algorithms for continuous-time deterministic optimal control problems with Lipschitz continuous controls. Our method is based on a new class of Hamilton-Jacobi-Bellman (HJB) equations derived from…

机器学习 · 计算机科学 2020-10-28 Jeongho Kim , Jaeuk Shin , Insoon Yang

This paper presents a two-stage framework for constrained near-optimal feedback control of input-affine nonlinear systems. An approximate value function for the unconstrained control problem is computed offline by solving the…

系统与控制 · 电气工程与系统科学 2026-03-18 Milad Alipour Shahraki , Laurent Lessard

In this paper we consider the optimal control of Hilbert space-valued infinite-dimensional Piecewise Deterministic Markov Processes (PDMP) and we prove that the corresponding value function can be represented via a Feynman-Kac type formula…

最优化与控制 · 数学 2019-06-07 Elena Bandini , Michele Thieullen
‹ 上一页 1 2 3 10 下一页 ›