中文
相关论文

相关论文: Learning-based Hamilton-Jacobi-Bellman Methods for…

200 篇论文

An off policy reinforcement learning based control strategy is developed for the optimal tracking control problem to achieve the prescribed performance of full states during the learning process. The optimal tracking control problem is…

系统与控制 · 电气工程与系统科学 2020-09-02 C. Li , Y. Wang , F. Liu , M. Buss

This paper proposes penalty schemes for a class of weakly coupled systems of Hamilton-Jacobi-Bellman quasi-variational inequalities (HJBQVIs) arising from stochastic hybrid control problems of regime-switching models with both continuous…

最优化与控制 · 数学 2020-01-06 Christoph Reisinger , Yufei Zhang

This paper is concerned with stochastic impulse control problems in which the running cost changes depending on the impulse control. Because of such a dependence, it brings several difficulties when the usual dynamic programming principle…

最优化与控制 · 数学 2025-11-11 Yuchen Cao , Jiongmin Yong

Learning-based driving solution, a new branch for autonomous driving, is expected to simplify the modeling of driving by learning the underlying mechanisms from data. To improve the tactical decision-making for learning-based driving…

机器人学 · 计算机科学 2020-05-11 Jingke Wang , Yue Wang , Dongkun Zhang , Yezhou Yang , Rong Xiong

Classically, the optimal control problem in the presence of an adversary is formulated as a two-player zero-sum differential game or an $H_\infty$ control problem. The solution to these problems can be obtained by solving the…

最优化与控制 · 数学 2022-04-26 Alexander Krolicki , Sarang Sutavani , Umesh Vaidya

In this paper, we aim to solve the high dimensional stochastic optimal control problem from the view of the stochastic maximum principle via deep learning. By introducing the extended Hamiltonian system which is essentially an FBSDE with a…

最优化与控制 · 数学 2021-06-23 Shaolin Ji , Shige Peng , Ying Peng , Xichuan Zhang

In this paper we study a first extension of the theory of mild solutions for HJB equations in Hilbert spaces to the case when the domain is not the whole space. More precisely, we consider a half-space as domain, and a semilinear…

最优化与控制 · 数学 2022-09-30 Alessandro Calvia , Gianluca Cappa , Fausto Gozzi , Enrico Priola

We study an optimal investment and consumption problem over a finite-time horizon, in which an individual invests in a risk-free asset and a risky asset, and evaluate utility using a general utility function that exhibits loss aversion with…

最优化与控制 · 数学 2025-07-08 Chonghu Guan , Xinfeng Gu , Wenhao Zhang , Xun Li

Stochastic optimal control control problems with merely measurable coefficients are not well understood. In this manuscript, we consider fully non-linear stochastic optimal control problems in infinite horizon with measurable coefficients…

最优化与控制 · 数学 2026-05-21 Filippo de Feo

Optimal control and the associated second-order path-dependent Hamilton-Jacobi-Bellman (PHJB) equation are studied for unbounded functional stochastic evolution systems in Hilbert spaces. The notion of viscosity solution without…

最优化与控制 · 数学 2024-02-27 Shanjian Tang , Jianjun Zhou

We present a framework to \emph{certify} Hamilton--Jacobi (HJ) reachability learned by reinforcement learning (RL). Building on a discounted initial time \emph{travel-cost} formulation that makes small-step RL value iteration provably…

系统与控制 · 电气工程与系统科学 2026-02-19 Prashant Solanki , Isabelle El-Hajj , Jasper J. van Beers , Erik-Jan van Kampen , Coen C. de Visser

Equipping approximate dynamic programming (ADP) with inputconstraints has a tremendous significance. This enables ADP to be applied tothe systems with actuator limitations, which is quite common for dynamicalsystems. In a conventional…

最优化与控制 · 数学 2018-05-24 Xuefeng Bao , Zhi-Hong Mao , Nitin Sharma

We address the problem of computing a control for a time-dependent nonlinear system to reach a target set in a minimal time. To solve this minimal time control problem, we introduce a hierarchy of linear semi-infinite programs, the values…

最优化与控制 · 数学 2023-07-04 Antoine Oustry , Matteo Tacchi

In this paper we consider an optimal investment and reinsurance problem with partially unknown model parameters which are allowed to be learned. The model includes multiple business lines and dependence between them. The aim is to maximize…

最优化与控制 · 数学 2025-10-16 Nicole Bäuerle , Gregor Leimcke

In this paper, further extensions of the result of the paper "A successive approximation method in functional spaces for hierarchical optimal control problems and its application to learning, arXiv:2410.20617 [math.OC], 2024" concerning a…

最优化与控制 · 数学 2024-11-26 Getachew K. Befekadu

In this article, a notion of viscosity solutions is introduced for second order path-dependent Hamilton-Jacobi-Bellman (PHJB) equations associated with optimal control problems for path-dependent stochastic evolution equations in Hilbert…

概率论 · 数学 2020-09-14 Jianjun Zhou

We investigate feedback control for infinite horizon optimal control problems for partial differential equations. The method is based on the coupling between Hamilton-Jacobi-Bellman (HJB) equations and model reduction techniques. It is…

最优化与控制 · 数学 2016-07-11 Alessandro Alla , Andreas Schmidt , Bernard Haasdonk

We study a time-optimal control problem of a two-peakon collision. First, we state the controllability. Next, we find the time-optimal strategy. This is done via the HamiltonJacobi-Bellman equation and the dynamic programming method. We…

最优化与控制 · 数学 2021-05-24 Tomasz Cieślak , Bidesh Das

In this work, we propose a class of numerical schemes for solving semilinear Hamilton-Jacobi-Bellman-Isaacs (HJBI) boundary value problems which arise naturally from exit time problems of diffusion processes with controlled drift. We…

数值分析 · 数学 2020-02-14 Kazufumi Ito , Christoph Reisinger , Yufei Zhang

Deep Reinforcement Learning (RL) has shown remarkable success in robotics with complex and heterogeneous dynamics. However, its vulnerability to unknown disturbances and adversarial attacks remains a significant challenge. In this paper, we…

机器人学 · 计算机科学 2024-10-01 Hanyang Hu , Xilun Zhang , Xubo Lyu , Mo Chen