中文
相关论文

相关论文: Policy Iteration for Stationary Discounted Hamilto…

200 篇论文

We study the convergence rates of policy iteration (PI) for nonconvex viscous Hamilton--Jacobi equations using a discrete space-time scheme, where both space and time variables are discretized. We analyze the case with an uncontrolled…

数值分析 · 数学 2025-03-05 Xiaoqin Guo , Hung Vinh Tran , Yuming Paul Zhang

Policy iteration (PI) is a widely used algorithm for synthesizing optimal feedback control policies across many engineering and scientific applications. When PI is deployed on infinite-horizon, nonlinear, autonomous optimal-control…

最优化与控制 · 数学 2025-07-15 Tobias Ehring , Behzad Azmi , Bernard Haasdonk

We treat infinite horizon optimal control problems by solving the associated stationary Hamilton-Jacobi-Bellman (HJB) equation numerically to compute the value function and an optimal feedback law. The dynamical systems under consideration…

最优化与控制 · 数学 2021-05-19 Mathias Oster , Leon Sallandt , Reinhold Schneider

This paper is concerned with the convergence rate of policy iteration for (deterministic) optimal control problems in continuous time. To overcome the problem of ill-posedness due to lack of regularity, we consider a semi-discrete scheme by…

最优化与控制 · 数学 2025-04-11 Wenpin Tang , Hung Vinh Tran , Yuming Paul Zhang

H{\infty} control of nonlinear continuous-time system depends on the solution of the Hamilton-Jacobi-Isaacs (HJI) equation, which has been proved impossible to obtain a closed-form solution due to the nonlinearity of HJI equation. In order…

系统与控制 · 电气工程与系统科学 2024-03-20 Qi Wang

The framework of deep operator network (DeepONet) has been widely exploited thanks to its capability of solving high dimensional partial differential equations. In this paper, we incorporate DeepONet with a recently developed policy…

最优化与控制 · 数学 2024-06-18 Jae Yong Lee , Yeoneung Kim

We present an accelerated algorithm for the solution of static Hamilton-Jacobi-Bellman equations related to optimal control problems. Our scheme is based on a classic policy iteration procedure, which is known to have superlinear…

最优化与控制 · 数学 2016-02-22 Alessandro Alla , Maurizio Falcone , Dante Kalise

Stochastic optimal principle leads to the resolution of a partial differential equation (PDE), namely the Hamilton-Jacobi-Bellman (HJB) equation. In general, this equation cannot be solved analytically, thus numerical algorithms are the…

数值分析 · 数学 2021-09-14 Christelle Dleuna Nyoumbi , Antoine Tambue

We study the policy iteration algorithm (PIA) for entropy-regularized stochastic control problems on an infinite time horizon with a large discount rate, focusing on two main scenarios. First, we analyze PIA with bounded coefficients where…

最优化与控制 · 数学 2025-05-28 Hung Vinh Tran , Zhenhua Wang , Yuming Paul Zhang

This paper addresses the model-free nonlinear optimal problem with generalized cost functional, and a data-based reinforcement learning technique is developed. It is known that the nonlinear optimal control problem relies on the solution of…

系统与控制 · 计算机科学 2013-11-20 Biao Luo , Huai-Ning Wu , Tingwen Huang , Derong Liu

This paper presents a {\delta}-PI algorithm which is based on damped Newton method for the H{\infty} tracking control problem of unknown continuous-time nonlinear system. A discounted performance function and an augmented system are used to…

机器学习 · 计算机科学 2024-01-24 Qi Wang

In this article we study a finite horizon optimal control problem with monotone controls. We consider the associated Hamilton-Jacobi-Bellman (HJB) equation which characterizes the value function. We consider the totally discretized problem…

最优化与控制 · 数学 2014-07-08 Eduardo A. Philipp , Laura S. Aragone , Lisandro A. Parente

For a general entropy-regularized stochastic control problem on an infinite horizon, we prove that a policy iteration algorithm (PIA) converges to an optimal relaxed control. Contrary to the standard stochastic control literature, classical…

最优化与控制 · 数学 2026-05-14 Yu-Jui Huang , Zhenhua Wang , Zhou Zhou

In this paper infinite horizon optimal control problems for nonlinear high-dimensional dynamical systems are studied. Nonlinear feedback laws can be computed via the value function characterized as the unique viscosity solution to the…

最优化与控制 · 数学 2016-02-22 Alessandro Alla , Maurizio Falcone , Stefan Volkwein

This paper investigates the optimal control problems for the finite-horizon continuous-time Markov decision processes with delay-dependent control policies. We develop compactification methods in decision processes, and show that the…

概率论 · 数学 2023-07-06 Zhong-Wei Liao , Jinghai Shao

Policy iteration is a widely used technique to solve the Hamilton Jacobi Bellman (HJB) equation, which arises from nonlinear optimal feedback control theory. Its convergence analysis has attracted much attention in the unconstrained case.…

最优化与控制 · 数学 2020-05-19 Sudeep Kundu , Karl Kunisch

We introduce a new and efficient numerical method for multicriterion optimal control and single criterion optimal control under integral constraints. The approach is based on extending the state space to include information on a "budget"…

最优化与控制 · 数学 2016-01-06 Ajeet Kumar , Alexander Vladimirsky

We address the problem of combined stochastic and impulse control for a market maker operating in a limit order book. The problem is formulated as a Hamilton-Jacobi-Bellman quasi-variational inequality (HJBQVI). We propose an implicit…

数理金融 · 定量金融 2025-12-25 Alexey Meteykin

We introduce a new numerical method to approximate the solution of a finite horizon deterministic optimal control problem. We exploit two Hamilton-Jacobi-Bellman PDE, arising by considering the dynamics in forward and backward time. This…

最优化与控制 · 数学 2023-04-21 Marianne Akian , Stéphane Gaubert , Shanqing Liu

This paper considers optimal control of dynamical systems which are represented by nonlinear stochastic differential equations. It is well-known that the optimal control policy for this problem can be obtained as a function of a value…

机器人学 · 计算机科学 2014-05-30 Oktay Arslan , Evangelos Theodorou , Panagiotis Tsiotras
‹ 上一页 1 2 3 10 下一页 ›