中文
相关论文

相关论文: Data-based approximate policy iteration for nonlin…

200 篇论文

Nonlinear optimal control is vital for numerous applications but remains challenging for unknown systems due to the difficulties in accurately modelling dynamics and handling computational demands, particularly in high-dimensional settings.…

系统与控制 · 电气工程与系统科学 2024-12-03 Zhexuan Zeng , Ruikun Zhou , Yiming Meng , Jun Liu

We present a semi-real-time algorithm for minimal-time optimal path planning based on optimal control theory, dynamic programming, and Hamilton-Jacobi (HJ) equations. Partial differential equation (PDE) based optimal path planning methods…

最优化与控制 · 数学 2023-09-06 Christian Parkinson , Kyle Polage

A new approach to feedback control design based on optimal control is proposed. Instead of expensive computations of the value function for different penalties on the states and inputs, we use a control Lyapunov function that amounts to be…

最优化与控制 · 数学 2021-11-22 Taouba Jouini , Anders Rantzer

In this paper time-driven learning refers to the machine learning method that updates parameters in a prediction model continuously as new data arrives. Among existing approximate dynamic programming (ADP) and reinforcement learning (RL)…

系统与控制 · 电气工程与系统科学 2020-06-17 Qingtao Zhao , Jennie Si , Jian Sun

In this paper, we propose a generalized successive approximation method (SAM), called invariantly admissible policy iteration (PI), for finding the solution to a class of input-affine nonlinear optimal control problems by iterations. Unlike…

最优化与控制 · 数学 2014-05-28 Jae Youg Lee , Jin Bae Park , Yoon Ho Choi

We propose a novel formulation for approximating reachable sets through a minimum discounted reward optimal control problem. The formulation yields a continuous solution that can be obtained by solving a Hamilton-Jacobi equation.…

最优化与控制 · 数学 2018-09-05 Anayo K. Akametalu , Shromona Ghosh , Jaime F. Fisac , Claire J. Tomlin

This paper proposes efficient policy iteration and value iteration algorithms for the continuous-time linear quadratic regulator problem with unmeasurable states and unknown system dynamics, from the perspective of direct data-driven…

系统与控制 · 电气工程与系统科学 2026-03-17 Jun Xie , Yuan-Hua Ni , Yiqin Yang , Bo Xu

Two key challenges in optimal control include efficiently solving high-dimensional problems and handling optimal control problems with state-dependent running costs. In this paper, we consider a class of optimal control problems whose…

最优化与控制 · 数学 2023-05-16 Paula Chen , Jérôme Darbon , Tingwei Meng

Model-based reinforcement learning methods learn a dynamics model with real data sampled from the environment and leverage it to generate simulated data to derive an agent. However, due to the potential distribution mismatch between…

机器学习 · 计算机科学 2020-10-29 Jian Shen , Han Zhao , Weinan Zhang , Yong Yu

We address the problem of combined stochastic and impulse control for a market maker operating in a limit order book. The problem is formulated as a Hamilton-Jacobi-Bellman quasi-variational inequality (HJBQVI). We propose an implicit…

数理金融 · 定量金融 2025-12-25 Alexey Meteykin

Adaptive optimal control of nonlinear dynamic systems with deterministic and known dynamics under a known undiscounted infinite-horizon cost function is investigated. Policy iteration scheme initiated using a stabilizing initial control is…

系统与控制 · 计算机科学 2015-05-21 Ali Heydari

We consider the problem of learning a linear control policy for a linear dynamical system, from demonstrations of an expert regulating the system. The standard approach to this problem is policy fitting, which fits a linear policy by…

最优化与控制 · 数学 2020-01-22 Malayandi Palan , Shane Barratt , Alex McCauley , Dorsa Sadigh , Vikas Sindhwani , Stephen Boyd

This paper studies approximate policy iteration (API) methods which use least-squares Bellman error minimization for policy evaluation. We address several of its enhancements, namely, Bellman error minimization using instrumental variables,…

最优化与控制 · 数学 2014-01-07 Warren R. Scott , Warren B. Powell , Somayeh Moazehi

This work addresses stochastic optimal control problems where the unknown state evolves in continuous time while partial, noisy, and possibly controllable measurements are only available in discrete time. We develop a framework for…

最优化与控制 · 数学 2025-08-19 Christian Bayer , Boualem Djehiche , Eliza Rezvanova , Raul Fidel Tempone

This article approaches deterministic filtering via an application of the min-plus linearity of the corresponding dynamic programming operator. This filter design method yields a set-valued state estimator for discrete-time nonlinear…

最优化与控制 · 数学 2012-03-14 Abhijit G. Kallapur , Srinivas Sridharan , William M. McEneaney , Ian R. Petersen

The ergodic control problem for a non-degenerate controlled diffusion controlled through its drift is considered under a uniform stability condition that ensures the well-posedness of the associated Hamilton-Jacobi-Bellman (HJB) equation. A…

最优化与控制 · 数学 2019-03-20 Ari Arapostathis , Vivek S. Borkar

This paper studies the statistical theory of batch data reinforcement learning with function approximation. Consider the off-policy evaluation problem, which is to estimate the cumulative value of a new target policy from logged history…

机器学习 · 计算机科学 2020-02-25 Yaqi Duan , Mengdi Wang

Many applications require solving non-linear control problems that are classically not well behaved. This paper develops a simple and efficient chattering algorithm that learns near optimal decision policies through an open-loop feedback…

机器学习 · 计算机科学 2017-03-21 Peeyush Kumar , Wolf Kohn , Zelda B. Zabinsky

In this paper, we propose a new policy iteration algorithm to compute the value function and the optimal controls of continuous time stochastic control problems. The algorithm relies on successive approximations using linear-quadratic…

最优化与控制 · 数学 2024-09-09 Dylan Possamaï , Ludovic Tangpi

Hamilton-Jacobi (HJ) reachability analysis is a widely used method for ensuring the safety of robotic systems. Traditional approaches compute reachable sets by numerically solving an HJ Partial Differential Equation (PDE) over a grid, which…

机器人学 · 计算机科学 2025-05-08 Zeyuan Feng , Le Qiu , Somil Bansal