English
Related papers

Related papers: Data-based approximate policy iteration for nonlin…

200 papers

Nonlinear optimal control is vital for numerous applications but remains challenging for unknown systems due to the difficulties in accurately modelling dynamics and handling computational demands, particularly in high-dimensional settings.…

Systems and Control · Electrical Eng. & Systems 2024-12-03 Zhexuan Zeng , Ruikun Zhou , Yiming Meng , Jun Liu

We present a semi-real-time algorithm for minimal-time optimal path planning based on optimal control theory, dynamic programming, and Hamilton-Jacobi (HJ) equations. Partial differential equation (PDE) based optimal path planning methods…

Optimization and Control · Mathematics 2023-09-06 Christian Parkinson , Kyle Polage

A new approach to feedback control design based on optimal control is proposed. Instead of expensive computations of the value function for different penalties on the states and inputs, we use a control Lyapunov function that amounts to be…

Optimization and Control · Mathematics 2021-11-22 Taouba Jouini , Anders Rantzer

In this paper time-driven learning refers to the machine learning method that updates parameters in a prediction model continuously as new data arrives. Among existing approximate dynamic programming (ADP) and reinforcement learning (RL)…

Systems and Control · Electrical Eng. & Systems 2020-06-17 Qingtao Zhao , Jennie Si , Jian Sun

In this paper, we propose a generalized successive approximation method (SAM), called invariantly admissible policy iteration (PI), for finding the solution to a class of input-affine nonlinear optimal control problems by iterations. Unlike…

Optimization and Control · Mathematics 2014-05-28 Jae Youg Lee , Jin Bae Park , Yoon Ho Choi

We propose a novel formulation for approximating reachable sets through a minimum discounted reward optimal control problem. The formulation yields a continuous solution that can be obtained by solving a Hamilton-Jacobi equation.…

Optimization and Control · Mathematics 2018-09-05 Anayo K. Akametalu , Shromona Ghosh , Jaime F. Fisac , Claire J. Tomlin

This paper proposes efficient policy iteration and value iteration algorithms for the continuous-time linear quadratic regulator problem with unmeasurable states and unknown system dynamics, from the perspective of direct data-driven…

Systems and Control · Electrical Eng. & Systems 2026-03-17 Jun Xie , Yuan-Hua Ni , Yiqin Yang , Bo Xu

Two key challenges in optimal control include efficiently solving high-dimensional problems and handling optimal control problems with state-dependent running costs. In this paper, we consider a class of optimal control problems whose…

Optimization and Control · Mathematics 2023-05-16 Paula Chen , Jérôme Darbon , Tingwei Meng

Model-based reinforcement learning methods learn a dynamics model with real data sampled from the environment and leverage it to generate simulated data to derive an agent. However, due to the potential distribution mismatch between…

Machine Learning · Computer Science 2020-10-29 Jian Shen , Han Zhao , Weinan Zhang , Yong Yu

We address the problem of combined stochastic and impulse control for a market maker operating in a limit order book. The problem is formulated as a Hamilton-Jacobi-Bellman quasi-variational inequality (HJBQVI). We propose an implicit…

Mathematical Finance · Quantitative Finance 2025-12-25 Alexey Meteykin

Adaptive optimal control of nonlinear dynamic systems with deterministic and known dynamics under a known undiscounted infinite-horizon cost function is investigated. Policy iteration scheme initiated using a stabilizing initial control is…

Systems and Control · Computer Science 2015-05-21 Ali Heydari

We consider the problem of learning a linear control policy for a linear dynamical system, from demonstrations of an expert regulating the system. The standard approach to this problem is policy fitting, which fits a linear policy by…

Optimization and Control · Mathematics 2020-01-22 Malayandi Palan , Shane Barratt , Alex McCauley , Dorsa Sadigh , Vikas Sindhwani , Stephen Boyd

This paper studies approximate policy iteration (API) methods which use least-squares Bellman error minimization for policy evaluation. We address several of its enhancements, namely, Bellman error minimization using instrumental variables,…

Optimization and Control · Mathematics 2014-01-07 Warren R. Scott , Warren B. Powell , Somayeh Moazehi

This work addresses stochastic optimal control problems where the unknown state evolves in continuous time while partial, noisy, and possibly controllable measurements are only available in discrete time. We develop a framework for…

Optimization and Control · Mathematics 2025-08-19 Christian Bayer , Boualem Djehiche , Eliza Rezvanova , Raul Fidel Tempone

This article approaches deterministic filtering via an application of the min-plus linearity of the corresponding dynamic programming operator. This filter design method yields a set-valued state estimator for discrete-time nonlinear…

Optimization and Control · Mathematics 2012-03-14 Abhijit G. Kallapur , Srinivas Sridharan , William M. McEneaney , Ian R. Petersen

The ergodic control problem for a non-degenerate controlled diffusion controlled through its drift is considered under a uniform stability condition that ensures the well-posedness of the associated Hamilton-Jacobi-Bellman (HJB) equation. A…

Optimization and Control · Mathematics 2019-03-20 Ari Arapostathis , Vivek S. Borkar

This paper studies the statistical theory of batch data reinforcement learning with function approximation. Consider the off-policy evaluation problem, which is to estimate the cumulative value of a new target policy from logged history…

Machine Learning · Computer Science 2020-02-25 Yaqi Duan , Mengdi Wang

Many applications require solving non-linear control problems that are classically not well behaved. This paper develops a simple and efficient chattering algorithm that learns near optimal decision policies through an open-loop feedback…

Machine Learning · Computer Science 2017-03-21 Peeyush Kumar , Wolf Kohn , Zelda B. Zabinsky

In this paper, we propose a new policy iteration algorithm to compute the value function and the optimal controls of continuous time stochastic control problems. The algorithm relies on successive approximations using linear-quadratic…

Optimization and Control · Mathematics 2024-09-09 Dylan Possamaï , Ludovic Tangpi

Hamilton-Jacobi (HJ) reachability analysis is a widely used method for ensuring the safety of robotic systems. Traditional approaches compute reachable sets by numerically solving an HJ Partial Differential Equation (PDE) over a grid, which…

Robotics · Computer Science 2025-05-08 Zeyuan Feng , Le Qiu , Somil Bansal