中文
相关论文

相关论文: Data-based approximate policy iteration for nonlin…

200 篇论文

We present a numerical method for generating the state-feedback control policy associated with general undiscounted, constant-setpoint, infinite-horizon, nonlinear optimal control problems with continuous state variables. The method is…

最优化与控制 · 数学 2021-04-23 Jonathan Lock , Tomas McKelvey

We study the policy iteration algorithm (PIA) for entropy-regularized stochastic control problems on an infinite time horizon with a large discount rate, focusing on two main scenarios. First, we analyze PIA with bounded coefficients where…

最优化与控制 · 数学 2025-05-28 Hung Vinh Tran , Zhenhua Wang , Yuming Paul Zhang

In this paper, we introduce Hamilton-Jacobi-Bellman (HJB) equations for Q-functions in continuous time optimal control problems with Lipschitz continuous controls. The standard Q-function used in reinforcement learning is shown to be the…

最优化与控制 · 数学 2020-05-05 Jeongho Kim , Insoon Yang

In this paper, we study the delayed stochastic recursive optimal control problem with a non-Lipschitz generator, in which both the dynamics of the control system and the recursive cost functional depend on the past path segment of the state…

最优化与控制 · 数学 2023-12-27 Jiaqiang Wen , Zhen Wu , Qi Zhang

In this paper, we investigate the distributed optimal control problem for a kind of nonlinear multi-agent systems. In particular,both the state and the system dynamic structures of each agent are private and can only be shared among…

最优化与控制 · 数学 2026-04-08 Ruixue Li , Wenjing Yang , Zhaorong Zhang , Xun Li , Juanjuan Xu

We propose a variant of consensus-based optimization (CBO) algorithms, controlled-CBO, which introduces a feedback control term to improve convergence towards global minimizers of non-convex functions in multiple dimensions. The feedback…

最优化与控制 · 数学 2025-07-29 Yuyang Huang , Michael Herty , Dante Kalise , Nikolas Kantas

Stochastic optimal control control problems with merely measurable coefficients are not well understood. In this manuscript, we consider fully non-linear stochastic optimal control problems in infinite horizon with measurable coefficients…

最优化与控制 · 数学 2026-05-21 Filippo de Feo

We consider a class of learning problem of point estimation for modeling high-dimensional nonlinear functions, whose learning dynamics is guided by model training dataset, while the estimated parameter in due course provides an acceptable…

最优化与控制 · 数学 2024-10-29 Getachew K. Befekadu

Neural network approaches that parameterize value functions have succeeded in approximating high-dimensional optimal feedback controllers when the Hamiltonian admits explicit formulas. However, many practical problems, such as the space…

最优化与控制 · 数学 2025-10-08 Eric Gelphman , Deepanshu Verma , Nicole Tianjiao Yang , Stanley Osher , Samy Wu Fung

We consider the infinite-horizon discounted optimal control problem formalized by Markov Decision Processes. We focus on several approximate variations of the Policy Iteration algorithm: Approximate Policy Iteration, Conservative Policy…

人工智能 · 计算机科学 2014-05-13 Bruno Scherrer

In deterministic systems, reinforcement learning-based online approximate optimal control methods typically require a restrictive persistence of excitation (PE) condition for convergence. This paper presents a concurrent learning-based…

系统与控制 · 计算机科学 2017-07-25 Rushikesh Kamalapurkar , Patrick Walters , Warren Dixon

This paper proposes a fully data-driven approach for optimal control of nonlinear control-affine systems represented by a stochastic diffusion. The focus is on the scenario where both the nonlinear dynamics and stage cost functions are…

最优化与控制 · 数学 2025-11-03 Nicolas Hoischen , Petar Bevanda , Stefan Sosnowski , Sandra Hirche , Boris Houska

We present a neural network approach for approximating the value function of high-dimensional stochastic control problems. Our training process simultaneously updates our value function estimate and identifies the part of the state space…

最优化与控制 · 数学 2024-05-08 Xingjian Li , Deepanshu Verma , Lars Ruthotto

This paper presents SIMPOL (Simplified Policy Iteration), a modular numerical framework for solving continuous-time heterogeneous agent models. The core economic problem, the optimization of consumption and savings under idiosyncratic…

计算金融 · 定量金融 2025-09-30 Ricardo Alonzo Fernández Salguero

We study the problem of learning the optimal control policy for fine-tuning a given diffusion process, using general value function approximation. We develop a new class of algorithms by solving a variational inequality problem based on the…

机器学习 · 计算机科学 2025-09-03 Wenlong Mou

Reachability analysis is important for studying optimal control problems and differential games, which are powerful theoretical tools for analyzing and modeling many practical problems in robotics, aircraft control, among other application…

最优化与控制 · 数学 2016-03-22 Mo Chen , Claire J. Tomlin

We are motivated by the real challenges presented in a human-robot system to develop new designs that are efficient at data level and with performance guarantees such as stability and optimality at systems level. Existing…

系统与控制 · 电气工程与系统科学 2021-01-19 Xiang Gao , Jennie Si , Yue Wen , Minhan Li , He , Huang

We present a simple and easy to implement method for the numerical solution of a rather general class of Hamilton-Jacobi-Bellman (HJB) equations. In many cases, the considered problems have only a viscosity solution, to which, fortunately,…

计算金融 · 定量金融 2011-02-17 Jan Hendrik Witte , Christoph Reisinger

This paper presents a new methodology to craft navigation functions for nonlinear systems with stochastic uncertainty. The method relies on the transformation of the Hamilton-Jacobi-Bellman (HJB) equation into a linear partial differential…

机器人学 · 计算机科学 2014-09-23 Matanya B. Horowitz , Joel W. Burdick

This work proposes a novel numerical scheme for solving the high-dimensional Hamilton-Jacobi-Bellman equation with a functional hierarchical tensor ansatz. We consider the setting of stochastic control, whereby one applies control to a…

数值分析 · 数学 2025-07-01 Xun Tang , Nan Sheng , Lexing Ying