中文
相关论文

相关论文: Data-based approximate policy iteration for nonlin…

200 篇论文

Efforts in this paper seek to combine graph theory with adaptive dynamic programming (ADP) as a reinforcement learning (RL) framework to determine forward-in-time, real-time, approximate optimal controllers for distributed multi-agent…

系统与控制 · 计算机科学 2017-07-25 Rushikesh Kamalapurkar , Huyen Dinh , Patrick Walters , Warren Dixon

Trajectory following is one of the complicated control problems when its dynamics are nonlinear, stochastic and include a large number of parameters. The problem has significant difficulties including a large number of trials required for…

机器人学 · 计算机科学 2019-02-14 Ali Lenjani

A new stochastic control model for the long-run environmental management of rivers is mathematically and numerically analyzed, focusing on a modern sediment replenishment problem with unique nonsmooth and nonlinear properties. Rational…

最优化与控制 · 数学 2022-03-11 Hidekazu Yoshioka , Motoh Tsujimura

This article investigates the core mechanisms of indirect data-driven control for unknown systems, focusing on the application of policy iteration (PI) within the context of the linear quadratic regulator (LQR) optimal control problem.…

系统与控制 · 电气工程与系统科学 2025-04-14 Bowen Song , Andrea Iannelli

A common pipeline in learning-based control is to iteratively estimate a model of system dynamics, and apply a trajectory optimization algorithm - e.g.~$\mathtt{iLQR}$ - on the learned model to minimize a target cost. This paper conducts a…

机器学习 · 计算机科学 2023-05-17 Daniel Pfrommer , Max Simchowitz , Tyler Westenbroek , Nikolai Matni , Stephen Tu

In this paper we investigate a dynamic stochastic portfolio optimization problem involving both the expected terminal utility and intertemporal utility maximization. We solve the problem by means of a solution to a fully nonlinear…

投资组合管理 · 定量金融 2019-03-26 Sona Kilianova , Daniel Sevcovic

In this paper, we propose a model-free adaptive learning solution for a model-following control problem. This approach employs policy iteration, to find an optimal adaptive control solution. It utilizes a moving finite-horizon of…

系统与控制 · 电气工程与系统科学 2023-02-07 Mohammed I. Abouheaf , Hashim A. Hashim , Mohammad A. Mayyas , Kyriakos G. Vamvoudakis

In this paper, further extensions of the result of the paper "A successive approximation method in functional spaces for hierarchical optimal control problems and its application to learning, arXiv:2410.20617 [math.OC], 2024" concerning a…

最优化与控制 · 数学 2024-11-26 Getachew K. Befekadu

Recent results in the study of the Hamilton Jacobi Bellman (HJB) equation have led to the discovery of a formulation of the value function as a linear Partial Differential Equation (PDE) for stochastic nonlinear systems with a mild…

最优化与控制 · 数学 2014-02-13 Matanya B. Horowitz , Joel W. Burdick

This paper presents a new method for solving a class of nonlinear optimal control problems with a quadratic performance index. In this method, first the original optimal control problem is transformed into a nonlinear two-point boundary…

最优化与控制 · 数学 2014-09-18 Amin Jajarmi , Hamidreza Ramezanpour , Arman Sargolzaei , Pouyan Shafaei

Recently, a novel class of Approximate Policy Iteration (API) algorithms have demonstrated impressive practical performance (e.g., ExIt from [2], AlphaGo-Zero from [27]). This new family of algorithms maintains, and alternately optimizes,…

机器学习 · 计算机科学 2019-04-09 Wen Sun , Geoffrey J. Gordon , Byron Boots , J. Andrew Bagnell

This paper considers linear-quadratic control of a non-linear dynamical system subject to arbitrary cost. I show that for this class of stochastic control problems the non-linear Hamilton-Jacobi-Bellman equation can be transformed into a…

综合物理 · 物理学 2009-11-11 H. J. Kappen

A learning based method for obtaining feedback laws for nonlinear optimal control problems is proposed. The learning problem is posed such that the open loop value function is its optimal solution. This infinite dimensional, function space,…

最优化与控制 · 数学 2022-10-26 Karl Kunisch , Donato Vásquez-Varas , Daniel Walter

This paper investigates a class of optimal control problems associated with Markov processes with local state information. The decision-maker has only local access to a subset of a state vector information as often encountered in…

系统与控制 · 电气工程与系统科学 2020-05-12 Guanze Peng , Veeraruna Kavitha , Qunayan Zhu

In this paper we consider a broad class of infinite horizon discrete-time optimal control models that involve a nonnegative cost function and an affine mapping in their dynamic programming equation. They include as special cases classical…

最优化与控制 · 数学 2017-11-29 Dimitri Bertsekas

One of the most crucial challenges faced by the Li-ion battery community concerns the search for the minimum time charging without irreversibly damaging the cells. This can fall into solving large-scale nonlinear optimal control problems…

系统与控制 · 电气工程与系统科学 2020-06-26 Saehong Park , Andrea Pozzi , Michael Whitmeyer , Hector Perez , Won Tae Joe , Davide M Raimondo , Scott Moura

We study an inverse problem of the stochastic optimal control of general diffusions with performance index having the quadratic penalty term of the control process. Under mild conditions on the system dynamics, the cost functions, and the…

最优化与控制 · 数学 2022-11-17 Yumiharu Nakano

The numerical realization of the dynamic programming principle for continuous-time optimal control leads to nonlinear Hamilton-Jacobi-Bellman equations which require the minimization of a nonlinear mapping over the set of admissible…

最优化与控制 · 数学 2015-02-26 Dante Kalise , Axel Kröner , Karl Kunisch

We present a novel approach to control design for nonlinear systems which leverages model-free policy optimization techniques to learn a linearizing controller for a physical plant with unknown dynamics. Feedback linearization is a…

We study the properties of the value function associated with an optimal control problem with uncertainties, known as average or Riemann-Stieltjes problem. Uncertainties are assumed to belong to a compact metric probability space, and…

最优化与控制 · 数学 2024-07-19 M. Soledad Aronna , Michele Palladino , Oscar Sierra
‹ 上一页 1 8 9 10 下一页 ›