中文
相关论文

相关论文: Model-free Value Iteration Algorithm for Continuou…

200 篇论文

We investigate the problem of learning an $\epsilon$-approximate solution for the discrete-time Linear Quadratic Regulator (LQR) problem via a Stochastic Variance-Reduced Policy Gradient (SVRPG) approach. Whilst policy gradient methods have…

最优化与控制 · 数学 2023-09-20 Leonardo F. Toso , Han Wang , James Anderson

This paper is concerned with a stochastic linear quadratic (LQ, for short) control problem with a recursive cost functional. It involves BSDEs in $L^1$ whose well-posedness is a subtle issue. A suitable framework has been adopted so that…

最优化与控制 · 数学 2026-01-30 Lin Li , Jiongmin Yong

We study linear-quadratic optimal control problems for Voterra systems, and problems that are linear-quadratic in the control but generally nonlinear in the state. In the case of linear-quadratic Volterra control, we obtain sharp necessary…

最优化与控制 · 数学 2021-01-14 S. A. Belbas

Iterative linear quadratic regulator (iLQR) has gained wide popularity in addressing trajectory optimization problems with nonlinear system models. However, as a model-based shooting method, it relies heavily on an accurate system model to…

机器学习 · 计算机科学 2022-09-16 Zilong Cheng , Yulin Li , Kai Chen , Jun Ma , Tong Heng Lee

This communication presents a longitudinal model-free control approach for computing the wheel torque command to be applied on a vehicle. This setting enables us to overcome the problem of unknown vehicle parameters for generating a…

系统与控制 · 计算机科学 2017-04-06 Philip Polack , Brigitte d'Andréa-Novel , Michel Fliess , Arnaud de la Fortelle , Lghani Menhour

A mixed linear quadratic (MLQ, for short) optimal control problem is considered. The controlled stochastic system consists of two diffusion processes which are in different time horizons. There are two control actions: a standard control…

最优化与控制 · 数学 2012-12-05 Jianhui Huang , Xun Li , Jiongmin Yong

This paper studies the stochastic optimal control problem for systems with unknown dynamics. First, an open-loop deterministic trajectory optimization problem is solved without knowing the explicit form of the dynamical system. Next, a…

系统与控制 · 计算机科学 2017-05-30 Dan Yu , Mohammadhussein Rafieisakhaei , Suman Chakravorty

In this paper, we investigate a sparse optimal control of continuous-time stochastic systems. We adopt the dynamic programming approach and analyze the optimal control via the value function. Due to the non-smoothness of the $L^0$ cost…

最优化与控制 · 数学 2021-09-17 Kaito Ito , Takuya Ikeda , Kenji Kashima

This paper offers a unified perspective on different approaches to the solution of optimal control problems through the lens of constrained sequential quadratic programming. In particular, it allows us to find the relationships between…

最优化与控制 · 数学 2025-10-07 Abhijeet , Suman Chakravorty

Linear-Quadratic (LQ) problems that arise in systems and controls include the classical optimal control problems of the Linear Quadratic Regulator (LQR) in both its deterministic and stochastic forms, as well as $H^\infty$-analysis (the…

系统与控制 · 电气工程与系统科学 2024-01-04 Bassam Bamieh

This paper introduces a generalization of the well-known Riccati recursion for solving the discrete-time equality-constrained linear quadratic optimal control problem. The recursion can be used to compute the solutions as well as optimal…

最优化与控制 · 数学 2024-12-31 Lander Vanroye , Joris De Schutter , Wilm Decré

The linear quadratic regulator is the fundamental problem of optimal control. Its state feedback version was set and solved in the early 1960s. However the static output feedback problem has no explicit-form solution. It is suggested to…

最优化与控制 · 数学 2020-11-03 Ilyas Fatkhullin , Boris Polyak

This paper considers a stochastic linear quadratic problem for discrete-time systems with multiplicative noises over an infinite horizon. To obtain the optimal solution, we propose an online iterative algorithm of reinforcement learning…

最优化与控制 · 数学 2023-11-22 Hongdan Li , Lucky Qiaofeng Li , Xun Li , Zhaorong Zhang

We explore reinforcement learning methods for finding the optimal policy in the linear quadratic regulator (LQR) problem. In particular, we consider the convergence of policy gradient methods in the setting of known and unknown parameters.…

机器学习 · 计算机科学 2021-06-25 Ben Hambly , Renyuan Xu , Huining Yang

We consider the problem of controlling an unknown linear dynamical system under a stochastic convex cost and full feedback of both the state and cost function. We present a computationally efficient algorithm that attains an optimal…

最优化与控制 · 数学 2022-06-23 Asaf Cassel , Alon Cohen , Tomer Koren

A self-learning optimal control algorithm for episodic fixed-horizon manufacturing processes with time-discrete control actions is proposed and evaluated on a simulated deep drawing process. The control model is built during consecutive…

系统与控制 · 计算机科学 2020-01-07 Johannes Dornheim , Norbert Link , Peter Gumbsch

We present one of the first algorithms on model based reinforcement learning and trajectory optimization with free final time horizon. Grounded on the optimal control theory and Dynamic Programming, we derive a set of backward differential…

系统与控制 · 计算机科学 2015-09-04 Wei Sun , Evangelos Theodorou , Panagiotis Tsiotras

We consider an optimal stochastic impulse control problem over an infinite time horizon motivated by a model of irreversible investment choices with fixed adjustment costs. By employing techniques of viscosity solutions and relying on…

最优化与控制 · 数学 2019-02-05 Salvatore Federico , Mauro Rosestolato , Elisa Tacconi

A method is presented for solving the discrete-time finite-horizon Linear Quadratic Regulator (LQR) problem subject to auxiliary linear equality constraints, such as fixed end-point constraints. The method explicitly determines an affine…

系统与控制 · 计算机科学 2018-09-18 Forrest Laine , Claire Tomlin

We propose a machine learning algorithm for solving finite-horizon stochastic control problems based on a deep neural network representation of the optimal policy functions. The algorithm has three features: (1) It can solve…

综合经济学 · 经济学 2024-12-09 Xianhua Peng , Steven Kou , Lekang Zhang
‹ 上一页 1 8 9 10 下一页 ›