中文
相关论文

相关论文: Model-free Value Iteration Algorithm for Continuou…

200 篇论文

In this paper a novel model-free algorithm is proposed. This algorithm can learn the nearly optimal control law of constrained-input systems from online data without requiring any a priori knowledge of system dynamics. Based on the concept…

系统与控制 · 电气工程与系统科学 2022-05-03 Han Zhao , Lei Guo

We solve a linear quadratic optimal control problem for sampled-data systems with stochastic delays. The delays are stochastically determined by the last few delays. The proposed optimal controller can be efficiently computed by iteratively…

最优化与控制 · 数学 2018-05-18 Masashi Wakaiki , Masaki Ogura , Joao P. Hespanha

The Sequential Linear Quadratic (SLQ) algorithm is a continuous-time variant of the well-known Differential Dynamic Programming (DDP) technique with a Gauss-Newton Hessian approximation. This family of methods has gained popularity in the…

机器人学 · 计算机科学 2021-03-29 Jean-Pierre Sleiman , Farbod Farshidian , Marco Hutter

Monotone variational inequalities (VIs) provide a unifying framework for convex minimization, equilibrium computation, and convex-concave saddle-point problems. Extragradient-type methods are among the most effective first-order algorithms…

最优化与控制 · 数学 2026-04-16 Lingqing Shen , Fatma Kılınç-Karzan

Linear time-invariant control systems can be considered as finitely generated modules over the commutative principal ideal ring $\mathbb{R}[\frac{d}{dt}]$ of linear differential operators with respect to the time derivative. The Kalman…

最优化与控制 · 数学 2025-12-15 Cédric Join , Emmanuel Delaleau , Michel Fliess

This paper deals with a class of time inconsistent stochastic linear quadratic (SLQ) optimal control problems in Markovian framework. Three notions, i.e., closed-loop equilibrium controls/strategies, open-loop equilibrium controls and their…

最优化与控制 · 数学 2018-02-06 Tianxiao Wang

An optimal control law for networked control systems with a discrete-time linear time-invariant (LTI) system as plant and networks between sensor and controller as well as between controller and actuator is proposed. This controller is…

系统与控制 · 电气工程与系统科学 2021-07-09 Marijan Palmisano , Martin Steinberger , Martin Horn

The present work addresses a finite-horizon linear-quadratic optimal control problem for uncertain systems driven by piecewise constant controls. The precise values of the system parameters are unknown, but assumed to belong to a finite set…

系统与控制 · 计算机科学 2021-08-05 Félix A. Miranda , Fernando Castaños , Alexander Poznyak

A classical approach for solving discrete time nonlinear control on a finite horizon consists in repeatedly minimizing linear quadratic approximations of the original problem around current candidate solutions. While widely popular in many…

最优化与控制 · 数学 2025-07-08 Vincent Roulet , Siddhartha Srinivasa , Maryam Fazel , Zaid Harchaoui

This work introduces a novel control strategy called Iterative Linear Quadratic Regulator for Iterative Tasks (i2LQR), which aims to improve closed-loop performance with local trajectory optimization for iterative tasks in a dynamic…

系统与控制 · 电气工程与系统科学 2023-09-08 Yifan Zeng , Suiyi He , Han Hoang Nguyen , Yihan Li , Zhongyu Li , Koushil Sreenath , Jun Zeng

This paper presents an algorithm to solve the infinite horizon constrained linear quadratic regulator (CLQR) problem using operator splitting methods. First, the CLQR problem is reformulated as a (finite-time) model predictive control (MPC)…

最优化与控制 · 数学 2016-09-20 L. Ferranti , G. Stathopoulos , C. N. Jones , T. Keviczky

This article studies inverse reinforcement learning (IRL) for the stochastic linear-quadratic optimal control problem, where two agents are considered. A learner agent does not know the expert agent's performance cost function, but it…

最优化与控制 · 数学 2024-05-28 Zhongshi Sun , Guangyan Jia

Optimal control problem is typically solved by first finding the value function through Hamilton-Jacobi equation (HJE) and then taking the minimizer of the Hamiltonian to obtain the control. In this work, instead of focusing on the value…

最优化与控制 · 数学 2021-09-10 Alain Bensoussan , Jiayue Han , Sheung Chi Phillip Yam , Xiang Zhou

We investigate the problem of learning linear quadratic regulators (LQR) in a multi-task, heterogeneous, and model-free setting. We characterize the stability and personalization guarantees of a policy gradient-based (PG) model-agnostic…

最优化与控制 · 数学 2024-06-04 Leonardo F. Toso , Donglin Zhan , James Anderson , Han Wang

In this paper, we investigate dynamic optimization problems featuring both stochastic control and optimal stopping in a finite time horizon. The paper aims to develop new methodologies, which are significantly different from those of mixed…

投资组合管理 · 定量金融 2014-06-27 Xiongfei Jian , Xun Li , Fahuai Yi

Q-learning is a promising method for solving optimal control problems for uncertain systems without the explicit need for system identification. However, approaches for continuous-time Q-learning have limited provable safety guarantees,…

系统与控制 · 电气工程与系统科学 2024-01-30 Soutrik Bandyopadhyay , Shubhendu Bhasin

This paper investigates numerical methods for solving stochastic linear quadratic (SLQ) optimal control problems governed by stochastic partial differential equations (SPDEs). Two distinct approaches, the open-loop and closed-loop ones, are…

最优化与控制 · 数学 2024-11-19 Andreas Prohl , Yanqing Wang

We consider the problem of finite-horizon optimal control of a discrete linear time-varying system subject to a stochastic disturbance and fully observable state. The initial state of the system is drawn from a known Gaussian distribution,…

最优化与控制 · 数学 2017-11-08 Maxim Goldshtein , Panagiotis Tsiotras

In this paper, we investigate a model-free optimal control design that minimizes an infinite horizon average expected quadratic cost of states and control actions subject to a probabilistic risk or chance constraint using input-output data.…

系统与控制 · 电气工程与系统科学 2024-11-11 Arunava Naha , Subhrakanti Dey

Model-free reinforcement learning attempts to find an optimal control action for an unknown dynamical system by directly searching over the parameter space of controllers. The convergence behavior and statistical properties of these…

最优化与控制 · 数学 2021-03-17 Hesameddin Mohammadi , Armin Zare , Mahdi Soltanolkotabi , Mihailo R. Jovanović