English
Related papers

Related papers: Convergence of Policy Gradient for Stochastic Line…

200 papers

Domain randomization is a simple, effective, and flexible scheme for obtaining robust feedback policies aimed at reducing the sim-to-real gap due to model mismatch. While domain randomization methods have yielded impressive demonstrations…

Systems and Control · Electrical Eng. & Systems 2026-03-17 Alex Nguyen-Le , Nikolai Matni

A linear control system with quadratic cost functional over infinite time horizon is considered without assuming controllability/stabilizability condition and the global integrability condition for the nonhomogeneous term of the state…

Optimization and Control · Mathematics 2020-08-25 Jianping Huang , Jiongmin Yong , Hua-Cheng Zhou

This paper studies indefinite stochastic linear-quadratic (LQ) optimal control for jump-diffusion systems with random coefficients. We construct an algebraic inverse flow from the zero-control base system, extract the semimartingale kernel…

Optimization and Control · Mathematics 2026-05-14 Xinyu Ma , Qingxin Meng

In this contribution, we present a full overview of the continuous stochastic gradient (CSG) method, including convergence results, step size rules and algorithmic insights. We consider optimization problems in which the objective function…

Optimization and Control · Mathematics 2023-03-23 Max Grieshammer , Lukas Pflug , Michael Stingl , Andrian Uihlein

We propose a Model Predictive Control (MPC) with a single-step prediction horizon to approximate the solution of infinite horizon optimal control problems with the expected sum of convex stage costs for constrained linear uncertain systems.…

Optimization and Control · Mathematics 2025-04-24 Eunhyek Joa , Francesco Borrelli

In this paper we consider a control system of the form $\dot x = F(x)u$, linear in the control variable $u$. Given a fixed starting point, we study a finite-horizon optimal control problem, where we want to minimize a weighted sum of an…

Optimization and Control · Mathematics 2023-11-20 Alessandro Scagliotti

This paper is concerned with a kind of linear-quadratic (LQ) optimal control problem of backward stochastic differential equation (BSDE) with partial information. The cost functional includes cross terms between the state and control, and…

Optimization and Control · Mathematics 2025-09-03 Jialong Li , Zhiyong Yu , Wanying Yue

Motivated by recent advances of reinforcement learning and direct data-driven control, we propose policy gradient adaptive control (PGAC) for the linear quadratic regulator (LQR), which uses online closed-loop data to improve the control…

Optimization and Control · Mathematics 2025-06-16 Feiran Zhao , Alessandro Chiuso , Florian Dörfler

We study the infinite-horizon distributionally robust (DR) control of linear systems with quadratic costs, where disturbances have unknown, possibly time-correlated distribution within a Wasserstein-2 ambiguity set. We aim to minimize the…

Optimization and Control · Mathematics 2024-06-12 Taylan Kargin , Joudi Hajar , Vikrant Malik , Babak Hassibi

Despite its nonconvexity, policy optimization for the Linear Quadratic Regulator (LQR) admits a favorable structural property known as gradient dominance, which facilitates linear convergence of policy gradient methods to the globally…

Optimization and Control · Mathematics 2026-02-27 Yuto Watanabe , Yang Zheng

The behaviour of a stochastic dynamical system may be largely influenced by those low-probability, yet extreme events. To address such occurrences, this paper proposes an infinite-horizon risk-constrained Linear Quadratic Regulator (LQR)…

Optimization and Control · Mathematics 2021-03-30 Feiran Zhao , Keyou You , Tamer Basar

Constrained Reinforcement Learning (CRL) addresses sequential decision-making problems where agents are required to achieve goals by maximizing the expected return while meeting domain-specific constraints. In this setting, policy-based…

Machine Learning · Computer Science 2025-06-09 Alessandro Montenegro , Leonardo Cesani , Marco Mussi , Matteo Papini , Alberto Maria Metelli

Off-policy learning refers to the problem of learning the value function of a way of behaving, or policy, while following a different policy. Gradient-based off-policy learning algorithms, such as GTD and TDC/GQ, converge even when using…

Artificial Intelligence · Computer Science 2015-12-15 Lucas Lehnert , Doina Precup

We develop a model-free learning algorithm for the infinite-horizon linear quadratic regulator (LQR) problem. Specifically, (risk) constraints and structured feedback are considered, in order to reduce the state deviation while allowing for…

Optimization and Control · Mathematics 2022-04-06 Kyung-bin Kwon , Lintao Ye , Vijay Gupta , Hao Zhu

The Sequential Linear Quadratic (SLQ) algorithm is a continuous-time variant of the well-known Differential Dynamic Programming (DDP) technique with a Gauss-Newton Hessian approximation. This family of methods has gained popularity in the…

Robotics · Computer Science 2021-03-29 Jean-Pierre Sleiman , Farbod Farshidian , Marco Hutter

This paper is concerned with coherent quantum linear quadratic Gaussian (CQLQG) control. The problem is to find a stabilizing measurement-free quantum controller for a quantum plant so as to minimize a mean square cost for the fully quantum…

Quantum Physics · Physics 2016-09-27 Arash Kh. Sichani , Igor G. Vladimirov , Ian R. Petersen

The closed-loop stability and infinite-horizon performance of receding-horizon approximations are studied for non-stationary linear-quadratic regulator (LQR) problems. The approach is based on a lifted reformulation of the optimal control…

Systems and Control · Electrical Eng. & Systems 2023-09-06 Jintao Sun , Michael Cantoni

This work explores generalizations of the Polyak-Lojasiewicz inequality (PLI) and their implications for the convergence behavior of gradient flows in optimization problems. Motivated by the continuous-time linear quadratic regulator…

Optimization and Control · Mathematics 2025-04-01 Arthur Castello B. de Oliveira , Leilei Cui , Eduardo D. Sontag

This paper considers the stochastic linear quadratic optimal control problem in which the control domain is nonconvex. By the functional analysis and convex perturbation methods, we establish a novel maximum principle. The application of…

Optimization and Control · Mathematics 2017-11-01 Shaolin Ji , Xiaole Xue

We analyze a sequential quadratic programming algorithm for solving a class of abstract optimization problems. Assuming that the initial point is in an $L^2$ neighborhood of a local solution that satisfies no-gap second-order sufficient…

Optimization and Control · Mathematics 2026-05-19 Eduardo Casas , Mariano Mateos