English
Related papers

Related papers: A model-free first-order method for linear quadrat…

200 papers

A sequential quadratic programming method is designed for solving general smooth nonlinear stochastic optimization problems subject to expectation equality constraints. We consider the setting where the objective and constraint function…

Optimization and Control · Mathematics 2026-03-17 Haoming Shen , Yang Zeng , Baoyu Zhou

The paper studies a class of quadratic optimal control problems for partially observable linear dynamical systems. In contrast to the full information case, the control is required to be adapted to the filtration generated by the…

Optimization and Control · Mathematics 2022-03-01 Jingrui Sun , Jie Xiong

Robust optimization (RO) is one of the key paradigms for solving optimization problems affected by uncertainty. Two principal approaches for RO, the robust counterpart method and the adversarial approach, potentially lead to excessively…

Optimization and Control · Mathematics 2024-09-05 Krzysztof Postek , Shimrit Shtern

We study the sample complexity of approximate policy iteration (PI) for the Linear Quadratic Regulator (LQR), building on a recent line of work using LQR as a testbed to understand the limits of reinforcement learning (RL) algorithms on…

Machine Learning · Computer Science 2019-05-31 Karl Krauth , Stephen Tu , Benjamin Recht

``Sim2real gap", in which the system learned in simulations is not the exact representation of the real system, can lead to loss of stability and performance when controllers learned using data from the simulated system are used on the real…

Systems and Control · Electrical Eng. & Systems 2025-05-15 Shivam Bajaj , Prateek Jaiswal , Vijay Gupta

We study a new two-time-scale stochastic gradient method for solving optimization problems, where the gradients are computed with the aid of an auxiliary variable under samples generated by time-varying MDPs controlled by the underlying…

Optimization and Control · Mathematics 2024-08-27 Sihan Zeng , Thinh T. Doan , Justin Romberg

This paper studies a class of continuous-time scalar-state stochastic Linear-Quadratic (LQ) optimal control problem with the linear control constraints. Applying the state separation theorem induced from its special structure, we develop…

Portfolio Management · Quantitative Finance 2018-06-12 Weiping Wu , Jianjun Gao , Junguo Lu , Xun Li

We consider the problem of nonstochastic control with a sequence of quadratic losses, i.e., LQR control. We provide an efficient online algorithm that achieves an optimal dynamic (policy) regret of $\tilde{O}(\text{max}\{n^{1/3}…

Machine Learning · Computer Science 2022-06-22 Dheeraj Baby , Yu-Xiang Wang

We study the problem of policy estimation for the Linear Quadratic Regulator (LQR) in discrete-time linear time-invariant uncertain dynamical systems. We propose a Moreau Envelope-based surrogate LQR cost, built from a finite set of…

Optimization and Control · Mathematics 2024-10-08 Ashwin Aravind , Mohammad Taha Toghani , César A. Uribe

This article examines a linear-quadratic elliptic optimal control problem in which the cost functional and the state equation involve a highly oscillatory periodic coefficient $A^\varepsilon$. The small parameter $\varepsilon>0$ denotes the…

Optimization and Control · Mathematics 2020-10-12 Agnes Lamacz-Keymling , Irwin Yousept

Direct data-driven design methods for the linear quadratic regulator (LQR) mainly use offline or episodic data batches, and their online adaptation has been acknowledged as an open problem. In this paper, we propose a direct adaptive method…

Optimization and Control · Mathematics 2024-10-07 Feiran Zhao , Florian Dörfler , Alessandro Chiuso , Keyou You

We study the adaptive control of an unknown linear system with a quadratic cost function subject to safety constraints on both the states and actions. The challenges of this problem arise from the tension among safety, exploration,…

Systems and Control · Electrical Eng. & Systems 2021-11-02 Yingying Li , Subhro Das , Jeff Shamma , Na Li

A scheme for generating a family of convex variational principles is developed, the Euler- Lagrange equations of each member of the family formally corresponding to the necessary conditions of optimal control of a given system of ordinary…

Optimization and Control · Mathematics 2025-06-13 Amit Acharya , Janusz Ginster

An algorithm is proposed for solving optimization problems with stochastic objective and deterministic equality and inequality constraints. This algorithm is objective-function-free in the sense that it only uses the objective's gradient…

Optimization and Control · Mathematics 2026-04-01 S. Gratton , Ph. L. Toint

In this paper, we propose a structured linear parameterization of a feedback policy to solve the model-free stochastic optimal control problem. This parametrization is corroborated by a decoupling principle that is shown to be near-optimal…

Optimization and Control · Mathematics 2020-02-19 Karthikeya S Parunandi , Aayushman Sharma , Suman Chakravorty , Dileep Kalathil

We consider the problem of controlling a Linear Quadratic Regulator (LQR) system over a finite horizon $T$ with fixed and known cost matrices $Q,R$, but unknown and non-stationary dynamics $\{A_t, B_t\}$. The sequence of dynamics matrices…

Machine Learning · Computer Science 2022-03-21 Yuwei Luo , Varun Gupta , Mladen Kolar

We study the policy gradient method (PGM) for the linear quadratic Gaussian (LQG) dynamic output-feedback control problem using an input-output-history (IOH) representation of the closed-loop system. First, we show that any dynamic…

Optimization and Control · Mathematics 2025-10-23 Tomonori Sadamoto , Takashi Tanaka

We study the problem of adaptive control of a high dimensional linear quadratic (LQ) system. Previous work established the asymptotic convergence to an optimal controller for various adaptive control schemes. More recently, for the average…

Machine Learning · Statistics 2013-03-26 Morteza Ibrahimi , Adel Javanmard , Benjamin Van Roy

We study the task of learning state representations from potentially high-dimensional observations, with the goal of controlling an unknown partially observable system. We pursue a cost-driven approach, where a dynamic model in some latent…

Machine Learning · Computer Science 2026-03-10 Yi Tian , Kaiqing Zhang , Russ Tedrake , Suvrit Sra

This paper presents a pioneering approach to solving the linear quadratic regulation (LQR) and linear quadratic tracking (LQT) problems with constrained inputs using a novel off-policy continuous-time Q-learning framework. The proposed…

Systems and Control · Electrical Eng. & Systems 2025-09-23 Duc Cuong Nguyen , Quang Huy Dao , Phuong Nam Dao