English
Related papers

Related papers: On the Effect of Quadratic Regularization in Direc…

200 papers

In this paper, we study the irregular output feedback linear quadratic (LQ) control problem, which is a continuous work of previous works for irregular LQ control [33] where the state is assumed to be exactly known priori. Different from…

Optimization and Control · Mathematics 2019-05-17 Juanjuan Xu , Huanshui Zhang

In this paper, we study the dynamic regret of online linear quadratic regulator (LQR) control with time-varying cost functions and disturbances. We consider the case where a finite look-ahead window of cost functions and disturbances is…

Optimization and Control · Mathematics 2021-02-03 Runyu Zhang , Yingying Li , Na Li

In this work we study the convergence of gradient methods for nonconvex optimization problems -- specifically the effect of the problem formulation to the convergence behavior of the solution of a gradient flow. We show through a simple…

Optimization and Control · Mathematics 2025-10-03 Moh Kamalul Wafi , Arthur Castello B. de Oliveira , Eduardo D. Sontag

System stabilization via policy gradient (PG) methods has drawn increasing attention in both control and machine learning communities. In this paper, we study their convergence and sample complexity for stabilizing linear time-invariant…

Optimization and Control · Mathematics 2023-09-15 Feiran Zhao , Xingyun Fu , Keyou You

We propose a new risk-constrained formulation of the classical Linear Quadratic (LQ) stochastic control problem for general partially-observed systems. Our framework is motivated by the fact that the risk-neutral LQ controllers, although…

Optimization and Control · Mathematics 2021-12-15 Anastasios Tsiamis , Dionysios S. Kalogerias , Alejandro Ribeiro , George J. Pappas

The goal of this paper is to develop data-driven control design and evaluation strategies based on linear matrix inequalities (LMIs) and dynamic programming. We consider deterministic discrete-time LTI systems, where the system model is…

Optimization and Control · Mathematics 2021-06-17 Donghwan Lee , Do Wan Kim

We formulate and solve an optimal control problem with cooperative, mean-field coupled linear-quadratic subsystems and additional risk-aware costs depending on the covariance and skew of the disturbance. This problem quantifies the…

Systems and Control · Electrical Eng. & Systems 2024-06-11 Dhairya Patel , Margaret Chapman

Distributional reinforcement learning (DRL) enhances the understanding of the effects of the randomness in the environment by letting agents learn the distribution of a random return, rather than its expected value as in standard RL. At the…

Optimization and Control · Mathematics 2023-03-27 Zifan Wang , Yulong Gao , Siyi Wang , Michael M. Zavlanos , Alessandro Abate , Karl H. Johansson

While differentiable control has emerged as a powerful paradigm combining model-free flexibility with model-based efficiency, the iterative Linear Quadratic Regulator (iLQR) remains underexplored as a differentiable component. The…

Robotics · Computer Science 2025-06-24 Shuyuan Wang , Philip D. Loewen , Michael Forbes , Bhushan Gopaluni , Wei Pan

We study the problem of policy estimation for the Linear Quadratic Regulator (LQR) in discrete-time linear time-invariant uncertain dynamical systems. We propose a Moreau Envelope-based surrogate LQR cost, built from a finite set of…

Optimization and Control · Mathematics 2024-10-08 Ashwin Aravind , Mohammad Taha Toghani , César A. Uribe

In this paper we consider new regularization methods for linear inverse problems of dynamic type. These methods are based on dynamic programming techniques for linear quadratic optimal control problems. Two different approaches are…

Numerical Analysis · Mathematics 2021-01-26 S. Kindermann , A. Leitao

Linear Quadratic Regulators (LQR) achieve enormous successful real-world applications. Very recently, people have been focusing on efficient learning algorithms for LQRs when their dynamics are unknown. Existing results effectively learn to…

Machine Learning · Computer Science 2021-02-15 Tianyu Wang , Lin F. Yang

We address the challenge of estimation in the context of constant linear effect models with dense functional responses. In this framework, the conditional expectation of the response curve is represented by a linear combination of…

Methodology · Statistics 2024-10-07 Pratim Guha Niyogi , Ping-Shou Zhong

Dropout is a widely-used regularization technique, often required to obtain state-of-the-art for a number of architectures. This work demonstrates that dropout introduces two distinct but entangled regularization effects: an explicit effect…

Machine Learning · Computer Science 2020-10-16 Colin Wei , Sham Kakade , Tengyu Ma

We present a novel direct data-driven algorithm that learns an optimal control policy for the Bilinear Biquadratic Regulator (BBR) for an unknown bilinear system. The BBR is difficult to solve owing to the presence of the nonlinear…

Systems and Control · Electrical Eng. & Systems 2024-10-10 Shanelle G. Clarke , Omanshu Thapliyal , Inseok Hwang

We study derivative-free methods for policy optimization over the class of linear policies. We focus on characterizing the convergence rate of these methods when applied to linear-quadratic systems, and study various settings of driving…

Machine Learning · Computer Science 2020-05-19 Dhruv Malik , Ashwin Pananjady , Kush Bhatia , Koulik Khamaru , Peter L. Bartlett , Martin J. Wainwright

We introduce a new problem setting for continuous control called the LQR with Rich Observations, or RichLQR. In our setting, the environment is summarized by a low-dimensional continuous latent state with linear dynamics and quadratic…

We approach the problem of implicit regularization in deep learning from a geometrical viewpoint. We highlight a regularization effect induced by a dynamical alignment of the neural tangent features introduced by Jacot et al, along a small…

We study the exploration-exploitation dilemma in the linear quadratic regulator (LQR) setting. Inspired by the extended value iteration algorithm used in optimistic algorithms for finite MDPs, we propose to relax the optimistic optimization…

Machine Learning · Statistics 2020-07-14 Marc Abeille , Alessandro Lazaric

Estimating natural effects is a core task in causal mediation analysis. Existing triply robust (TR) frameworks (Tchetgen Tchetgen & Shpitser 2012) and their extensions have been developed to estimate the natural effects. In this work, we…

Methodology · Statistics 2026-02-16 Zhen Qi , Yuqian Zhang