English
Related papers

Related papers: Policy iteration using Q-functions: Linear dynamic…

200 papers

In this paper, we address an extension of the Loewner framework for learning quadratic control systems from input-output data. The proposed method first constructs a reduced-order linear model from measurements of the classical transfer…

Optimization and Control · Mathematics 2020-12-04 Ion Victor Gosea , Dimitrios S. Karachalios , Athanasios C. Antoulas

We focus on variational inference in dynamical systems where the discrete time transition function (or evolution rule) is modelled by a Gaussian process. The dominant approach so far has been to use a factorised posterior distribution,…

Machine Learning · Statistics 2018-12-17 Alessandro Davide Ialongo , Mark van der Wilk , James Hensman , Carl Edward Rasmussen

Iterative optimization algorithms depend on access to information about the objective function. In a differentiable programming framework, this information, such as gradients, can be automatically derived from the computational graph. We…

Optimization and Control · Mathematics 2025-07-08 Vincent Roulet , Siddhartha Srinivasa , Maryam Fazel , Zaid Harchaoui

We present an end-to-end quantum algorithm for simulating nonlinear dynamics described by a system of stochastic dissipative differential equations with a quadratic nonlinearity. The stochastic part of the system is modeled by a Gaussian…

Quantum Physics · Physics 2025-10-07 Sergey Bravyi , Robert Manson-Sawko , Mykhaylo Zayats , Sergiy Zhuk

We present a measurement noise reduction scheme based on information flow of a chaotic system. This scheme operates on conditions of chaoticity and well-defined noise level, not depending on other detailed characteristics of noise. Starting…

Chaotic Dynamics · Physics 2007-05-23 Seung Ki Baek

Many of the recent trajectory optimization algorithms alternate between linear approximation of the system dynamics around the mean trajectory and conservative policy update. One way of constraining the policy change is by bounding the…

Machine Learning · Computer Science 2018-07-03 Riad Akrour , Abbas Abdolmaleki , Hany Abdulsamad , Jan Peters , Gerhard Neumann

We consider the optimal control problem for a linear conditional McKean-Vlasov equation with quadratic cost functional. The coefficients of the system and the weigh-ting matrices in the cost functional are allowed to be adapted processes…

Probability · Mathematics 2017-03-09 Huyên Pham

Q-learning is a stochastic approximation version of the classic value iteration. The literature has established that Q-learning suffers from both maximization bias and slower convergence. Recently, multi-step algorithms have shown practical…

Machine Learning · Computer Science 2024-07-03 Antony Vijesh , Shreyas S R

When applying Dynamic Power Management (DPM) technique to pervasively deployed embedded systems, the technique needs to be very efficient so that it is feasible to implement the technique on low end processor and tight-budget memory.…

Other Computer Science · Computer Science 2011-11-09 Min Li , Xiaobo Wu , Richard Yao , Xiaolang Yan

Policy gradient methods in reinforcement learning update policy parameters by taking steps in the direction of an estimated gradient of policy value. In this paper, we consider the statistically efficient estimation of policy gradients from…

Machine Learning · Statistics 2020-02-21 Nathan Kallus , Masatoshi Uehara

This paper investigates the state estimation problem for a class of complex networks, in which the dynamics of each node is subject to Gaussian noise, system uncertainties and nonlinearities. Based on a regularized least-squares approach,…

Systems and Control · Electrical Eng. & Systems 2021-03-16 Peihu Duan , Qishao Wang , Zhisheng Duan , Guanrong Chen

We analyze a simple prefiltered variation of the least squares estimator for the problem of estimation with biased, semi-parametric noise, an error model studied more broadly in causal statistics and active learning. We prove an oracle…

Machine Learning · Computer Science 2019-02-05 Max Simchowitz , Ross Boczar , Benjamin Recht

This work introduces a novel control strategy called Iterative Linear Quadratic Regulator for Iterative Tasks (i2LQR), which aims to improve closed-loop performance with local trajectory optimization for iterative tasks in a dynamic…

Systems and Control · Electrical Eng. & Systems 2023-09-08 Yifan Zeng , Suiyi He , Han Hoang Nguyen , Yihan Li , Zhongyu Li , Koushil Sreenath , Jun Zeng

The problem of numerical differentiation can be thought of as an inverse problem by considering it as solving a Volterra equation. It is well known that such inverse integral problems are ill-posed and one requires regularization methods to…

Numerical Analysis · Mathematics 2020-04-15 Abinash Nayak

We propose controller synthesis for state regulation problems in which a human operator shares control with an autonomy system, running in parallel. The autonomy system continuously improves over human action, with minimal intervention, and…

Systems and Control · Computer Science 2019-09-23 Murad Abu-Khalaf , Sertac Karaman , Daniela Rus

This paper aims at the study of controllability properties and induced controllability metrics on complex networks governed by a class of (discrete time) linear decision processes with mul-tiplicative noise. The dynamics are given by a…

Optimization and Control · Mathematics 2016-12-15 Tidiane Diallo , Dan Goreac

Policy optimization has drawn increasing attention in reinforcement learning, particularly in the context of derivative-free methods for linear quadratic regulator (LQR) problems with unknown dynamics. This paper focuses on characterizing…

Optimization and Control · Mathematics 2025-06-17 Weijian Li , Panagiotis Kounatidis , Zhong-Ping Jiang , Andreas A. Malikopoulos

The problem of interest is the minimization of a nonlinear function subject to nonlinear equality constraints using a sequential quadratic programming (SQP) method. The minimization must be performed while observing only noisy evaluations…

Optimization and Control · Mathematics 2021-10-12 Figen Oztoprak , Richard Byrd , Jorge Nocedal

This paper addresses the average cost minimization problem for discrete-time systems with multiplicative and additive noises via reinforcement learning. By using Q-function, we propose an online learning scheme to estimate the kernel matrix…

Systems and Control · Electrical Eng. & Systems 2020-10-14 Jing Lai , Junlin Xiong

This paper proposes a novel informativity-based data-driven synthesis method for a sub-optimal linear quadratic (LQ) regulator for linear input-delay systems from noisy input-state data. Exploiting the augmented state structure of…

Systems and Control · Electrical Eng. & Systems 2026-05-05 Kohei Ayaka , Takumi Namba , Kiyotsugu Takaba