English
Related papers

Related papers: Policy Gradient Methods for Discrete Time Linear Q…

200 papers

We study reinforcement learning (RL) for a class of continuous-time linear-quadratic (LQ) control problems for diffusions, where states are scalar-valued and running control rewards are absent but volatilities of the state processes depend…

Machine Learning · Computer Science 2025-07-25 Yilie Huang , Yanwei Jia , Xun Yu Zhou

We present one of the first algorithms on model based reinforcement learning and trajectory optimization with free final time horizon. Grounded on the optimal control theory and Dynamic Programming, we derive a set of backward differential…

Systems and Control · Computer Science 2015-09-04 Wei Sun , Evangelos Theodorou , Panagiotis Tsiotras

This paper proposes a receding horizon active learning and control problem for dynamical systems in which Gaussian Processes (GPs) are utilized to model the system dynamics. The active learning objective in the optimization problem is…

Systems and Control · Electrical Eng. & Systems 2021-05-13 Viet-Anh Le , Truong X. Nghiem

Model-free algorithms are brought into the control system's research with the emergence of reinforcement learning algorithms. However, there are two practical challenges of reinforcement learning-based methods. First, learning by…

Systems and Control · Electrical Eng. & Systems 2024-09-18 Mi Zhou , Erik Verriest , Chaouki Abdallah

We explore the infinite-horizon Distributionally Robust (DR) linear-quadratic control. While the probability distribution of disturbances is unknown and potentially correlated over time, it is confined within a Wasserstein-2 ball of a…

Optimization and Control · Mathematics 2024-08-13 Joudi Hajar , Taylan Kargin , Vikrant Malik , Babak Hassibi

The inverse linear-quadratic optimal control problem is a system identification problem whose aim is to recover the quadratic cost function and hence the closed-loop system matrices based on observations of optimal trajectories. In this…

Optimization and Control · Mathematics 2022-09-22 Han Zhang , Axel Ringh

Direct policy optimization in reinforcement learning is usually solved with policy-gradient algorithms, which optimize policy parameters via stochastic gradient ascent. This paper provides a new theoretical interpretation and justification…

Machine Learning · Computer Science 2023-10-24 Adrien Bolland , Gilles Louppe , Damien Ernst

The goal of this paper is to investigate new and simple convergence analysis of dynamic programming for linear quadratic regulator problem of discrete-time linear time-invariant systems. In particular, bounds on errors are given in terms of…

Optimization and Control · Mathematics 2021-06-18 Donghwan Lee

Policy gradient methods hold great potential for solving complex continuous control tasks. Still, their training efficiency can be improved by exploiting structure within the optimization problem. Recent work indicates that supervised…

Machine Learning · Computer Science 2024-03-19 Jan Schneider , Pierre Schumacher , Simon Guist , Le Chen , Daniel Häufle , Bernhard Schölkopf , Dieter Büchler

In this paper we develop a novel, discrete-time optimal control framework for mechanical systems with uncertain model parameters. We consider finite-horizon problems where the performance index depends on the statistical moments of the…

Optimization and Control · Mathematics 2017-05-17 George I. Boutselis , Yunpeng Pan , Gerardo De La Tore , Evangelos A. Theodorou

We study the time-inconsistent linear quadratic optimal control problem for forward-backward stochastic differential equations with potentially indefinite cost weighting matrices for both the state and the control variables. Our research…

Optimization and Control · Mathematics 2023-12-15 Qi Lü , Bowen Ma

Q-learning is a promising method for solving optimal control problems for uncertain systems without the explicit need for system identification. However, approaches for continuous-time Q-learning have limited provable safety guarantees,…

Systems and Control · Electrical Eng. & Systems 2024-01-30 Soutrik Bandyopadhyay , Shubhendu Bhasin

We study a generalization of the classical discrete-time, Linear-Quadratic-Gaussian (LQG) control problem where the noise distributions affecting the states and observations are unknown and chosen adversarially from divergence-based…

Optimization and Control · Mathematics 2025-09-30 Bahar Taşkesen , Dan A. Iancu , Çağıl Koçyiğit , Daniel Kuhn

This paper is concerned with the distributed control and stabilization problems for linear discrete-time large scale systems with imposed constraints. The main contributions of this paper are: Firstly, by using the maximum principle…

Optimization and Control · Mathematics 2018-01-03 Qingyuan Qi , Huanshui Zhang , Peijun Ju

We prove a general existence result in stochastic optimal control in discrete time where controls take values in conditional metric spaces, and depend on the current state and the information of past decisions through the evolution of a…

Optimization and Control · Mathematics 2018-12-19 Asgar Jamneshan , Michael Kupper , José Miguel Zapata

We propose a method for designing policies for convex stochastic control problems characterized by random linear dynamics and convex stage cost. We consider policies that employ quadratic approximate value functions as a substitute for the…

Optimization and Control · Mathematics 2023-11-10 Alan Yang , Stephen Boyd

Recently, policy optimization for control purposes has received renewed attention due to the increasing interest in reinforcement learning. In this paper, we investigate the convergence of policy optimization for quadratic control of…

Optimization and Control · Mathematics 2020-02-12 Joao Paulo Jansch-Porto , Bin Hu , Geir Dullerud

A linear control system with quadratic cost functional over infinite time horizon is considered without assuming controllability/stabilizability condition and the global integrability condition for the nonhomogeneous term of the state…

Optimization and Control · Mathematics 2020-08-25 Jianping Huang , Jiongmin Yong , Hua-Cheng Zhou

We present differentiable predictive control (DPC), a method for learning constrained neural control policies for linear systems with probabilistic performance guarantees. We employ automatic differentiation to obtain direct policy…

Systems and Control · Electrical Eng. & Systems 2022-01-28 Jan Drgona , Aaron Tuor , Draguna Vrabie

We consider a finite-horizon linear-quadratic optimal control problem where only a limited number of control messages are allowed for sending from the controller to the actuator. To restrict the number of control actions computed and…

Systems and Control · Computer Science 2017-01-19 Burak Demirel , Euhanna Ghadimi , Daniel E. Quevedo , Mikael Johansson
‹ Prev 1 8 9 10 Next ›