English
Related papers

Related papers: Inverse Optimal Control with Discount Factor for C…

200 papers

This paper employs a policy iteration reinforcement learning (RL) method to study continuous-time linear-quadratic mean-field control problems in infinite horizon. The drift and diffusion terms in the dynamics involve the states, the…

Optimization and Control · Mathematics 2024-11-05 Na Li , Xun Li , Zuo Quan Xu

This paper presents a pioneering approach to solving the linear quadratic regulation (LQR) and linear quadratic tracking (LQT) problems with constrained inputs using a novel off-policy continuous-time Q-learning framework. The proposed…

Systems and Control · Electrical Eng. & Systems 2025-09-23 Duc Cuong Nguyen , Quang Huy Dao , Phuong Nam Dao

This paper studies a {\it reversible} investment problem where a social planner aims to control its capacity production in order to fit optimally the random demand of a good. Our model allows for general diffusion dynamics on the demand as…

Probability · Mathematics 2013-07-08 Salvatore Federico , Huyen Pham

The optimal objective is a fundamental aspect of reinforcement learning (RL), as it determines how policies are evaluated and optimized. While total return maximization is the ideal objective in RL, discounted return maximization is the…

Machine Learning · Computer Science 2025-03-19 Shuyu Yin , Fei Wen , Peilin Liu , Tao Luo

This paper is concerned with a stochastic linear-quadratic optimal control problem in a finite time horizon, where the coefficients of the control system are allowed to be random, and the weighting matrices in the cost functional are…

Optimization and Control · Mathematics 2019-11-12 Jingrui Sun , Jie Xiong , Jiongmin Yong

We consider the problem of reinforcement learning when provided with (1) a baseline control policy and (2) a set of constraints that the learner must satisfy. The baseline policy can arise from demonstration data or a teacher agent and may…

Machine Learning · Computer Science 2021-07-13 Tsung-Yen Yang , Justinian Rosca , Karthik Narasimhan , Peter J. Ramadge

Optimal control of stochastic nonlinear dynamical systems is a major challenge in the domain of robot learning. Given the intractability of the global control problem, state-of-the-art algorithms focus on approximate sequential optimization…

Machine Learning · Computer Science 2020-04-23 Joe Watson , Hany Abdulsamad , Jan Peters

The naive application of Reinforcement Learning algorithms to continuous control problems -- such as locomotion and manipulation -- often results in policies which rely on high-amplitude, high-frequency control signals, known colloquially…

Robotics · Computer Science 2019-02-14 Steven Bohez , Abbas Abdolmaleki , Michael Neunert , Jonas Buchli , Nicolas Heess , Raia Hadsell

We obtain a probabilistic solution to linear-quadratic optimal control problems with state constraints. Given a closed set $\mathcal{D}\subseteq [0,T]\times\mathbb{R}^d$, a diffusion $X$ in $\mathbb{R}^d$ must be linearly controlled in…

Optimization and Control · Mathematics 2026-03-06 Tiziano De Angelis , Erik Ekström

An optimal control problem with a time-parameter is considered. The functional to be optimized includes the maximum over time-horizon reached by a function of the state variable, and so an $L^\infty$-term. In addition to the classical…

Optimization and Control · Mathematics 2018-11-01 Sébastien Court , Karl Kunisch , Laurent Pfeiffer

We present a method for optimal control with respect to a linear cost function for positive linear systems with coupled input constraints. We show that the optimal cost function and resulting sparse state feedback for these systems can be…

Optimization and Control · Mathematics 2023-11-07 David Ohlin , Emma Tegling , Anders Rantzer

This paper is concerned with a linear quadratic (LQ, for short) optimal control problem with fixed terminal states and integral quadratic constraints. A Riccati equation with infinite terminal value is introduced, which is uniquely solvable…

Optimization and Control · Mathematics 2017-05-11 Jingrui Sun

In this paper, two Q-learning (QL) methods are proposed and their convergence theories are established for addressing the model-free optimal control problem of general nonlinear continuous-time systems. By introducing the Q-function for…

Systems and Control · Computer Science 2014-10-14 Biao Luo , Derong Liu , Tingwen Huang

It has been recently established that a deterministic infinite horizon discounted optimal control problem in discrete time is closely related to a certain infinite dimensional linear programming problem and its dual. In the present paper,…

Optimization and Control · Mathematics 2018-02-19 Vladimir Gaitsgory , Alex Parkinson , Ilya Shvartsman

Model predictive control can optimally deal with nonlinear systems under consideration of constraints. The control performance depends on the model accuracy and the prediction horizon. Recent advances propose to use reinforcement learning…

Machine Learning · Computer Science 2024-11-01 Dean Brandner , Sergio Lucia

An optimal control problem with an infinite horizon quadratic cost functional for a linear system with a known additive disturbance is considered. The feature of this problem is that a weight matrix of the control cost in the cost…

Optimization and Control · Mathematics 2016-03-08 Valery Y. Glizer , Oleg Kelis

We study the discrete-time linear-quadratic (LQ) control model using reinforcement learning (RL). Using entropy to measure the cost of exploration, we prove that the optimal feedback policy for the problem must be Gaussian type. Then, we…

Machine Learning · Statistics 2025-02-05 Lucky Li

A linear control system with quadratic cost functional over infinite time horizon is considered without assuming controllability/stabilizability condition and the global integrability condition for the nonhomogeneous term of the state…

Optimization and Control · Mathematics 2020-08-25 Jianping Huang , Jiongmin Yong , Hua-Cheng Zhou

Though switched dynamical systems have shown great utility in modeling a variety of physical phenomena, the construction of an optimal control of such systems has proven difficult since it demands some type of optimal mode scheduling. In…

Optimization and Control · Mathematics 2014-02-04 Ramanarayan Vasudevan , Humberto Gonzalez , Ruzena Bajcsy , S. Shankar Sastry

We consider the problem of estimating the possibly non-convex cost of an agent by observing its interactions with a nonlinear, non-stationary and stochastic environment. For this inverse problem, we give a result that allows to estimate the…

Optimization and Control · Mathematics 2023-07-24 Émiland Garrabé , Hozefa Jesawada , Carmen Del Vecchio , Giovanni Russo