English
Related papers

Related papers: Robustness of Online Identification-based Policy I…

200 papers

Learning, say through direct policy updates, often requires assumptions such as knowing a priori that the initial policy (gain) is stabilizing, or persistently exciting (PE) input-output data, is available. In this paper, we examine online…

Systems and Control · Electrical Eng. & Systems 2022-01-21 Shahriar Talebi , Siavash Alemzadeh , Niyousha Rahimi , Mehran Mesbahi

Data-driven controllers design is an important research problem, in particular when data is corrupted by the noise. In this paper, we propose a data-driven min-max model predictive control (MPC) scheme using noisy input-state data for…

Systems and Control · Electrical Eng. & Systems 2025-01-31 Yifan Xie , Julian Berberich , Frank Allgöwer

In this contribution, we derive ILEG, an iterative algorithm to find risk sensitive solutions to nonlinear, stochastic optimal control problems. The algorithm is based on a linear quadratic approximation of an exponential risk sensitive…

Systems and Control · Computer Science 2015-12-23 Farbod Farshidian , Jonas Buchli

Policy iteration is one of the classical frameworks of reinforcement learning, which requires a known initial stabilizing control. However, finding the initial stabilizing control depends on the known system model. To relax this requirement…

Systems and Control · Electrical Eng. & Systems 2025-03-20 Dongdong Li , Jiuxiang Dong

In a recent work, we proposed Reliable Policy Iteration (RPI), that restores policy iteration's monotonicity-of-value-estimates property to the function approximation setting. Here, we assess the robustness of RPI's empirical performance on…

Artificial Intelligence · Computer Science 2025-12-16 S. R. Eshwar , Aniruddha Mukherjee , Kintan Saha , Krishna Agarwal , Gugan Thoppe , Aditya Gopalan , Gal Dalal

In this paper we propose an end-to-end algorithm for indirect data-driven control for bilinear systems with stability guarantees. We consider the case where the collected i.i.d. data is affected by probabilistic noise with possibly…

Systems and Control · Electrical Eng. & Systems 2026-03-23 Nicolas Chatzikiriakos , Robin Strässer , Frank Allgöwer , Andrea Iannelli

This note proposes a data-driven output-feedback stabilizing policy iteration for unknown linear discrete-time systems with unmeasurable states. Existing policy iteration methods for optimal control must start from a stabilizing control…

Systems and Control · Electrical Eng. & Systems 2025-12-01 Dongdong Li , Jiuxiang Dong

The goal of robust reinforcement learning (RL) is to learn a policy that is robust against the uncertainty in model parameters. Parameter uncertainty commonly occurs in many real-world RL applications due to simulator modeling errors,…

Machine Learning · Computer Science 2022-10-19 Kishan Panaganti , Zaiyan Xu , Dileep Kalathil , Mohammad Ghavamzadeh

This paper studies finite-horizon robust tracking control for discrete-time linear systems, based on input-output data. We leverage behavioral theory to represent system trajectories through a set of noiseless historical data, instead of…

Optimization and Control · Mathematics 2021-02-25 Liang Xu , Mustafa Sahin Turan , Baiwei Guo , Giancarlo Ferrari-Trecate

One of the fundamental challenges for offline reinforcement learning (RL) is ensuring robustness to data distribution. Whether the data originates from a near-optimal policy or not, we anticipate that an algorithm should demonstrate its…

Machine Learning · Computer Science 2023-10-18 Xiaohan Hu , Yi Ma , Chenjun Xiao , Yan Zheng , Jianye Hao

Output regulation is a fundamental problem in control theory, extensively studied since the 1970s. Traditionally, research has primarily addressed scenarios where the system model is explicitly known, leaving the problem in the absence of a…

Systems and Control · Electrical Eng. & Systems 2025-05-15 Wenjie Liu , Yifei Li , Jian Sun , Gang Wang , Keyou You , Lihua Xie , Jie Chen

Direct data-driven design methods for the linear quadratic regulator (LQR) mainly use offline or episodic data batches, and their online adaptation has been acknowledged as an open problem. In this paper, we propose a direct adaptive method…

Optimization and Control · Mathematics 2024-10-07 Feiran Zhao , Florian Dörfler , Alessandro Chiuso , Keyou You

This paper studies the robustness of reinforcement learning algorithms to errors in the learning process. Specifically, we revisit the benchmark problem of discrete-time linear quadratic regulation (LQR) and study the long-standing open…

Optimization and Control · Mathematics 2021-03-16 Bo Pang , Zhong-Ping Jiang

Policy iteration (PI) is a recursive process of policy evaluation and improvement for solving an optimal decision-making/control problem, or in other words, a reinforcement learning (RL) problem. PI has also served as the fundamental for…

Artificial Intelligence · Computer Science 2021-04-06 Jaeyoung Lee , Richard S. Sutton

This paper studies data-driven iterative learning control (ILC) for linear time-invariant (LTI) systems with unknown dynamics, output disturbances and input box-constraints. Our main contributions are: 1) using a non-parametric data-driven…

Systems and Control · Electrical Eng. & Systems 2023-12-25 Jia Wang , Leander Hemelhof , Ivan Markovsky , Panagiotis Patrinos

Control of networked systems, comprised of interacting agents, is often achieved through modeling the underlying interactions. Constructing accurate models of such interactions--in the meantime--can become prohibitive in applications.…

Systems and Control · Electrical Eng. & Systems 2023-11-17 Siavash Alemzadeh , Shahriar Talebi , Mehran Mesbahi

We present an approach to compute stabilizing controllers for continuous-time linear time-invariant systems directly from an input-output trajectory affected by process and measurement noise. The proposed output-feedback design combines (i)…

Systems and Control · Electrical Eng. & Systems 2025-11-17 Alessandro Bosso , Marco Borghesi , Andrea Iannelli , Bowen Yi , Giuseppe Notarstefano

In this paper, we deal with data-driven predictive control of linear time-invariant (LTI) systems. Specifically, we show for the first time how explicit predictive laws can be learnt directly from data, without needing to identify the…

Systems and Control · Electrical Eng. & Systems 2021-09-14 Andrea Sassella , Valentina Breschi , Simone Formentin

This paper proposes efficient policy iteration and value iteration algorithms for the continuous-time linear quadratic regulator problem with unmeasurable states and unknown system dynamics, from the perspective of direct data-driven…

Systems and Control · Electrical Eng. & Systems 2026-03-17 Jun Xie , Yuan-Hua Ni , Yiqin Yang , Bo Xu

We are motivated by the real challenges presented in a human-robot system to develop new designs that are efficient at data level and with performance guarantees such as stability and optimality at systems level. Existing…

Systems and Control · Electrical Eng. & Systems 2021-01-19 Xiang Gao , Jennie Si , Yue Wen , Minhan Li , He , Huang