English
Related papers

Related papers: An efficient data-based off-policy Q-learning algo…

200 papers

This paper investigates the linear output regulation problem with both the exosystem and the plant fully unknown. A data-driven regulator is proposed to achieve asymptotic regulation and closed-loop stability without performing model…

Systems and Control · Electrical Eng. & Systems 2025-12-08 Shangkun Liu , Lei Wang , Bowen Yi

This paper studies an infinite horizon optimal control problem for discrete-time linear system and quadratic criteria, both with random parameters which are independent and identically distributed with respect to time. In this general…

Optimization and Control · Mathematics 2024-03-04 Deyue Li

An optimal control law for networked control systems with a discrete-time linear time-invariant (LTI) system as plant and networks between sensor and controller as well as between controller and actuator is proposed. This controller is…

Systems and Control · Electrical Eng. & Systems 2021-07-09 Marijan Palmisano , Martin Steinberger , Martin Horn

From a multi-input-multi-output (MIMO) discrete-time linear system, we collect input-output data affected by noise in the form of an unknown exosignal and, from these data points (without knowledge of the system model), we design a feedback…

Optimization and Control · Mathematics 2026-05-22 Andrea Bisoffi , Wenjie Liu , Zhongjie Hu , Claudio De Persis

A finite horizon linear quadratic(LQ) optimal control problem is studied for a class of discrete-time linear fractional systems (LFSs) affected by multiplicative, independent random perturbations. Based on the dynamic programming technique,…

Optimization and Control · Mathematics 2016-07-01 J. J. Trujillo , V. M. Ungureanu

In this paper we design a novel class of online distributed optimization algorithms leveraging control theoretical techniques. We start by focusing on quadratic costs, and assuming to know an internal model of their variation. In this…

Optimization and Control · Mathematics 2026-01-21 Wouter J. A. van Weerelt , Nicola Bastianello

This paper studies distributed Q-learning for Linear Quadratic Regulator (LQR) in a multi-agent network. The existing results often assume that agents can observe the global system state, which may be infeasible in large-scale systems due…

Multiagent Systems · Computer Science 2020-12-24 Hang Wang , Sen Lin , Hamid Jafarkhani , Junshan Zhang

In a sequential decision-making problem, off-policy evaluation estimates the expected cumulative reward of a target policy using logged trajectory data generated from a different behavior policy, without execution of the target policy.…

Machine Learning · Computer Science 2022-11-04 Jie Wang , Rui Gao , Hongyuan Zha

Consider a discrete-time Linear Quadratic Regulator (LQR) problem solved using policy gradient descent when the system matrices are unknown. The gradient is transmitted across a noisy channel over a finite time horizon using analog…

Optimization and Control · Mathematics 2025-07-22 Ashwin Verma , Aritra Mitra , Lintao Ye , Vijay Gupta

Behavior constrained policy optimization has been demonstrated to be a successful paradigm for tackling Offline Reinforcement Learning. By exploiting historical transitions, a policy is trained to maximize a learned value function while…

Machine Learning · Computer Science 2023-07-25 Jiachen Li , Edwin Zhang , Ming Yin , Qinxun Bai , Yu-Xiang Wang , William Yang Wang

We propose a data-driven online convex optimization algorithm for controlling dynamical systems. In particular, the control scheme makes use of an initially measured input-output trajectory and behavioral systems theory which enable it to…

Optimization and Control · Mathematics 2021-11-03 Marko Nonhoff , Matthias A. Müller

This paper introduces an innovative approach based on policy iteration (PI), a reinforcement learning (RL) algorithm, to obtain an optimal observer with a quadratic cost function. This observer is designed for systems with a given…

Systems and Control · Electrical Eng. & Systems 2023-11-29 Soroush Asri , Luis Rodrigues

Given the recent surge of interest in data-driven control, this paper proposes a two-step method to study robust data-driven control for a parameter-unknown linear time-invariant (LTI) system that is affected by energy-bounded noises.…

Systems and Control · Electrical Eng. & Systems 2022-03-15 Jiabao He , Xuan Zhang , Feng Xu , Junbo Tan , Xueqian Wang

Continuous monitoring and real-time control of high-dimensional distributed systems are often crucial in applications to ensure a desired physical behavior, without degrading stability and system performances. Traditional feedback control…

Optimization and Control · Mathematics 2024-12-16 Matteo Tomasetto , Francesco Braghin , Andrea Manzoni

We study the problem of learning safe control policies that are also effective; i.e., maximizing the probability of satisfying a linear temporal logic (LTL) specification of a task, and the discounted reward capturing the (classic) control…

Robotics · Computer Science 2026-04-07 Alper Kamil Bozkurt , Yu Wang , Miroslav Pajic

The paper deals with the problem of output regulation of nonlinear systems by presenting a learning-based adaptive internal model-based design strategy. We borrow from the adaptive internal model design technique recently proposed in [1]…

Systems and Control · Electrical Eng. & Systems 2022-06-27 Lorenzo Gentilini , Michelangelo Bin , Lorenzo Marconi

We study whether optimal state-feedback laws for a family of heterogeneous Multiple-Input, Multiple-Output (MIMO) Linear Time-Invariant (LTI) systems can be captured by a single learned controller. We train one transformer policy on…

Systems and Control · Electrical Eng. & Systems 2026-03-17 Turki Bin Mohaya , Maitham F. AL-Sunni , John M. Dolan , Peter Seiler

The goal of this paper is to develop data-driven control design and evaluation strategies based on linear matrix inequalities (LMIs) and dynamic programming. We consider deterministic discrete-time LTI systems, where the system model is…

Optimization and Control · Mathematics 2021-06-17 Donghwan Lee , Do Wan Kim

The field of quickest change detection (QCD) focuses on the design and analysis of online algorithms that estimate the time at which a significant event occurs. In this paper, design and analysis are cast in a Bayesian framework, where QCD…

Optimization and Control · Mathematics 2025-12-30 Austin Cooper , Sean Meyn

The training of autonomous agents often requires expensive and unsafe trial-and-error interactions with the environment. Nowadays several data sets containing recorded experiences of intelligent agents performing various tasks, spanning…

Machine Learning · Computer Science 2020-10-06 Giorgio Angelotti , Nicolas Drougard , Caroline Ponzoni Carvalho Chanel