中文
相关论文

相关论文: Stabilization of uncertain linear dynamics: an off…

200 篇论文

We analyze offline designs of linear quadratic regulator (LQR) strategies with uncertain disturbances. First, we consider the scenario where the exogenous variable can be estimated in a controlled environment, and subsequently, consider a…

系统与控制 · 电气工程与系统科学 2025-09-26 Sayak Mukherjee , Ramij R. Hossain , Mahantesh Halappanavar

To provide rigorous uncertainty quantification for online learning models, we develop a framework for constructing uncertainty sets that provably control risk -- such as coverage of confidence intervals, false negative rate, or F1 score --…

机器学习 · 计算机科学 2023-01-30 Shai Feldman , Liran Ringel , Stephen Bates , Yaniv Romano

Equipped with the trained environmental dynamics, model-based offline reinforcement learning (RL) algorithms can often successfully learn good policies from fixed-sized datasets, even some datasets with poor quality. Unfortunately, however,…

机器学习 · 计算机科学 2023-07-27 Junjie Zhang , Jiafei Lyu , Xiaoteng Ma , Jiangpeng Yan , Jun Yang , Le Wan , Xiu Li

We present a novel technique for solving the problem of safe control for a general class of nonlinear, control-affine systems subject to parametric model uncertainty. Invoking Lyapunov analysis and the notion of fixed-time stability (FxTS),…

最优化与控制 · 数学 2020-11-26 Mitchell Black , Ehsan Arabi , Dimitra Panagou

Recent work by Mania et al. has proved that certainty equivalent control achieves nearly optimal regret for linear systems with quadratic costs. However, when parameter uncertainty is large, certainty equivalence cannot be relied upon to…

最优化与控制 · 数学 2020-01-01 Jack Umenberger , Thomas B. Schon

The typical offline protocol to evaluate recommendation algorithms is to collect a dataset of user-item interactions and then use a part of this dataset to train a model, and the remaining data to measure how closely the model…

信息检索 · 计算机科学 2026-05-15 Maria João Lavoura , Robert Jungnickel , João Vinagre

We study offline-online reinforcement learning in linear mixture Markov decision processes (MDPs) under environment shift. In the offline phase, data are collected by an unknown behavior policy and may come from a mismatched environment,…

机器学习 · 计算机科学 2026-04-15 Zhongjun Zhang , Sean R. Sinclair

This paper proposes a framework for adaptively learning a feedback linearization-based tracking controller for an unknown system using discrete-time model-free policy-gradient parameter update rules. The primary advantage of the scheme over…

In recent years, stabilizing unknown dynamical systems has became a critical problem in control systems engineering. Addressing this for linear time-invariant (LTI) systems is an essential fist step towards solving similar problems for more…

最优化与控制 · 数学 2025-08-08 Xinpei Zhang , Guangyan Jia

Stabilizing unstable periodic orbits in a chaotic invariant set not only reveals information about its structure but also leads to various interesting applications. For the successful application of a chaos control scheme, convergence speed…

适应与自组织系统 · 物理学 2016-10-10 Christian Bick , Marc Timme , Christoph Kolodziejski

Current Reinforcement Learning (RL) is often limited by the large amount of data needed to learn a successful policy. Offline RL aims to solve this issue by using transitions collected by a different behavior policy. We address a novel…

机器学习 · 计算机科学 2024-05-29 Johannes Ackermann , Takayuki Osa , Masashi Sugiyama

This paper investigates adaptive model predictive control (MPC) for a class of constrained linear systems with unknown model parameters. This is also posed as the dual control problem consisting of system identification and regulation. We…

最优化与控制 · 数学 2020-11-24 Kunwu Zhang , Yang Shi

We provide two solutions to the heretofore open problem of stabilization of systems with arbitrarily long delays at the input and output of a nonlinear system using output feedback only. Both of our solutions are global, employ the…

最优化与控制 · 数学 2011-08-24 Iasson Karafyllis , Miroslav Krstic

This paper addresses the problem of online inverse reinforcement learning for nonlinear systems with modeling uncertainties while in the presence of unknown disturbances. The developed approach observes state and input trajectories for an…

系统与控制 · 电气工程与系统科学 2021-07-07 Ryan Self , Moad Abudia , Rushikesh Kamalapurkar

Model-based reinforcement learning techniques accelerate the learning task by employing a transition model to make predictions. In this paper, a model-based learning approach is presented that iteratively computes the optimal value function…

最优化与控制 · 数学 2020-10-22 Milad Farsi , Jun Liu

We study feedback stabilization of continuous-time linear systems under finite data-rate constraints in the presence of unknown disturbances. A communication and control strategy based on sampled and quantized state measurements is…

系统与控制 · 电气工程与系统科学 2026-03-31 Mahmoud Zamani , Guosong Yang

An autonomous and resilient controller is proposed for leader-follower multi-agent systems under uncertainties and cyber-physical attacks. The leader is assumed non-autonomous with a nonzero control input, which allows changing the team…

多智能体系统 · 计算机科学 2018-04-10 Rohollah Moghadam , Hamidreza Modares

In this work, we consider the problem of online (real-time, single-shot) estimation of static or slow-varying parameters along quantum trajectories in quantum dynamical systems. Based on the measurement signal of a continuously-monitored…

量子物理 · 物理学 2024-06-19 Henrik Glavind Clausen , Pierre Rouchon , Rafal Wisniewski

The prescribed-time stabilization problem for a general class of nonlinear systems with unknown input gain and appended dynamics (with unmeasured state) is addressed. Unlike the asymptotic stabilization problem, the prescribed-time…

最优化与控制 · 数学 2021-08-10 Prashanth Krishnamurthy , Farshad Khorrami

Stability and control of a non-linear system represent an important system configuration that frequently arises in practical engineering. Stability covers a vast range of systems that do not obey the superposition principle and applies to…

系统与控制 · 电气工程与系统科学 2022-02-04 Asifa Yousaf