中文
相关论文

相关论文: Safe Output Feedback Improvement with Baselines

200 篇论文

We consider the problem of selecting a subset of points from a dataset of $n$ unlabeled examples for labeling, with the goal of training a multiclass classifier. To address this, we build upon the regret minimization framework introduced by…

机器学习 · 计算机科学 2026-02-27 Youguang Chen , George Biros

A high-gain observer is used for a class of feedback linearisable nonlinear systems to synthesize safety-preserving controllers over the observer output. A bound on the distance between trajectories under state and output feedback is…

系统与控制 · 计算机科学 2016-03-23 Kendra Lesser , Alessandro Abate

This paper derives an optimal control strategy for a simple stochastic dynamical system with constant drift and an additive control input. Motivated by the example of a physical system with an unexpected change in its dynamics, we take the…

最优化与控制 · 数学 2022-02-09 Daniel Gurevich , Debdipta Goswami , Charles L. Fefferman , Clarence W. Rowley

Feedback optimization is a control paradigm that enables physical systems to autonomously reach efficient operating points. Its central idea is to interconnect optimization iterations in closed-loop with the physical plant. Since iterative…

最优化与控制 · 数学 2024-07-16 Zhiyu He , Saverio Bolognani , Jianping He , Florian Dörfler , Xinping Guan

Over the recent past data-driven algorithms for solving stochastic optimal control problems in face of model uncertainty have become an increasingly active area of research. However, for singular controls and underlying diffusion dynamics…

最优化与控制 · 数学 2024-10-15 Sören Christensen , Asbjørn Holk Thomsen , Lukas Trottner

This paper studies online solutions for regret-optimal control in partially observable systems over an infinite-horizon. Regret-optimal control aims to minimize the difference in LQR cost between causal and non-causal controllers while…

系统与控制 · 电气工程与系统科学 2023-11-15 Joudi Hajar , Oron Sabag , Babak Hassibi

Feedback control actively dissipates uncertainty from a dynamical system by means of actuation. We develop a notion of "control capacity" that gives a fundamental limit (in bits) on the rate at which a controller can dissipate the…

信息论 · 计算机科学 2017-01-17 Gireeja Ranade , Anant Sahai

We study the problem of adaptively controlling a known discrete-time nonlinear system subject to unmodeled disturbances. We prove the first finite-time regret bounds for adaptive nonlinear control with matched uncertainty in the stochastic…

机器学习 · 计算机科学 2020-11-30 Nicholas M. Boffi , Stephen Tu , Jean-Jacques E. Slotine

This paper proposes a data-driven framework to solve time-varying optimization problems associated with unknown linear dynamical systems. Making online control decisions to regulate a dynamical system to the solution of an optimization…

最优化与控制 · 数学 2021-09-08 Gianluca Bianchin , Miguel Vaquero , Jorge Cortes , Emiliano Dall'Anese

Minimum variance controllers have been employed in a wide-range of industrial applications. A key challenge experienced by many adaptive controllers is their poor empirical performance in the initial stages of learning. In this paper, we…

系统与控制 · 电气工程与系统科学 2023-05-29 Rahul Singh , Akshay Mete , Avik Kar , P. R. Kumar

This paper studies bandit convex optimization with constraints, where the learner aims to generate a sequence of decisions under partial information of loss functions such that the cumulative loss is reduced as well as the cumulative…

机器学习 · 计算机科学 2023-10-18 Yasunari Hikima

Motivated by perception-based control problems in autonomous systems, this paper addresses the problem of developing feedback controllers to regulate the inputs and the states of a dynamical system to optimal solutions of an optimization…

系统与控制 · 电气工程与系统科学 2023-10-17 Liliaokeawawa Cothren , Gianluca Bianchin , Sarah Dean , Emiliano Dall'Anese

We consider robust counterparts of uncertain combinatorial optimization problems, where the difference to the best possible solution over all scenarios is to be minimized. Such minmax regret problems are typically harder to solve than their…

最优化与控制 · 数学 2016-06-06 A. Chassein , M. Goerigk

This paper develops a method to learn optimal controls from data for bilinear systems without a priori knowledge of the system dynamics. Given an unknown bilinear system, we first characterize when the available data is suitable to solve…

最优化与控制 · 数学 2023-10-13 Zhenyi Yuan , Jorge Cortes

We consider a safe optimization problem with bandit feedback in which an agent sequentially chooses actions and observes responses from the environment, with the goal of maximizing an arbitrary function of the response while respecting…

机器学习 · 计算机科学 2023-05-02 Spencer Hutchinson , Berkay Turan , Mahnoosh Alizadeh

Fundamental limits on the performance of feedback controllers are essential for benchmarking algorithms, guiding sensor selection, and certifying task feasibility -- yet few general-purpose tools exist for computing them. Existing…

最优化与控制 · 数学 2026-05-26 Vincent Pacelli , Evangelos A. Theodorou

We introduce a method to deal with the data-driven control design of nonlinear systems. We derive conditions to design controllers via (approximate) nonlinearity cancellation. These conditions take the compact form of data-dependent…

系统与控制 · 电气工程与系统科学 2022-01-26 Claudio De Persis , Monica Rotulo , Pietro Tesi

We address the problem of learning to control an unknown nonlinear dynamical system through sequential interactions. Motivated by high-stakes applications in which mistakes can be catastrophic, such as robotics and healthcare, we study…

机器学习 · 计算机科学 2025-04-14 James Wang , Bruce D. Lee , Ingvar Ziemann , Nikolai Matni

Feedback optimization enables autonomous optimality seeking of a dynamical system through its closed-loop interconnection with iterative optimization algorithms. Among various iteration structures, model-based approaches require the…

最优化与控制 · 数学 2026-05-26 Zhiyu He , Saverio Bolognani , Michael Muehlebach , Florian Dörfler

We consider the problem of using observational bandit feedback data from multiple heterogeneous data sources to learn a personalized decision policy that robustly generalizes across diverse target settings. To achieve this, we propose a…

机器学习 · 计算机科学 2024-10-14 Aldo Gael Carranza , Susan Athey