中文
相关论文

相关论文: Online switching control with stability and regret…

200 篇论文

We consider the problem of controlling an unknown linear dynamical system in the presence of (nonstochastic) adversarial perturbations and adversarial convex loss functions. In contrast to classical control, the a priori determination of an…

机器学习 · 计算机科学 2020-01-22 Elad Hazan , Sham M. Kakade , Karan Singh

We study online conformal prediction for non-stationary data streams subject to unknown distribution drift. While most prior work studied this problem under adversarial settings and/or assessed performance in terms of gaps of time-averaged…

统计理论 · 数学 2026-03-06 Jiadong Liang , Zhimei Ren , Yuxin Chen

We present online prediction methods for time series that let us explicitly handle nonstationary artifacts (e.g. trend and seasonality) present in most real time series. Specifically, we show that applying appropriate transformations to…

机器学习 · 统计学 2018-08-28 Christopher Xie , Avleen Bijral , Juan Lavista Ferres

In this paper we focus on the solution of online problems with time-varying, linear equality and inequality constraints. Our approach is to design a novel online algorithm by leveraging the tools of control theory. In particular, for the…

最优化与控制 · 数学 2025-09-04 Umberto Casti , Nicola Bastianello , Ruggero Carli , Sandro Zampieri

In online selective conformal inference, data arrives sequentially, and prediction intervals are constructed only when an online selection rule is met. Since online selections may break the exchangeability between the selected test datum…

机器学习 · 统计学 2025-03-24 Yusuf Sale , Aaditya Ramdas

This article is concerned with stability analysis and stabilization of randomly switched systems under a class of switching signals. The switching signal is modeled as a jump stochastic (not necessarily Markovian) process independent of the…

最优化与控制 · 数学 2011-10-04 Debasish Chatterjee , Daniel Liberzon

This paper proposes a modular approach that combines the online convex optimization framework and reference governors to solve a constrained control problem featuring time-varying and a priori unknown cost functions. Compared to existing…

系统与控制 · 电气工程与系统科学 2025-07-14 Marko Nonhoff , Johannes Köhler , Matthias A. Müller

Stabilizing unstable periodic orbits in a chaotic invariant set not only reveals information about its structure but also leads to various interesting applications. For the successful application of a chaos control scheme, convergence speed…

适应与自组织系统 · 物理学 2016-10-10 Christian Bick , Marc Timme , Christoph Kolodziejski

Motivated by the fact that humans like some level of unpredictability or novelty, and might therefore get quickly bored when interacting with a stationary policy, we introduce a novel non-stationary bandit problem, where the expected reward…

机器学习 · 计算机科学 2022-03-08 Pierre Laforgue , Giulia Clerici , Nicolò Cesa-Bianchi , Ran Gilad-Bachrach

Soft robots manufactured with flexible materials can be highly compliant and adaptive to their surroundings, which facilitates their application in areas such as dexterous manipulation and environmental exploration. This paper aims at…

系统与控制 · 电气工程与系统科学 2025-07-01 Renjie Ma , Ziyao Qu , Zhijian Hu , Dong Zhao , Marios M. Polycarpou

We investigate the problem of online learning, which has gained significant attention in recent years due to its applicability in a wide range of fields from machine learning to game theory. Specifically, we study the online optimization of…

机器学习 · 计算机科学 2021-09-29 Kaan Gokcesu , Hakan Gokcesu

The assignment game models a housing market where buyers and sellers are matched, and transaction prices are set so that the resulting allocation is stable. Shapley and Shubik showed that every stable allocation is necessarily built on a…

计算机科学与博弈论 · 计算机科学 2026-02-23 Emile Martinez , Felipe Garrido-Lucero , Umberto Grandi

We analyze the convergence properties of a robust adaptive model predictive control algorithm used to control an unknown nonlinear system. We show that by employing a standard quadratic stabilizing cost function, and by recursively updating…

最优化与控制 · 数学 2024-05-31 Riccardo Zuliani , Raffaele Soloperto , John Lygeros

In online learning, the data is provided in a sequential order, and the goal of the learner is to make online decisions to minimize overall regrets. This note is concerned with continuous-time models and algorithms for several online…

机器学习 · 统计学 2024-05-20 Lexing Ying

We introduce algorithms for online, full-information prediction that are competitive with contextual tree experts of unknown complexity, in both probabilistic and adversarial settings. We show that by incorporating a probabilistic framework…

机器学习 · 计算机科学 2018-05-23 Vidya Muthukumar , Mitas Ray , Anant Sahai , Peter L. Bartlett

As the relevance of control systems capable of dealing with multiple objectives rises (e.g. being economic while maintaining a certain performance), multi-objective Switched Model Predictive Control combines all the advantages of Model…

最优化与控制 · 数学 2025-11-18 Elias Niepötter , Adrian Grimm , Torbjørn Cunis

Continuous-time adaptive controllers for systems with a matched uncertainty often comprise an online parameter estimator and a corresponding parameterized controller to cancel the uncertainty. However, such methods are often impossible to…

系统与控制 · 电气工程与系统科学 2025-03-18 Aren Karapetyan , Efe C. Balta , Anastasios Tsiamis , Andrea Iannelli , John Lygeros

We develop a new framework for designing online policies given access to an oracle providing statistical information about an offline benchmark. Having access to such prediction oracles enables simple and natural Bayesian selection…

数据结构与算法 · 计算机科学 2020-02-28 Alberto Vera , Siddhartha Banerjee

We study control problems in the context of matching under preferences: We examine how a central authority, called the controller, can manipulate an instance of the Stable Marriage or Stable Roommates problems in order to achieve certain…

计算机科学与博弈论 · 计算机科学 2025-02-04 Jiehua Chen , Ildikó Schlotter

We study the prediction with expert advice setting, where the aim is to produce a decision by combining the decisions generated by a set of experts, e.g., independently running algorithms. We achieve the min-max optimal dynamic regret under…

机器学习 · 计算机科学 2022-08-09 Hakan Gokcesu , Suleyman S. Kozat