English
Related papers

Related papers: Learn and Control while Switching: with Guaranteed…

200 papers

We provide an algorithm for the simultaneous system identification and model predictive control of nonlinear systems. The algorithm has finite-time near-optimality guarantees and asymptotically converges to the optimal (non-causal)…

Robotics · Computer Science 2025-11-04 Hongyu Zhou , Vasileios Tzoumas

Autonomous agents often require multiple strategies to solve complex tasks, but determining when to switch between strategies remains challenging. This research introduces a reinforcement learning technique to learn switching thresholds…

Machine Learning · Computer Science 2025-12-09 Chris Tava

We present an iterative approach for planning and controlling motions of underactuated robots with uncertain dynamics. At its core, there is a learning process which estimates the perturbations induced by the model uncertainty on the active…

We consider the problem of online adaptive control of the linear quadratic regulator, where the true system parameters are unknown. We prove new upper and lower bounds demonstrating that the optimal regret scales as…

Machine Learning · Computer Science 2023-10-05 Max Simchowitz , Dylan J. Foster

We study the problem of learning-augmented predictive linear quadratic control. Our goal is to design a controller that balances \textit{"consistency"}, which measures the competitive ratio when predictions are accurate, and…

Systems and Control · Electrical Eng. & Systems 2025-04-08 Tongxin Li , Ruixiao Yang , Guannan Qu , Guanya Shi , Chenkai Yu , Adam Wierman , Steven H. Low

In software-defined networking (SDN) systems, it is a common practice to adopt a multi-controller design and control devolution techniques to improve the performance of the control plane. However, in such systems, the decision-making for…

Networking and Internet Architecture · Computer Science 2020-08-05 Xi Huang , Yinxu Tang , Ziyu Shao , Yang Yang , Hong Xu

We investigate online convex optimization in non-stationary environments and choose dynamic regret as the performance measure, defined as the difference between cumulative loss incurred by the online algorithm and that of any feasible…

Machine Learning · Computer Science 2024-04-09 Peng Zhao , Yu-Jie Zhang , Lijun Zhang , Zhi-Hua Zhou

We consider the problem of online learning where the sequence of actions played by the learner must adhere to an unknown safety constraint at every round. The goal is to minimize regret with respect to the best safe action in hindsight…

Machine Learning · Computer Science 2024-03-08 Karthik Sridharan , Seung Won Wilson Yoo

We consider systems that require timely monitoring of sources over a communication network, where the cost of delayed information is unknown, time-varying and possibly adversarial. For the single source monitoring problem, we design…

Networking and Internet Architecture · Computer Science 2021-05-31 Vishrant Tripathi , Eytan Modiano

We study the problem of online learning in predictive control of an unknown linear dynamical system with time varying cost functions which are unknown apriori. Specifically, we study the online learning problem where the control algorithm…

Machine Learning · Computer Science 2022-11-01 Deepan Muthirayan , Jianjun Yuan , Dileep Kalathil , Pramod P. Khargonekar

Though switched dynamical systems have shown great utility in modeling a variety of physical phenomena, the construction of an optimal control of such systems has proven difficult since it demands some type of optimal mode scheduling. In…

Optimization and Control · Mathematics 2014-02-04 Ramanarayan Vasudevan , Humberto Gonzalez , Ruzena Bajcsy , S. Shankar Sastry

This paper presents a novel approach to sustain transient chaos in the Lorenz system through the estimation of safety functions using a transformer-based model. Unlike classical methods that rely on iterative computations, the proposed…

Chaotic Dynamics · Physics 2025-04-01 David Valle , Rubén Capeans , Alexandre Wagemakers , Miguel A. F. Sanjuán

A key challenge in the field of reinforcement learning is to develop agents that behave cautiously in novel situations. It is generally impossible to anticipate all situations that an autonomous system may face or what behavior would best…

Artificial Intelligence · Computer Science 2025-10-14 Montaser Mohammedalamen , Dustin Morrill , Alexander Sieusahai , Yash Satsangi , Michael Bowling

A wide range of sustainability and grid-integration strategies depend on workload shifting, which aligns the timing of energy consumption with external signals such as grid curtailment events, carbon intensity, or time-of-use electricity…

Data Structures and Algorithms · Computer Science 2025-10-01 Ezra Johnson , Adam Lechowicz , Mohammad Hajiesmaili

Quadruped robots have strong adaptability to extreme environments but may also experience faults. Once these faults occur, robots must be repaired before returning to the task, reducing their practical feasibility. One prevalent concern…

Robotics · Computer Science 2024-01-01 Xinyuan Wu , Wentao Dong , Hang Lai , Yong Yu , Ying Wen

The integration of distributed energy resources (DERs) into sub-transmission systems has enabled new opportunities for flexibility provision in ancillary services such as frequency and voltage support, as well as congestion management. This…

Systems and Control · Electrical Eng. & Systems 2025-06-17 Florian Klein-Helmkamp , Tina Möllemann , Irina Zettl , Andreas Ulbig

In this paper we provide a set of stability conditions for linear time-varying networked control systems with arbitrary topologies using a piecewise quadratic switching stabilization approach with multiple quadratic Lyapunov functions. We…

Optimization and Control · Mathematics 2018-04-04 Mohammad Razeghi-Jahromi , Saeed Manaffam , Alireza Seyedi , Azadeh Vosoughi

We study the problem of adaptive control of the linear quadratic regulator for systems in very high, or even infinite dimension. We demonstrate that while sublinear regret requires finite dimensional inputs, the ambient state dimension of…

Optimization and Control · Mathematics 2021-07-16 Juan C. Perdomo , Max Simchowitz , Alekh Agarwal , Peter Bartlett

In this paper, we investigate the fixed-time behavioral control problem for a team of second-order nonlinear agents, aiming to achieve a desired formation with collision/obstacle~avoidance. In the proposed approach, the two behaviors(tasks)…

Optimization and Control · Mathematics 2021-03-12 Ning Zhou , Xiaodong Cheng , Zhongqi Sun , Yuanqing Xia

In performative prediction, the deployment of a predictive model triggers a shift in the data distribution. As these shifts are typically unknown ahead of time, the learner needs to deploy a model to get feedback about the distribution it…

Machine Learning · Computer Science 2022-07-19 Meena Jagadeesan , Tijana Zrnic , Celestine Mendler-Dünner