中文
相关论文

相关论文: Regret-Guaranteed Safe Switching with Minimum Cost…

200 篇论文

Sustained research efforts have been devoted to learning optimal controllers for linear stochastic dynamical systems with unknown parameters, but due to the corruption of noise, learned controllers are usually uncertified in the sense that…

系统与控制 · 电气工程与系统科学 2023-02-07 Yiwen Lu , Yilin Mo

Bandit-style algorithms have been studied extensively in stochastic and adversarial settings. Such algorithms have been shown to be useful in multiplayer settings, e.g. to solve the wireless network selection problem, which can be…

网络与互联网体系结构 · 计算机科学 2019-04-30 Shunhao Oh , Anuja Meetoo Appavoo , Seth Gilbert

Learning to control an unknown dynamical system with respect to high-level temporal specifications is an important problem in control theory. We present the first regret-free online algorithm for learning a controller for linear temporal…

人工智能 · 计算机科学 2025-06-09 Rupak Majumdar , Mahmoud Salamati , Sadegh Soudjani

This paper investigates the reinforcement learning (RL) based disturbance rejection control for uncertain nonlinear systems having non-simple nominal models. An extended state observer (ESO) is first designed to estimate the system state…

动力系统 · 数学 2020-11-25 Maopeng Ran , Juncheng Li , Lihua Xie

Extra-Sensory Perception (ESP) cheats, which reveal hidden in-game information such as enemy locations, are difficult to detect because their effects are not directly observable in player behavior. The lack of observable evidence makes it…

机器学习 · 计算机科学 2026-01-01 Inkyu Park , Jeong-Gwan Lee , Taehwan Kwon , Juheon Choi , Seungku Kim , Junsu Kim , Kimin Lee

The quality of control (QoC) of a resource-constrained embedded control system may be jeopardized in dynamic environments with variable workload. This gives rise to the increasing demand of co-design of control and scheduling. To deal with…

其他计算机科学 · 计算机科学 2008-12-18 Feng Xia , Youxian Sun , Yu-Chu Tian , Moses Tade , Jinxiang Dong

The standard supervised learning paradigm works effectively when training data shares the same distribution as the upcoming testing samples. However, this stationary assumption is often violated in real-world applications, especially when…

机器学习 · 计算机科学 2023-01-18 Yong Bai , Yu-Jie Zhang , Peng Zhao , Masashi Sugiyama , Zhi-Hua Zhou

Recent developments in digital platforms have highlighted the prevalence of open systems, where agents can arrive and depart over time. While bandit learning in open systems has recently received initial attention, existing work imposes…

机器学习 · 计算机科学 2026-05-08 Mengfan Xu

We consider a type of optimal switching problems with non-uniform execution delays and ramping. Such problems frequently occur in the operation of economical and engineering systems. We first provide a solution to the problem by applying a…

最优化与控制 · 数学 2017-02-15 Magnus Perninge

Continuous-time adaptive controllers for systems with a matched uncertainty often comprise an online parameter estimator and a corresponding parameterized controller to cancel the uncertainty. However, such methods are often impossible to…

系统与控制 · 电气工程与系统科学 2025-03-18 Aren Karapetyan , Efe C. Balta , Anastasios Tsiamis , Andrea Iannelli , John Lygeros

The increase in renewable energy sources (RESs), like wind or solar power, results in growing uncertainty also in transmission grids. This affects grid stability through fluctuating energy supply and an increased probability of overloaded…

系统与控制 · 电气工程与系统科学 2022-04-13 Rebecca Bauer , Tillmann Mühlpfordt , Nicole Ludwig , Veit Hagenmeyer

This paper investigates a class of games with large strategy spaces, motivated by challenges in AI alignment and language games. We introduce the hidden game problem, where for each player, an unknown subset of strategies consistently…

人工智能 · 计算机科学 2025-10-07 Gon Buzaglo , Noah Golowich , Elad Hazan

This paper proposes a novel and interpretable recurrent neural-network structure using the echo-state network (ESN) paradigm for time-series prediction. While the traditional ESNs perform well for dynamical systems prediction, it needs a…

机器学习 · 计算机科学 2024-04-01 Debdipta Goswami

In this work, we focus on the design of optimal controllers that must comply with an information structure. State-of-the-art approaches do so based on the H2 or Hinfty norm to minimize the expected or worst-case cost in the presence of…

系统与控制 · 电气工程与系统科学 2025-11-24 Daniele Martinelli , Andrea Martin , Giancarlo Ferrari-Trecate , Luca Furieri

Automotive FMCW radars are indispensable to modern ADAS and autonomous-driving systems, but their increasing density has intensified the risk of mutual interference. Existing mitigation techniques, including reactive receiver-side…

系统与控制 · 电气工程与系统科学 2026-01-01 Yunian Pan , Jun Li , Lifan Xu , Shunqiao Sun , Quanyan Zhu

High penetrations of intermittent renewable energy resources in the power system require large balancing reserves for reliable operations. Aggregated and coordinated behind-the-meter loads can provide these fast reserves, but represent…

最优化与控制 · 数学 2019-10-03 Mahraz Amini , Mads Almassalkhi

We study the problem of reinforcement learning (RL) with low (policy) switching cost - a problem well-motivated by real-life RL applications in which deployments of new policies are costly and the number of policy updates must be low. In…

机器学习 · 计算机科学 2022-06-07 Dan Qiao , Ming Yin , Ming Min , Yu-Xiang Wang

This study develops an inverse portfolio optimization framework for recovering latent investor preferences including risk aversion, transaction cost sensitivity, and ESG orientation from observed portfolio allocations. Using controlled…

综合金融 · 定量金融 2025-10-14 Jinho Cha , Long Pham , Thi Le Hoa Vo , Jaeyoung Cho , Jaejin Lee

The dynamical behavior of switched affine systems is known to be more intricate than that of the well-studied switched linear systems, essentially due to the existence of distinct equilibrium points for each subsystem. First, under…

系统与控制 · 电气工程与系统科学 2022-03-15 Matteo Della Rossa , Lucas N. Egidio , Raphaël M. Jungers

Error feedback (EF), also known as error compensation, is an immensely popular convergence stabilization mechanism in the context of distributed training of supervised machine learning models enhanced by the use of contractive communication…

机器学习 · 计算机科学 2021-06-10 Peter Richtárik , Igor Sokolov , Ilyas Fatkhullin