中文
相关论文

相关论文: Technical Report: A New Decision-Theory-Based Fram…

200 篇论文

Offline safe reinforcement learning (RL) aims to train a constraint satisfaction policy from a fixed dataset. Current state-of-the-art approaches are based on supervised learning with a conditioned policy. However, these approaches fall…

机器学习 · 计算机科学 2025-01-28 Zijian Guo , Weichao Zhou , Wenchao Li

This paper is concerned with the design of a linear control law for linear systems with stationary additive disturbances. The objective is to find a state feedback gain that minimizes a quadratic stage cost function, while observing chance…

最优化与控制 · 数学 2025-10-03 Georg Schildbach , Paul Goulart , Manfred Morari

Many processes in nature such as conformal changes in biomolecules and clusters of interacting particles, genetic switches, mechanical or electromechanical oscillators with added noise, and many others are modeled using stochastic…

最优化与控制 · 数学 2023-10-13 Jiaxin Yuan , Amar Shah , Channing Bentz , Maria Cameron

Conventional test-time adaptation (TTA) approaches typically adapt the model using only a small fraction of test samples, often those with low-entropy predictions, thereby failing to fully leverage the available information in the test…

计算机视觉与模式识别 · 计算机科学 2026-04-21 Nam Nguyen Phuong , Duc Nguyen The Minh , Phi Le Nguyen , Ehsan Abbasnejad , Minh Hoai

We propose a new method for the problem of controlling linear dynamical systems under partial observation and adversarial disturbances. Our new algorithm, Double Spectral Control (DSC), matches the best known regret guarantees while…

机器学习 · 计算机科学 2025-05-28 Anand Brahmbhatt , Gon Buzaglo , Sofiia Druchyna , Elad Hazan

This paper considers controlled scalar systems relying on a lossy wireless feedback channel. In contrast with the existing literature, the focus is not on the system controller but on the wireless transmit power controller that is…

系统与控制 · 电气工程与系统科学 2022-05-30 Yifei Sun , Samson Lasaulce , Michel Kieffer , Romain Postoyan , Dragan Nešić

Decision Transformers (DT) play a crucial role in modern reinforcement learning, leveraging offline datasets to achieve impressive results across various domains. However, DT requires high-quality, comprehensive data to perform optimally.…

人工智能 · 计算机科学 2025-05-15 Minh Hoang Nguyen , Linh Le Pham Van , Thommen George Karimpanal , Sunil Gupta , Hung Le

Recent advances have shown how decision trees are apt data structures for concisely representing strategies (or controllers) satisfying various objectives. Moreover, they also make the strategy more explainable. The recent tool dtControl…

Clause Learning is one of the most important components of a conflict driven clause learning (CDCL) SAT solver that is effective on industrial instances. Since the number of learned clauses is proved to be exponential in the worse case, it…

人工智能 · 计算机科学 2017-06-01 Jerry Lonlac , Engelbert Mephu Nguifo

Differential balancing theory for nonlinear model reduction relies on differential controllability and observability functions. In this paper, we further investigate them from two different perspectives. First, we establish novel…

系统与控制 · 电气工程与系统科学 2025-04-07 Yu Kawano , Bart Besselink , Jacquelien M. A. Scherpen

Due to its rapid response time and a high degree of robustness, the selective fixed-filter active noise control (SFANC) method appears to be a viable candidate for widespread use in a variety of practical active noise control (ANC) systems.…

机器学习 · 计算机科学 2022-08-19 Zhengding Luo , Dongyuan Shi , Woon-Seng Gan

Estimating and detecting faults is crucial in ensuring safe and efficient automated systems. In the presence of disturbances, noise or varying system dynamics, such estimation is even more challenging. To address this challenge, this…

最优化与控制 · 数学 2021-12-13 Chris van der Ploeg , Emilia Silvas , Nathan van de Wouw , Peyman Mohajerin Esfahani

Analyzing and controlling system entropy is a powerful tool for regulating predictability of control systems. Applications benefiting from such approaches range from reinforcement learning and data security to human-robot collaboration. In…

系统与控制 · 电气工程与系统科学 2026-03-06 Menno van Zutphen , Giannis Delimpaltadakis , Duarte J. Antunes

For the application of MPC design in on-line regulation or tracking control problems, several studies have attempted to develop an accurate model, and realize adequate uncertainty description of linear or non-linear plants of the processes.…

最优化与控制 · 数学 2019-04-03 Yuanqiang Zhou , Dewei Li , Yugeng Xi , Zhongxue Gan

Construction project management requires dynamic mitigation control ensuring the project's timely completion by a best fit for common purpose strategy for all stakeholders. Current mitigation approaches are usually performed by an iterative…

适应与自组织系统 · 物理学 2022-06-22 Nina Prins , Omar Kammouh , A. R. M. Wolfert

Monte Carlo dropout may effectively capture model uncertainty in deep learning, where a measure of uncertainty is obtained by using multiple instances of dropout at test time. However, Monte Carlo dropout is applied across the whole network…

信号处理 · 电气工程与系统科学 2020-02-03 Liangping Ma , John Kaewell

This article proposes an improved trajectory optimization approach for stochastic optimal control of dynamical systems affected by measurement noise by combining optimal control with maximum likelihood techniques to improve the reduction of…

系统与控制 · 电气工程与系统科学 2023-12-25 Prakash Mallick , Zhiyong Chen

This paper addresses the distributed optimal frequency control of power systems considering a network-preserving model with nonlinear power flows and excitation voltage dynamics. Salient features of the proposed distributed control strategy…

系统与控制 · 计算机科学 2018-02-14 Zhaojian Wang , Feng Liu , John Z. F. Pang , Steven Low , Shengwei Mei

Large language models can exhibit emergent reasoning behaviors, often manifested as recurring lexical patterns (e.g., "wait," indicating verification). However, complex reasoning trajectories remain sparse in unconstrained sampling, and…

人工智能 · 计算机科学 2026-03-03 Po-Nien Kung , Zhen Yang , Jeffrey Luo , Cheng-Fu Yang , Haikang Deng , Zi-Yi Dou , Yinfei Yang , Nanyun Peng , Zhe Gan , Kai-Wei Chang

Agents acting in the natural world aim at selecting appropriate actions based on noisy and partial sensory observations. Many behaviors leading to decision mak- ing and action selection in a closed loop setting are naturally phrased within…

机器学习 · 统计学 2014-06-30 Alex Susemihl , Ron Meir , Manfred Opper