中文
相关论文

相关论文: Robust Online Residual Refinement via Koopman-Guid…

200 篇论文

Data-driven control methods need to be sample-efficient and lightweight, especially when data acquisition and computational resources are limited -- such as during learning on hardware. Most modern data-driven methods require large datasets…

机器人学 · 计算机科学 2025-09-11 Zixin Zhang , James Avtges , Todd D. Murphey

The accurate modeling of dynamics in interactive environments is critical for successful long-range prediction. Such a capability could advance Reinforcement Learning (RL) and Planning algorithms, but achieving it is challenging.…

机器学习 · 计算机科学 2024-05-14 Arnab Kumar Mondal , Siba Smarak Panigrahi , Sai Rajeswar , Kaleem Siddiqi , Siamak Ravanbakhsh

Reinforcement Learning (RL) has made significant strides in various domains, and policy gradient methods like Proximal Policy Optimization (PPO) have gained popularity due to their balance in performance, training stability, and…

机器学习 · 计算机科学 2025-05-21 Andrei Cozma , Landon Harris , Hairong Qi

This paper presents a model-based reinforcement learning (RL) framework for optimal closed-loop control of nonlinear robotic systems. The proposed approach learns linear lifted dynamics through Koopman operator theory and integrates the…

机器人学 · 计算机科学 2026-04-23 Wenjian Hao , Yuxuan Fang , Zehui Lu , Shaoshuai Mou

Reinforcement learning (RL) models have shown the capability of learning complex behaviors, but quantitatively assessing those behaviors - which is critical for safety assurance and the discovery of novel strategies - is challenging. By…

最优化与控制 · 数学 2026-03-23 William T. Redman

This paper proposes a Koopman-based framework for modeling, prediction, and control of unknown nonlinear time-varying systems. We present a novel Koopman-based learning method for predicting the state of unknown nonlinear time-varying…

系统与控制 · 电气工程与系统科学 2026-01-30 Hengde Zhang , Yunxiao Ren , Zhisheng Duan , Zhiyong Sun , Guanrong Chen

Offline reinforcement learning leverages large datasets to train policies without interactions with the environment. The learned policies may then be deployed in real-world settings where interactions are costly or dangerous. Current…

机器学习 · 计算机科学 2022-06-29 Matthias Weissenbacher , Samarth Sinha , Animesh Garg , Yoshinobu Kawahara

A learning method is proposed for Koopman operator-based models with the goal of improving closed-loop control behavior. A neural network-based approach is used to discover a space of observables in which nonlinear dynamics is linearly…

最优化与控制 · 数学 2023-03-23 Daisuke Uchida , Karthik Duraisamy

We study a problem of simultaneous system identification and model predictive control of nonlinear systems. Particularly, we provide an algorithm for systems with unknown residual dynamics that can be expressed by Koopman operators. Such…

系统与控制 · 电气工程与系统科学 2025-12-11 Hongyu Zhou , Vasileios Tzoumas

Time-dependent structural reliability analysis of nonlinear dynamical systems is non-trivial; subsequently, scope of most of the structural reliability analysis methods is limited to time-independent reliability analysis only. In this work,…

机器学习 · 统计学 2024-09-21 Navaneeth N. , Souvik Chakraborty

Many machine learning approaches for decision making, such as reinforcement learning, rely on simulators or predictive models to forecast the time-evolution of quantities of interest, e.g., the state of an agent or the reward of a policy.…

机器学习 · 计算机科学 2024-01-17 Petar Bevanda , Max Beier , Armin Lederer , Stefan Sosnowski , Eyke Hüllermeier , Sandra Hirche

This study presents an innovative approach to Model Predictive Control (MPC) by leveraging the powerful combination of Koopman theory and Deep Reinforcement Learning (DRL). By transforming nonlinear dynamical systems into a…

系统与控制 · 电气工程与系统科学 2025-05-22 Md Nur-A-Adam Dony

Approaches based on Koopman operators have shown great promise in forecasting time series data generated by complex nonlinear dynamical systems (NLDS). Although such approaches are able to capture the latent state representation of a NLDS,…

机器学习 · 计算机科学 2024-10-01 Ashutosh Singh , Ashish Singh , Tales Imbiriba , Deniz Erdogmus , Ricardo Borsoi

We introduce Conformal Online Learning of Koopman embeddings (COLoKe), a novel framework for adaptively updating Koopman-invariant representations of nonlinear dynamical systems from streaming data. Our modeling approach combines deep…

机器学习 · 计算机科学 2026-01-28 Ben Gao , Jordan Patracone , Stéphane Chrétien , Olivier Alata

Imitation learning (IL) is a general learning paradigm for tackling sequential decision-making problems. Interactive imitation learning, where learners can interactively query for expert demonstrations, has been shown to achieve provably…

机器学习 · 计算机科学 2022-09-27 Yichen Li , Chicheng Zhang

Iterative refinement has emerged as an effective paradigm for enhancing the capabilities of large language models (LLMs) on complex tasks. However, existing approaches typically implement iterative refinement at the application or prompting…

计算与语言 · 计算机科学 2024-10-15 Yuxi Xie , Anirudh Goyal , Xiaobao Wu , Xunjian Yin , Xiao Xu , Min-Yen Kan , Liangming Pan , William Yang Wang

Koopman operator theory provides a framework for nonlinear dynamical system analysis and time-series forecasting by mapping dynamics to a space of real-valued measurement functions, enabling a linear operator representation. Despite the…

机器学习 · 计算机科学 2025-06-18 Yitian Zhang , Liheng Ma , Antonios Valkanas , Boris N. Oreshkin , Mark Coates

In many real-world settings, reinforcement learning systems suffer performance degradation when the environment encountered at deployment differs from that observed during training. Distributionally robust reinforcement learning (DR-RL)…

机器学习 · 计算机科学 2026-03-05 Debamita Ghosh , George K. Atia , Yue Wang

We study a class of dynamical systems modelled as Markov chains that admit an invariant distribution via the corresponding transfer, or Koopman, operator. While data-driven algorithms to reconstruct such operators are well known, their…

Long-horizon dynamical prediction is fundamental in robotics and control, underpinning canonical methods like model predictive control. Yet, many systems and disturbance phenomena are difficult to model due to effects like nonlinearity,…

机器人学 · 计算机科学 2025-12-04 Albert H. Li , Ivan Dario Jimenez Rodriguez , Joel W. Burdick , Yisong Yue , Aaron D. Ames
‹ 上一页 1 2 3 10 下一页 ›