中文
相关论文

相关论文: An Adaptive Data-Enabled Policy Optimization Appro…

200 篇论文

PID control architectures are widely used in industrial applications. Despite their low number of open parameters, tuning multiple, coupled PID controllers can become tedious in practice. In this paper, we extend PILCO, a model-based policy…

机器学习 · 计算机科学 2017-03-09 Andreas Doerr , Duy Nguyen-Tuong , Alonso Marco , Stefan Schaal , Sebastian Trimpe

This paper focuses on adaptive control of the discrete-time linear quadratic regulator (adaptive LQR). Recent literature has made significant contributions in proving non-asymptotic convergence rates, but existing approaches have a few…

系统与控制 · 电气工程与系统科学 2026-04-27 Peter A. Fisher , Anuradha M. Annaswamy

Delays endanger safety of autonomous systems operating in a rapidly changing environment, such as nondeterministic surrounding traffic participants in autonomous driving and high-speed racing. Unfortunately, delays are typically not…

机器人学 · 计算机科学 2022-08-31 Dvij Kalaria , Qin Lin , John M. Dolan

Reinforcement learning (RL) has become central to enhancing reasoning in large language models (LLMs). Yet on-policy algorithms such as Group Relative Policy Optimization (GRPO) often suffer in early training: noisy gradients from…

机器学习 · 计算机科学 2026-03-19 Ziyan Wang , Zheng Wang , Xingwei Qu , Qi Cheng , Jie Fu , Shengpu Tang , Minjia Zhang , Xiaoming Huo

We design dynamic routing policies for an overlay network which meet delay requirements of real-time traffic being served on top of an underlying legacy network, where the overlay nodes do not know the underlay characteristics. We pose the…

网络与互联网体系结构 · 计算机科学 2019-04-19 Rahul Singh , Eytan Modiano

In many control systems, tracking accuracy can be enhanced by combining (data-driven) feedforward (FF) control with feedback (FB) control. However, designing effective data-driven FF controllers typically requires large amounts of…

机器学习 · 计算机科学 2026-03-25 Jakob Weber , Markus Gurtner , Benedikt Alt , Adrian Trachte , Andreas Kugi

Direct Preference Optimization (DPO) demonstrates the advantage of aligning a large language model with human preference using only an offline dataset. However, DPO has the limitation that the KL penalty, which prevents excessive deviation…

机器学习 · 计算机科学 2025-10-28 Sangkyu Lee , Janghoon Han , Hosung Song , Stanley Jungkyu Choi , Honglak Lee , Youngjae Yu

In this paper, an adaptive controller is designed for the synchronization of the trajectory of a robot with unknown kinematics and dynamics to that of the current human trajectory in the task space using the delayed human trajectory…

机器人学 · 计算机科学 2025-09-30 Rounak Bhattacharya , Vrithik R. Guthikonda , Ashwin P. Dani

Group Relative Policy Optimization (GRPO) effectively scales LLM reasoning but incurs prohibitive computational costs due to its extensive group-based sampling requirement. While recent selective data utilization methods can mitigate this…

机器学习 · 计算机科学 2026-03-05 Haodong Zhu , Yangyang Ren , Yanjing Li , Mingbao Lin , Linlin Yang , Xuhui Liu , Xiantong Zhen , Haiguang Liu , Baochang Zhang

Data-enabled predictive control (DeePC) has emerged as a powerful technique to control complex systems without the need for extensive modeling efforts. However, relying solely on offline collected data trajectories to represent the system…

系统与控制 · 电气工程与系统科学 2025-08-06 Sebastian Zieglmeier , Mathias Hudoba de Badyn , Narada D. Warakagoda , Thomas R. Krogstad , Paal Engelstad

We propose a novel flexible-step model predictive control algorithm for unknown linear time-invariant discrete-time systems. The goal is to asymptotically stabilize the system without relying on a pre-collected dataset that describes its…

最优化与控制 · 数学 2025-10-02 Markus Pietschner , Christian Ebenbauer , Bahman Gharesifard , Raik Suttner

In this paper, we propose the reduced model for the full dynamics of a bicycle and analyze its nonlinear behavior under a proportional control law for steering. Based on the Gibbs-Appell equations for the Whipple bicycle, we obtain a…

系统与控制 · 电气工程与系统科学 2021-03-31 Jiaming Xiong , Bo Li , Ruihan Yu , Daolin Ma , Wei Wang , Caishan Liu

Control barrier function (CBF)-based safety filters provide a systematic way to enforce state constraints, but they can significantly alter the closed-loop dynamics induced by a nominal, stabilizing controller. In particular, the resulting…

系统与控制 · 电气工程与系统科学 2026-04-03 Yiting Chen , Pol Mestres , Emiliano Dall'Anese , Jorge Cortés

Factors like improved data availability and increasing system complexity have sparked interest in data-driven predictive control (DDPC) methods like Data-enabled Predictive Control (DeePC). However, closed-loop identification bias arises in…

系统与控制 · 电气工程与系统科学 2024-02-23 Rogier Dinkla , Sebastiaan Mulders , Tom Oomen , Jan-Willem van Wingerden

We demonstrate that time-delayed feedback control can be improved by adaptively tuning the feedback gain. This adaptive controller is applied to the stabilization of an unstable fixed point and an unstable periodic orbit embedded in a…

适应与自组织系统 · 物理学 2016-08-10 Judith Lehnert , Philipp Hövel , Valentin Flunkert , Peter Yu. Guzenko , Alexander L. Fradkov , Eckehard Schöll

This paper discusses the systematic design of an adaptive feedback linearizing neurocontroller for a high-order model of the synchronous machine/infinite bus power system. The power system is first modelled as an input-output nonlinear…

最优化与控制 · 数学 2007-05-23 Kingsley Fregene , Diane Kennedy

Motion planning for autonomous vehicles requires spatio-temporal motion plans (i.e. state trajectories) to account for dynamic obstacles. This requires a trajectory tracking control process which faithfully tracks planned trajectories. In…

机器人学 · 计算机科学 2018-11-13 Peng Liu , Brian Paden , Umit Ozguner

Developing a universal and versatile embodied intelligence system presents two primary challenges: the critical embodied data bottleneck, where real-world data is scarce and expensive, and the algorithmic inefficiency of existing methods,…

The convergence of policy gradient algorithms in reinforcement learning hinges on the optimization landscape of the underlying optimal control problem. Theoretical insights into these algorithms can often be acquired from analyzing those of…

机器学习 · 计算机科学 2023-11-01 Jingliang Duan , Wenhan Cao , Yang Zheng , Lin Zhao

This paper establishes new sufficient conditions for Mittag-Leffler stability of Caputo fractional-order nonlinear systems with state-dependent delays. The central analytical tool is a class of Lyapunov-Krasovskii functionals that…

动力系统 · 数学 2026-02-10 Abdallah Alsammani , Gassan Farah