English
Related papers

Related papers: An Adaptive Data-Enabled Policy Optimization Appro…

200 papers

PID control architectures are widely used in industrial applications. Despite their low number of open parameters, tuning multiple, coupled PID controllers can become tedious in practice. In this paper, we extend PILCO, a model-based policy…

Machine Learning · Computer Science 2017-03-09 Andreas Doerr , Duy Nguyen-Tuong , Alonso Marco , Stefan Schaal , Sebastian Trimpe

This paper focuses on adaptive control of the discrete-time linear quadratic regulator (adaptive LQR). Recent literature has made significant contributions in proving non-asymptotic convergence rates, but existing approaches have a few…

Systems and Control · Electrical Eng. & Systems 2026-04-27 Peter A. Fisher , Anuradha M. Annaswamy

Delays endanger safety of autonomous systems operating in a rapidly changing environment, such as nondeterministic surrounding traffic participants in autonomous driving and high-speed racing. Unfortunately, delays are typically not…

Robotics · Computer Science 2022-08-31 Dvij Kalaria , Qin Lin , John M. Dolan

Reinforcement learning (RL) has become central to enhancing reasoning in large language models (LLMs). Yet on-policy algorithms such as Group Relative Policy Optimization (GRPO) often suffer in early training: noisy gradients from…

Machine Learning · Computer Science 2026-03-19 Ziyan Wang , Zheng Wang , Xingwei Qu , Qi Cheng , Jie Fu , Shengpu Tang , Minjia Zhang , Xiaoming Huo

We design dynamic routing policies for an overlay network which meet delay requirements of real-time traffic being served on top of an underlying legacy network, where the overlay nodes do not know the underlay characteristics. We pose the…

Networking and Internet Architecture · Computer Science 2019-04-19 Rahul Singh , Eytan Modiano

In many control systems, tracking accuracy can be enhanced by combining (data-driven) feedforward (FF) control with feedback (FB) control. However, designing effective data-driven FF controllers typically requires large amounts of…

Machine Learning · Computer Science 2026-03-25 Jakob Weber , Markus Gurtner , Benedikt Alt , Adrian Trachte , Andreas Kugi

Direct Preference Optimization (DPO) demonstrates the advantage of aligning a large language model with human preference using only an offline dataset. However, DPO has the limitation that the KL penalty, which prevents excessive deviation…

Machine Learning · Computer Science 2025-10-28 Sangkyu Lee , Janghoon Han , Hosung Song , Stanley Jungkyu Choi , Honglak Lee , Youngjae Yu

In this paper, an adaptive controller is designed for the synchronization of the trajectory of a robot with unknown kinematics and dynamics to that of the current human trajectory in the task space using the delayed human trajectory…

Robotics · Computer Science 2025-09-30 Rounak Bhattacharya , Vrithik R. Guthikonda , Ashwin P. Dani

Group Relative Policy Optimization (GRPO) effectively scales LLM reasoning but incurs prohibitive computational costs due to its extensive group-based sampling requirement. While recent selective data utilization methods can mitigate this…

Machine Learning · Computer Science 2026-03-05 Haodong Zhu , Yangyang Ren , Yanjing Li , Mingbao Lin , Linlin Yang , Xuhui Liu , Xiantong Zhen , Haiguang Liu , Baochang Zhang

Data-enabled predictive control (DeePC) has emerged as a powerful technique to control complex systems without the need for extensive modeling efforts. However, relying solely on offline collected data trajectories to represent the system…

Systems and Control · Electrical Eng. & Systems 2025-08-06 Sebastian Zieglmeier , Mathias Hudoba de Badyn , Narada D. Warakagoda , Thomas R. Krogstad , Paal Engelstad

We propose a novel flexible-step model predictive control algorithm for unknown linear time-invariant discrete-time systems. The goal is to asymptotically stabilize the system without relying on a pre-collected dataset that describes its…

Optimization and Control · Mathematics 2025-10-02 Markus Pietschner , Christian Ebenbauer , Bahman Gharesifard , Raik Suttner

In this paper, we propose the reduced model for the full dynamics of a bicycle and analyze its nonlinear behavior under a proportional control law for steering. Based on the Gibbs-Appell equations for the Whipple bicycle, we obtain a…

Systems and Control · Electrical Eng. & Systems 2021-03-31 Jiaming Xiong , Bo Li , Ruihan Yu , Daolin Ma , Wei Wang , Caishan Liu

Control barrier function (CBF)-based safety filters provide a systematic way to enforce state constraints, but they can significantly alter the closed-loop dynamics induced by a nominal, stabilizing controller. In particular, the resulting…

Systems and Control · Electrical Eng. & Systems 2026-04-03 Yiting Chen , Pol Mestres , Emiliano Dall'Anese , Jorge Cortés

Factors like improved data availability and increasing system complexity have sparked interest in data-driven predictive control (DDPC) methods like Data-enabled Predictive Control (DeePC). However, closed-loop identification bias arises in…

Systems and Control · Electrical Eng. & Systems 2024-02-23 Rogier Dinkla , Sebastiaan Mulders , Tom Oomen , Jan-Willem van Wingerden

We demonstrate that time-delayed feedback control can be improved by adaptively tuning the feedback gain. This adaptive controller is applied to the stabilization of an unstable fixed point and an unstable periodic orbit embedded in a…

Adaptation and Self-Organizing Systems · Physics 2016-08-10 Judith Lehnert , Philipp Hövel , Valentin Flunkert , Peter Yu. Guzenko , Alexander L. Fradkov , Eckehard Schöll

This paper discusses the systematic design of an adaptive feedback linearizing neurocontroller for a high-order model of the synchronous machine/infinite bus power system. The power system is first modelled as an input-output nonlinear…

Optimization and Control · Mathematics 2007-05-23 Kingsley Fregene , Diane Kennedy

Motion planning for autonomous vehicles requires spatio-temporal motion plans (i.e. state trajectories) to account for dynamic obstacles. This requires a trajectory tracking control process which faithfully tracks planned trajectories. In…

Robotics · Computer Science 2018-11-13 Peng Liu , Brian Paden , Umit Ozguner

Developing a universal and versatile embodied intelligence system presents two primary challenges: the critical embodied data bottleneck, where real-world data is scarce and expensive, and the algorithmic inefficiency of existing methods,…

The convergence of policy gradient algorithms in reinforcement learning hinges on the optimization landscape of the underlying optimal control problem. Theoretical insights into these algorithms can often be acquired from analyzing those of…

Machine Learning · Computer Science 2023-11-01 Jingliang Duan , Wenhan Cao , Yang Zheng , Lin Zhao

This paper establishes new sufficient conditions for Mittag-Leffler stability of Caputo fractional-order nonlinear systems with state-dependent delays. The central analytical tool is a class of Lyapunov-Krasovskii functionals that…

Dynamical Systems · Mathematics 2026-02-10 Abdallah Alsammani , Gassan Farah