English
Related papers

Related papers: An Adaptive Data-Enabled Policy Optimization Appro…

200 papers

The paper presents a technique using reinforcement learning (RL) to adapt the control gains of a quadcopter controller. Specifically, we employed Proximal Policy Optimization (PPO) to train a policy which adapts the gains of a cascaded…

Systems and Control · Electrical Eng. & Systems 2024-03-13 Mike Timmerman , Aryan Patel , Tim Reinhart

This paper focuses on a novel feedback linearization control (FLC) law based on a self-learning disturbance observer (SLDO) to counteract mismatched uncertainties. The FLC based on BNDO (FLC-BNDO) demonstrates robust control performance…

Systems and Control · Electrical Eng. & Systems 2021-03-23 Erkan Kayacan , Thor I. Fossen

Many real-world systems are governed by the time-dependent, nonlinear differential equations. Dynamics of an electrical system are also best described using the very equations. Being one of the preferred machines when using advanced control…

Systems and Control · Computer Science 2017-12-05 Srikanth Peetha , Michael L. McIntyre

Model-based policy optimization often struggles with inaccurate system dynamics models, leading to suboptimal closed-loop performance. This challenge is especially evident in Model Predictive Control (MPC) policies, which rely on the model…

Systems and Control · Electrical Eng. & Systems 2026-04-21 Riccardo Zuliani , Efe C. Balta , John Lygeros

This paper proposes a data-driven framework to solve time-varying optimization problems associated with unknown linear dynamical systems. Making online control decisions to regulate a dynamical system to the solution of an optimization…

Optimization and Control · Mathematics 2021-09-08 Gianluca Bianchin , Miguel Vaquero , Jorge Cortes , Emiliano Dall'Anese

This work studies data-driven switched controller design for discrete-time switched linear systems. Instead of having access to the full system dynamics, an initialization phase is performed, during which noiseless measurements of the state…

Optimization and Control · Mathematics 2022-09-13 Jaap Eising , Shenyu Liu , Sonia Martinez , Jorge Cortes

The alignment of language models~(LMs) with human preferences is critical for building reliable AI systems. The problem is typically framed as optimizing an LM policy to maximize the expected reward that reflects human preferences.…

Artificial Intelligence · Computer Science 2026-01-28 Zetian Sun , Dongfang Li , Xuhui Chen , Baotian Hu , Min Zhang

Federated Learning (FL) is a distributed machine learning (ML) paradigm, aiming to train a global model by exploiting the decentralized data across millions of edge devices. Compared with centralized learning, FL preserves the clients'…

Distributed, Parallel, and Cluster Computing · Computer Science 2023-08-15 Bocheng Chen , Nikolay Ivanov , Guangjing Wang , Qiben Yan

Data Enabled Predictive Control (DeePC) is an established model free approach to predictive control, but it faces two open challenges: computational complexity that scales cubically with dataset size and performance degradation when data…

Systems and Control · Electrical Eng. & Systems 2026-03-25 Jiachen Li , Shihao Li , Jian Chu , Dongmei Chen

Optimal state-feedback controllers, capable of changing between different objective functions, are advantageous to systems in which unexpected situations may arise. However, synthesising such controllers, even for a single objective, is a…

Systems and Control · Computer Science 2020-10-13 Christopher Iliffe Sprague , Dario Izzo , Petter Ögren

On-policy reinforcement learning (RL) algorithms are widely used for their strong asymptotic performance and training stability, but they struggle to scale with larger batch sizes, as additional parallel environments yield redundant data…

Machine Learning · Computer Science 2025-11-13 Jianren Wang , Yifan Su , Abhinav Gupta , Deepak Pathak

Networked control strategies based on limited information about the plant model usually results in worse closed-loop performance than optimal centralized control with full plant model information. Recently, this fact has been established by…

Optimization and Control · Mathematics 2014-07-23 Farhad Farokhi , Karl H. Johansson

Reliable control and state estimation of differential drive robots (DDR) operating in dynamic and uncertain environments remains a challenge, particularly when system dynamics are partially unknown and sensor measurements are prone to…

Systems and Control · Electrical Eng. & Systems 2026-03-17 Amos Alwala , Yuchen Hu , Gabriel da Silva Lima , Wallace Moreira Bessa

In-phase synchronization is a special case of synchronous behavior when coupled oscillators have the same phases for any time moments. Such behavior appears naturally for nearly identical coupled limit-cycle oscillators when the coupling…

Adaptation and Self-Organizing Systems · Physics 2019-09-24 Viktor Novičenko , Irmantas Ratas

Direct Preference Optimization (DPO) has emerged as a more computationally efficient alternative to Reinforcement Learning from Human Feedback (RLHF) with Proximal Policy Optimization (PPO), eliminating the need for reward models and online…

Computation and Language · Computer Science 2024-10-28 Xin Mao , Feng-Lin Li , Huimin Xu , Wei Zhang , Wang Chen , Anh Tuan Luu

Aligning large language models (LLMs) with human preferences in federated learning (FL) is challenging due to decentralized, privacy-sensitive, and highly non-IID preference data. Direct Preference Optimization (DPO) offers an efficient…

Machine Learning · Computer Science 2026-03-23 Kewen Zhu , Liping Yi , Zhiming Zhao , Zhuang Qi , Han Yu , Qinghua Hu

Modular reconfigurable robots suit task-specific space operations, but the combinatorial growth of morphologies hinders unified control. We propose a decentralized reinforcement learning (Dec-RL) scheme where each module learns its own…

In this paper, we propose an Expectation-Maximization-based (EM) Personalized Federated Learning (PFL) framework for multi-objective optimization (MOO) in Integrated Sensing and Communication (ISAC) systems. In contrast to standard…

Signal Processing · Electrical Eng. & Systems 2025-10-09 Zhou Ni , Sravan Reddy Chintareddy , Peiyuan Guan , Morteza Hashemi

This paper is concerned with the data-driven stabilization of unknown boundary controlled semilinear parabolic systems. The nonlinear dynamics of the system are lifted using a finite number of eigenfunctionals of the Koopman operator…

Systems and Control · Electrical Eng. & Systems 2025-04-17 Joachim Deutscher , Tarik Enderes

We propose KFCPO, a novel Safe Reinforcement Learning (Safe RL) algorithm that combines scalable Kronecker-Factored Approximate Curvature (K-FAC) based second-order policy optimization with safety-aware gradient manipulation. KFCPO…

Machine Learning · Computer Science 2025-11-04 Joonyoung Lim , Younghwan Yoo
‹ Prev 1 8 9 10 Next ›