中文
相关论文

相关论文: A Model-Based Reinforcement Learning Approach for …

200 篇论文

Lagrangian methods are widely used algorithms for constrained optimization problems, but their learning dynamics exhibit oscillations and overshoot which, when applied to safe reinforcement learning, leads to constraint-violating behavior…

最优化与控制 · 数学 2020-07-09 Adam Stooke , Joshua Achiam , Pieter Abbeel

Safety is essential for reinforcement learning (RL) applied in the real world. Adding chance constraints (or probabilistic constraints) is a suitable way to enhance RL safety under uncertainty. Existing chance-constrained RL methods like…

机器学习 · 计算机科学 2021-08-27 Baiyu Peng , Jingliang Duan , Jianyu Chen , Shengbo Eben Li , Genjin Xie , Congsheng Zhang , Yang Guan , Yao Mu , Enxin Sun

"Model-free control" and the corresponding "intelligent" PID controllers (iPIDs), which already had many successful concrete applications, are presented here for the first time in an unified manner, where the new advances are taken into…

最优化与控制 · 数学 2013-11-21 Michel Fliess , Cédric Join

The proportional-integral-derivative (PID) controller and its variants are widely used in control engineering, but they often rely on linearization around equilibrium points and empirical parameter tuning, making them ineffective for…

系统与控制 · 电气工程与系统科学 2026-01-13 Zimao Sheng

In control applications there is often a compromise that needs to be made with regards to the complexity and performance of the controller and the computational resources that are available. For instance, the typical hardware platform in…

系统与控制 · 电气工程与系统科学 2020-11-30 Eivind Bøhn , Sebastien Gros , Signe Moe , Tor Arne Johansen

Decision transformer based sequential policies have emerged as a powerful paradigm in offline reinforcement learning (RL), yet their efficacy remains constrained by the quality of static datasets and inherent architectural limitations.…

机器学习 · 计算机科学 2026-03-05 Yihao Qin , Yuanfei Wang , Hang Zhou , Peiran Liu , Hao Dong , Yiding Ji

Classical PID control is widely applied in an engineering system, with parameter regulation relying on a method like Trial - Error Tuning or the Ziegler - Nichols rule, mainly for a Single - Input Single - Output (SISO) system. However, the…

系统与控制 · 电气工程与系统科学 2025-04-22 Zimao Sheng , Hong'an Yang

This paper proposes a framework for adaptively learning a feedback linearization-based tracking controller for an unknown system using discrete-time model-free policy-gradient parameter update rules. The primary advantage of the scheme over…

Reinforcement Learning (RL) has recently impressed the world with stunning results in various applications. While the potential of RL is now well-established, many critical aspects still need to be tackled, including safety and stability…

系统与控制 · 电气工程与系统科学 2024-09-23 Mario Zanon , Sébastien Gros

Offline reinforcement learning (RL) learns effective policies from pre-collected datasets, offering a practical solution for applications where online interactions are risky or costly. Model-based approaches are particularly advantageous…

机器学习 · 计算机科学 2026-05-14 Xuyang Chen , Keyu Yan , Guojian Wang , Lin Zhao

In recent years, reinforcement learning (RL) has gained increasing attention in control engineering. Especially, policy gradient methods are widely used. In this work, we improve the tracking performance of proximal policy optimization…

In this thesis, advanced design technique in sliding mode control (SMC) is presented with focus on PID (Proportional-Integral-Derivative) type Sliding surfaces based Sliding mode control with improved power rate exponential reaching law for…

系统与控制 · 电气工程与系统科学 2022-07-25 Kirtiman Singh

While originally developed for continuous control problems, Proximal Policy Optimization (PPO) has emerged as the work-horse of a variety of reinforcement learning (RL) applications, including the fine-tuning of generative models.…

In this paper, a new model based nonlinear control technique, called PID (Proportional-Integral-Derivative) type sliding surface based sliding mode control is designed using improved reaching law. To improve the performance of the second…

系统与控制 · 电气工程与系统科学 2022-09-20 Kirtiman Singh , Prabin Kumar Padhy

During a multi-speed transmission development process, the final calibration of the gearshift controller parameters is usually performed on a physical test bench. Engineers typically treat the mapping from the controller parameters to the…

系统与控制 · 电气工程与系统科学 2021-12-02 Marc-Antoine Beaudoin , Benoit Boulet

Proportional-Integrator-Derivative (PID) controller is used in a wide range of industrial and experimental processes. There are a couple of offline methods for tuning PID gains. However, due to the uncertainty of model parameters and…

系统与控制 · 电气工程与系统科学 2025-08-19 Iman Sharifi , Aria Alasty

In many practical control applications, the performance level of a closed-loop system degrades over time due to the change of plant characteristics. Thus, there is a strong need for redesigning a controller without going through the system…

系统与控制 · 电气工程与系统科学 2023-12-01 Mei Minami , Yuka Masumoto , Yoshihiro Okawa , Tomotake Sasaki , Yutaka Hori

The aim of this paper is to present a comprehensive range of design techniques for the synthesis of the standard compensators (Lead and Lag networks as well as PID controllers) that in the last twenty years have proved to be of great…

系统与控制 · 计算机科学 2012-10-16 Lorenzo Ntogramatzidis , Roberto Zanasi , Stefania Cuoghi

Reinforcement learning (RL) is used to directly design a control policy using data collected from the system. This paper considers the robustness of controllers trained via model-free RL. The discussion focuses on the standard model-based…

系统与控制 · 计算机科学 2019-04-09 Harish K. Venkataraman , Peter J. Seiler

In the era of Industry 4.0 and smart manufacturing, process systems engineering must adapt to digital transformation. While reinforcement learning offers a model-free approach to process control, its applications are limited by the…

系统与控制 · 电气工程与系统科学 2025-05-28 Runze Lin , Junghui Chen , Biao Huang , Lei Xie , Hongye Su