English
Related papers

Related papers: Proximal Policy Optimization Learning based Contro…

200 papers

Reinforcement learning (RL) holds significant promise for adaptive traffic signal control. While existing RL-based methods demonstrate effectiveness in reducing vehicular congestion, their predominant focus on vehicle-centric optimization…

Machine Learning · Computer Science 2025-07-24 Bibek Poudel , Xuan Wang , Weizi Li , Lei Zhu , Kevin Heaslip

A reliable controller is critical for execution of safe and smooth maneuvers of an autonomous vehicle. The controller must be robust to external disturbances, such as road surface, weather, wind conditions, and so on. It also needs to deal…

Systems and Control · Electrical Eng. & Systems 2024-09-23 Pin Wang , Tianyu Shi , Chonghao Zou , Long Xin , Ching-Yao Chan

In recent years, autonomous driving has become a popular field of study. As control at tire grip limit is essential during emergency situations, algorithms developed for racecars are useful for road cars too. This paper examines the use of…

Robotics · Computer Science 2025-04-15 Gergely Bári , László Palkovics

Efficient mobility management and load balancing are critical to sustaining Quality of Service (QoS) in dense, highly dynamic 5G radio access networks. We present a deep reinforcement learning framework based on Proximal Policy Optimization…

Networking and Internet Architecture · Computer Science 2026-05-13 Mehrshad Eskandarpour , Hossein Soleimani

Model-free and reinforcement learning-based adaptive filtering methods are gaining traction for denoising in dynamic, non-stationary environments such as wireless signal channels. Traditional filters like LMS, RLS, Wiener, and Kalman are…

Signal Processing · Electrical Eng. & Systems 2025-06-10 Abdullah Burkan Bereketoglu

Contemporary autopilot systems for unmanned aerial vehicles (UAVs) are far more limited in their flight envelope as compared to experienced human pilots, thereby restricting the conditions UAVs can operate in and the types of missions they…

Robotics · Computer Science 2019-11-14 Eivind Bøhn , Erlend M. Coates , Signe Moe , Tor Arne Johansen

We present a reinforcement learning method for training neuro-fuzzy controllers using Proximal Policy Optimization (PPO). Unlike prior approaches that used Deep Q-Networks (DQN) with Adaptive Neuro-Fuzzy Inference Systems (ANFIS), our…

Machine Learning · Computer Science 2025-07-08 Kaaustaaub Shankar , Wilhelm Louw , Kelly Cohen

The key challenges in design of predictor-based control laws for switched systems with arbitrary switching and long input delay are the potential unavailability of the future values of the switching signal (at current time) and the fact…

Systems and Control · Electrical Eng. & Systems 2025-03-20 Andreas Katsanikakis , Nikolaos Bekiaris-Liberis

This paper proposes modifications to the data-enabled policy optimization (DeePO) algorithm to mitigate state perturbations. DeePO is an adaptive, data-driven approach designed to iteratively compute a feedback gain equivalent to the…

Systems and Control · Electrical Eng. & Systems 2025-07-29 Mojtaba Kaheni , Niklas Persson , Vittorio De Iuliis , Costanzo Manes , Alessandro V. Papadopoulos

In this work, we propose a rigorous method for implementing predictor feedback controllers in nonlinear systems with unknown and arbitrarily long actuator delays. To address the analytically intractable nature of the predictor, we…

Systems and Control · Electrical Eng. & Systems 2025-08-29 Luke Bhan , Miroslav Krstic , Yuanyuan Shi

Previous studies on automatic berthing systems based on artificial neural network (ANN) showed great berthing performance by training the ANN with ship berthing data as training data. However, because the ANN requires a large amount of…

Machine Learning · Computer Science 2021-12-06 Daesoo Lee

The optimal information feedback has a significant effect on many socioeconomic systems like stock market and traffic systems aiming to make full use of resources. In this paper, we studied dynamics of traffic flow with real-time…

Data Analysis, Statistics and Probability · Physics 2009-09-29 Dong Chuan-Fei , Ma Xu , Wang Guan-Wen , Sun Xiao-Yan , Wang Bing-Hong

This paper addresses the challenge of edge caching in dynamic environments, where rising traffic loads strain backhaul links and core networks. We propose a Proximal Policy Optimization (PPO)-based caching strategy that fully incorporates…

Networking and Internet Architecture · Computer Science 2024-11-18 Farnaz Niknia , Ping Wang

In traffic signal control, flow-based (optimizing the overall flow) and pressure-based methods (equalizing and alleviating congestion) are commonly used but often considered separately. This study introduces a unified framework using…

Systems and Control · Electrical Eng. & Systems 2024-01-18 Chaolun Ma , Bruce Wang , Zihao Li , Ahmadreza Mahmoudzadeh , Yunlong Zhang

Proximal policy optimization (PPO) has yielded state-of-the-art results in policy search, a subfield of reinforcement learning, with one of its key points being the use of a surrogate objective function to restrict the step size at each…

Machine Learning · Computer Science 2020-12-07 Wangshu Zhu , Andre Rosendo

Proximal Policy Optimization (PPO) is a highly popular model-free reinforcement learning (RL) approach. However, we observe that in a continuous action space, PPO can prematurely shrink the exploration variance, which leads to slow progress…

Machine Learning · Computer Science 2020-11-04 Perttu Hämäläinen , Amin Babadi , Xiaoxiao Ma , Jaakko Lehtinen

A reliable controller is critical and essential for the execution of safe and smooth maneuvers of an autonomous vehicle.The controller must be robust to external disturbances, such as road surface, weather, and wind conditions, and so on.It…

Robotics · Computer Science 2019-05-01 Tianyu Shi , Pin Wang , Ching-Yao Chan , Chonghao Zou

On-policy deep reinforcement learning algorithms have low data utilization and require significant experience for policy improvement. This paper proposes a proximal policy optimization algorithm with prioritized trajectory replay (PTR-PPO)…

Machine Learning · Computer Science 2021-12-09 Xingxing Liang , Yang Ma , Yanghe Feng , Zhong Liu

This paper deals with traffic control at motorway bottlenecks assuming the existence of an unknown, time-varying, Fundamental Diagram (FD). The FD may change over time due to different traffic compositions, e.g., light and heavy vehicles,…

Systems and Control · Electrical Eng. & Systems 2023-08-02 Farzam Tajdari , Claudio Roncoli

Effective frequency control in power grids has become increasingly important with the increasing demand for renewable energy sources. Here, we propose a novel strategy for resolving this challenge using graph convolutional proximal policy…