中文
相关论文

相关论文: Reducing Learning Difficulties: One-Step Two-Criti…

200 篇论文

Intelligent reflecting surfaces (IRSs) technology has been considered a promising solution in visible light communication (VLC) systems due to its potential to overcome the line-of-sight (LoS) blockage issue and enhance coverage. Moreover,…

信号处理 · 电气工程与系统科学 2025-04-29 Ahrar N. Hamad , Ahmad Adnan Qidan , Taisir E. H. El-Gorashi , Jaafar M. H. Elmirghani

In response to the trade-off between control performance and computational burden hindering the deployment of Deep Reinforcement Learning (DRL) in power inverters, this paper presents a novel model-free control framework leveraging policy…

系统与控制 · 电气工程与系统科学 2026-03-10 Yang Yang , Chenggang Cui , Xitong Niu , Jiaming Liu , Chuanlin Zhang

A comparative assessment of machine learning (ML) methods for active flow control is performed. The chosen benchmark problem is the drag reduction of a two-dimensional K\'arm\'an vortex street past a circular cylinder at a low Reynolds…

流体动力学 · 物理学 2022-04-25 R. Castellanos , G. Y. Cornejo Maceda , I. de la Fuente , B. R. Noack , A. Ianiro , S. Discetti

The adaptive traffic signal control (ATSC) problem can be modeled as a multiagent cooperative game among urban intersections, where intersections cooperate to optimize their common goal. Recently, reinforcement learning (RL) has achieved…

机器学习 · 计算机科学 2021-04-23 Chengwei Zhang , Shan Jin , Wanli Xue , Xiaofei Xie , Shengyong Chen , Rong Chen

This paper is a study of reinforcement learning (RL) as an optimal-control strategy for control of nonlinear valves. It is evaluated against the PID (proportional-integral-derivative) strategy, using a unified framework. RL is an autonomous…

机器学习 · 计算机科学 2021-02-05 Rajesh Siraskar

We consider the problem of training machine learning models on distributed data in a decentralized way. For finite-sum problems, fast single-machine algorithms for large datasets rely on stochastic updates combined with variance reduction.…

最优化与控制 · 数学 2020-06-26 Hadrien Hendrikx , Francis Bach , Laurent Massoulié

Transmission interface power flow adjustment is a critical measure to ensure the security and economy operation of power systems. However, conventional model-based adjustment schemes are limited by the increasing variations and…

系统与控制 · 电气工程与系统科学 2024-05-28 Shunyu Liu , Wei Luo , Yanzhen Zhou , Kaixuan Chen , Quan Zhang , Huating Xu , Qinglai Guo , Mingli Song

In this paper, we propose reconfigurable intelligent surface (RIS)-assisted unmanned aerial vehicles (UAVs) networks that can utilise both advantages of UAV's agility and RIS's reflection for enhancing the network's performance. To aim at…

信号处理 · 电气工程与系统科学 2021-08-09 Khoi Khac Nguyen , Saeed Khosravirad , Daniel Benevides da Costa , Long D. Nguyen , Trung Q. Duong

Although deep reinforcement learning (DRL) algorithms have made important achievements in many control tasks, they still suffer from the problems of sample inefficiency and unstable training process, which are usually caused by sparse…

机器人学 · 计算机科学 2020-02-28 Ke Lin , Liang Gong , Xudong Li , Te Sun , Binhao Chen , Chengliang Liu , Zhengfeng Zhang , Jian Pu , Junping Zhang

Under voltage load shedding (UVLS) for power grid emergency control builds the last defensive perimeter to prevent cascade outages and blackouts in case of contingencies. This letter proposes a novel cooperative multi-agent deep…

系统与控制 · 电气工程与系统科学 2023-10-23 Ying Zhang , Meng Yue

One-step offline RL actors are attractive because they avoid backpropagating through long iterative samplers and keep inference cheap, but they still have to improve under a critic without drifting away from actions that the dataset can…

机器学习 · 计算机科学 2026-04-27 Zhancun Mu , Guangyu Zhao , Yiwu Zhong , Chi Zhang

Preference-based reinforcement learning (PBRL) in the offline setting has succeeded greatly in industrial applications such as chatbots. A two-step learning framework where one applies a reinforcement learning step after a reward modeling…

Operational constraint violations may occur when deep reinforcement learning (DRL) agents interact with real-world active distribution systems to learn their optimal policies during training. This letter presents a universal…

系统与控制 · 电气工程与系统科学 2023-08-22 Hoang Tien Nguyen , Dae-Hyun Choi

The increasing penetration of renewable energy resources in distribution systems necessitates high-speed monitoring and control of voltage for ensuring reliable system operation. However, existing voltage control algorithms often make…

系统与控制 · 电气工程与系统科学 2024-10-03 Mohammad Golgol , Anamitra Pal

The stability and convergence of an Iterative Learning Controller (ILC) may be assessed either by directly iterating the equations for a variety of inputs, or by finding the eigenvalues of the iterated system, or by forming the Z-transform…

系统与控制 · 电气工程与系统科学 2022-11-22 Shane Rupert Koscielniak

We introduce Iterative Dual Reinforcement Learning (IDRL), a new method that takes an optimal discriminator-weighted imitation view of solving RL. Our method is motivated by a simple experiment in which we find training a discriminator…

机器学习 · 计算机科学 2025-04-21 Haoran Xu , Shuozhe Li , Harshit Sikchi , Scott Niekum , Amy Zhang

Network Slice placement with the problem of allocation of resources from a virtualized substrate network is an optimization problem which can be formulated as a multiobjective Integer Linear Programming (ILP) problem. However, to cope with…

网络与互联网体系结构 · 计算机科学 2021-05-17 Jose Jurandir Alves Esteves , Amina Boubendir , Fabrice Guillemin , Pierre Sens

Visible light communication (VLC) has been widely applied as a promising solution for modern short range communication. When it comes to the deployment of LED arrays in VLC networks, the emerging ultra-dense network (UDN) technology can be…

信号处理 · 电气工程与系统科学 2023-03-10 Xiao Tang , Sicong Liu

The rapid advancement of electric vertical takeoff and landing (eVTOL) aircraft offers a promising opportunity to alleviate urban traffic congestion but is still limited by excessive power demands, especially during the takeoff phase. Thus,…

机器学习 · 计算机科学 2026-05-06 Nathan M. Roberts , Xiaosong Du

Dynamic distribution network reconfiguration (DNR) algorithms perform hourly status changes of remotely controllable switches to improve distribution system performance. The problem is typically solved by physical model-based control…

系统与控制 · 电气工程与系统科学 2020-06-24 Yuanqi Gao , Wei Wang , Jie Shi , Nanpeng Yu