中文
相关论文

相关论文: Function Approximation for Reinforcement Learning …

200 篇论文

This paper proposes a deep learning-based optimal battery management scheme for frequency regulation (FR) by integrating model predictive control (MPC), supervised learning (SL), reinforcement learning (RL), and high-fidelity battery…

系统与控制 · 电气工程与系统科学 2022-01-05 Yun Li , Yixiu Wang , Yifu Chen , Kaixun Hua , Jiayang Ren , Ghazaleh Mozafari , Qiugang Lu , Yankai Cao

This paper presents the network load balancing problem, a challenging real-world task for multi-agent reinforcement learning (MARL) methods. Traditional heuristic solutions like Weighted-Cost Multi-Path (WCMP) and Local Shortest Queue (LSQ)…

分布式、并行与集群计算 · 计算机科学 2022-08-23 Zhiyuan Yao , Zihan Ding , Thomas Clausen

We develop a novel multi-objective reinforcement learning (MORL) framework to jointly optimize wireless network selection and autonomous driving policies in a multi-band vehicular network operating on conventional sub-6GHz spectrum and…

机器学习 · 计算机科学 2025-06-17 Zijiang Yan , Hina Tabassum

With the exponential growth of smart devices connected to wireless networks, data production is increasing rapidly, requiring machine learning (ML) techniques to unlock its value. However, the centralized ML paradigm raises concerns over…

分布式、并行与集群计算 · 计算机科学 2025-07-15 Xiangwang Hou , Jingjing Wang , Jun Du , Chunxiao Jiang , Yong Ren , Dusit Niyato

As power systems are undergoing a significant transformation with more uncertainties, less inertia and closer to operation limits, there is increasing risk of large outages. Thus, there is an imperative need to enhance grid emergency…

机器学习 · 计算机科学 2022-02-08 Renke Huang , Yujiao Chen , Tianzhixi Yin , Qiuhua Huang , Jie Tan , Wenhao Yu , Xinya Li , Ang Li , Yan Du

Load frequency control (LFC) is a key factor to maintain the stable frequency in multi-area power systems. As the modern power systems evolve from centralized to distributed paradigm, LFC needs to consider the peer-to-peer (P2P) based…

最优化与控制 · 数学 2022-09-27 Kyung-bin Kwon , Sayak Mukherjee , Hao Zhu , Thanh Long Vu

The transformer architecture and variants presented remarkable success across many machine learning tasks in recent years. This success is intrinsically related to the capability of handling long sequences and the presence of…

机器学习 · 计算机科学 2022-06-15 Luckeciano C. Melo

The stringent requirements of mobile edge computing (MEC) applications and functions fathom the high capacity and dense deployment of MEC hosts to the upcoming wireless networks. However, operating such high capacity MEC hosts can…

机器学习 · 计算机科学 2021-02-11 Md. Shirajum Munir , Nguyen H. Tran , Walid Saad , Choong Seon Hong

Restoring power distribution systems (PDSs) after large-scale outages requires sequential switching actions that reconfigure feeder topology and coordinate distributed energy resources (DERs) under nonlinear constraints, including power…

人工智能 · 计算机科学 2026-02-03 Parya Dolatyabi , Ali Farajzadeh Bavil , Mahdi Khodayar

As a main use case of 5G and Beyond wireless network, the ever-increasing machine type communications (MTC) devices pose critical challenges over MTC network in recent years. It is imperative to support massive MTC devices with limited…

网络与互联网体系结构 · 计算机科学 2022-02-23 Ziru Chen , Ran Zhang , Lin X. Cai , Yu Cheng , Yong Liu

This paper studies a distributed policy gradient in collaborative multi-agent reinforcement learning (MARL), where agents over a communication network aim to find the optimal policy to maximize the average of all agents' local returns. Due…

多智能体系统 · 计算机科学 2022-12-06 Xiaoxiao Zhao , Jinlong Lei , Li Li , Jie Chen

This paper develops an efficient multi-agent deep reinforcement learning algorithm for cooperative controls in powergrids. Specifically, we consider the decentralized inverter-based secondary voltage control problem in distributed…

系统与控制 · 电气工程与系统科学 2021-08-03 Dong Chen , Kaian Chen. Zhaojian Li , Tianshu Chu , Rui Yao , Feng Qiu , Kaixiang Lin

A major challenge of reinforcement learning (RL) in real-world applications is the variation between environments, tasks or clients. Meta-RL (MRL) addresses this issue by learning a meta-policy that adapts to new tasks. Standard MRL methods…

机器学习 · 计算机科学 2023-10-03 Ido Greenberg , Shie Mannor , Gal Chechik , Eli Meirom

Power control in decentralized wireless networks poses a complex stochastic optimization problem when formulated as the maximization of the average sum rate for arbitrary interference graphs. Recent work has introduced data-driven design…

信息论 · 计算机科学 2021-05-04 Ivana Nikoloska , Osvaldo Simeone

The electric grid is undergoing a major transition from fossil fuel-based power generation to renewable energy sources, typically interfaced to the grid via power electronics. The future power systems are thus expected to face increased…

系统与控制 · 电气工程与系统科学 2020-07-13 Ognjen Stanojev , Ognjen Kundacina , Uros Markovic , Evangelos Vrettos , Petros Aristidou , Gabriela Hug

Despite rapid advancements in sensor networks, conventional battery-powered sensor networks suffer from limited operational lifespans and frequent maintenance requirements that severely constrain their deployment in remote and inaccessible…

网络与互联网体系结构 · 计算机科学 2025-10-27 Bowei Tong , Hui Kang , Jiahui Li , Geng Sun , Jiacheng Wang , Yaoqi Yang , Bo Xu , Dusit Niyato

This paper demonstrates that continual relearning of control policies using incremental deep reinforcement learning (RL) can improve policy learning for non-stationary processes. We demonstrate this approach for a data-driven 'smart…

机器学习 · 计算机科学 2020-08-06 Avisek Naug , Marcos Quiñones-Grueiro , Gautam Biswas

The Diffusion Transformer plays a pivotal role in advancing text-to-image and text-to-video generation, owing primarily to its inherent scalability. However, existing controlled diffusion transformer methods incur significant parameter and…

计算机视觉与模式识别 · 计算机科学 2026-02-27 Ke Cao , Jing Wang , Ao Ma , Jiasong Feng , Xuanhua He , Run Ling , Haowei Liu , Jian Lu , Wei Feng , Haozhe Wang , Hongjuan Pei , Yihua Shao , Zhanjie Zhang , Jie Zhang

With large-scale integration of renewable generation and distributed energy resources, modern power systems are confronted with new operational challenges, such as growing complexity, increasing uncertainty, and aggravating volatility.…

机器学习 · 计算机科学 2022-02-28 Xin Chen , Guannan Qu , Yujie Tang , Steven Low , Na Li

Judicious resource allocation can effectively enhance federated learning (FL) training performance in wireless networks by addressing both system and statistical heterogeneity. However, existing strategies typically rely on block fading…

机器学习 · 计算机科学 2025-05-07 Jiacheng Wang , Le Liang , Hao Ye , Chongtao Guo , Shi Jin