中文
相关论文

相关论文: A Hierarchical DRL Approach for Resource Optimizat…

200 篇论文

Deep reinforcement learning (DRL) is capable of learning high-performing policies on a variety of complex high-dimensional tasks, ranging from video games to robotic manipulation. However, standard DRL methods often suffer from poor sample…

机器学习 · 计算机科学 2020-03-04 Caleb Chuck , Supawit Chockchowwat , Scott Niekum

Hierarchical Reinforcement Learning (HRL) exploits temporally extended actions, or options, to make decisions from a higher-dimensional perspective to alleviate the sparse reward problem, one of the most challenging problems in…

机器学习 · 计算机科学 2019-05-15 Libo Xing

Hierarchical Reinforcement Learning (HRL) has made notable progress in complex control tasks by leveraging temporal abstraction. However, previous HRL algorithms often suffer from serious data inefficiency as environments get large. The…

机器学习 · 计算机科学 2022-11-22 Seungjae Lee , Jigang Kim , Inkyu Jang , H. Jin Kim

Decision-making in military aviation Prognostics and Health Management (PHM) faces significant challenges due to the "curse of dimensionality" in large-scale fleet operations, combined with sparse feedback and stochastic mission profiles.…

机器学习 · 计算机科学 2026-04-09 Yong Si , Mingfei Lu , Jing Li , Yang Hu , Guijiang Li , Yueheng Song , Zhaokui Wang

Coverage and capacity are the important metrics for performance evaluation in wireless networks, while the coverage and capacity have several conflicting relationships, e.g. high transmit power contributes to large coverage but high…

信息论 · 计算机科学 2022-04-14 Xinyu Gao , Wenqiang Yi , Alexandros Agapitos , Hao Wang , Yuanwei Liu

Owing to the unique advantages of low cost and controllability, reconfigurable intelligent surface (RIS) is a promising candidate to address the blockage issue in millimeter wave (mmWave) communication systems, consequently has captured…

信息论 · 计算机科学 2022-02-24 Yuqian Zhu , Zhu Bo , Ming Li , Yang Liu , Qian Liu , Zheng Chang , Yulin Hu

Multi-Agent Proximal Policy Optimization (MAPPO) is a variant of the Proximal Policy Optimization (PPO) algorithm, specifically tailored for multi-agent reinforcement learning (MARL). MAPPO optimizes cooperative multi-agent settings by…

机器学习 · 计算机科学 2026-05-14 Changha Lee , Gyusang Cho

Retrieval-Augmented Generation (RAG) has revolutionized natural language processing by dynamically integrating external knowledge into Large Language Models (LLMs), addressing their limitation of static training datasets. Recent…

计算与语言 · 计算机科学 2024-09-05 Krish Goel , Mahek Chandak

In order to improve reproducibility, deep reinforcement learning (RL) has been adopting better scientific practices such as standardized evaluation metrics and reporting. However, the process of hyperparameter optimization still varies…

机器学习 · 计算机科学 2023-06-05 Theresa Eimer , Marius Lindauer , Roberta Raileanu

The emergent technology of Reconfigurable Intelligent Surfaces (RISs) has the potential to transform wireless environments into controllable systems, through programmable propagation of information-bearing signals. Techniques stemming from…

Indoor multi-robot communications face two key challenges: one is the severe signal strength degradation caused by blockages (e.g., walls) and the other is the dynamic environment caused by robot mobility. To address these issues, we…

机器人学 · 计算机科学 2022-07-19 Ruyu Luo , Wanli Ni , Hui Tian , Julian Cheng

Resource allocation and task prioritisation are key problem domains in the fields of autonomous vehicles, networking, and cloud computing. The challenge in developing efficient and robust algorithms comes from the dynamic nature of these…

人工智能 · 计算机科学 2021-02-17 Niall Creech , Natalia Criado Pacheco , Simon Miles

Deep reinforcement learning (DRL) is one of the promising approaches for introducing robots into complicated environments. The recent remarkable progress of DRL stands on regularization of policy, which allows the policy to improve stably…

机器学习 · 计算机科学 2023-07-04 Taisuke Kobayashi

The Integrated Process Planning and Scheduling (IPPS) problem combines process route planning and shop scheduling to achieve high efficiency in manufacturing and maximize resource utilization, which is crucial for modern manufacturing…

最优化与控制 · 数学 2024-09-04 Hongpei Li , Han Zhang , Ziyan He , Yunkai Jia , Bo Jiang , Xiang Huang , Dongdong Ge

Taking advantage of their data-driven and model-free features, Deep Reinforcement Learning (DRL) algorithms have the potential to deal with the increasing level of uncertainty due to the introduction of renewable-based generation. To deal…

系统与控制 · 电气工程与系统科学 2022-08-02 Hou Shengren , Edgar Mauricio Salazar , Pedro P. Vergara , Peter Palensky

Managing disruptions in railway traffic management is a major challenge. Rising traffic density and infrastructure limits increase complexity, making the Vehicle Routing and Scheduling Problem (VRSP) difficult to solve reliably and in real…

Due to the development of communication technology and the rise of user network demand, a reasonable resource allocation for wireless networks is the key to guaranteeing regular operation and improving system performance. Various frequency…

信息论 · 计算机科学 2023-08-21 Yuanyuan Qiao , Yong Niu , Zhu Han , Shiwen Mao , Ruisi He , Ning Wang , Zhangdui Zhong , Bo Ai

As an essential resource management problem in network virtualization, virtual network embedding (VNE) aims to allocate the finite resources of physical network to sequentially arriving virtual network requests (VNRs) with different…

网络与互联网体系结构 · 计算机科学 2024-06-26 Tianfu Wang , Li Shen , Qilin Fan , Tong Xu , Tongliang Liu , Hui Xiong

Instability and slowness are two main problems in deep reinforcement learning. Even if proximal policy optimization (PPO) is the state of the art, it still suffers from these two problems. We introduce an improved algorithm based on…

机器学习 · 计算机科学 2019-10-01 Zhenyu Zhang , Xiangfeng Luo , Tong Liu , Shaorong Xie , Jianshu Wang , Wei Wang , Yang Li , Yan Peng

Terahertz (THz) communications and reconfigurable intelligent surfaces (RISs) have been recently proposed to enable various powerful indoor applications, such as wireless virtual reality (VR). For an efficient servicing of VR users, an…

网络与互联网体系结构 · 计算机科学 2022-12-26 Mounir Bensalem , Anna Engelmann , Admela Jukan
‹ 上一页 1 8 9 10 下一页 ›