中文
相关论文

相关论文: A Reinforcement Learning Framework with Region-Awa…

200 篇论文

Reinforcement learning (RL) has emerged as a promising strategy for finetuning small language models (SLMs) to solve targeted tasks such as math and coding. However, RL algorithms tend to be resource-intensive, taking a significant amount…

机器学习 · 计算机科学 2025-10-07 Lianghuan Huang , Sagnik Anupam , Insup Lee , Shuo Li , Osbert Bastani

This paper presents an optimization framework for routing in software-defined elastic optical networks using reinforcement learning algorithms. We specifically implement and compare the epsilon-greedy bandit, upper confidence bound (UCB)…

网络与互联网体系结构 · 计算机科学 2024-10-21 Ryan McCann , Arash Rezaee , Vinod M. Vokkarane

The Internet of Things (IoT) has been increasingly used in our everyday lives as well as in numerous industrial applications. However, due to limitations in computing and power capabilities, IoT devices need to send their respective tasks…

网络与互联网体系结构 · 计算机科学 2025-07-01 Ziad Qais Al Abbasi , Khaled M. Rabie , Senior Member , Xingwang Li , Senior Member , Wali Ullah Khan , Asma Abu Samah

In this article, for the first time, we propose a transformer network-based reinforcement learning (RL) method for power distribution network (PDN) optimization of high bandwidth memory (HBM). The proposed method can provide an optimal…

Congestion Control (CC), as the core networking task to efficiently utilize network capacity, received great attention and widely used in various Internet communication applications such as 5G, Internet-of-Things, UAN, and more. Various CC…

网络与互联网体系结构 · 计算机科学 2022-06-07 Jianing Bai , Tianhao Zhang , Guangming Xie

Managing mixed traffic comprising human-driven and robot vehicles (RVs) across large-scale networks presents unique challenges beyond single-intersection control. This paper proposes a reinforcement learning framework for coordinating mixed…

机器学习 · 计算机科学 2024-12-18 Iftekharul Islam , Weizi Li

In the forthcoming 6G era, extend reality (XR) has been regarded as an emerging application for ultra-reliable and low latency communications (URLLC) with new traffic characteristics and more stringent requirements. In addition to the…

网络与互联网体系结构 · 计算机科学 2025-04-08 Luyuan Zhang , An Liu , Kexuan Wang

Reinforcement learning (RL) enables agents to learn optimal policies through environmental interaction. However, RL suffers from reduced learning efficiency due to the curse of dimensionality in high-dimensional spaces. Quantum…

机器学习 · 计算机科学 2025-07-02 Seok Bin Son , Joongheon Kim

SMART NoC, which transmits unconflicted flits to distant processing elements (PEs) in one cycle through the express bypass, is a high-performance NoC design proposed recently. However, if contention occurs, flits with low priority would not…

硬件体系结构 · 计算机科学 2020-11-19 Hui Chen , Peng Chen , Jun Zhou , Duong H. K. Luan , Weichen Liu

The growing complexity and capacity demands for mobile networks necessitate innovative techniques for optimizing resource usage. Meanwhile, recent breakthroughs have brought Reinforcement Learning (RL) into the domain of continuous control…

网络与互联网体系结构 · 计算机科学 2022-10-28 Vegard Edvardsen , Gard Spreemann , Jeriek Van den Abeele

In recent years, autonomous networks have been designed with Predictive Quality of Service (PQoS) in mind, as a means for applications operating in the industrial and/or automotive sectors to predict unanticipated Quality of Service (QoS)…

网络与互联网体系结构 · 计算机科学 2022-02-07 Federico Mason , Matteo Drago , Tommaso Zugno , Marco Giordani , Mate Boban , Michele Zorzi

The increasing integration of electric vehicles (EVs) into the grid can pose a significant risk to the distribution system operation in the absence of coordination. In response to the need for effective coordination of EVs within the…

系统与控制 · 电气工程与系统科学 2024-03-21 Jiarong Fan , Ariel Liebman , Hao Wang

Reinforcement learning (RL) faces substantial challenges when applied to real-life problems, primarily stemming from the scarcity of available data due to limited interactions with the environment. This limitation is exacerbated by the fact…

神经与进化计算 · 计算机科学 2024-04-10 Cristiano Capone , Paolo Muratore

Reinforcement learning (RL) emerges as a promising data-driven approach for adaptive traffic signal control (ATSC) in complex urban traffic networks, with deep neural networks substantially augmenting its learning capabilities. However,…

人工智能 · 计算机科学 2025-02-25 Yuli Zhang , Shangbo Wang , Dongyao Jia , Pengfei Fan , Ruiyuan Jiang , Hankang Gu , Andy H. F. Chow

District cooling energy plants (DCEPs) consisting of chillers, cooling towers, and thermal energy storage (TES) systems consume a considerable amount of electricity. Optimizing the scheduling of the TES and chillers to take advantage of…

系统与控制 · 电气工程与系统科学 2022-03-16 Zhong Guo , Austin R. Coffman , Prabir Barooah

Next generation cellular networks will have to leverage large cell densifications to accomplish the ambitious goals for aggregate multi-user sum rates, for which CRAN architecture is a favored network design. This shifts the attention back…

信息论 · 计算机科学 2017-03-31 Sahar Imtiaz , Hadi Ghauch , Muhammad Mahboob Ur Rahman , George Koudouridis , James Gross

Offline reinforcement learning (RL) aims to learn a policy from a static dataset without further interactions with the environment. Collecting sufficiently large datasets for offline RL is exhausting since this data collection requires…

人工智能 · 计算机科学 2025-10-22 Jongchan Park , Mingyu Park , Donghwan Lee

A reconfigurable intelligent surface (RIS) enhanced non-orthogonal multiple access assisted backscatter communication (RIS-NOMABC) system is considered. A joint optimization problem over power reflection coefficients and phase shifts is…

信息论 · 计算机科学 2020-12-21 Jiakuo Zuo , Yuanwei Liu , Liang Yang , Lingyang Song , Ying-Chang Liang

We introduce QuARC, Quantum Adaptive Routing using Clusters, a novel clustering-based entanglement routing protocol that leverages redundant, multi-path routing through multi-particle projective quantum measurements to enable…

量子物理 · 物理学 2024-10-31 Connor Clayton , Xiaodi Wu , Bobby Bhattacharjee

This study presents a novel computer system performance optimization and adaptive workload management scheduling algorithm based on Q-learning. In modern computing environments, characterized by increasing data volumes, task complexity, and…

机器学习 · 计算机科学 2024-11-11 Pochun Li , Yuyang Xiao , Jinghua Yan , Xuan Li , Xiaoye Wang