中文
相关论文

相关论文: Structured Reinforcement Learning for Delay-Optima…

200 篇论文

The integration of satellite communication networks with next-generation (NG) technologies is a promising approach towards global connectivity. However, the quality of services is highly dependant on the availability of accurate channel…

This paper addresses a novel multi-agent deep reinforcement learning (MADRL)-based positioning algorithm for multiple unmanned aerial vehicles (UAVs) collaboration (i.e., UAVs work as mobile base stations). The primary objective of the…

机器学习 · 计算机科学 2023-07-03 Chanyoung Park , Soohyun Park , Soyi Jung , Carlos Cordeiro , Joongheon Kim

Demands for data traffic in high-speed railway (HSR) has increased drastically. The increasing entertainment needs of passengers, safety control information exchanges of trains, etc., make train-to-train (T2T) communications face the…

信息论 · 计算机科学 2023-08-22 Yunhan Ma , Yong Niu , Shiwen Mao , Zhu Han , Ruisi He , Zhangdui Zhong , Ning Wang , Bo Ai

Demand response providers (DRPs) are intermediaries between the upper-level distribution system operator and the lower-level participants in demand response (DR) programs. Usually, DRPs act as leaders and determine electricity pricing…

系统与控制 · 电气工程与系统科学 2025-09-04 Xin Li , Li Ding , Qiao Lin , Zhen-Wei Yu

In this article, for the first time, we propose a transformer network-based reinforcement learning (RL) method for power distribution network (PDN) optimization of high bandwidth memory (HBM). The proposed method can provide an optimal…

Millimeter-wave (mmWave) communications provide access to spectra with bandwidths and in abundance. However, the high susceptibility of mmWave to blockage imposes crucial challenges, especially to low-latency services. In this paper, a…

信号处理 · 电气工程与系统科学 2021-04-20 Yashuai Cao , Tiejun Lv , Zhipeng Lin , Wei Ni

In this paper, we study the scheduling problem for downlink transmission in a multi-channel (e.g., OFDM-based) wireless network. We focus on a single cell, with the aim of developing a unifying framework for designing low-complexity…

网络与互联网体系结构 · 计算机科学 2013-11-19 Bo Ji , Gagan R. Gupta , Xiaojun Lin , Ness B. Shroff

Deep reinforcement learning (DRL) algorithms have proven effective in robot navigation, especially in unknown environments, by directly mapping perception inputs into robot control commands. However, most existing methods ignore the local…

机器人学 · 计算机科学 2023-07-06 Yu'an Chen , Ruosong Ye , Ziyang Tao , Hongjian Liu , Guangda Chen , Jie Peng , Jun Ma , Yu Zhang , Jianmin Ji , Yanyong Zhang

Millimeter wave (mmWave) and terahertz MIMO systems rely on pre-defined beamforming codebooks for both initial access and data transmission. Being pre-defined, however, these codebooks are commonly not optimized for specific environments,…

信息论 · 计算机科学 2021-02-24 Yu Zhang , Muhammad Alrabeiah , Ahmed Alkhateeb

In this letter, we propose a novel Multi-Agent Deep Reinforcement Learning (MADRL) framework for Medium Access Control (MAC) protocol design. Unlike centralized approaches, which rely on a single entity for decision-making, MADRL empowers…

系统与控制 · 电气工程与系统科学 2024-11-25 Navid Keshtiarast , Oliver Renaldi , Marina Petrova

In many Cyber-Physical Systems, we encounter the problem of remote state estimation of geographically distributed and remote physical processes. This paper studies the scheduling of sensor transmissions to estimate the states of multiple…

系统与控制 · 计算机科学 2020-05-28 Alex S. Leong , Arunselvan Ramaswamy , Daniel E. Quevedo , Holger Karl , Ling Shi

Large transformer models trained on diverse datasets have shown a remarkable ability to learn in-context, achieving high few-shot performance on tasks they were not explicitly trained to solve. In this paper, we study the in-context…

机器学习 · 计算机科学 2023-06-27 Jonathan N. Lee , Annie Xie , Aldo Pacchiano , Yash Chandak , Chelsea Finn , Ofir Nachum , Emma Brunskill

We consider distributed caching of content across several small base stations (SBSs) in a wireless network, where the content is encoded using a maximum distance separable code. Specifically, we apply soft time-to-live (STTL) cache…

信息论 · 计算机科学 2021-04-15 Jesper Pedersen , Alexandre Graell i Amat , Fredrik Brännström , Eirik Rosnes

We present a model-free reinforcement learning algorithm to find an optimal policy for a finite-horizon Markov decision process while guaranteeing a desired lower bound on the probability of satisfying a signal temporal logic (STL)…

系统与控制 · 电气工程与系统科学 2021-09-29 Krishna C. Kalagarla , Rahul Jain , Pierluigi Nuzzo

The complex transmission mechanism of cross-packet hybrid automatic repeat request (XP-HARQ) hinders its optimal system design. To overcome this difficulty, this letter attempts to use the deep reinforcement learning (DRL) to solve the rate…

信息论 · 计算机科学 2023-08-07 Da Wu , Jiahui Feng , Zheng Shi , Hongjiang Lei , Guanghua Yang , Shaodan Ma

In this paper, we study a real-time monitoring system in which multiple source nodes are responsible for sending update packets to a common destination node in order to maintain the freshness of information at the destination. Since it may…

信息论 · 计算机科学 2019-08-20 Mohamed A. Abd-Elmagid , Harpreet S. Dhillon , Nikolaos Pappas

Contextual Reinforcement Learning (CRL) tackles the problem of solving a set of related Contextual Markov Decision Processes (CMDPs) that vary across different context variables. Traditional approaches--independent training and multi-task…

机器学习 · 计算机科学 2026-03-31 Tianyue Zhou , Jung-Hoon Cho , Cathy Wu

WiFi densification leads to the existence of multiple overlapping coverage areas, which allows user stations (STAs) to choose between different Access Points (APs). The standard WiFi association method makes the STAs select the AP with the…

网络与互联网体系结构 · 计算机科学 2019-03-04 Marc Carrascosa , Boris Bellalta

Millimeter-wave (mmWave) networks have the potential to support high throughput and low-latency requirements of 5G-and-beyond communication standards. But transmissions in this band are highly vulnerable to attenuation and blockages from…

最优化与控制 · 数学 2025-11-25 Manali Dutta , Gourav Saha , Rahul Singh , Ness B. Shroff

Densely deployed base stations are responsible for the majority of the energy consumed in Radio access network (RAN). While these deployments are crucial to deliver the required data rate in busy hours of the day, the network can save…

系统与控制 · 电气工程与系统科学 2026-04-02 Xuanyu Liang , Ahmed Al-Tahmeesschi , Swarna Chetty , Cicek Cavdar , Berk Canberk , Hamed Ahmadi