中文
相关论文

相关论文: Deep Reinforcement Learning Based Optimization for…

200 篇论文

Intelligent reflecting surface (IRS)-enabled backscatter communications can be enabled by an access point (AP) that splits its transmit signal into modulated and unmodulated parts. This letter integrates non-orthogonal multiple access…

信息论 · 计算机科学 2022-09-07 Azar Hakimi , Shayan Zargari , Chintha Tellambura , Sanjeewa Herath

In this paper, we consider an aerial reconfigurable intelligent surface (ARIS)-assisted wireless network, where multiple unmanned aerial vehicles (UAVs) collect data from ground users (GUs) by using the non-orthogonal multiple access (NOMA)…

网络与互联网体系结构 · 计算机科学 2024-12-31 Songhan Zhao , Shimin Gong , Bo Gu , Lanhua Li , Bin Lyu , Dinh Thai Hoang , Changyan Yi

This work adopts the very successful distributional perspective on reinforcement learning and adapts it to the continuous control setting. We combine this within a distributed framework for off-policy learning in order to develop what we…

Motion planning is an essential component in most of today's robotic applications. In this work, we consider the learning setting, where a set of solved motion planning problems is used to improve the efficiency of motion planning on…

机器人学 · 计算机科学 2019-06-04 Tom Jurgenson , Aviv Tamar

Considering intelligent reflecting surface (IRS), we study a multi-cluster multiple-input-single-output (MISO) non-orthogonal multiple access (NOMA) downlink communication network. In the network, an IRS assists the communication from the…

信息论 · 计算机科学 2019-09-17 Yiqing Li , Miao Jiang , Qi Zhang , Jiayin Qin

Non-orthogonal multiple access (NOMA) scheme enables serving users with the same resource block i.e. frequency or time by multiplexing the signal of the users. Intelligent reflecting surfaces (IRS) or reconfigurable intelligent surfaces…

信号处理 · 电气工程与系统科学 2022-11-24 Mobasshir Mahbub , Raed M. Shubair

Maze navigation is a fundamental challenge in robotics, requiring agents to traverse complex environments efficiently. While the Deep Deterministic Policy Gradient (DDPG) algorithm excels in control tasks, its performance in maze navigation…

机器人学 · 计算机科学 2025-08-08 Wenjie Hu , Ye Zhou , Hann Woei Ho

Millimeter-wave (mmWave) base station can offer abundant high capacity channel resources toward connected vehicles so that quality-of-service (QoS) of them in terms of downlink throughput can be highly improved. The mmWave base station can…

信号处理 · 电气工程与系统科学 2020-01-09 Dohyun Kwon , Joongheon Kim

The combination of multiple-input multiple-output (MIMO) systems and intelligent reflecting surfaces (IRSs) is foreseen as a critical enabler of beyond 5G (B5G) and 6G. In this work, two different approaches are considered for the joint…

信息论 · 计算机科学 2024-01-31 Dariel Pereira-Ruisánchez , Óscar Fresnedo , Darian Pérez-Adán , Luis Castedo

The promising coverage and spectral efficiency gains of intelligent reflecting surfaces (IRSs) are attracting increasing interest. In order to realize these surfaces in practice, however, several challenges need to be addressed. One of…

信息论 · 计算机科学 2020-02-26 Abdelrahman Taha , Yu Zhang , Faris B. Mismar , Ahmed Alkhateeb

By employing powerful edge servers for data processing, mobile edge computing (MEC) has been recognized as a promising technology to support emerging computation-intensive applications. Besides, non-orthogonal multiple access (NOMA)-aided…

信号处理 · 电气工程与系统科学 2022-09-22 Jiadong Yu , Yang Li , Xiaolan Liu , Bo Sun , Yuan Wu , Danny H. K. Tsang

Intelligent reflecting surfaces (IRSs) have emerged as a promising solution to mitigate line-of-sight (LoS) blockages and enhance signal coverage in optical wireless communication (OWC) systems with minimal additional power. In this work,…

系统与控制 · 电气工程与系统科学 2025-10-15 Ahrar N. Hamad , Ahmad Adnan Qidan , Taisir E. H. El-Gorashi , Jaafar M. H. Elmirghani

In this paper, we consider the application of intelligent reflecting surface (IRS) in unmanned aerial vehicle (UAV)-based orthogonal frequency division multiple access (OFDMA) communication systems, which exploits both the significant…

信息论 · 计算机科学 2020-10-13 Zhiqiang Wei , Yuanxin Cai , Zhuo Sun , Derrick Wing Kwan Ng , Jinhong Yuan , Mingyu Zhou , Lixin Sun

In this paper, we investigate joint vehicle association and multi-dimensional resource management in a vehicular network assisted by multi-access edge computing (MEC) and unmanned aerial vehicle (UAV). To efficiently manage the available…

网络与互联网体系结构 · 计算机科学 2020-09-09 Haixia Peng , Xuemin Shen

Model-free reinforcement learning algorithms such as Deep Deterministic Policy Gradient (DDPG) often require additional exploration strategies, especially if the actor is of deterministic nature. This work evaluates the use of model-based…

机器学习 · 计算机科学 2019-11-19 Kevin Sebastian Luck , Mel Vecerik , Simon Stepputtis , Heni Ben Amor , Jonathan Scholz

Contemporary autopilot systems for unmanned aerial vehicles (UAVs) are far more limited in their flight envelope as compared to experienced human pilots, thereby restricting the conditions UAVs can operate in and the types of missions they…

机器人学 · 计算机科学 2019-11-14 Eivind Bøhn , Erlend M. Coates , Signe Moe , Tor Arne Johansen

This paper tackles the challenge of learning non-Markovian optimal execution strategies in dynamic financial markets. We introduce a novel actor-critic algorithm based on Deep Deterministic Policy Gradient (DDPG) to address this issue, with…

机器学习 · 计算机科学 2024-10-18 Alessandro Micheli , Mélodie Monod

This paper investigates a multi-Unmanned Aerial Vehicle (UAV) joint base station-assisted Internet of Vehicles (IoV) task offloading system in dense urban environments. To minimize system delay and energy consumption under strict coupling…

网络与互联网体系结构 · 计算机科学 2026-05-07 Maoxin Ji , Qiong Wu , Pingyi Fan , Cui Zhang , Nan Cheng , Wen Chen , Khaled B. Letaief

In this paper, we study an unmanned aerial vehicle (UAV) communication system, where a ground node (GN) communicate with a UAV assisted by intelligent reflecting surface (IRS) in the presence of a jammer with imperfect location information.…

信息论 · 计算机科学 2022-01-25 Zhi Ji , Xinrong Guan , Jia Tu , Qingqing Wu , Wendong Yang

This paper presents a robust reinforcement learning algorithm called robust deterministic policy gradient (RDPG), which reformulates the H-infinity control problem as a two-player zero-sum dynamic game between a user and an adversary. The…

机器人学 · 计算机科学 2025-12-04 Taeho Lee , Donghwan Lee