中文
相关论文

相关论文: Q-NAV: NAV Setting Method based on Reinforcement L…

200 篇论文

Aligning large language models (LLMs) with human preferences through reinforcement learning (RLHF) can lead to reward hacking, where LLMs exploit failures in the reward model (RM) to achieve seemingly high rewards without meeting the…

With the rapid advance of information technology, network systems have become increasingly complex and hence the underlying system dynamics are often unknown or difficult to characterize. Finding a good network control policy is of…

性能 · 计算机科学 2022-04-08 Bai Liu , Qiaomin Xie , Eytan Modiano

Wireless networks are evolving from radio resource providers to complex systems that also involve computing, with the latter being distributed across edge and cloud facilities. Also, their optimization is shifting more and more from a…

计算机科学与博弈论 · 计算机科学 2025-08-08 Mandar Datar , Mattia Merluzzi

Recently, Optical wireless communication (OWC) have been considered as a key element in the next generation of wireless communications due to its potential in supporting unprecedented communication speeds. In this paper, infrared lasers…

系统与控制 · 电气工程与系统科学 2023-01-20 Ahmad Adnan Qidan , Taisir El-Gorashi , Jaafar M. H. Elmirghani

To satisfy the high data rate requirement andreliable transmission demands in underwater scenarios, it isdesirable to construct an efficient hybrid underwater opticalacoustic network (UWOAN) architecture by considering the keyfeatures and…

网络与互联网体系结构 · 计算机科学 2021-10-01 Yuanhao Liu , Fen Zhou , Tao Shang

This paper considers a set of multiple independent control systems that are each connected over a non-stationary wireless channel. The goal is to maximize control performance over all the systems through the allocation of transmitting power…

最优化与控制 · 数学 2019-03-27 Mark Eisen , Konstantinos Gatsis , George J. Pappas , Alejandro Ribeiro

Social navigation has been gaining attentions with the growth in machine intelligence. Since reinforcement learning can select an action in the prediction phase at a low computational cost, it has been formulated in a social navigation…

机器人学 · 计算机科学 2021-04-15 Takato Okudo , Seiji Yamada

In this paper, we for the first time investigate the random access problem for a delay-constrained heterogeneous wireless network. We begin with a simple two-device problem where two devices deliver delay-constrained traffic to an access…

网络与互联网体系结构 · 计算机科学 2022-05-05 Lei Deng , Danzhou Wu , Zilong Liu , Yijin Zhang , Yunghsiang S. Han

This paper introduces the deployment of unmanned aerial vehicles (UAVs) as lightweight wireless access points that leverage the fixed infrastructure in the context of the emerging open radio access network (O-RAN). More precisely, we…

系统与控制 · 电气工程与系统科学 2022-11-22 Hossein Mohammadi , Vuk Marojevic , Bodong Shang

Uncrewed autonomous vehicles (UAVs) have made significant contributions to reconnaissance and surveillance missions in past US military campaigns. As the prevalence of UAVs increases, there has also been improvements in counter-UAV…

机器学习 · 计算机科学 2021-12-02 Yixuan Liu , Chrysafis Vogiatzis , Ruriko Yoshida , Erich Morman

Unmanned aerial vehicles (UAVs) are expected to be a key component of the next-generation wireless systems. Due to their deployment flexibility, UAVs are being considered as an efficient solution for collecting information data from ground…

信息论 · 计算机科学 2019-08-20 Mohamed A. Abd-Elmagid , Aidin Ferdowsi , Harpreet S. Dhillon , Walid Saad

With the rapid advancement of technology, the recognition of underwater acoustic signals in complex environments has become increasingly crucial. Currently, mainstream underwater acoustic signal recognition relies primarily on…

声音 · 计算机科学 2024-01-08 Minghao Chen

For Industrial Wireless Sensor Networks, it is essential to reliably sense and deliver the environmental data on time to avoid system malfunction. While energy harvesting is a promising technique to extend the lifetime of sensor nodes, it…

网络与互联网体系结构 · 计算机科学 2016-05-12 Lei Lei , Yiru Kuang , Xuemin , Shen , Kan Yang , Jian Qiao , Zhangdui Zhong

We consider reinforcement learning (RL) methods in offline domains without additional online data collection, such as mobile health applications. Most of existing policy optimization algorithms in the computer science literature are…

机器学习 · 统计学 2022-07-28 Chengchun Shi , Shikai Luo , Yuan Le , Hongtu Zhu , Rui Song

Autonomous underwater vehicles are required to perform multiple tasks adaptively and in an explainable manner under dynamic, uncertain conditions and limited sensing, challenges that classical controllers struggle to address. This demands…

机器学习 · 计算机科学 2026-04-24 Yi-Ling Liu , Melvin Laux , Mariela De Lucas Alvarez , Frank Kirchner , Rebecca Adam

This paper presents a new strategy for simultaneously reducing energy consumption, transmission delays, and bit error rate in Unmanned Aerial Vehicle UAV networks. A UAV is fitted with a wireless Bidirectional Relay BR to enable coverage…

信号处理 · 电气工程与系统科学 2022-05-17 Muhammad Khalil

We propose a technique to authenticate received packets in underwater acoustic networks based on the physical layer features of the underwater acoustic channel (UWAC). Several sensors a) locally estimate features (e.g., the number of taps…

信号处理 · 电气工程与系统科学 2024-05-24 Francesco Ardizzon , Roee Diamant , Paolo Casari , Stefano Tomasin

To convey desired behavior to a Reinforcement Learning (RL) agent, a designer must choose a reward function for the environment, arguably the most important knob designers have in interacting with RL agents. Although many reward functions…

机器学习 · 计算机科学 2022-06-01 Henry Sowerby , Zhiyuan Zhou , Michael L. Littman

Since the application of Deep Q-Learning to the continuous action domain in Atari-like games, Deep Reinforcement Learning (Deep-RL) techniques for motion control have been qualitatively enhanced. Nowadays, modern Deep-RL can be successfully…

Deep learning (DL) has made notable progress in addressing complex radio access network control challenges that conventional analytic methods have struggled to solve. However, DL has shown limitations in solving constrained NP-hard problems…

系统与控制 · 电气工程与系统科学 2025-02-05 Hyeonho Noh , Byonghyo Shim , Hyun Jong Yang