中文
相关论文

相关论文: Deep Contextual Bandit and Reinforcement Learning …

200 篇论文

In this paper, we propose reconfigurable intelligent surface (RIS)-assisted unmanned aerial vehicles (UAVs) networks that can utilise both advantages of UAV's agility and RIS's reflection for enhancing the network's performance. To aim at…

信号处理 · 电气工程与系统科学 2021-08-09 Khoi Khac Nguyen , Saeed Khosravirad , Daniel Benevides da Costa , Long D. Nguyen , Trung Q. Duong

We study the corruption-robustness of in-context reinforcement learning (ICRL), focusing on the Decision-Pretrained Transformer (DPT, Lee et al., 2023). To address the challenge of reward poisoning attacks targeting the DPT, we propose a…

机器学习 · 计算机科学 2025-09-30 Paulius Sasnauskas , Yiğit Yalın , Goran Radanović

We tackle tag-based query refinement as a mobile-friendly alternative to standard facet search. We approach the inference challenge with reinforcement learning, and propose a deep contextual bandit that can be efficiently scaled in a…

机器学习 · 计算机科学 2020-10-20 Bingqing Yu , Jacopo Tagliabue

In a practical massive MIMO (multiple-input multiple-output) system, the number of antennas at a base station (BS) is constrained by the space and cost factors, which limits the throughput gain promised by theoretical analysis. This paper…

信息论 · 计算机科学 2020-05-06 Zhaorui Wang , Liang Liu , Shuguang Cui

In this work, we examine an intelligent reflecting surface (IRS) assisted downlink non-orthogonal multiple access (NOMA) scenario with the aim of maximizing the sum rate of users. The optimization problem at the IRS is quite complicated,…

信息论 · 计算机科学 2021-04-06 Muhammad Shehab , Bekir S. Ciftler , Tamer Khattab , Mohamed Abdallah , Daniele Trinchero

Embodied agents, such as robots and virtual characters, must continuously select actions to execute tasks effectively, solving complex sequential decision-making problems. Given the difficulty of designing such controllers manually,…

机器人学 · 计算机科学 2026-05-18 Pedro Santana

This paper explores the feasibility of leveraging concepts from deep reinforcement learning (DRL) to enable dynamic resource management in Wi-Fi networks implementing distributed multi-user MIMO (D-MIMO). D-MIMO is a technique by which a…

In this paper, we consider jointly optimizing cell load balance and network throughput via a reinforcement learning (RL) approach, where inter-cell handover (i.e., user association assignment) and massive MIMO antenna tilting are configured…

机器学习 · 计算机科学 2020-12-03 Zhou Zhou , Yan Xin , Hao Chen , Charlie Zhang , Lingjia Liu

Deep reinforcement learning (DRL) has been applied in financial portfolio management to improve returns in changing market conditions. However, unlike most fields where DRL is widely used, the stock market is more volatile and dynamic as it…

机器学习 · 计算机科学 2025-02-12 Fengchen Gu , Angelos Stefanidis , Ángel García-Fernández , Jionglong Su , Huakang Li

Millimeter-wave (mmWave) communication systems, particularly those leveraging multi-user multiple-input and multiple-output (MU-MIMO) with hybrid beamforming, face challenges in optimizing user throughput and minimizing latency due to the…

信息论 · 计算机科学 2026-03-04 Ramin Hashemi , Vismika Ranasinghe , Teemu Veijalainen , Petteri Kela , Risto Wichman

This paper defines the problem of optimizing the downlink multi-user multiple input, single output (MU-MISO) sum-rate for ground users served by an aerial reconfigurable intelligent surface (ARIS) that acts as a relay to the terrestrial…

信号处理 · 电气工程与系统科学 2022-07-14 Aly Sabri Abdalla , Vuk Marojevic

This paper investigates the application of deep deterministic policy gradient (DDPG) to intelligent reflecting surface (IRS) based unmanned aerial vehicles (UAV) assisted non-orthogonal multiple access (NOMA) downlink networks. The…

信号处理 · 电气工程与系统科学 2023-04-06 Shiyu Jiao , Ximing Xie , Zhiguo Ding

Contextual bandits provide an effective way to model the dynamic data problem in ML by leveraging online (incremental) learning to continuously adjust the predictions based on changing environment. We explore details on contextual bandits,…

机器学习 · 计算机科学 2020-09-24 Dattaraj Rao

In the evolving landscape of the Internet of Things (IoT), integrating cognitive radio (CR) has become a practical solution to address the challenge of spectrum scarcity, leading to the development of cognitive IoT (CIoT). However, the…

信号处理 · 电气工程与系统科学 2025-12-18 Nadia Abdolkhani , Nada Abdel Khalek , Walaa Hamouda

Beamforming enhances signal strength and quality by focusing energy in specific directions. This capability is particularly crucial in cell-free integrated sensing and communication (ISAC) systems, where multiple distributed access points…

新兴技术 · 计算机科学 2026-01-21 Jiexin Zhang , Shu Xu , Chunguo Li , Yongming Huang , Luxi Yang

The proliferation of Internet of Things (IoT) devices and the advent of 6G technologies have introduced computationally intensive tasks that often surpass the processing capabilities of user devices. Efficient and secure resource allocation…

机器学习 · 计算机科学 2025-01-22 Jianfei Sun , Qiang Gao , Cong Wu , Yuxian Li , Jiacheng Wang , Dusit Niyato

Context, the embedding of previous collected trajectories, is a powerful construct for Meta-Reinforcement Learning (Meta-RL) algorithms. By conditioning on an effective context, Meta-RL policies can easily generalize to new tasks within a…

机器学习 · 计算机科学 2020-12-16 Haotian Fu , Hongyao Tang , Jianye Hao , Chen Chen , Xidong Feng , Dong Li , Wulong Liu

In this paper, we investigate resource allocation algorithm design for intelligent reflecting surface (IRS)-assisted multiuser cognitive radio (CR) systems. In particular, an IRS is deployed to mitigate the interference caused by the…

信息论 · 计算机科学 2020-03-03 Dongfang Xu , Xianghao Yu , Robert Schober

We study how representation learning can improve the learning efficiency of contextual bandit problems. We study the setting where we play T contextual linear bandits with dimension d simultaneously, and these T bandit tasks collectively…

机器学习 · 计算机科学 2025-01-08 Jiabin Lin , Shana Moothedath , Namrata Vaswani

Due to its static protocol design, IEEE 802.11 (aka Wi-Fi) channel access lacks adaptability to address dynamic network conditions, resulting in inefficient spectrum utilization, unnecessary contention, and packet collisions. This paper…