English
Related papers

Related papers: Deep Contextual Bandit and Reinforcement Learning …

200 papers

In this paper, we propose reconfigurable intelligent surface (RIS)-assisted unmanned aerial vehicles (UAVs) networks that can utilise both advantages of UAV's agility and RIS's reflection for enhancing the network's performance. To aim at…

Signal Processing · Electrical Eng. & Systems 2021-08-09 Khoi Khac Nguyen , Saeed Khosravirad , Daniel Benevides da Costa , Long D. Nguyen , Trung Q. Duong

We study the corruption-robustness of in-context reinforcement learning (ICRL), focusing on the Decision-Pretrained Transformer (DPT, Lee et al., 2023). To address the challenge of reward poisoning attacks targeting the DPT, we propose a…

Machine Learning · Computer Science 2025-09-30 Paulius Sasnauskas , Yiğit Yalın , Goran Radanović

We tackle tag-based query refinement as a mobile-friendly alternative to standard facet search. We approach the inference challenge with reinforcement learning, and propose a deep contextual bandit that can be efficiently scaled in a…

Machine Learning · Computer Science 2020-10-20 Bingqing Yu , Jacopo Tagliabue

In a practical massive MIMO (multiple-input multiple-output) system, the number of antennas at a base station (BS) is constrained by the space and cost factors, which limits the throughput gain promised by theoretical analysis. This paper…

Information Theory · Computer Science 2020-05-06 Zhaorui Wang , Liang Liu , Shuguang Cui

In this work, we examine an intelligent reflecting surface (IRS) assisted downlink non-orthogonal multiple access (NOMA) scenario with the aim of maximizing the sum rate of users. The optimization problem at the IRS is quite complicated,…

Information Theory · Computer Science 2021-04-06 Muhammad Shehab , Bekir S. Ciftler , Tamer Khattab , Mohamed Abdallah , Daniele Trinchero

Embodied agents, such as robots and virtual characters, must continuously select actions to execute tasks effectively, solving complex sequential decision-making problems. Given the difficulty of designing such controllers manually,…

Robotics · Computer Science 2026-05-18 Pedro Santana

This paper explores the feasibility of leveraging concepts from deep reinforcement learning (DRL) to enable dynamic resource management in Wi-Fi networks implementing distributed multi-user MIMO (D-MIMO). D-MIMO is a technique by which a…

In this paper, we consider jointly optimizing cell load balance and network throughput via a reinforcement learning (RL) approach, where inter-cell handover (i.e., user association assignment) and massive MIMO antenna tilting are configured…

Machine Learning · Computer Science 2020-12-03 Zhou Zhou , Yan Xin , Hao Chen , Charlie Zhang , Lingjia Liu

Deep reinforcement learning (DRL) has been applied in financial portfolio management to improve returns in changing market conditions. However, unlike most fields where DRL is widely used, the stock market is more volatile and dynamic as it…

Machine Learning · Computer Science 2025-02-12 Fengchen Gu , Angelos Stefanidis , Ángel García-Fernández , Jionglong Su , Huakang Li

Millimeter-wave (mmWave) communication systems, particularly those leveraging multi-user multiple-input and multiple-output (MU-MIMO) with hybrid beamforming, face challenges in optimizing user throughput and minimizing latency due to the…

Information Theory · Computer Science 2026-03-04 Ramin Hashemi , Vismika Ranasinghe , Teemu Veijalainen , Petteri Kela , Risto Wichman

This paper defines the problem of optimizing the downlink multi-user multiple input, single output (MU-MISO) sum-rate for ground users served by an aerial reconfigurable intelligent surface (ARIS) that acts as a relay to the terrestrial…

Signal Processing · Electrical Eng. & Systems 2022-07-14 Aly Sabri Abdalla , Vuk Marojevic

This paper investigates the application of deep deterministic policy gradient (DDPG) to intelligent reflecting surface (IRS) based unmanned aerial vehicles (UAV) assisted non-orthogonal multiple access (NOMA) downlink networks. The…

Signal Processing · Electrical Eng. & Systems 2023-04-06 Shiyu Jiao , Ximing Xie , Zhiguo Ding

Contextual bandits provide an effective way to model the dynamic data problem in ML by leveraging online (incremental) learning to continuously adjust the predictions based on changing environment. We explore details on contextual bandits,…

Machine Learning · Computer Science 2020-09-24 Dattaraj Rao

In the evolving landscape of the Internet of Things (IoT), integrating cognitive radio (CR) has become a practical solution to address the challenge of spectrum scarcity, leading to the development of cognitive IoT (CIoT). However, the…

Signal Processing · Electrical Eng. & Systems 2025-12-18 Nadia Abdolkhani , Nada Abdel Khalek , Walaa Hamouda

Beamforming enhances signal strength and quality by focusing energy in specific directions. This capability is particularly crucial in cell-free integrated sensing and communication (ISAC) systems, where multiple distributed access points…

Emerging Technologies · Computer Science 2026-01-21 Jiexin Zhang , Shu Xu , Chunguo Li , Yongming Huang , Luxi Yang

The proliferation of Internet of Things (IoT) devices and the advent of 6G technologies have introduced computationally intensive tasks that often surpass the processing capabilities of user devices. Efficient and secure resource allocation…

Machine Learning · Computer Science 2025-01-22 Jianfei Sun , Qiang Gao , Cong Wu , Yuxian Li , Jiacheng Wang , Dusit Niyato

Context, the embedding of previous collected trajectories, is a powerful construct for Meta-Reinforcement Learning (Meta-RL) algorithms. By conditioning on an effective context, Meta-RL policies can easily generalize to new tasks within a…

Machine Learning · Computer Science 2020-12-16 Haotian Fu , Hongyao Tang , Jianye Hao , Chen Chen , Xidong Feng , Dong Li , Wulong Liu

In this paper, we investigate resource allocation algorithm design for intelligent reflecting surface (IRS)-assisted multiuser cognitive radio (CR) systems. In particular, an IRS is deployed to mitigate the interference caused by the…

Information Theory · Computer Science 2020-03-03 Dongfang Xu , Xianghao Yu , Robert Schober

We study how representation learning can improve the learning efficiency of contextual bandit problems. We study the setting where we play T contextual linear bandits with dimension d simultaneously, and these T bandit tasks collectively…

Machine Learning · Computer Science 2025-01-08 Jiabin Lin , Shana Moothedath , Namrata Vaswani

Due to its static protocol design, IEEE 802.11 (aka Wi-Fi) channel access lacks adaptability to address dynamic network conditions, resulting in inefficient spectrum utilization, unnecessary contention, and packet collisions. This paper…

Networking and Internet Architecture · Computer Science 2025-11-14 Miguel Casasnovas , Francesc Wilhelmi , Richard Combes , Maksymilian Wojnar , Katarzyna Kosek-Szott , Szymon Szott , Anders Jonsson , Luis Esteve , Boris Bellalta
‹ Prev 1 3 4 5 6 7 10 Next ›