English
Related papers

Related papers: Linear Jamming Bandits: Learning to Jam 5G-based C…

200 papers

In this paper, we investigate the stochastic contextual bandit with general function space and graph feedback. We propose an algorithm that addresses this problem by adapting to both the underlying graph structures and reward gaps. To the…

Machine Learning · Computer Science 2024-01-09 Xueping Gong , Jiheng Zhang

We study contextual bandit (CB) problems, where the user can sometimes respond with the best action in a given context. Such an interaction arises, for example, in text prediction or autocompletion settings, where a poor suggestion is…

Machine Learning · Computer Science 2023-02-09 Alekh Agarwal , Claudio Gentile , Teodor V. Marinov

This letter proposes a linear bandit-based beam training framework for near-field communication under multi-path channels. By leveraging Thompson Sampling (TS), the framework adaptively balances exploration and exploitation to maximize…

Signal Processing · Electrical Eng. & Systems 2026-03-11 Junchi Liu , Zijun Wang , Rui Zhang

Deep reinforcement learning (DRL) has recently been used to perform efficient resource allocation in wireless communications. In this paper, the vulnerabilities of such DRL agents to adversarial attacks is studied. In particular, we…

Machine Learning · Computer Science 2021-05-13 Feng Wang , M. Cenk Gursoy , Senem Velipasalar

We investigate the problem of stochastic, combinatorial multi-armed bandits where the learner only has access to bandit feedback and the reward function can be non-linear. We provide a general framework for adapting discrete offline…

Machine Learning · Computer Science 2023-10-13 Guanyu Nie , Yididiya Y Nadew , Yanhui Zhu , Vaneet Aggarwal , Christopher John Quinn

Network slicing is a pivotal paradigm in wireless networks enabling customized services to users and applications. Yet, intelligent jamming attacks threaten the performance of network slicing. In this paper, we focus on the security aspect…

Networking and Internet Architecture · Computer Science 2024-10-08 Shavbo Salehi , Hao Zhou , Medhat Elsayed , Majid Bavand , Raimundas Gaigalas , Yigit Ozcan , Melike Erol-Kantarci

The deployment of Unmanned Aerial Vehicle (UAV) swarms as dynamic communication relays is critical for next-generation tactical networks. However, operating in contested environments requires solving a complex trade-off, including…

Networking and Internet Architecture · Computer Science 2025-12-10 Thai Duong Nguyen , Ngoc-Tan Nguyen , Thanh-Dao Nguyen , Nguyen Van Huynh , Dinh-Hieu Tran , Symeon Chatzinotas

We study reliable communication in uncoordinated vehicular communication from the perspective of Shannon theory. Our system model for the information transmission is that of an Arbitrarily Varying Channel (AVC): One sender-receiver pair…

Information Theory · Computer Science 2019-06-06 Christian Arendt , Janis Nötzel , Holger Boche

Reinforcement algorithms refer to the schemes where the results of the previous trials and a reward-punishment rule are used for parameter setting in the next steps. In this paper, we use the concept of reinforcement algorithms to develop…

Information Theory · Computer Science 2014-09-12 Behrooz Makki , Tommy Svensson , Merouane Debbah

In this paper, a novel covert semantic communication framework is investigated. Within this framework, a server extracts and transmits the semantic information, i.e., the meaning of image data, to a user over several time slots. An attacker…

Artificial Intelligence · Computer Science 2025-08-12 Wenjing Zhang , Ye Hu , Tao Luo , Zhilong Zhang , Mingzhe Chen

Channel coding from 2G to 5G has assumed the inputs bits at the physical layer to be uniformly distributed. However, hybrid automatic repeat request acknowledgement (HARQ-ACK) bits transmitted in the uplink are inherently non-uniformly…

Signal Processing · Electrical Eng. & Systems 2026-02-20 Akash Doshi , Pinar Sen , Kirill Ivanov , Wei Yang , June Namgoong , Runxin Wang , Rachel Wang , Taesang Yoo , Jing Jiang , Tingfang Ji

We study a multi-armed bandit problem where the rewards exhibit regime switching. Specifically, the distributions of the random rewards generated from all arms are modulated by a common underlying state modeled as a finite-state Markov…

Machine Learning · Computer Science 2021-02-02 Xiang Zhou , Yi Xiong , Ningyuan Chen , Xuefeng Gao

The dominating waveform in 5G is orthogonal frequency division multiplexing (OFDM). OFDM will remain a promising waveform candidate for joint communication and sensing (JCAS) in 6G since OFDM can provide excellent data transmission…

Information Theory · Computer Science 2023-01-10 Yi Geng

We study two-sided matching markets in which one side of the market (the players) does not have a priori knowledge about its preferences for the other side (the arms) and is required to learn its preferences from experience. Also, we assume…

Machine Learning · Computer Science 2021-06-23 Lydia T. Liu , Feng Ruan , Horia Mania , Michael I. Jordan

We consider contextual linear bandits over networks, a class of sequential decision-making problems where learning occurs simultaneously across multiple locations and the reward distributions share structural similarities while also…

Machine Learning · Computer Science 2025-08-26 Chuyun Deng , Huiwen Jia

In this paper, we investigate the problem of beam alignment in millimeter wave (mmWave) systems, and design an optimal algorithm to reduce the overhead. Specifically, due to directional communications, the transmitter and receiver beams…

Information Theory · Computer Science 2017-12-29 Morteza Hashemi , Ashutosh Sabharwal , C. Emre Koksal , Ness B. Shroff

Wireless communications are vulnerable against radio frequency (RF) jamming which might be caused either intentionally or unintentionally. A particular subset of wireless networks, vehicular ad-hoc networks (VANET) which incorporate a…

Cryptography and Security · Computer Science 2019-01-01 Dimitrios Kosmanos , Dimitrios Karagiannis , Antonios Argyriou , Spyros Lalis , Leandros Maglaras

We propose and examine the idea of continuously adapting state-of-the-art neural network (NN)-based orthogonal frequency division multiplex (OFDM) receivers to current channel conditions. This online adaptation via retraining is mainly…

We study the problem of online learning in adversarial bandit problems under a partial observability model called off-policy feedback. In this sequential decision making problem, the learner cannot directly observe its rewards, but instead…

Machine Learning · Computer Science 2022-07-20 Germano Gabbianelli , Matteo Papini , Gergely Neu

We study the constrained variant of the \emph{multi-armed bandit} (MAB) problem, in which the learner aims not only at minimizing the total loss incurred during the learning dynamic, but also at controlling the violation of multiple…

Machine Learning · Computer Science 2026-02-17 Francesco Emanuele Stradi , Kalana Kalupahana , Matteo Castiglioni , Alberto Marchesi , Nicola Gatti
‹ Prev 1 8 9 10 Next ›