English
Related papers

Related papers: Optimal and Low-Complexity Dynamic Spectrum Access…

200 papers

We consider a reinforcement learning (RL) setting in which the agent interacts with a sequence of episodic MDPs. At the start of each episode the agent has access to some side-information or context that determines the dynamics of the MDP…

Machine Learning · Statistics 2019-10-24 Aditya Modi , Nan Jiang , Satinder Singh , Ambuj Tewari

Federated learning (FL) necessitates that edge devices conduct local training and communicate with a parameter server, resulting in significant energy consumption. A key challenge in practical FL systems is the rapid depletion of…

Machine Learning · Computer Science 2025-06-24 Kai Zhang , Xuanyu Cao , Khaled B. Letaief

Intelligent reflecting surface (IRS) has recently been emerged as an effective way for improving the performance of wireless networks by reconfiguring the propagation environment through a large number of passive reflecting elements. This…

Information Theory · Computer Science 2021-12-07 Parisa Ramezani , Abbas Jamalipour

We optimize finite horizon multi-agent reach-avoid Markov decision process (MDP) via \emph{local feedback policies}. The global feedback policy solution yields global optimality but its communication complexity, memory usage and computation…

Systems and Control · Electrical Eng. & Systems 2026-04-10 Adam Casselman , Abraham P. Vinod , Sarah H. Q. Li

Despite rapid progress in AI agents for enterprise automation and decision-making, their real-world deployment and further performance gains remain constrained by limited data quality and quantity, complex real-world reasoning demands,…

Artificial Intelligence · Computer Science 2026-03-24 Xi Yang , Aurelie Lozano , Naoki Abe , Bhavya , Saurabh Jha , Noah Zheutlin , Rohan R. Arora , Yu Deng , Daby M. Sow

In the context of railway systems, the application performance can be very critical and the radio conditions not advantageous. Hence, the communication problem parameters include both a survival time stemming from the application layer and…

Information Theory · Computer Science 2023-03-22 Vincent Corlay , Jean-Christophe Sibel

Dynamic spectrum sharing can provide many benefits to wireless networks operators. However, its efficiency requires sophisticated control mechanisms. The more context information is used by it, the higher performance of networks is…

Networking and Internet Architecture · Computer Science 2018-11-08 Paweł Kryszkiewicz , Adrian Kliks , Łukasz Kułacz , Hanna Bogucka , Georgios P. Koudouridis , Marcin Dryjański

Most reinforcement learning methods are based upon the key assumption that the transition dynamics and reward functions are fixed, that is, the underlying Markov decision process is stationary. However, in many real-world applications, this…

Machine Learning · Computer Science 2020-09-23 Yash Chandak , Georgios Theocharous , Shiv Shankar , Martha White , Sridhar Mahadevan , Philip S. Thomas

Multi-objective Markov decision processes are a special kind of multi-objective optimization problem that involves sequential decision making while satisfying the Markov property of stochastic processes. Multi-objective reinforcement…

Machine Learning · Computer Science 2023-08-22 Sherif Abdelfattah , Kathryn Kasmarik , Jiankun Hu

We consider a multichannel random access system in which each user accesses a single channel at each time slot to communicate with an access point (AP). Users arrive to the system at random and be activated for a certain period of time…

Signal Processing · Electrical Eng. & Systems 2021-05-11 Muhammad Sohaib , Jongjin Jeong , Sang-Woon Jeon

Ambient backscatter communication has shown great potential in the development of future wireless networks. It enables a backscatter transmitter (BTx) to send information directly to an adjacent receiver by modulating over ambient radio…

Information Theory · Computer Science 2019-08-16 Shaoqing Zhou , Wei Xu , Kezhi Wang , Cunhua Pan , Mohamed-Slim Alouini , Arumugam Nallanathan

In this paper, we employ deep reinforcement learning to develop a novel radio resource allocation and packet scheduling scheme for different Quality of Service (QoS) requirements applicable to LTEadvanced and 5G networks. In addition,…

Signal Processing · Electrical Eng. & Systems 2020-08-18 Mahdi Nouri Boroujerdi , Mohammad Akbari , Roghayeh Joda , Mohammad Ali Maddah-Ali , Babak Hossein Khalaj

Optical camera communications (OCC) has emerged as a key enabling technology for the seamless operation of future autonomous vehicles. In this paper, we introduce a spectral efficiency optimization approach in vehicular OCC. Specifically,…

Machine Learning · Computer Science 2022-05-06 Amirul Islam , Leila Musavian , Nikolaos Thomos

This paper proposes a formal approach to online learning and planning for agents operating in a priori unknown, time-varying environments. The proposed method computes the maximally likely model of the environment, given the observations…

Machine Learning · Computer Science 2021-02-09 Melkior Ornik , Ufuk Topcu

The problem of offline reinforcement learning focuses on learning a good policy from a log of environment interactions. Past efforts for developing algorithms in this area have revolved around introducing constraints to online reinforcement…

Machine Learning · Computer Science 2022-04-27 Ian Char , Viraj Mehta , Adam Villaflor , John M. Dolan , Jeff Schneider

In this work, we introduce a stochastic maximum principle (SMP) approach for solving the reinforcement learning problem with the assumption that the unknowns in the environment can be parameterized based on physics knowledge. For the…

Optimization and Control · Mathematics 2023-06-14 Richard Archibald , Feng Bao , Jiongmin Yong

One of the major challenges in Dynamic Spectrum Access (DSA) systems is to guarantee a required level of Quality of Service (QoS) to secondary users of the spectrum. In this paper, we propose efficient algorithms for deriving optimal…

Information Theory · Computer Science 2016-07-08 Spyridon Vassilaras , George C. Alexandropoulos

This article investigates the problem of dynamic spectrum access for canonical wireless networks, in which the channel states are time-varying. In the most existing work, the commonly used optimization objective is to maximize the…

Information Theory · Computer Science 2017-07-31 Yuhua Xu , Jinlong Wang , Qihui Wu , Jianchao Zheng , Liang Shen , Alagan Anpalagan

In this paper, we consider an information update system where wireless sensor sends timely updates to the destination over a random blocking terahertz channel with the supply of harvested energy and reliable energy backup. The paper aims to…

Information Theory · Computer Science 2021-10-19 Lixin Wang , Fuzhou Peng , Xiang Chen , Shidong Zhou

Wirelessly powered backscatter communication (WPBC) has been identified as a promising technology for low-power communication systems, which can reap the benefits of energy beamforming to improve energy transfer efficiency. Existing studies…

Networking and Internet Architecture · Computer Science 2019-07-30 Wenyuan Ma , Wei Wang , Tao Jiang