English
Related papers

Related papers: Multi-Agent Q-Learning for Minimizing Demand-Suppl…

200 papers

We study a class of sequential decision-making problems with augmented predictions, potentially provided by a machine learning algorithm. In this setting, the decision-maker receives prediction intervals for unknown parameters that become…

Machine Learning · Computer Science 2025-05-05 Xin Chen , Yuze Chen , Yuan Zhou

It is challenging for a security analyst to detect or defend against cyber-attacks. Moreover, traditional defense deployment methods require the security analyst to manually enforce the defenses in the presence of uncertainties about the…

Cryptography and Security · Computer Science 2022-07-14 Xiaofan Zhou , Simon Yusuf Enoch , Dong Seong Kim

This paper addresses the energy management of a grid-connected renewable generation plant coupled with a battery energy storage device in the capacity firming market, designed to promote renewable power generation facilities in small…

Reinforcement learning usually assumes a given or sometimes even fixed environment in which an agent seeks an optimal policy to maximize its long-term discounted reward. In contrast, we consider agents that are not limited to passive…

Machine Learning · Computer Science 2025-10-20 Ziqing Lu , Babak Hassibi , Lifeng Lai , Weiyu Xu

Advances in microgrids powered by Distributed Energy Resources (DERs) make them an attractive response capability for improving the resilience of electricity distribution networks (DNs). This paper presents an approach to evaluate the value…

Optimization and Control · Mathematics 2019-05-13 Devendra Shelar , Saurabh Amin , Ian Hiskens

The deployment of ultra-dense networks is one of the main methods to meet the 5G data rate requirements. However, high density of independent small base stations (SBSs) will increase the interference within the network. To circumvent this…

Signal Processing · Electrical Eng. & Systems 2018-12-27 Roohollah Amiri , Hani Mehrpouyan , David Matolak , Maged Elkashlan

In this paper, we propose a distributed reinforcement learning (RL) technique called distributed power control using Q-learning (DPC-Q) to manage the interference caused by the femtocells on macro-users in the downlink. The DPC-Q leverages…

Machine Learning · Computer Science 2012-03-20 Hussein Saad , Amr Mohamed , Tamer ElBatt

In device-to-device (D2D) communication under a cell with resource sharing mode the spectrum resource utilization of the system will be improved. However, if the interference generated by the D2D user is not controlled, the performance of…

Networking and Internet Architecture · Computer Science 2025-11-04 Shi Gengtian , Takashi Koshimizu , Megumi Saito , Pan Zhenni , Liu Jiang , Shigeru Shimamoto

This work advocates the use of deep learning to perform max-min and max-prod power allocation in the downlink of Massive MIMO networks. More precisely, a deep neural network is trained to learn the map between the positions of user…

Signal Processing · Electrical Eng. & Systems 2019-06-04 Luca Sanguinetti , Alessio Zappone , Merouane Debbah

Battery energy storage systems are providing increasing level of benefits to power grid operations by decreasing the resource uncertainty and supporting frequency regulation. Thus, it is crucial to obtain the optimal policy for battery to…

Optimization and Control · Mathematics 2022-06-06 Kyung-bin Kwon , Hao Zhu

We apply Reinforcement Learning algorithms to solve the classic quantitative finance Market Making problem, in which an agent provides liquidity to the market by placing buy and sell orders while maximizing a utility function. The optimal…

Machine Learning · Computer Science 2021-04-12 Matias Selser , Javier Kreiner , Manuel Maurette

In the paper the joint optimization of uplink multiuser power and resource block (RB) allocation are studied, where each user has quality of service (QoS) constraints on both long- and short-blocklength transmissions. The objective is to…

Signal Processing · Electrical Eng. & Systems 2025-03-11 Manru Yin , Shengqian Han , Chenyang Yang

General purpose intelligent learning agents cycle through (complex,non-MDP) sequences of observations, actions, and rewards. On the other hand, reinforcement learning is well-developed for small finite state Markov Decision Processes…

Artificial Intelligence · Computer Science 2009-12-30 Marcus Hutter

The extraordinary electric vehicle (EV) popularization in the recent years has facilitated research studies in alleviating EV energy charging demand. Previous studies primarily focused on the optimizations over charging stations (CS) profit…

Multiagent Systems · Computer Science 2024-06-18 Tianhao Bu , Hang Li , Guojie Li

The integration of renewable energy resources in rural areas, such as dairy farming communities, enables decentralized energy management through Peer-to-Peer (P2P) energy trading. This research highlights the role of P2P trading in…

Artificial Intelligence · Computer Science 2025-12-01 Mian Ibad Ali Shah , Marcos Eduardo Cruz Victorio , Maeve Duffy , Enda Barrett , Karl Mason

We present a framework to address a class of sequential decision making problems. Our framework features learning the optimal control policy with robustness to noisy data, determining the unknown state and action parameters, and performing…

Machine Learning · Computer Science 2022-01-20 Amber Srivastava , Srinivasa M Salapaka

Achieving the economical and stable operation of Multi-microgrids (MMG) systems is vital. However, there are still some challenging problems to be solved. Firstly, from the perspective of stable operation, it is necessary to minimize the…

Systems and Control · Electrical Eng. & Systems 2023-07-03 Yijian Wang , Yang Cui , Yang Li , Yang Xu

Q-learning is a popular reinforcement learning algorithm. This algorithm has however been studied and analysed mainly in the infinite horizon setting. There are several important applications which can be modeled in the framework of finite…

Machine Learning · Computer Science 2022-08-09 Vivek VP , Dr. Shalabh Bhatnagar

In this work we solve the day-ahead unit commitment (UC) problem, by formulating it as a Markov decision process (MDP) and finding a low-cost policy for generation scheduling. We present two reinforcement learning algorithms, and devise a…

Artificial Intelligence · Computer Science 2016-11-17 Gal Dalal , Shie Mannor

Despite high reliability, modern power systems with growing renewable penetration face an increasing risk of cascading outages. Real-time cascade mitigation requires fast, complex operational decisions under uncertainty. In this work, we…

Physics and Society · Physics 2025-06-11 Kai Zhou , Youbiao He , Chong Zhong , Yifu Wu