English
Related papers

Related papers: Deep Multi-Agent Reinforcement Learning for Cost E…

200 papers

The fog radio access network (F-RAN) is a promising technology in which the user mobile devices (MDs) can offload computation tasks to the nearby fog access points (F-APs). Due to the limited resource of F-APs, it is important to design an…

Machine Learning · Computer Science 2022-06-14 Lingling Zhang , Yanxiang Jiang , Fu-Chun Zheng , Mehdi Bennis , Xiaohu You

This work adopts the very successful distributional perspective on reinforcement learning and adapts it to the continuous control setting. We combine this within a distributed framework for off-policy learning in order to develop what we…

The paper considers a class of multi-agent Markov decision processes (MDPs), in which the network agents respond differently (as manifested by the instantaneous one-stage random costs) to a global controlled state and the control actions of…

Machine Learning · Statistics 2015-06-04 Soummya Kar , Jose' M. F. Moura , H. Vincent Poor

Unmanned Aerial Vehicles (UAVs) promise to become an intrinsic part of next generation communications, as they can be deployed to provide wireless connectivity to ground users to supplement existing terrestrial networks. The majority of the…

Signal Processing · Electrical Eng. & Systems 2021-11-04 Boris Galkin , Babatunji Omoniwa , Ivana Dusparic

Traditional methods plan feasible paths for multiple agents in the stochastic environment. However, the methods' iterations with the changes in the environment result in computation complexities, especially for the decentralized agents…

Robotics · Computer Science 2024-10-28 Qizhen Wu , Kexin Liu , Lei Chen , Jinhu Lü

Planning future operational scenarios of bulk power systems that meet security and economic constraints typically requires intensive labor efforts in performing massive simulations. To automate this process and relieve engineers' burden, a…

Machine Learning · Computer Science 2021-02-18 Xiumin Shang , Jinping Yang , Bingquan Zhu , Lin Ye , Jing Zhang , Jianping Xu , Qin Lyu , Ruisheng Diao

The combination of energy harvesting (EH), cognitive radio (CR), and non-orthogonal multiple access (NOMA) is a promising solution to improve energy efficiency and spectral efficiency of the upcoming beyond fifth generation network (B5G),…

Information Theory · Computer Science 2021-09-21 Zhaoyuan Shi , Xianzhong Xie , Huabing Lu , Helin Yang , Jun Cai , Zhiguo Ding

In modern power systems, frequency regulation is a fundamental prerequisite for ensuring system reliability and assessing the robustness of expansion projects. Conventional feedback control schemes, however, exhibit limited accuracy under…

Systems and Control · Electrical Eng. & Systems 2025-12-05 Amin Masoumi , Mert Korkali

In order to collaborate efficiently with unknown partners in cooperative control settings, adaptation of the partners based on online experience is required. The rather general and widely applicable control setting, where each cooperation…

Multiagent Systems · Computer Science 2019-10-30 Florian Köpf , Samuel Tesfazgi , Michael Flad , Sören Hohmann

The increasing integration of renewable energy sources (RESs) is transforming traditional power grid networks, which require new approaches for managing decentralized energy production and consumption. Microgrids (MGs) provide a promising…

Machine Learning · Computer Science 2025-11-19 Davide Salaorni , Federico Bianchi , Francesco Trovò , Marcello Restelli

The explosive growth of dynamic and heterogeneous data traffic brings great challenges for 5G and beyond mobile networks. To enhance the network capacity and reliability, we propose a learning-based dynamic time-frequency division duplexing…

Machine Learning · Computer Science 2023-03-22 Ziyan Yin , Zhe Wang , Jun Li , Ming Ding , Wen Chen , Shi Jin

Learning in multi-agent systems is highly challenging due to several factors including the non-stationarity introduced by agents' interactions and the combinatorial nature of their state and action spaces. In particular, we consider the…

Machine Learning · Statistics 2023-05-10 Barna Pásztor , Ilija Bogunovic , Andreas Krause

Classical paradigms for distributed learning, such as federated or decentralized gradient descent, employ consensus mechanisms to enforce homogeneity among agents. While these strategies have proven effective in i.i.d. scenarios, they can…

Machine Learning · Computer Science 2023-04-18 Shreya Wadehra , Roula Nassif , Stefan Vlaski

In mobile edge computing systems, an edge node may have a high load when a large number of mobile devices offload their tasks to it. Those offloaded tasks may experience large processing delay or even be dropped when their deadlines expire.…

Networking and Internet Architecture · Computer Science 2020-05-07 Ming Tang , Vincent W. S. Wong

Heave compensation is an essential part in various offshore operations. It is used in various applications, which include on-loading or off-loading systems, offshore drilling, landing helicopter on oscillating structures, and deploying and…

Systems and Control · Electrical Eng. & Systems 2021-07-26 Shrenik Zinage , Abhilash Somayajula

Underfrequency load shedding (UFLS) is a critical control strategy in power systems aimed at maintaining system stability and preventing blackouts during severe frequency drops. Traditional UFLS schemes often rely on predefined rules and…

Systems and Control · Electrical Eng. & Systems 2024-10-08 Glory Justin , Santiago Paternain

Mapping deep neural networks (DNNs) to hardware is critical for optimizing latency, energy consumption, and resource utilization, making it a cornerstone of high-performance accelerator design. Due to the vast and complex mapping space,…

Manipulate and control of the complex quantum system with high precision are essential for achieving universal fault tolerant quantum computing. For a physical system with restricted control resources, it is a challenge to control the…

Quantum Physics · Physics 2021-01-20 Zheng An , Qi-Kai He , Hai-Jing Song , D. L. Zhou

This paper proposes a reinforcement learning-based approach for optimal transient frequency control in power systems with stability and safety guarantees. Building on Lyapunov stability theory and safety-critical control, we derive…

Systems and Control · Electrical Eng. & Systems 2024-02-22 Zhenyi Yuan , Changhong Zhao , Jorge Cortes

We study the process of multi-agent reinforcement learning in the context of load balancing in a distributed system, without use of either central coordination or explicit communication. We first define a precise framework in which to study…

Artificial Intelligence · Computer Science 2014-11-17 A. Schaerf , Y. Shoham , M. Tennenholtz