English
Related papers

Related papers: Muti-Agent Proximal Policy Optimization For Data F…

200 papers

Due to flexibility, autonomy and low operational cost, unmanned aerial vehicles (UAVs), as fixed aerial base stations, are increasingly being used as \textit{relays} to collect time-sensitive information (i.e., status updates) from IoT…

Networking and Internet Architecture · Computer Science 2021-09-28 Biplav Choudhury , Vijay K. Shah , Aidin Ferdowsi , Jeffrey H. Reed , Y. Thomas Hou

In recent years, there has been a growing interest in using networks of Unmanned Aerial Vehicles (UAV) that collectively perform complex tasks for diverse applications. An important challenge in realizing UAV networks is the need for a…

Systems and Control · Computer Science 2017-11-01 Abolfazl Razi , Fatemeh Afghah , Jacob Chakareski

Proximal Policy Optimization (PPO) is among the most widely used algorithms in reinforcement learning, which achieves state-of-the-art performance in many challenging problems. The keys to its success are the reliable policy updates through…

Machine Learning · Computer Science 2021-07-02 Mónika Farsang , Luca Szegletes

This paper summarizes in depth the state of the art of aerial swarms, covering both classical and new reinforcement-learning-based approaches for their management. Then, it proposes a hybrid AI system, integrating deep reinforcement…

Artificial Intelligence · Computer Science 2025-01-16 Raúl Arranz , David Carramiñana , Gonzalo de Miguel , Juan A. Besada , Ana M. Bernardos

A novel deep multi-agent reinforcement learning framework is proposed to identify and resolve conflicts among a variable number of aircraft in a high-density, stochastic, and dynamic sector. Currently the sector capacity is constrained by…

Machine Learning · Computer Science 2020-08-28 Marc Brittain , Xuxi Yang , Peng Wei

To obtain a near-optimal policy with fewer interactions in Reinforcement Learning (RL), a promising approach involves the combination of offline RL, which enhances sample efficiency by leveraging offline datasets, and online RL, which…

Machine Learning · Computer Science 2024-11-18 Xiaoyu Wen , Xudong Yu , Rui Yang , Haoyuan Chen , Chenjia Bai , Zhen Wang

Recently, Unmanned Aerial Vehicles (UAVs) have attracted the attention of researchers in academia and industry for providing wireless services to ground users in diverse scenarios like festivals, large sporting events, natural and man-made…

Signal Processing · Electrical Eng. & Systems 2024-06-26 Marwan Dhuheir , Aiman Erbad , Ala Al-Fuqaha , Mohsen Guizani

The exponential growth of Low Earth Orbit (LEO) satellites has revolutionised Earth Observation (EO) missions, addressing challenges in climate monitoring, disaster management, and more. However, autonomous coordination in multi-satellite…

Artificial Intelligence · Computer Science 2025-11-06 Mohamad A. Hady , Siyi Hu , Mahardhika Pratama , Jimmy Cao , Ryszard Kowalczyk

Offline reinforcement learning (RL), also known as batch RL, aims to optimize policy from a large pre-recorded dataset without interaction with the environment. This setting offers the promise of utilizing diverse, pre-collected datasets to…

Machine Learning · Computer Science 2021-01-05 Qiang He , Xinwen Hou

In this paper, a novel machine learning (ML) framework is proposed for enabling a predictive, efficient deployment of unmanned aerial vehicles (UAVs), acting as aerial base stations (BSs), to provide on-demand wireless service to cellular…

Signal Processing · Electrical Eng. & Systems 2018-05-02 Qianqian Zhang , Mohammad Mozaffari , Walid Saad , Mehdi Bennis , Merouane Debbah

Multi-Agent Reinforcement Learning (MARL) has become a powerful framework for numerous real-world applications, modeling distributed decision-making and learning from interactions with complex environments. Resource Allocation Optimization…

Multiagent Systems · Computer Science 2025-05-01 Mohamad A. Hady , Siyi Hu , Mahardhika Pratama , Jimmy Cao , Ryszard Kowalczyk

The integration of Unmanned Aerial Vehicles (UAVs) into Open Radio Access Networks (O-RAN) enhances communication in disaster management and Search and Rescue (SAR) operations by ensuring connectivity when infrastructure fails. However, SAR…

Cryptography and Security · Computer Science 2025-10-22 Zaineh Abughazzah , Emna Baccour , Loay Ismail , Amr Mohamed , Mounir Hamdi

In this letter, a novel framework to deliver critical spread out URLLC services deploying unmanned aerial vehicles (UAVs) in an out-of-coverage area is developed. To this end, the resource optimization problem, i.e., resource blocks (RBs)…

Networking and Internet Architecture · Computer Science 2021-04-13 Shashi Raj Pandey , Kitae Kim , Madyan Alsenwi , Yan Kyaw Tun , Zhu Han , Choong Seon Hong

Large language models (LLMs) increasingly rely on multi-turn tool-integrated planning for knowledge-intensive and complex reasoning tasks. Existing implementations typically rely on a single agent, but they suffer from limited context…

Computation and Language · Computer Science 2025-10-07 Zhanfeng Mo , Xingxuan Li , Yuntao Chen , Lidong Bing

Training multiple agents to coordinate is an essential problem with applications in robotics, game theory, economics, and social sciences. However, most existing Multi-Agent Reinforcement Learning (MARL) methods are online and thus…

Machine Learning · Computer Science 2024-01-19 Paul Barde , Jakob Foerster , Derek Nowrouzezahrai , Amy Zhang

Decentralized cooperative pursuit in cluttered environments is challenging for autonomous aerial swarms, especially under partial and noisy perception. Existing methods often rely on abstracted geometric features or privileged ground-truth…

Robotics · Computer Science 2026-03-26 Yude Li , Zhexuan Zhou , Huizhe Li , Yanke Sun , Yenan Wu , Yichen Lai , Yiming Wang , Youmin Gong , Jie Mei

This paper investigates the problem of age of information (AoI) aware radio resource management for a platooning system. Multiple autonomous platoons exploit the cellular wireless vehicle-to-everything (C-V2X) communication technology to…

Signal Processing · Electrical Eng. & Systems 2021-05-11 Mohammad Parvini , Mohammad Reza Javan , Nader Mokari , Bijan Abbasi , Eduard A. Jorswieck

With the rapid growth of the low-altitude economy, there is increasing demand for real-time data collection using UAV-assisted wireless sensor networks. This paper investigates the problem of minimizing the age of information (AoI) in…

Networking and Internet Architecture · Computer Science 2025-07-10 Bisheng Wei , Ruichen Zhang , Ruihong Jiang , Mugen Peng , Dusit Niyato

Large language models (LLMs) are versatile, yet their deployment in complex real-world settings is limited by static knowledge cutoffs and the difficulty of producing controllable behavior within a single inference. Multi-agent search…

Machine Learning · Computer Science 2026-04-21 Guanzhong Chen , Shaoxiong Yang , Chao Li , Wei Liu , Jian Luan , Zenglin Xu

Offline reinforcement learning (RL) aims to find optimal policies in dynamic environments in order to maximize the expected total rewards by leveraging pre-collected data. Learning from heterogeneous data is one of the fundamental…

Machine Learning · Statistics 2026-03-10 Rui Miao , Babak Shahbaba , Annie Qu