中文
相关论文

相关论文: Reinforcement Learning based Multi-connectivity Re…

200 篇论文

The development of urban-air-mobility (UAM) is rapidly progressing with spurs, and the demand for efficient transportation management systems is a rising need due to the multifaceted environmental uncertainties. Thus, this paper proposes a…

多智能体系统 · 计算机科学 2023-06-08 Chanyoung Park , Gyu Seon Kim , Soohyun Park , Soyi Jung , Joongheon Kim

In typical multi-agent reinforcement learning (MARL) problems, communication is important for agents to share information and make the right decisions. However, due to the complexity of training multi-agent communication, existing methods…

多智能体系统 · 计算机科学 2025-05-01 Xuyan Ma , Yawen Wang , Junjie Wang , Xiaofei Xie , Boyu Wu , Shoubin Li , Fanjiang Xu , Qing Wang

Reinforcement Learning (RL) is a potent tool for sequential decision-making and has achieved performance surpassing human capabilities across many challenging real-world tasks. As the extension of RL in the multi-agent system domain,…

Search and Rescue (SAR) missions in remote environments often employ autonomous multi-robot systems that learn, plan, and execute a combination of local single-robot control actions, group primitives, and global mission-oriented…

In this paper, we introduce an alternative approach to enhancing Multi-Agent Reinforcement Learning (MARL) through the integration of domain knowledge and attention-based policy mechanisms. Our methodology focuses on the incorporation of…

机器学习 · 计算机科学 2025-04-04 Andre R Kuroswiski , Annie S Wu , Angelo Passaro

In multi-agent deep reinforcement learning (MADRL), agents can communicate with one another to perform a task in a coordinated manner. When multiple tasks are involved, agents can also leverage knowledge from one task to improve learning in…

多智能体系统 · 计算机科学 2025-11-07 Changxi Zhu , Mehdi Dastani , Shihan Wang

Purpose of review: Recent advances in sensing, actuation, and computation have opened the door to multi-robot systems consisting of hundreds/thousands of robots, with promising applications to automated manufacturing, disaster relief,…

机器人学 · 计算机科学 2022-04-08 Yutong Wang , Mehul Damani , Pamela Wang , Yuhong Cao , Guillaume Sartoretti

The Intelligent Transportation System (ITS) environment is known to be dynamic and distributed, where participants (vehicle users, operators, etc.) have multiple, changing and possibly conflicting objectives. Although Reinforcement Learning…

机器学习 · 计算机科学 2024-03-19 Jing Tan , Ramin Khalili , Holger Karl

This paper investigates an interference-aware joint path planning and power allocation mechanism for a cellular-connected unmanned aerial vehicle (UAV) in a sparse suburban environment. The UAV's goal is to fly from an initial point and…

机器学习 · 计算机科学 2023-06-21 Alireza Shamsoshoara , Fatemeh Lotfi , Sajad Mousavi , Fatemeh Afghah , Ismail Guvenc

Traffic optimization challenges, such as load balancing, flow scheduling, and improving packet delivery time, are difficult online decision-making problems in wide area networks (WAN). Complex heuristics are needed for instance to find…

网络与互联网体系结构 · 计算机科学 2021-12-01 Shan Sun , Mariam Kiran , Wei Ren

Next-generation autonomous and networked industrial systems (i.e., robots, vehicles, drones) have driven advances in ultra-reliable, low latency communications (URLLC) and computing. These networked multi-agent systems require fast,…

机器学习 · 计算机科学 2021-04-20 Stefano Savazzi , Monica Nicoli , Mehdi Bennis , Sanaz Kianoush , Luca Barbieri

Reinforcement Learning (RL) in Traffic Signal Control (TSC) faces significant hurdles in real-world deployment due to limited generalization to dynamic traffic flow variations. Existing approaches often overfit static patterns and use…

Harvesting data from distributed Internet of Things (IoT) devices with multiple autonomous unmanned aerial vehicles (UAVs) is a challenging problem requiring flexible path planning methods. We propose a multi-agent reinforcement learning…

多智能体系统 · 计算机科学 2021-06-04 Harald Bayerlein , Mirco Theile , Marco Caccamo , David Gesbert

Traffic signal control (TSC) is a challenging problem within intelligent transportation systems and has been tackled using multi-agent reinforcement learning (MARL). While centralized approaches are often infeasible for large-scale TSC…

多智能体系统 · 计算机科学 2023-10-05 Rohit Bokade , Xiaoning Jin , Christopher Amato

In the field of precision manufacturing in complex constrained environments, the role of soft robots is increasingly prominent, and the realization of anti-winding control based on multi-intelligent body reinforcement learning has become a…

机器人学 · 计算机科学 2026-05-08 Haoyang Le , Shengxuan Wang , Mohan Chen , Shuo Feng

Multi-Agent Reinforcement Learning (MARL) has become a powerful framework for numerous real-world applications, modeling distributed decision-making and learning from interactions with complex environments. Resource Allocation Optimization…

多智能体系统 · 计算机科学 2025-05-01 Mohamad A. Hady , Siyi Hu , Mahardhika Pratama , Jimmy Cao , Ryszard Kowalczyk

Efficient load balancing is crucial in cloud computing environments to ensure optimal resource utilization, minimize response times, and prevent server overload. Traditional load balancing algorithms, such as round-robin or least…

分布式、并行与集群计算 · 计算机科学 2024-09-10 Kavish Chawla

The huge research interest in cellular vehicle-to-everything (C-V2X) communications in recent days is attributed to their ability to schedule multiple access more efficiently as compared to its predecessor technology, i.e., dedicated…

网络与互联网体系结构 · 计算机科学 2021-01-27 Seungmo Kim , Byung-Jun Kim , B. Brian Park

Effective solutions for intelligent data collection in terrestrial cellular networks are crucial, especially in the context of Internet of Things applications. The limited spectrum and coverage area of terrestrial base stations pose…

系统与控制 · 电气工程与系统科学 2024-06-04 Abhishek Mondal , Deepak Mishra , Ganesh Prasad , George C. Alexandropoulos , Azzam Alnahari , Riku Jantti

In URLLC, short packet transmission is adopted to reduce latency, such that conventional Shannon's capacity formula is no longer applicable, and the achievable data rate in finite blocklength becomes a complex expression with respect to the…

信号处理 · 电气工程与系统科学 2019-12-02 Hong Ren , Cunhua Pan , Yansha Deng , Maged Elkashlan , Arumugam Nallanathan