English
Related papers

Related papers: Multi-Agent Pointer Transformer: Seq-to-Seq Reinfo…

200 papers

In this paper a deep reinforcement based multi-agent path planning approach is introduced. The experiments are realized in a simulation environment and in this environment different multi-agent path planning problems are produced. The…

Machine Learning · Computer Science 2021-10-05 Mert Çetinkaya

Many real-world multi-agent systems exhibit nonlinear dynamics and complex inter-agent interactions. As these systems increase in scale, the main challenges arise from achieving scalability and handling nonconvexity. To address these…

Optimization and Control · Mathematics 2025-10-22 Taehyun Yoon , Augustinos D. Saravanos , Evangelos A. Theodorou

We present a fully decentralized routing framework for multi-robot exploration missions operating under the constraints of a Lunar Delay-Tolerant Network (LDTN). In this setting, autonomous rovers must relay collected data to a lander under…

Machine Learning · Statistics 2025-10-24 Federico Lozano-Cuadra , Beatriz Soret , Marc Sanchez Net , Abhishek Cauligi , Federico Rossi

Transmission interface power flow adjustment is a critical measure to ensure the security and economy operation of power systems. However, conventional model-based adjustment schemes are limited by the increasing variations and…

Systems and Control · Electrical Eng. & Systems 2024-05-28 Shunyu Liu , Wei Luo , Yanzhen Zhou , Kaixuan Chen , Quan Zhang , Huating Xu , Qinglai Guo , Mingli Song

General-purpose robotic systems must master a large repertoire of diverse skills to be useful in a range of daily tasks. While reinforcement learning provides a powerful framework for acquiring individual behaviors, the time needed to…

We present a novel deep reinforcement learning method to learn construction heuristics for vehicle routing problems. In specific, we propose a Multi-Decoder Attention Model (MDAM) to train multiple diverse policies, which effectively…

Machine Learning · Computer Science 2020-12-22 Liang Xin , Wen Song , Zhiguang Cao , Jie Zhang

Multi Agent Path Finding (MAPF) requires identification of conflict free paths for agents which could be point-sized or with dimensions. In this paper, we propose an approach for MAPF for spatially-extended agents. These find application in…

Multiagent Systems · Computer Science 2021-06-10 Shyni Thomas , M. Narasimha Murty

The coordination of large-scale, decentralised systems, such as a fleet of Electric Vehicles (EVs) in a Vehicle-to-Grid (V2G) network, presents a significant challenge for modern control systems. While collaborative Digital Twins have been…

Distributed, Parallel, and Cluster Computing · Computer Science 2026-04-01 Zhengchang Hua , Panagiotis Oikonomou , Karim Djemame , Nikos Tziritas , Georgios Theodoropoulos

The profiled vehicle routing problem (PVRP) is a generalization of the heterogeneous capacitated vehicle routing problem (HCVRP) in which the objective is to optimize the routes of vehicles to serve client demands subject to different…

Multiagent Systems · Computer Science 2025-02-05 Chuanbo Hua , Federico Berto , Jiwoo Son , Seunghyun Kang , Changhyun Kwon , Jinkyoo Park

In this paper, we study a courier dispatching problem (CDP) raised from an online pickup-service platform of Alibaba. The CDP aims to assign a set of couriers to serve pickup requests with stochastic spatial and temporal arrival rate among…

Artificial Intelligence · Computer Science 2019-03-08 Yujie Chen , Yu Qian , Yichen Yao , Zili Wu , Rongqi Li , Yinzhi Zhou , Haoyuan Hu , Yinghui Xu

The Entrance Dependent Vehicle Routing Problem (EDVRP) is a variant of the Vehicle Routing Problem (VRP) where the scale of cities influences routing outcomes, necessitating consideration of their entrances. This paper addresses EDVRP in…

Robotics · Computer Science 2025-03-05 Yixuan Fan , Haotian Xu , Mengqiao Liu , Qing Zhuo , Tao Zhang

This paper focuses on dynamic origin-destination matrix estimation (DODE), a crucial calibration process necessary for the effective application of microscopic traffic simulations. The fundamental challenge of the DODE problem in…

Machine Learning · Computer Science 2026-03-26 Donggyu Min , Seongjin Choi , Dong-Kyu Kim

Recent advances in multi-agent reinforcement learning have been largely limited in training one model from scratch for every new task. The limitation is due to the restricted model architecture related to fixed input and output dimensions.…

Machine Learning · Computer Science 2021-02-09 Siyi Hu , Fengda Zhu , Xiaojun Chang , Xiaodan Liang

Motivated by the promising advances of deep-reinforcement learning (DRL) applied to cooperative multi-agent systems we propose a model and learning procedure to solve the Capacitated Multi-Vehicle Routing Problem (CMVRP) with fixed fleet…

Neural and Evolutionary Computing · Computer Science 2019-12-10 Jose Manuel Vera , Andres G. Abad

Multi-Agent Pickup and Delivery (MAPD) is a fundamental problem in robotics, particularly in applications such as warehouse automation and logistics. Existing solutions often face challenges in scalability, adaptability, and efficiency,…

Robotics · Computer Science 2025-04-22 Kushal Shah , Jihyun Park , Seung-Kyum Choi

Active tracking of space noncooperative object that merely relies on vision camera is greatly significant for autonomous rendezvous and debris removal. Considering its Partial Observable Markov Decision Process (POMDP) property, this paper…

Robotics · Computer Science 2023-01-02 Dong Zhou , Guanghui Sun , Zhao Zhang , Ligang Wu

In automated warehouses, teams of mobile robots fulfill the packaging process by transferring inventory pods to designated workstations while navigating narrow aisles formed by tightly packed pods. This problem is typically modeled as a…

Artificial Intelligence · Computer Science 2023-05-22 David Vainshtein , Yaakov Sherma , Kiril Solovey , Oren Salzman

Although multi-tier vehicular Metaverse promises to transform vehicles into essential nodes -- within an interconnected digital ecosystem -- using efficient resource allocation and seamless vehicular twin (VT) migration, this can hardly be…

Networking and Internet Architecture · Computer Science 2025-02-27 Nahom Abishu Hayla , A. Mohammed Seid , Aiman Erbad , Tilahun M. Getu , Ala Al-Fuqaha , Mohsen Guizani

Sequential Recommendation (SR) captures users' dynamic preferences by modeling how users transit among items. However, SR models that utilize only single type of behavior interaction data encounter performance degradation when the sequences…

Information Retrieval · Computer Science 2024-02-23 Jiajie Su , Chaochao Chen , Zibin Lin , Xi Li , Weiming Liu , Xiaolin Zheng

This paper proposes a multi-agent reinforcement learning based medium access framework for wireless networks. The access problem is formulated as a Markov Decision Process (MDP), and solved using reinforcement learning with every network…

Machine Learning · Computer Science 2021-04-30 Hrishikesh Dutta , Subir Biswas