中文
相关论文

相关论文: VIL2C: Value-of-Information Aware Low-Latency Comm…

200 篇论文

Multi-agent reinforcement learning (MARL) has recently received considerable attention due to its applicability to a wide range of real-world applications. However, achieving efficient communication among agents has always been an…

机器学习 · 计算机科学 2019-11-04 Sai Qian Zhang , Qi Zhang , Jieyu Lin

Communication is supposed to improve multi-agent collaboration and overall performance in cooperative Multi-agent reinforcement learning (MARL). However, such improvements are prevalently limited in practice since most existing…

多智能体系统 · 计算机科学 2022-12-06 Tingting Yuan , Hwei-Ming Chung , Jie Yuan , Xiaoming Fu

This paper investigates the problem of age of information (AoI) aware radio resource management for a platooning system. Multiple autonomous platoons exploit the cellular wireless vehicle-to-everything (C-V2X) communication technology to…

信号处理 · 电气工程与系统科学 2021-05-11 Mohammad Parvini , Mohammad Reza Javan , Nader Mokari , Bijan Abbasi , Eduard A. Jorswieck

Large Language Model (LLM) agents deployed for real-world tasks face a fundamental dilemma: user requests are underspecified, yet agents must decide whether to act on incomplete information or interrupt users for clarification. Existing…

计算与语言 · 计算机科学 2026-01-13 Yijiang River Dong , Tiancheng Hu , Zheng Hui , Caiqi Zhang , Ivan Vulić , Andreea Bobu , Nigel Collier

Communication is essential for the collective execution of complex tasks by human agents, motivating interest in communication mechanisms for multi-agent reinforcement learning (MARL). However, existing communication protocols in MARL are…

There is a prevalence of multiagent reinforcement learning (MARL) methods that engage in centralized training. But, these methods involve obtaining various types of information from the other agents, which may not be feasible in competitive…

机器学习 · 计算机科学 2023-05-10 Keyang He , Prashant Doshi , Bikramjit Banerjee

In this paper, we study the trade-off between the transmission cost and the control performance of the multi-loop networked control system subject to network-induced delay. Within the linear-quadratic-Gaussian (LQG) framework, the joint…

系统与控制 · 电气工程与系统科学 2021-12-30 Siyi Wang , Qingchen Liu , Precious Ugo Abara , John S. Baras , Sandra Hirche

In recent years, large language models have shown exceptional performance in fulfilling diverse human needs. However, their training data can introduce harmful content, underscoring the necessity for robust value alignment. Mainstream…

人工智能 · 计算机科学 2024-12-19 Rui Zou , Mengqi Wei , Jintian Feng , Qian Wan , Jianwen Sun , Sannyuya Liu

Multi-agent systems (MAS) solve complex problems through coordinated autonomous entities with individual decision-making capabilities. While Multi-Agent Reinforcement Learning (MARL) enables these agents to learn intelligent strategies, it…

多智能体系统 · 计算机科学 2025-10-10 Xinren Zhang , Sixi Cheng , Zixin Zhong , Jiadong Yu

Multi-Agent Systems (MAS) have emerged as a powerful paradigm for modeling complex interactions among autonomous entities in distributed environments. In Multi-Agent Reinforcement Learning (MARL), communication enables coordination but can…

多智能体系统 · 计算机科学 2025-11-13 Xinren Zhang , Jiadong Yu , Zixin Zhong

Learning communication strategies in cooperative multi-agent reinforcement learning (MARL) has recently attracted intensive attention. Early studies typically assumed a fully-connected communication topology among agents, which induces high…

多智能体系统 · 计算机科学 2023-05-24 Xuefeng Wang , Xinran Li , Jiawei Shao , Jun Zhang

The 5G Phase-2 and beyond wireless systems will focus more on vertical applications such as autonomous driving and industrial Internet-of-things, many of which are categorized as ultra-Reliable Low-Latency Communications (uRLLC). In this…

信息论 · 计算机科学 2019-12-04 Zhiyuan Jiang , Siyu Fu , Sheng Zhou , Zhisheng Niu , Shunqing Zhang , Shugong Xu

With the rapid advancement of vehicular communication facilities and autonomous driving technologies, connected vehicle platooning has emerged as a promising approach to improve traffic efficiency and driving safety. Reliable…

系统与控制 · 电气工程与系统科学 2025-08-22 Yaqi Xu , Yan Shi , Jin Tian , Fanzeng Xia , Tongxin Li , Shanzhi Chen , Yuming Ge

Taking inspiration from linguistics, the communications theoretical community has recently shown a significant recent interest in pragmatic , or goal-oriented, communication. In this paper, we tackle the problem of pragmatic communication…

网络与互联网体系结构 · 计算机科学 2023-06-07 Josefine Holm , Federico Chiariotti , Anders E. Kalør , Beatriz Soret , Torben Bach Pedersen , Petar Popovski

Communication is a key component in multi-agent reinforcement learning (MARL) for mitigating partial observability, yet prior approaches often rely on inefficient information exchange or fail to transmit sufficient state information. To…

人工智能 · 计算机科学 2026-05-19 Sangjun Bae , Yisak Park , Sanghyeon Lee , Seungyul Han

Value-based methods of multi-agent reinforcement learning (MARL), especially the value decomposition methods, have been demonstrated on a range of challenging cooperative tasks. However, current methods pay little attention to the…

机器学习 · 计算机科学 2021-02-12 Xiaoteng Ma , Yiqin Yang , Chenghao Li , Yiwen Lu , Qianchuan Zhao , Yang Jun

In this paper, we propose a maximum mutual information (MMI) framework for multi-agent reinforcement learning (MARL) to enable multiple agents to learn coordinated behaviors by regularizing the accumulated return with the mutual information…

多智能体系统 · 计算机科学 2020-06-05 Woojun Kim , Whiyoung Jung , Myungsik Cho , Youngchul Sung

This paper introduces a novel Multi-Agent Reinforcement Learning (MARL) framework to enhance integrated sensing and communication (ISAC) networks using unmanned aerial vehicle (UAV) swarms as sensing radars. By framing the positioning and…

信号处理 · 电气工程与系统科学 2025-01-14 Obed Morrison Atsu , Salmane Naoumi , Roberto Bomfin , Marwa Chafii

Real-world multi-agent reinforcement learning (MARL) systems must often operate under stale observations, stochastic communication delays, and intermittent packet loss. Policies trained under idealized synchronous conditions frequently…

多智能体系统 · 计算机科学 2026-05-27 Maxim Mednikov , Oren Gal

In this paper, a general value of information (VoI) framework is formalised for latent variable models. In particular, the mutual information between the current status at the source node and the observed noisy measurements at the…

信息论 · 计算机科学 2020-08-21 Zijing Wang , Mihai-Alin Badiu , Justin P. Coon
‹ 上一页 1 2 3 10 下一页 ›