中文
相关论文

相关论文: Learning to Guide and to Be Guided in the Architec…

200 篇论文

In this work, we propose a novel memory-based multi-agent meta-learning architecture and learning procedure that allows for learning of a shared communication policy that enables the emergence of rapid adaptation to new and unseen…

Previous research into agent communication has shown that a pre-trained guide can speed up the learning process of an imitation learning agent. The guide achieves this by providing the agent with discrete messages in an emerged language…

In numerous artificial intelligence applications, the collaborative efforts of multiple intelligent agents are imperative for the successful attainment of target objectives. To enhance coordination among these agents, a distributed…

机器学习 · 计算机科学 2024-05-15 Shengchao Hu , Li Shen , Ya Zhang , Dacheng Tao

We examine the problem of learning to cooperate in the context of wireless communication. In our setting, two agents must learn modulation schemes that enable them to communicate across a power-constrained additive white Gaussian noise…

信号处理 · 电气工程与系统科学 2020-04-03 Anant Sahai , Joshua Sanz , Vignesh Subramanian , Caryn Tran , Kailas Vodrahalli

Today's AI models learn primarily through mimicry and refining, so it is not surprising that they struggle to solve problems beyond the limits set by existing data. To solve novel problems, agents should acquire skills for exploring and…

In numerous artificial intelligence applications, the collaborative efforts of multiple intelligent agents are imperative for the successful attainment of target objectives. To enhance coordination among these agents, a distributed…

机器学习 · 计算机科学 2024-11-04 Shengchao Hu , Li Shen , Ya Zhang , Dacheng Tao

An increasing number of emerging applications, e.g., internet of things, vehicular communications, augmented reality, and the growing complexity due to the interoperability requirements of these systems, lead to the need to change the tools…

多智能体系统 · 计算机科学 2019-01-16 Merim Dzaferagic , M. Majid Butt , Maria Murphy , Nicholas Kaminski , Nicola Marchetti

Robust coordination is critical for effective decision-making in multi-agent systems, especially under partial observability. A central question in Multi-Agent Reinforcement Learning (MARL) is whether to engineer communication protocols or…

多智能体系统 · 计算机科学 2025-11-25 Brennen A. Hill , Mant Koh En Wei , Thangavel Jishnuanandh

Albrecht and Stone (2018) state that modeling of changing behaviors remains an open problem "due to the essentially unconstrained nature of what other agents may do". In this work we evaluate the adaptability of neural artificial agents…

计算与语言 · 计算机科学 2024-02-08 Philipp Sadler , Sherzod Hakimov , David Schlangen

Traditional AI reasoning techniques have been used successfully in many domains, including logistics, scheduling and game playing. This paper is part of a project aimed at investigating how such techniques can be extended to coordinate…

人工智能 · 计算机科学 2014-05-07 Marcello Balduccini , William C. Regli , Duc N. Nguyen

We propose an active learning architecture for robots, capable of organizing its learning process to achieve a field of complex tasks by learning sequences of motor policies, called Intrinsically Motivated Procedure Babbling (IM-PB). The…

人机交互 · 计算机科学 2019-02-18 Nicolas Duminy , Sao Mai Nguyen , Dominique Duhaut

The growing availability of building operational data motivates the use of reinforcement learning (RL), which can learn control policies directly from data and cope with the complexity and uncertainty of large-scale building clusters.…

人工智能 · 计算机科学 2026-03-30 Borui Zhang , Nariman Mahdavi , Subbu Sethuvenkatraman , Shuang Ao , Flora Salim

Traditional AI reasoning techniques have been used successfully in many domains, including logistics, scheduling and game playing. This paper is part of a project aimed at investigating how such techniques can be extended to coordinate…

人工智能 · 计算机科学 2014-05-22 Marcello Balduccini , William C. Regli , Duc N. Nguyen

Emergent communication offers insight into how agents develop shared structured representations, yet most research assumes homogeneous modalities or aligned representational spaces, overlooking the perceptual heterogeneity of real-world…

多智能体系统 · 计算机科学 2026-01-30 Naomi Pitzer , Daniela Mihai

Intelligent tutoring systems (ITS) are effective for improving students' learning outcomes. However, their development is often complex, time-consuming, and requires specialized programming and tutor design knowledge, thus hindering their…

人机交互 · 计算机科学 2024-04-12 Glen Smith , Adit Gupta , Christopher MacLellan

The human-agent team, which is a problem in which humans and autonomous agents collaborate to achieve one task, is typical in human-AI collaboration. For effective collaboration, humans want to have an effective plan, but in realistic…

人工智能 · 计算机科学 2021-09-02 Ryo Nakahashi , Seiji Yamada

Many real-world problems require the coordination of multiple autonomous agents. Recent work has shown the promise of Graph Neural Networks (GNNs) to learn explicit communication strategies that enable complex multi-agent coordination.…

机器人学 · 计算机科学 2020-11-05 Jan Blumenkamp , Amanda Prorok

An assignment problem arises when there exists a set of tasks that must be allocated to a set of agents. The bottleneck assignment problem (BAP) has the objective of minimising the most costly allocation of a task to an agent. Under certain…

最优化与控制 · 数学 2020-08-26 Mitchell Khoo , Tony A. Wood , Chris Manzie , Iman Shames

Many real-world tasks require agents to coordinate their behavior to achieve shared goals. Successful collaboration requires not only adopting the same communicative conventions, but also grounding these conventions in the same…

计算与语言 · 计算机科学 2021-07-02 William P. McCarthy , Robert D. Hawkins , Haoliang Wang , Cameron Holdaway , Judith E. Fan

Ensuring artificial intelligence behaves in such a way that is aligned with human values is commonly referred to as the alignment challenge. Prior work has shown that rational agents, behaving in such a way that maximizes a utility…

人工智能 · 计算机科学 2024-02-16 Paulo Garcia
‹ 上一页 1 2 3 10 下一页 ›