中文
相关论文

相关论文: K-level Reasoning for Zero-Shot Coordination in Ha…

200 篇论文

One of the fundamental quests of AI is to produce agents that coordinate well with humans. This problem is challenging, especially in domains that lack high quality human behavioral data, because multi-agent reinforcement learning (RL)…

人工智能 · 计算机科学 2023-06-13 Hengyuan Hu , Dorsa Sadigh

Zero-shot coordination (ZSC) is a popular setting for studying the ability of reinforcement learning (RL) agents to coordinate with novel partners. Prior ZSC formulations assume the $\textit{problem setting}$ is common knowledge: each agent…

机器学习 · 计算机科学 2024-11-08 Usman Anwar , Ashish Pandian , Jia Wan , David Krueger , Jakob Foerster

AI agents hold the potential to transform everyday life by helping humans achieve their goals. To do this successfully, agents need to be able to coordinate with novel partners without prior interaction, a setting known as zero-shot…

人工智能 · 计算机科学 2025-03-25 Tobias Gessler , Tin Dizdarevic , Ani Calinescu , Benjamin Ellis , Andrei Lupu , Jakob Nicolaus Foerster

We present the task of "Social Rearrangement", consisting of cooperative everyday tasks like setting up the dinner table, tidying a house or unpacking groceries in a simulated multi-agent environment. In Social Rearrangement, two robots…

机器学习 · 计算机科学 2023-06-02 Andrew Szot , Unnat Jain , Dhruv Batra , Zsolt Kira , Ruta Desai , Akshara Rai

Zero-shot coordination (ZSC) is a new cooperative multi-agent reinforcement learning (MARL) challenge that aims to train an ego agent to work with diverse, unseen partners during deployment. The significant difference between the…

人工智能 · 计算机科学 2024-09-27 Xihuai Wang , Shao Zhang , Wenhao Zhang , Wentao Dong , Jingxiao Chen , Ying Wen , Weinan Zhang

While we would like agents that can coordinate with humans, current algorithms such as self-play and population-based training create agents that can coordinate with themselves. Agents that assume their partner to be optimal or similar to…

机器学习 · 计算机科学 2020-01-10 Micah Carroll , Rohin Shah , Mark K. Ho , Thomas L. Griffiths , Sanjit A. Seshia , Pieter Abbeel , Anca Dragan

Zero-shot coordination (ZSC) is a significant challenge in multi-agent collaboration, aiming to develop agents that can coordinate with unseen partners they have not encountered before. Recent cutting-edge ZSC methods have primarily focused…

机器人学 · 计算机科学 2024-10-02 Yang Li , Dengyu Zhang , Junfan Chen , Ying Wen , Qingrui Zhang , Shaoshuai Mou , Wei Pan

The partially observable card game Hanabi has recently been proposed as a new AI challenge problem due to its dependence on implicit communication conventions and apparent necessity of theory of mind reasoning for efficient play. In this…

人工智能 · 计算机科学 2021-01-26 Andrew Fuchs , Michael Walton , Theresa Chadwick , Doug Lange

Hanabi is a cooperative game that brings the problem of modeling other players to the forefront. In this game, coordinated groups of players can leverage pre-established conventions to great effect, but playing in an ad-hoc setting requires…

人工智能 · 计算机科学 2022-08-31 Rodrigo Canaan , Xianbo Gao , Julian Togelius , Andy Nealen , Stefan Menzel

Existing language agents often encounter difficulties in dynamic adversarial games due to poor strategic reasoning. To mitigate this limitation, a promising approach is to allow agents to learn from game interactions automatically, without…

计算与语言 · 计算机科学 2025-10-21 Yikai Zhang , Ye Rong , Siyu Yuan , Jiangjie Chen , Jian Xie , Yanghua Xiao

Learning in games provides a powerful framework to design control policies for self-interested agents that may be coupled through their dynamics, costs, or constraints. We consider the case where the dynamics of the coupled system can be…

系统与控制 · 电气工程与系统科学 2024-09-18 Mostafa M. Shibl , Vijay Gupta

Traditional multi-agent reinforcement learning (MARL) systems can develop cooperative strategies through repeated interactions. However, these systems are unable to perform well on any other setting than the one they have been trained on,…

多智能体系统 · 计算机科学 2025-03-20 Arjun V Sudhakar , Hadi Nekoei , Mathieu Reymond , Miao Liu , Janarthanan Rajendran , Sarath Chandar

Humans can quickly adapt to new partners in collaborative tasks (e.g. playing basketball), because they understand which fundamental skills of the task (e.g. how to dribble, how to shoot) carry over across new partners. Humans can also…

机器学习 · 计算机科学 2021-04-08 Andy Shih , Arjun Sawhney , Jovana Kondic , Stefano Ermon , Dorsa Sadigh

Zero-shot human-AI coordination holds the promise of collaborating with humans without human data. Prevailing methods try to train the ego agent with a population of partners via self-play. However, these methods suffer from two problems:…

人工智能 · 计算机科学 2023-05-23 Xingzhou Lou , Jiaxian Guo , Junge Zhang , Jun Wang , Kaiqi Huang , Yali Du

Synchronizing expectations and knowledge about the state of the world is an essential capability for effective collaboration. For robots to effectively collaborate with humans and other autonomous agents, it is critical that they be able to…

机器人学 · 计算机科学 2021-01-07 Aaquib Tabrez , Ryan Leonard , Bradley Hayes

As a schematic model of the complexity economic agents are confronted with, we introduce the ``SK-game'', a discrete time binary choice model inspired from mean-field spin-glasses. We show that even in a completely static environment,…

统计力学 · 物理学 2024-08-27 Jerome Garnier-Brun , Michael Benzaquen , Jean-Philippe Bouchaud

Many real-world applications involve teams of agents that have to coordinate their actions to reach a common goal against potential adversaries. This paper focuses on zero-sum games where a team of players faces an opponent, as is the case,…

人工智能 · 计算机科学 2019-12-18 Andrea Celli , Marco Ciccone , Raffaele Bongo , Nicola Gatti

Effective communication is an important skill for enabling information exchange in multi-agent settings and emergent communication is now a vibrant field of research, with common settings involving discrete cheap-talk channels. Since, by…

多智能体系统 · 计算机科学 2021-06-23 Kalesha Bullard , Douwe Kiela , Franziska Meier , Joelle Pineau , Jakob Foerster

Within the context of video games the notion of perfectly rational agents can be undesirable as it leads to uninteresting situations, where humans face tough adversarial decision makers. Current frameworks for stochastic games and…

人工智能 · 计算机科学 2019-01-09 Jordi Grau-Moya , Felix Leibfried , Haitham Bou-Ammar

In pursuit of enhanced multi-agent collaboration, we analyze several on-policy deep reinforcement learning algorithms in the recently published Hanabi benchmark. Our research suggests a perhaps counter-intuitive finding, where Proximal…

机器学习 · 计算机科学 2022-03-23 Bram Grooten , Jelle Wemmenhove , Maurice Poot , Jim Portegies