English

gComm: An environment for investigating generalization in Grounded Language Acquisition

Computation and Language 2021-05-21 v2 Artificial Intelligence

Abstract

gComm is a step towards developing a robust platform to foster research in grounded language acquisition in a more challenging and realistic setting. It comprises a 2-d grid environment with a set of agents (a stationary speaker and a mobile listener connected via a communication channel) exposed to a continuous array of tasks in a partially observable setting. The key to solving these tasks lies in agents developing linguistic abilities and utilizing them for efficiently exploring the environment. The speaker and listener have access to information provided in different modalities, i.e. the speaker's input is a natural language instruction that contains the target and task specifications and the listener's input is its grid-view. Each must rely on the other to complete the assigned task, however, the only way they can achieve the same, is to develop and use some form of communication. gComm provides several tools for studying different forms of communication and assessing their generalization.

Keywords

Cite

@article{arxiv.2105.03943,
  title  = {gComm: An environment for investigating generalization in Grounded Language Acquisition},
  author = {Rishi Hazra and Sonu Dixit},
  journal= {arXiv preprint arXiv:2105.03943},
  year   = {2021}
}

Comments

Accepted in NAACL 2021 workshop: Visually Grounded Interaction and Language (ViGIL). arXiv admin note: substantial text overlap with arXiv:2012.05011

R2 v1 2026-06-24T01:55:06.651Z