English

Results of the NeurIPS 2023 Neural MMO Competition on Multi-task Reinforcement Learning

Machine Learning 2025-08-19 v1

Abstract

We present the results of the NeurIPS 2023 Neural MMO Competition, which attracted over 200 participants and submissions. Participants trained goal-conditional policies that generalize to tasks, maps, and opponents never seen during training. The top solution achieved a score 4x higher than our baseline within 8 hours of training on a single 4090 GPU. We open-source everything relating to Neural MMO and the competition under the MIT license, including the policy weights and training code for our baseline and for the top submissions.

Keywords

Cite

@article{arxiv.2508.12524,
  title  = {Results of the NeurIPS 2023 Neural MMO Competition on Multi-task Reinforcement Learning},
  author = {Joseph Suárez and Kyoung Whan Choe and David Bloomin and Jianming Gao and Yunkun Li and Yao Feng and Saidinesh Pola and Kun Zhang and Yonghui Zhu and Nikhil Pinnaparaju and Hao Xiang Li and Nishaanth Kanna and Daniel Scott and Ryan Sullivan and Rose S. Shuman and Lucas de Alcântara and Herbie Bradley and Kirsty You and Bo Wu and Yuhao Jiang and Qimai Li and Jiaxin Chen and Louis Castricato and Xiaolong Zhu and Phillip Isola},
  journal= {arXiv preprint arXiv:2508.12524},
  year   = {2025}
}