信条在多智能体学习中的重要性
人工智能
2023-04-13 v2
摘要
我们为系统中被配置为多个组(即团队)的智能体提出了一种多目标优化模型——信条(credo)。我们的 credo 模型调节智能体如何为其所属组优化行为。我们在具有强化学习智能体的挑战性社会困境背景下评估 credo。我们的结果表明,为实现全局有益结果,队友或整个系统的利益并非必须完全对齐。我们识别出两种无完全共同利益的情景,其实现了高平等性,且相比所有智能体利益对齐时显著更高的种群平均奖励。
引用
@article{arxiv.2204.07471,
title = {The Importance of Credo in Multiagent Learning},
author = {David Radke and Kate Larson and Tim Brecht},
journal= {arXiv preprint arXiv:2204.07471},
year = {2023}
}
备注
12 pages, 8 figures, Proceedings of the 22nd International Conference on Autonomous Agents and Multiagent Systems (AAMAS 2023)