以人类为中心的 Democratic AI 机制设计
人工智能
2022-01-28 v1 人机交互
多智能体系统
综合经济学
经济学
摘要
构建与人类价值观对齐的人工智能(AI)是一个尚未解决的问题。在此,我们开发了一种称为 Democratic AI 的人机协同研究流程,其中利用强化学习设计一种受人类多数偏好的社会机制。一大群人类参与了一款在线投资游戏,需决定是否保留一笔货币禀赋或将其与他人共享以获得集体利益。共享收益在两种不同再分配机制下返还给玩家,一种由 AI 设计,另一种由人类设计。AI 发现了一种机制,可纠正初始财富不平衡、制裁搭便车者,并成功赢得多数投票。通过优化人类偏好,Democratic AI 可能成为一种有前景的价值观对齐政策创新方法。
引用
@article{arxiv.2201.11441,
title = {Human-centered mechanism design with Democratic AI},
author = {Raphael Koster and Jan Balaguer and Andrea Tacchetti and Ari Weinstein and Tina Zhu and Oliver Hauser and Duncan Williams and Lucy Campbell-Gillingham and Phoebe Thacker and Matthew Botvinick and Christopher Summerfield},
journal= {arXiv preprint arXiv:2201.11441},
year = {2022}
}
备注
18 pages, 4 figures, 54 pages including supplemental materials