为正确原因而行动:创建 reason-sensitive 人工道德智能体
人工智能
2025-06-17 v2 计算机与社会
机器学习
摘要
我们提出了一种延伸强化学习架构的方案,使强化学习智能体能够基于规范原因进行道德决策。核心在于一种 reason-based 护盾生成器,产生道德护盾,将智能体绑定于符合被认可的规范原因的行动,从而使整体架构限制智能体仅执行(内部上)道德合理的行动。此外,我们描述了一种算法,允许通过来自道德裁判的案例反馈迭代改进 reason-based 护盾生成器。
引用
@article{arxiv.2409.15014,
title = {Acting for the Right Reasons: Creating Reason-Sensitive Artificial Moral Agents},
author = {Kevin Baum and Lisa Dargasz and Felix Jahn and Timo P. Gros and Verena Wolf},
journal= {arXiv preprint arXiv:2409.15014},
year = {2025}
}
备注
8 pages, 2 figures, Workshop paper accepted to FEAR24 (IFM Workshop)