English

Acting for the Right Reasons: Creating Reason-Sensitive Artificial Moral Agents

Artificial Intelligence 2025-06-17 v2 Computers and Society Machine Learning

Abstract

We propose an extension of the reinforcement learning architecture that enables moral decision-making of reinforcement learning agents based on normative reasons. Central to this approach is a reason-based shield generator yielding a moral shield that binds the agent to actions that conform with recognized normative reasons so that our overall architecture restricts the agent to actions that are (internally) morally justified. In addition, we describe an algorithm that allows to iteratively improve the reason-based shield generator through case-based feedback from a moral judge.

Keywords

Cite

@article{arxiv.2409.15014,
  title  = {Acting for the Right Reasons: Creating Reason-Sensitive Artificial Moral Agents},
  author = {Kevin Baum and Lisa Dargasz and Felix Jahn and Timo P. Gros and Verena Wolf},
  journal= {arXiv preprint arXiv:2409.15014},
  year   = {2025}
}

Comments

8 pages, 2 figures, Workshop paper accepted to FEAR24 (IFM Workshop)

R2 v1 2026-06-28T18:53:43.257Z