English

On Feasible Rewards in Multi-Agent Inverse Reinforcement Learning

Machine Learning 2025-11-26 v4

Abstract

Multi-agent Inverse Reinforcement Learning (MAIRL) aims to recover agent reward functions from expert demonstrations. We characterize the feasible reward set in Markov games, identifying all reward functions that rationalize a given equilibrium. However, equilibrium-based observations are often ambiguous: a single Nash equilibrium can correspond to many reward structures, potentially changing the game's nature in multi-agent systems. We address this by introducing entropy-regularized Markov games, which yield a unique equilibrium while preserving strategic incentives. For this setting, we provide a sample complexity analysis detailing how errors affect learned policy performance. Our work establishes theoretical foundations and practical insights for MAIRL.

Keywords

Cite

@article{arxiv.2411.15046,
  title  = {On Feasible Rewards in Multi-Agent Inverse Reinforcement Learning},
  author = {Till Freihaut and Giorgia Ramponi},
  journal= {arXiv preprint arXiv:2411.15046},
  year   = {2025}
}

Comments

Currently under review

R2 v1 2026-06-28T20:09:11.200Z