中文

有序偏好游戏的近似解

系统与控制 2025-07-16 v1 计算机科学与博弈论 系统与控制

摘要

自动驾驶车辆必须在降低旅行时间、确保安全和与交通协调等排名目标之间进行权衡。有序偏好游戏有效地建模了这些相互作用,但随着时间范围、玩家数量或偏好层级数量的增加,变得难以计算。虽然滚动时域框架通过解决顺序更短的游戏来缓解长时域的不可计算性,往往进行热启动,但并不能解决现有方法固有的复杂度增长。本文引入了一种解决方案策略,通过在滚动时域中使用 lexicographic iterated best response (IBR) 的近似方法来避免过度的复杂度增长,称为“ lexicographic IBR over time”。 lexicographic IBR over time 使用过去的信息来加速收敛。我们通过仿真交通情景表明, lexicographic IBR over time 有效地计算了滚动时域有序偏好游戏的近似最优解,收敛至广义纳什均衡。

关键词

引用

@article{arxiv.2507.11021,
  title  = {Approximate solutions to games of ordered preference},
  author = {Pau de las Heras Molins and Eric Roy-Almonacid and Dong Ho Lee and Lasse Peters and David Fridovich-Keil and Georgios Bakirtzis},
  journal= {arXiv preprint arXiv:2507.11021},
  year   = {2025}
}