利用强盗反馈计算势博弈纳什均衡的贝叶斯优化方法
计算机科学与博弈论
2018-11-16 v1 多智能体系统
摘要
对于昂贵的黑箱系统,计算策略多智能体系统的纳什均衡是一项挑战。受涉及公共资源开发的博弈普遍存在的启发,本文针对势博弈考虑上述问题。我们利用势博弈的结构,采用贝叶斯优化框架,获得了求解有限(离散动作空间)与无限(实区间动作空间)势博弈的新算法。数值结果说明了该方法在计算静态势博弈的纳什均衡与动态势博弈的线性纳什均衡时的效率。
引用
@article{arxiv.1811.06503,
title = {A Bayesian optimization approach to compute the Nash equilibria of potential games using bandit feedback},
author = {Anup Aprem and Stephen J. Roberts},
journal= {arXiv preprint arXiv:1811.06503},
year = {2018}
}