中文
相关论文

相关论文: From Poincar\'e Recurrence to Convergence in Imper…

200 篇论文

Under what conditions do the behaviors of players, who play a game repeatedly, converge to a Nash equilibrium? If one assumes that the players' behavior is a discrete-time or continuous-time rule whereby the current mixed strategy profile…

计算机科学与博弈论 · 计算机科学 2022-03-29 Jason Milionis , Christos Papadimitriou , Georgios Piliouras , Kelly Spendlove

We show that the problem of deciding whether in a multi-player perfect information recursive game (i.e. a stochastic game with terminal rewards) there exists a stationary Nash equilibrium ensuring each player a certain payoff is Existential…

计算机科学与博弈论 · 计算机科学 2020-08-19 Kristoffer Arnsfelt Hansen , Steffan Christ Sølvsten

We introduce DREAM, a deep reinforcement learning algorithm that finds optimal strategies in imperfect-information games with multiple agents. Formally, DREAM converges to a Nash Equilibrium in two-player zero-sum games and to an…

机器学习 · 计算机科学 2020-12-01 Eric Steinberger , Adam Lerer , Noam Brown

In this paper, we examine the Nash equilibrium convergence properties of no-regret learning in general N-player games. For concreteness, we focus on the archetypal follow the regularized leader (FTRL) family of algorithms, and we consider…

计算机科学与博弈论 · 计算机科学 2021-02-05 Angeliki Giannou , Emmanouil-Vasileios Vlatakis-Gkaragkounis , Panayotis Mertikopoulos

We study the problem of implementing equilibria of complete information games in settings of incomplete information, and address this problem using "recommender mechanisms." A recommender mechanism is one that does not have the power to…

计算机科学与博弈论 · 计算机科学 2015-12-11 Michael Kearns , Mallesh M. Pai , Aaron Roth , Jonathan Ullman

We investigate the existence of certain types of equilibria (Nash, $\varepsilon$-Nash, subgame perfect, $\varepsilon$-subgame perfect, Pareto-optimal) in multi-player multi-outcome infinite sequential games. We use two fundamental…

计算机科学中的逻辑 · 计算机科学 2016-03-18 Stéphane Le Roux , Arno Pauly

Secure equilibrium is a refinement of Nash equilibrium, which provides some security to the players against deviations when a player changes his strategy to another best response strategy. The concept of secure equilibrium is specifically…

计算机科学与博弈论 · 计算机科学 2014-05-08 Julie De Pril , János Flesch , Jeroen Kuipers , Gijs Schoenmakers , Koos Vrieze

This paper studies the implementation of Bayes correlated equilibria in symmetric Bayesian games with nonatomic players, using direct information structures and obedient strategies. The main results demonstrate full implementation in a…

理论经济学 · 经济学 2026-02-25 Frederic Koessler , Marco Scarsini , Tristan Tomala

Two fundamental problems in computational game theory are computing a Nash equilibrium and learning to exploit opponents given observations of their play (opponent exploitation). The latter is perhaps even more important than the former:…

计算机科学与博弈论 · 计算机科学 2018-06-29 Sam Ganzfried , Qingyun Sun

Follow the regularized leader FTRL is the premier algorithm for online optimization. However, despite decades of research on its convergence in constrained optimization -- and potential games in particular -- its behavior remained hitherto…

计算机科学与博弈论 · 计算机科学 2026-02-02 Ioannis Anagnostides , Ioannis Panageas , Nikolas Patris , Tuomas Sandholm

Search has played a fundamental role in computer game research since the very beginning. And while online search has been commonly used in perfect information games such as Chess and Go, online search methods for imperfect information games…

计算机科学与博弈论 · 计算机科学 2021-03-03 Michal Šustr , Martin Schmid , Matej Moravčík , Neil Burch , Marc Lanctot , Michael Bowling

In this paper, we examine the long-run behavior of regularized, no-regret learning in finite games. A well-known result in the field states that the empirical frequencies of no-regret play converge to the game's set of coarse correlated…

计算机科学与博弈论 · 计算机科学 2023-11-07 Victor Boone , Panayotis Mertikopoulos

We present a framework for computing approximate mixed-strategy Nash equilibria of continuous-action games. It is a modification of the traditional double oracle algorithm, extended to multiple players and continuous action spaces. Unlike…

计算机科学与博弈论 · 计算机科学 2024-06-14 Carlos Martin , Tuomas Sandholm

Consider a strongly monotone game where the players' utility functions include a reward function and a linear term for each dimension, with coefficients that are controlled by the manager. Gradient play converges to a unique Nash…

多智能体系统 · 计算机科学 2026-02-25 Siddharth Chandak , Ilai Bistritz , Nicholas Bambos

Counterfactual regret minimization (CFR) is an effective algorithm for solving extensive games with imperfect information (IIEGs). However, CFR is only allowed to be applied in known environments, where the transition function of the chance…

计算机科学与博弈论 · 计算机科学 2024-10-30 Chen Qiu , Xuan Wang , Tianzi Ma , Yaojun Wen , Jiajia Zhang

In this paper, we consider a differential stochastic zero-sum game in which two players intervene by adopting impulse controls in a finite time horizon. We provide a numerical solution as an approximation of the value function, which turns…

最优化与控制 · 数学 2024-10-14 Antoine Zolome , Brahim El Asri

A model of stochastic games where multiple controllers jointly control the evolution of the state of a dynamic system but have access to different information about the state and action processes is considered. The asymmetry of information…

计算机科学与博弈论 · 计算机科学 2012-09-18 Ashutosh Nayyar , Abhishek Gupta , Cédric Langbort , Tamer Başar

Creating strong agents for games with more than two players is a major open problem in AI. Common approaches are based on approximating game-theoretic solution concepts such as Nash equilibrium, which have strong theoretical guarantees in…

计算机科学与博弈论 · 计算机科学 2018-11-07 Sam Ganzfried , Austin Nowak , Joannier Pinales

Dominance is a fundamental concept in game theory. In normal-form games dominated strategies can be identified in polynomial time. As a consequence, iterative removal of dominated strategies can be performed efficiently as a preprocessing…

计算机科学与博弈论 · 计算机科学 2026-03-26 Sam Ganzfried

We consider learning to play multiplayer imperfect-information games with simultaneous moves and large state-action spaces. Previous attempts to tackle such challenging games have largely focused on model-free learning methods, often…

人工智能 · 计算机科学 2020-12-23 Rinu Boney , Alexander Ilin , Juho Kannala , Jarno Seppänen