中文
相关论文

相关论文: The turnpike theorems for Markov games

200 篇论文

We develop provably efficient reinforcement learning algorithms for two-player zero-sum finite-horizon Markov games with simultaneous moves. To incorporate function approximation, we consider a family of Markov games where the reward…

机器学习 · 计算机科学 2020-06-25 Qiaomin Xie , Yudong Chen , Zhaoran Wang , Zhuoran Yang

A large body of research is currently investigating on the connection between machine learning and game theory. In this work, game theory notions are injected into a preference learning framework. Specifically, a preference learning problem…

机器学习 · 计算机科学 2018-12-20 Mirko Polato , Fabio Aiolli

We introduce and study the turnpike property for time-varying shapes, within the viewpoint of optimal control. We focus here on second-order linear parabolic equations where the shape acts as a source term and we seek the optimal…

偏微分方程分析 · 数学 2020-06-23 Gontran Lance , Emmanuel Trélat , Enrique Zuazua

This paper develops a novel operator theoretic framework to study the contraction properties of Markov semigroups with respect to a general class of Kantorovich semi-distances, which notably includes Wasserstein distances. The rather simple…

概率论 · 数学 2026-03-04 Pierre Del Moral , Mathieu Gerber

The known results regarding two-player zero-sum games are naturally generalized in complex space and are presented through a complete compact theory. The payoff function is defined by the real part of the payoff function in the real case,…

最优化与控制 · 数学 2022-11-30 Nick Dimou

Reinforcement learning has been successful both empirically and theoretically in single-agent settings, but extending these results to multi-agent reinforcement learning in general-sum Markov games remains challenging. This paper studies…

机器学习 · 计算机科学 2026-04-07 Narim Jeong , Donghwan Lee

Game theory is the mathematical framework for analyzing strategic interactions in conflict and competition situations. In recent years quantum game theory has earned the attention of physicists, and has emerged as a branch of quantum…

量子物理 · 物理学 2015-05-30 Puya Sharif , Hoshang Heydari

The topics treated in this thesis are inherently two-fold. The first part considers the problem of a market maker optimally setting bid/ask quotes over a finite time horizon, to maximize her expected utility. The intensities of the orders…

最优化与控制 · 数学 2020-09-15 Diego Zabaljauregui

This paper analyzes the limiting behavior of stochastic linear-quadratic optimal control problems in finite time horizon $[0,T]$ as $T\rightarrow\infty$. The so-called turnpike properties are established for such problems, under…

最优化与控制 · 数学 2022-02-28 Jingrui Sun , Hanxiao Wang , Jiongmin Yong

We consider a two-player game in which the first player (the Guesser) tries to guess, edge-by-edge, the path that second player (the Chooser) takes through a directed graph. At each step, the Guesser makes a wager as to the correctness of…

概率论 · 数学 2009-07-14 Marcus Pendergrass

We study dynamic finite-player and mean-field stochastic games within the framework of Markov perfect equilibria (MPE). Our focus is on discrete time and space structures without monotonicity. Unlike their continuous-time analogues,…

最优化与控制 · 数学 2025-09-29 Felix Höfer , H. Mete Soner , Atilla Yılmaz

We consider multiplayer stochastic games in which the payoff of each player is a bounded and Borel-measurable function of the infinite play. By using a generalization of the technique of Martin (1998) and Maitra and Sudderth (1998), we show…

最优化与控制 · 数学 2022-08-26 János Flesch , Eilon Solan

We investigate the increasingly important and common game-solving setting where we do not have an explicit description of the game but only oracle access to it through gameplay, such as in financial or military simulations and computer…

人工智能 · 计算机科学 2020-02-26 Carlos Martin , Tuomas Sandholm

We study the convergence of an $N$-particle Markovian controlled system to the solution of a family of stochastic McKean-Vlasov control problems, either with a finite horizon or Schr\"odinger type cost functional. Specifically, under…

概率论 · 数学 2024-05-22 Francesco C. De Vecchi , Chiara Rigoni

n infinite two-player zero-sum game with a Borel winning set, in which the opponent's actions are monitored eventually but not necessarily immediately after they are played, is determined. The proof relies on a representation of the game as…

逻辑 · 数学 2011-07-06 Eran Shmaya

Markov random fields area popular model for high-dimensional probability distributions. Over the years, many mathematical, statistical and algorithmic problems on them have been studied. Until recently, the only known algorithms for…

机器学习 · 计算机科学 2017-06-01 Linus Hamilton , Frederic Koehler , Ankur Moitra

We revisit the problem of learning in two-player zero-sum Markov games, focusing on developing an algorithm that is uncoupled, convergent, and rational, with non-asymptotic convergence rates. We start from the case of stateless matrix game…

计算机科学与博弈论 · 计算机科学 2023-11-10 Yang Cai , Haipeng Luo , Chen-Yu Wei , Weiqiang Zheng

We study what dataset assumption permits solving offline two-player zero-sum Markov games. In stark contrast to the offline single-agent Markov decision process, we show that the single strategy concentration assumption is insufficient for…

机器学习 · 计算机科学 2022-10-17 Qiwen Cui , Simon S. Du

We construct subgame-perfect equilibria with mixed strategies for symmetric stochastic timing games with arbitrary strategic incentives. The strategies are qualitatively different for local first- or second-mover advantages, which we…

最优化与控制 · 数学 2018-05-23 Jan-Henrik Steg

Simple stochastic games are two-player zero-sum stochastic games with turn-based moves, perfect information, and reachability winning conditions. We present two new algorithms computing the values of simple stochastic games. Both of them…

计算机科学与博弈论 · 计算机科学 2015-07-01 Hugo Gimbert , Florian Horn