中文
相关论文

相关论文: The turnpike theorems for Markov games

200 篇论文

We study a two-player, zero-sum, stochastic game with incomplete information on one side in which the players are allowed to play more and more frequently. The informed player observes the realization of a Markov chain on which the payoffs…

最优化与控制 · 数学 2013-07-15 Pierre Cardaliaguet , Catherine Rainer , Dinah Rosenberg , Nicolas Vieille

In this paper, we introduce turnpike arguments in the context of optimal state estimation. In particular, we show that the optimal solution of the state estimation problem involving all available past data serves as turnpike for the…

最优化与控制 · 数学 2025-10-22 Julian D. Schiller , Lars Grüne , Matthias A. Müller

Semi-Markov model is one of the most general models for stochastic dynamic systems. This paper deals with a two-person zero-sum game for semi-Markov processes. We focus on the expected discounted payoff criterion with state-action-dependent…

计算机科学与博弈论 · 计算机科学 2021-03-09 Zhihui Yu , Xianping Guo , Li Xia

In this paper, we introduce and study different dissipativity notions and different turnpike properties for discrete-time stochastic nonlinear optimal control problems. The proposed stochastic dissipativity notions extend the classic notion…

最优化与控制 · 数学 2025-04-02 Jonas Schießl , Michael H. Baumann , Timm Faulwasser , Lars Grüne

We examine the problem of the existence of optimal deterministic stationary strategiesintwo-players antagonistic (zero-sum) perfect information stochastic games with finitely many states and actions.We show that the existenceof such…

计算机科学与博弈论 · 计算机科学 2016-11-28 Hugo Gimbert , Wieslaw Zielonka

We investigate zero-sum turn-based two-player stochastic games in which the objective of one player is to maximize the amount of rewards obtained during a play, while the other aims at minimizing it. We focus on games in which the minimizer…

计算机科学中的逻辑 · 计算机科学 2022-05-20 Pablo F. Castro , Pedro R. D'Argenio , Luciano Putruele , Ramiro Demasi

Mean field games is a recent area of study introduced by Lions and Lasry in a series of seminal papers in 2006. Mean field games model situations of competition between large number of rational agents that play non-cooperative dynamic games…

最优化与控制 · 数学 2011-03-18 Diogo A. Gomes , Joana Mohr , Rafael R. Souza

We study a model of two-player, zero-sum, stopping games with asymmetric information. We assume that the payoff depends on two continuous-time Markov chains (X, Y), where X is only observed by player 1 and Y only by player 2, implying that…

最优化与控制 · 数学 2017-12-06 Fabien Gensbittel , Christine Grün

This paper investigates value function approximation in the context of zero-sum Markov games, which can be viewed as a generalization of the Markov decision process (MDP) framework to the two-agent case. We generalize error bounds from MDPs…

人工智能 · 计算机科学 2013-01-07 Michail Lagoudakis , Ron Parr

We consider robust Markov Decision Processes with Borel state and action spaces, unbounded cost and finite time horizon. Our formulation leads to a Stackelberg game against nature. Under integrability, continuity and compactness assumptions…

最优化与控制 · 数学 2025-10-16 Nicole Bäuerle , Alexander Glauner

This paper presents analyses for the maximum hands-off control using the geometric methods developed for the theory of turnpike in optimal control. First, a sufficient condition is proved for the existence of the maximum hands-off control…

最优化与控制 · 数学 2020-05-01 Noboru Sakamoto , Masaaki Nagahara

In the paper we present a model of discrete-time mean-field game with several populations of players. Mean-field games with multiple populations of the players have only been studied in the literature in the continuous-time setting. The…

最优化与控制 · 数学 2023-04-07 Piotr Więcek

Using methods from the statistical mechanics of disordered systems we analyze the properties of bimatrix games with random payoffs in the limit where the number of pure strategies of each player tends to infinity. We analytically calculate…

无序系统与神经网络 · 物理学 2009-10-31 Johannes Berg

The turnpike principle is a fundamental concept in optimal control theory, stating that for a wide class of long-horizon optimal control problems, the optimal trajectory spends most of its time near a steady-state solution (the…

最优化与控制 · 数学 2025-03-27 Emmanuel Trélat , Enrique Zuazua

We study the turnpike phenomenon for optimal control problems with mean field dynamics that are obtained as the limit $N\rightarrow \infty$ of systems governed by a large number $N$ of ordinary differential equations. We show that the…

最优化与控制 · 数学 2024-11-20 Martin Gugat , Michael Herty , Chiara Segala

In this paper we consider two-person zero-sum risk-sensitive stochastic dynamic games with Borel state and action spaces and bounded reward. The term risk-sensitive refers to the fact that instead of the usual risk neutral optimization…

最优化与控制 · 数学 2021-07-21 Nicole Bäuerle , Ulrich Rieder

This paper proves several Tauberian theorems for general iterations of operators, and provides two applications to zero-sum stochastic games where the total payoff is a weighted sum of the stage payoffs. The first application is to provide…

最优化与控制 · 数学 2016-09-09 Bruno Ziliotto

In the paper we consider the controlled continuous-time Markov chain describing the interacting particles system with the finite number of types. The system is controlled by two players with the opposite purposes. The limiting game as the…

最优化与控制 · 数学 2014-12-02 Yurii Averboukh

In optimal stopping problems, a Markov structure guarantees Markovian optimal stopping times (first exit times). Surprisingly, there is no analogous result for Markovian stopping games once randomization is required. This paper addresses…

概率论 · 数学 2024-08-02 Sören Christensen , Boy Schultz

We study two-player games on finite graphs. Turn-based games have many nice properties, but concurrent games are harder to tame: e.g. turn-based stochastic parity games have positional optimal strategies, whereas even basic concurrent…

计算机科学与博弈论 · 计算机科学 2023-11-27 Benjamin Bordais , Patricia Bouyer , Stéphane Le Roux