中文
相关论文

相关论文: Stopping times in the game Rock-Paper-Scissors

200 篇论文

Matrix games constitute a fundamental problem of game theory and describe a situation of two players with completely conflicting interests. We show how methods from statistical mechanics can be used to investigate the statistical properties…

无序系统与神经网络 · 物理学 2009-10-31 J. Berg , A. Engel

We prove that every two-player nonzero-sum stopping game in discrete time admits an \epsilon-equilibrium in randomized strategies for every \epsilon >0. We use a stochastic variation of Ramsey's theorem, which enables us to reduce the…

概率论 · 数学 2007-05-23 Eran Shmaya , Eilon Solan

We develop the linear programming approach to mean-field games in a general setting. This relaxed control approach allows to prove existence results under weak assumptions, and lends itself well to numerical implementation. We consider…

最优化与控制 · 数学 2020-11-24 Roxana Dumitrescu , Marcos Leutscher , Peter Tankov

We consider a dynamic version of sender-receiver games, where the sequence of states follows an irreducible Markov chain observed by the sender. Under mild assumptions, we provide a simple characterization of the limit set of equilibrium…

概率论 · 数学 2012-04-03 Jerome Renault , Eilon Solan , Nicolas Vieille

We consider a sequential inspection game where an inspector uses a limited number of inspections over a larger number of time periods to detect a violation (an illegal act) of an inspectee. Compared with earlier models, we allow varying…

计算机科学与博弈论 · 计算机科学 2016-08-24 Bernhard von Stengel

Using methods from the statistical mechanics of disordered systems we analyze the properties of bimatrix games with random payoffs in the limit where the number of pure strategies of each player tends to infinity. We analytically calculate…

无序系统与神经网络 · 物理学 2009-10-31 Johannes Berg

We introduce and analyze a natural game formulated as follows. In this one-person game, the player is given a random permutation $A=(a_1,\dots, a_n)$ of a multiset $M$ of $n$ reals that sum up to $0$, where each of the $n!$ permutation…

离散数学 · 计算机科学 2024-11-21 Adrian Dumitrescu , Arsenii Sagdeev

This paper considers a multiple stopping time problem for a Markov chain observed in noise, where a decision maker chooses at most L stopping times to maximize a cumulative objective. We formulate the problem as a Partially Observed Markov…

系统与控制 · 计算机科学 2017-12-05 Vikram Krishnamurthy , Anup Aprem , Sujay Bhatt

A Dynkin game is a zero-sum, stochastic stopping game between two players where either player can stop the game at any time for an observable payoff. Typically the payoff process of the max-player is assumed to be smaller than the payoff…

概率论 · 数学 2020-08-18 Ivan Guo

In this paper, we investigate a new model of a linear-quadratic mean-field stochastic Stackelberg differential game with one leader and two followers, in which the leader is allowed to stop her strategy at a random time. Our overarching…

最优化与控制 · 数学 2021-06-08 Zhun Gou , Nan-jing Huang , Ming-hui Wang

We model the scrambling of a Rubik's cube by a Markov chain and introduce a stopping time $T$ which is a quite natural candidate to be a strong uniform time. This may pave the way for estimating the number of moves required to scramble a…

概率论 · 数学 2025-09-16 Thomas Fernique

We prove that zero-sum Dynkin games in continuous time with partial and asymmetric information admit a value in randomised stopping times when the stopping payoffs of the players are general \cadlag measurable processes. As a by-product of…

概率论 · 数学 2022-06-08 Tiziano De Angelis , Nikita Merkulov , Jan Palczewski

In this paper we introduce and solve a class of optimal stopping problems of recursive type. In particular, the stopping payoff depends directly on the value function of the problem itself. In a multi-dimensional Markovian setting we show…

最优化与控制 · 数学 2021-06-23 Katia Colaneri , Tiziano De Angelis

In the standard models for optimal multiple stopping problems it is assumed that between two exercises there is always a time period of deterministic length $\delta$, the so called refraction period. This prevents the optimal exercise times…

证券定价 · 定量金融 2013-10-17 Sören Christensen , Albrecht Irle , Stephan Jürgens

We study a game where one player selects a random function, and the other has to guess that function, and show that with high probability the second player can correctly guess most of the random function. We apply this analysis to…

最优化与控制 · 数学 2023-11-28 Catherine Rainer , Eilon Solan

We study online reinforcement learning in average-reward stochastic games (SGs). An SG models a two-player zero-sum game in a Markov environment, where state transitions and one-step payoffs are determined simultaneously by a learner and an…

机器学习 · 计算机科学 2017-12-05 Chen-Yu Wei , Yi-Te Hong , Chi-Jen Lu

A wide class of ``counting'' problems have been studied in Computer Science. Three typical examples are the estimation of - (i) the permanent of an $n\times n$ 0-1 matrix, (ii) the partition function of certain $n-$ particle Statistical…

概率论 · 数学 2007-05-23 Ravi Kannan

One of the proposed solutions to the equilibrium selection problem for agents learning in repeated games is obtained via the notion of stochastic stability. Learning algorithms are perturbed so that the Markov chain underlying the learning…

计算机科学与博弈论 · 计算机科学 2012-07-09 John Wicks , Amy Greenwald

Let $T\$ be a stopping time associated with a sequence of independent random variables $Z_{1},Z_{2},...$ . By applying a suitable change in the probability measure we present relations between the moment or probability generating functions…

统计理论 · 数学 2011-06-28 M. V. Boutsikas , A. C. Rakitzis , D. L. Antzoulakos

Each of two players, by turns, rolls a dice several times accumulating the successive scores until he decides to stop, or he rolls an ace. When stopping, the accumulated turn score is added to the player account and the dice is given to his…

概率论 · 数学 2009-12-31 Fabian Crocce , Ernesto Mordecki