English
Related papers

Related papers: Stopping times in the game Rock-Paper-Scissors

200 papers

In this paper, we present a family of a control-stopping games which arise naturally in equilibrium-based models of market microstructure, as well as in other models with strategic buyers and sellers. A distinctive feature of this family of…

Mathematical Finance · Quantitative Finance 2019-03-20 Roman Gayduk , Sergey Nadtochiy

This paper introduces a new class of Dynkin games, where the two players are allowed to make their stopping decisions at a sequence of exogenous Poisson arrival times. The value function and the associated optimal stopping strategy are…

Optimization and Control · Mathematics 2019-07-18 Gechun Liang , Haodong Sun

We compute the stationary distribution of a continuous-time Markov chain which is constructed by gluing together two finite, irreducible Markov chains by identifying a pair of states of one chain with a pair of states of the other and…

Probability · Mathematics 2015-10-22 Bence Mélykúti , Peter Pfaffelhuber

We consider reinforcement learning for continuous-time Markov decision processes (MDPs) in the infinite-horizon, average-reward setting. In contrast to discrete-time MDPs, a continuous-time process moves to a state and stays there for a…

Machine Learning · Computer Science 2024-07-03 Xuefeng Gao , Xun Yu Zhou

The existence of stationary Markov perfect equilibria in stochastic games is shown under a general condition called "(decomposable) coarser transition kernels". This result covers various earlier existence results on correlated equilibria,…

Optimization and Control · Mathematics 2017-01-24 Wei He , Yeneng Sun

In an iterated non-cooperative game, if all the players act to maximize their individual accumulated payoff, the system as a whole usually converges to a Nash equilibrium that poorly benefits any player. Here we show that such an…

Physics and Society · Physics 2015-06-22 Zedong Bi , Hai-Jun Zhou

In the classical theory of Markov chains, one may study the mean time to reach some chosen state, and it is well-known that in the irreducible, finite case, such quantity can be calculated in terms of the fundamental matrix of the walk, as…

Quantum Physics · Physics 2022-06-17 C. F. Lardizabal , L. Velázquez

In this paper, we provide a methodology for computing the probability distribution of sojourn times for a wide class of Markov chains. Our methodology consists in writing out linear systems and matrix equations for generating functions…

Probability · Mathematics 2018-01-09 Valentina Cammarota , Aimé Lachal

Standard Markovian optimal stopping problems are consistent in the sense that the first entrance time into the stopping set is optimal for each initial state of the process. Clearly, the usual concept of optimality cannot in a…

Optimization and Control · Mathematics 2018-12-05 Sören Christensen , Kristoffer Lindensjö

The mixing time of the Markov chain induced by a policy limits performance in real-world continual learning scenarios. Yet, the effect of mixing times on learning in continual reinforcement learning (RL) remains underexplored. In this…

Recent work has shown that pairwise interactions may not be sufficient to fully model ecological dynamics in the wild. In this letter, we consider a replicator dynamic that takes both pairwise and triadic interactions into consideration…

Adaptation and Self-Organizing Systems · Physics 2023-05-17 Christopher Griffin , Rongling Wu

How humans make decisions in non-cooperative strategic interactions is a challenging question. For the fundamental model system of Rock-Paper-Scissors (RPS) game, classic game theory of infinite rationality predicts the Nash equilibrium…

Physics and Society · Physics 2014-07-29 Zhijian Wang , Bin Xu , Hai-Jun Zhou

We study the phenomenon of cyclic dominance in the paradigmatic Rock--Paper--Scissors model, as occurring in both stochastic individual-based models of finite populations and in the deterministic replicator equations. The mean-field…

Populations and Evolution · Quantitative Biology 2017-08-03 Qian Yang , Tim Rogers , Jonathan H P Dawes

Stopping times are used in applications to model random arrivals. A standard assumption in many models is that they are conditionally independent, given an underlying filtration. This is a widely useful assumption, but there are…

Probability · Mathematics 2024-11-21 Philip Protter , Alejandra Quintos

The rank of a bimatrix game is the matrix rank of the sum of the two payoff matrices. This paper comprehensively analyzes games of rank one, and shows the following: (1) For a game of rank r, the set of its Nash equilibria is the…

Computer Science and Game Theory · Computer Science 2023-07-27 Bharat Adsul , Jugal Garg , Ruta Mehta , Milind Sohoni , Bernhard von Stengel

Spiro, Surya and Zeng (Electron. J. Combin. 2023; arXiv:2207.11272) recently studied a semi-restricted variant of the well-known game Rock, Paper, Scissors; in this variant the game is played for $3n$ rounds, but one of the two players is…

Probability · Mathematics 2024-05-03 Svante Janson

We propose a numerical method to approximate the value function for the optimal stopping problem of a piecewise deterministic Markov process (PDMP). Our approach is based on quantization of the post jump location---inter-arrival time Markov…

Probability · Mathematics 2016-08-14 Benoîte de Saporta , François Dufour , Karen Gonzalez

New algorithms for computing power moments of hitting times and accumulated rewards of hitting type for semi-Markov processes. The algorithms are based on special techniques of sequential phase space reduction and recurrence relations…

Probability · Mathematics 2016-03-21 Dmitrii Silvestrov , Raimondo Manca

A basic question for zero-sum repeated games consists in determining whether the mean payoff per time unit is independent of the initial state. In the special case of "zero-player" games, i.e., of Markov chains equipped with additive…

Optimization and Control · Mathematics 2015-10-20 Marianne Akian , Stéphane Gaubert , Antoine Hochart

We study the relationship between performance and practice by analyzing the activity of many players of a casual online game. We find significant heterogeneity in the improvement of player performance, given by score, and address this by…

Computers and Society · Computer Science 2017-03-16 Tushar Agarwal , Keith A. Burghardt , Kristina Lerman
‹ Prev 1 3 4 5 6 7 10 Next ›