English
Related papers

Related papers: When to Quit Gambling, if You Must!

200 papers

We study the computational complexity of basic decision problems for one-counter simple stochastic games (OC-SSGs), under various objectives. OC-SSGs are 2-player turn-based stochastic games played on the transition graph of classic…

Computer Science and Game Theory · Computer Science 2010-09-29 Tomáš Brázdil , Václav Brožek , Kousha Etessami

We consider a stochastic differential equation that is controlled by means of an additive finite-variation process. A singular stochastic controller, who is a minimizer, determines this finite-variation process, while a discretionary…

Probability · Mathematics 2015-01-20 Daniel Hernandez-Hernandez , Robert S. Simon , Mihail Zervos

We study the termination problem for nondeterministic recursive probabilistic programs. First, we show that a ranking-supermartingales-based approach is both sound and complete for bounded terminiation (i.e., bounded expected termination…

Programming Languages · Computer Science 2017-01-12 Krishnendu Chatterjee , Hongfei Fu

We consider a setting where in a known future time, a certain continuous random variable will be realized. There is a public prediction that gradually converges to its realized value, and an expert that has access to a more accurate…

Computer Science and Game Theory · Computer Science 2016-05-25 Amir Ban , Yossi Azar , Yishay Mansour

We bound expected capture time and throttling number for the cop versus gambler game on a connected graph with $n$ vertices, a variant of the cop versus robber game that is played in darkness, where the adversary hops between vertices using…

Discrete Mathematics · Computer Science 2019-06-03 Jesse Geneson , Carl Joshua Quines , Espen Slettnes , Shen-Fu Tsai

Probabilistic timed automata are a suitable formalism to model systems with real-time, nondeterministic and probabilistic behaviour. We study two-player zero-sum games on such automata where the objective of the game is specified as the…

Logic in Computer Science · Computer Science 2016-04-18 Vojtěch Forejt , Marta Kwiatkowska , Gethin Norman , Ashutosh Trivedi

We study the problem of selling an asset near its ultimate maximum in the minimax setting. The regret-based notion of a perfect stopping time is introduced. A perfect stopping time is uniquely characterized by its optimality properties and…

Portfolio Management · Quantitative Finance 2016-07-15 Dmitry B. Rokhlin

Quitting games are one of the simplest stochastic games in which at any stage each player has only two possible actions, continue and quit. The game ends as soon as at least one player chooses to quit. The players then receive a payoff,…

Probability · Mathematics 2011-01-13 Katharina Fischer

We consider a novel stochastic multi-armed bandit setting, where playing an arm makes it unavailable for a fixed number of time slots thereafter. This models situations where reusing an arm too often is undesirable (e.g. making the same…

Machine Learning · Computer Science 2024-07-31 Soumya Basu , Rajat Sen , Sujay Sanghavi , Sanjay Shakkottai

In this paper, we consider a simple discrete-time optimal betting problem using the celebrated Kelly criterion, which calls for maximization of the expected logarithmic growth of wealth. While the classical Kelly betting problem can be…

Optimization and Control · Mathematics 2021-03-11 Chung-Han Hsieh

In the gambling foundation of probability theory, rationality requires that a subject should always (never) find desirable all nonnegative (negative) gambles, because no matter the result of the experiment the subject never (always)…

Optimization and Control · Mathematics 2018-11-21 Alessio Benavoli , Alessandro Facchini , Dario Piga , Marco Zaffalon

We introduce a simple stochastic volatility model, whose novelty consists in taking into account hitting times of the asset price, and study the optimal stopping problem corresponding to a put option whose time horizon (after the asset…

Pricing of Securities · Quantitative Finance 2017-03-29 Sigurd Assing , Yufan Zhao

Discounted-sum games provide a formal model for the study of reinforcement learning, where the agent is enticed to get rewards early since later rewards are discounted. When the agent interacts with the environment, she may regret her…

Computer Science and Game Theory · Computer Science 2018-11-20 Michaël Cadilhac , Guillermo A. Pérez , Marie van den Bogaard

This paper considers the discounted criterion of nonzero-sum decentralized stochastic games with prospect players. The state and action spaces are finite. The state transition probability is nonstationary. Each player independently controls…

Optimization and Control · Mathematics 2024-05-16 Yiting Wu , Junyu Zhang

The iterated prisoner's dilemma is a game that produces many counter-intuitive and complex behaviors in a social environment, based on very simple basic rules. It illustrates that cooperation can be a good thing even in a competitive world,…

Computer Science and Game Theory · Computer Science 2020-09-07 Robert Prentner

We first study an optimal stopping problem in which a player (an agent) uses a discrete stopping time in order to stop optimally a payoff process whose risk is evaluated by a (non-linear) $g$-expectation. We then consider a non-zero-sum…

Probability · Mathematics 2017-05-11 Miryana Grigorova , Marie-Claire Quenez

The problem of sequential probability forecasting is considered in the most general setting: a model set C is given, and it is required to predict as well as possible if any of the measures (environments) in C is chosen to generate the…

Machine Learning · Computer Science 2019-10-25 Daniil Ryabko

We study a class of optimal stopping games (Dynkin games) of preemption type, with uncertainty about the existence of competitors. The set-up is well-suited to model, for example, real options in the context of investors who do not want to…

Probability · Mathematics 2019-05-17 Tiziano De Angelis , Erik Ekström

Consider the following probabilistic one-player game: The board is a graph with $n$ vertices, which initially contains no edges. In each step, a new edge is drawn uniformly at random from all non-edges and is presented to the player,…

Combinatorics · Mathematics 2009-11-20 Michael Belfrage , Torsten Mütze , Reto Spöhel

In $\mathcal{X}$-armed bandit problem an agent sequentially interacts with environment which yields a reward based on the vector input the agent provides. The agent's goal is to maximise the sum of these rewards across some number of time…

Machine Learning · Statistics 2021-01-19 Valeriy Avanesov
‹ Prev 1 4 5 6 7 8 10 Next ›