English
Related papers

Related papers: A Proof of the Bomber Problem's Spend-It-All Conje…

200 papers

We investigate the problem dependent regime in the stochastic Thresholding Bandit problem (TBP) under several shape constraints. In the TBP, the objective of the learner is to output, at the end of a sequential game, the set of arms whose…

Machine Learning · Statistics 2021-06-21 James Cheshire , Pierre Ménard , Alexandra Carpentier

This paper considers a multi-armed bandit game where the number of arms is much larger than the maximum budget and is effectively infinite. We characterize necessary and sufficient conditions on the total budget for an algorithm to return…

Machine Learning · Statistics 2019-01-15 Maryam Aziz , Kevin Jamieson , Javed Aslam

We study the airplane refueling problem which was introduced by the physicists Gamow and Stern in their classical book Puzzle-Math (1958). Sticking to the original story behind this problem, suppose we have to deliver a bomb in some distant…

Data Structures and Algorithms · Computer Science 2015-12-22 Iftah Gamzu , Danny Segev

We adapt ideas and concepts developed in optimal transport (and its martingale variant) to give a geometric description of optimal stopping times of Brownian motion subject to the constraint that the distribution of the stopping time is a…

Probability · Mathematics 2017-09-14 Mathias Beiglboeck , Manu Eder , Christiane Elgert , Uwe Schmock

We explore intertemporal preferences that are recursive and account for local intertemporal substitution. First, we establish a rigorous foundation for these preferences and analyze their properties. Next, we examine the associated optimal…

Optimization and Control · Mathematics 2024-09-13 Hanwu Li , Frank Riedel

We solve the problem of optimal stopping of a Brownian motion subject to the constraint that the stopping time's distribution is a given measure consisting of finitely-many atoms. In particular, we show that this problem can be converted to…

Optimization and Control · Mathematics 2017-07-07 Erhan Bayraktar , Christopher W. Miller

We obtain the first probabilistic proof of continuous differentiability of time-dependent optimal boundaries in optimal stopping problems. The underlying stochastic dynamics is a one-dimensional, time-inhomogeneous diffusion. The gain…

Probability · Mathematics 2024-05-28 Tiziano De Angelis , Damien Lamberton

This study investigates the problem of $K$-armed linear contextual bandits, an instance of the multi-armed bandit problem, under an adversarial corruption. At each round, a decision-maker observes an independent and identically distributed…

Machine Learning · Computer Science 2023-12-29 Masahiro Kato , Shinji Ito

The randomized $k$-number partitioning problem is the task to distribute $N$ i.i.d. random variables into $k$ groups in such a way that the sums of the variables in each group are as similar as possible. The restricted $k$-partitioning…

Disordered Systems and Neural Networks · Physics 2007-05-23 Anton Bovier , Irina Kurkova

We use probabilistic methods to characterise time dependent optimal stopping boundaries in a problem of multiple optimal stopping on a finite time horizon. Motivated by financial applications we consider a payoff of immediate stopping of…

Optimization and Control · Mathematics 2017-01-10 Tiziano De Angelis , Yerkin Kitapbayev

We consider a sequential decision-making problem where an agent can take one action at a time and each action has a stochastic temporal extent, i.e., a new action cannot be taken until the previous one is finished. Upon completion, the…

Machine Learning · Computer Science 2020-03-26 P Sharoff , Nishant A. Mehta , Ravi Ganti

Many discrete-time optimal stopping problems are known to have more tractable limit forms based on a planar Poisson process. Using this tool we find a solution to the optimal stopping problem for i.i.d. sequence of $n$ discrete uniform…

Probability · Mathematics 2026-01-09 Alexander Gnedin

The problem of detection time distribution concerns a quantum particle surrounded by detectors and consists of computing the probability distribution of where and when the particle will be detected. While the correct answer can be obtained…

Quantum Physics · Physics 2016-01-19 Roderich Tumulka

We consider a multi-hypothesis testing problem involving a K-armed bandit. Each arm's signal follows a distribution from a vector exponential family. The actual parameters of the arms are unknown to the decision maker. The decision maker…

Information Theory · Computer Science 2022-06-13 Gayathri R Prabhu , Srikrishna Bhashyam , Aditya Gopalan , Rajesh Sundaresan

The Prophet Inequality and Pandora's Box problems are fundamental stochastic problem with applications in Mechanism Design, Online Algorithms, Stochastic Optimization, Optimal Stopping, and Operations Research. A usual assumption in these…

Data Structures and Algorithms · Computer Science 2023-12-08 Khashayar Gatmiry , Thomas Kesselheim , Sahil Singla , Yifan Wang

A defender dispatches patrollers to circumambulate a perimeter to guard against potential attacks. The defender decides on the time points to dispatch patrollers and each patroller's direction and speed, as long as the long-run rate…

Optimization and Control · Mathematics 2020-11-10 Kyle Y Lin

This paper is concerned with the axiomatic foundation and explicit construction of a general class of optimality criteria that can be used for investment problems with multiple time horizons, or when the time horizon is not known in…

Portfolio Management · Quantitative Finance 2014-02-03 Sergey Nadtochiy , Michael Tehranchi

We study an optimal process control problem with multiple assignable causes. The process is initially in-control but is subject to random transition to one of multiple out-of-control states due to assignable causes. The objective is to find…

Optimization and Control · Mathematics 2012-12-12 Jue Wang , Chi-Guhn Lee

This paper provides a full characterization of the value function and solution(s) of an optimal stopping problem for a one-dimensional diffusion with an integral criterion. The results hold under very weak assumptions, namely, the diffusion…

Probability · Mathematics 2017-03-21 Manuel Guerra , Cláudia Nunes , Carlos Oliveira

The Colonel Blotto game is a renowned resource allocation problem with a long-standing literature in game theory (almost 100 years). However, its scope of application is still restricted by the lack of studies on the incomplete-information…

Computer Science and Game Theory · Computer Science 2019-09-12 Dong Quan Vu , Patrick Loiseau , Alonso Silva