English
Related papers

Related papers: Grab It Before It's Gone: Testing Uncertain Reward…

200 papers

We consider the Wiener process with drift $$ dX_t=\mu dt +\sigma d W_t $$ with initial value problem $X_0=x_0$, where $x_0 \in R$, $ \mu \in R$ and $\sigma > 0$ are parameters. By use values $(z_k)_{k \in N}$ of corresponding trajectories…

Statistics Theory · Mathematics 2016-11-08 Levan Labadze , Gimzer Saatashvili , Gogi Pantsulaia

The stochastic generalised linear bandit is a well-understood model for sequential decision-making problems, with many algorithms achieving near-optimal regret guarantees under immediate feedback. However, the stringent requirement for…

Machine Learning · Computer Science 2023-04-12 Benjamin Howson , Ciara Pike-Burke , Sarah Filippi

The problem of stochastic deadline scheduling is considered. A constrained Markov decision process model is introduced in which jobs arrive randomly at a service center with stochastic job sizes, rewards, and completion deadlines. The…

Optimization and Control · Mathematics 2017-07-10 Zhe Yu , Yunjian Xu , Lang Tong

We consider the Gittins index for a normal distribution with unknown mean $\theta$ and known variance where $\theta$ has a normal prior. In addition to presenting some monotonicity properties of the Gittins index, we derive an approximation…

Statistics Theory · Mathematics 2007-06-13 Yi-Ching Yao

We study an inverse first-passage-time problem for Wiener process $X(t)$ subject to hold and jump from a boundary $c.$ Let be given a threshold $S>X(0) \ge c,$ and a distribution function $F$ on $[0, + \infty ).$ The problem consists in…

Probability · Mathematics 2017-03-02 Mario Abundo

We study optimal stopping of Feller-Markov processes to maximise an undiscounted functional consisting of running and terminal rewards. In a finite-time horizon setting, we extend classical results to unbounded rewards. In infinite horizon,…

Optimization and Control · Mathematics 2016-07-21 Jan Palczewski , Lukasz Stettner

In this paper, we introduce a new method for applying the implicit function theorem to find nontrivial solutions to overdetermined problems with a fixed boundary (given) and a free boundary (to be determined). The novelty of this method…

Analysis of PDEs · Mathematics 2021-04-06 Lorenzo Cavallina

This paper is devoted to studying an infinite time horizon stochastic recursive control problem with jumps, where infinite time horizon stochastic differential equation and backward stochastic differential equation with jumps describe the…

Optimization and Control · Mathematics 2024-08-15 Sheng Luo , Xun Li , Qingmeng Wei

We provide sufficient conditions for the continuity of the free-boundary in a general class of finite-horizon optimal stopping problems arising for instance in finance and economics. The underlying process is a strong solution of one…

Optimization and Control · Mathematics 2013-05-07 Tiziano De Angelis

The Lipschitz bandit problem extends stochastic bandits to a continuous action set defined over a metric space, where the expected reward function satisfies a Lipschitz condition. In this work, we introduce a new problem of Lipschitz bandit…

Machine Learning · Computer Science 2026-02-12 Zhongxuan Liu , Yue Kang , Thomas C. M. Lee

In order to understand the impact of random influences at physical boundary on the evolution of multiscale systems, a stochastic partial differential equation model under a fast random dynamical boundary condition is investigated. The…

Dynamical Systems · Mathematics 2008-08-07 Wei Wang , Jinqiao Duan

We study the problem of scheduling periodic real-time tasks so as to meet their individual minimum reward requirements. A task generates jobs that can be given arbitrary service times before their deadlines. A task then obtains rewards…

Other Computer Science · Computer Science 2010-07-06 I-Hong Hou , P. R. Kumar

Control barrier functions are widely used to synthesize safety-critical controls. However, the presence of Gaussian-type noise in dynamical systems can generate unbounded signals and potentially result in severe consequences. Although…

Systems and Control · Electrical Eng. & Systems 2023-12-21 Chuanzheng Wang , Yiming Meng , Jun Liu , Stephen Smith

The random greedy algorithm for constructing a large partial Steiner-Triple-System is defined as follows. Begin with a complete graph on $n$ vertices and proceed to remove the edges of triangles one at a time, where each triangle removed is…

Combinatorics · Mathematics 2012-10-29 Tom Bohman , Alan Frieze , Eyal Lubetzky

In this paper, we address the stochastic contextual linear bandit problem, where a decision maker is provided a context (a random set of actions drawn from a distribution). The expected reward of each action is specified by the inner…

Machine Learning · Statistics 2023-05-30 Osama A. Hanna , Lin F. Yang , Christina Fragouli

We develop novel empirical Bernstein inequalities for the variance of bounded random variables. Our inequalities hold under constant conditional variance and mean, without further assumptions like independence or identical distribution of…

Statistics Theory · Mathematics 2026-05-28 Diego Martinez-Taboada , Aaditya Ramdas

We focus on a stochastic learning model where the learner observes a finite set of training examples and the output of the learning process is a data-dependent distribution over a space of hypotheses. The learned data-dependent distribution…

Machine Learning · Statistics 2020-12-29 Omar Rivasplata , Ilja Kuzborskij , Csaba Szepesvari , John Shawe-Taylor

The input to the stochastic orienteering problem consists of a budget $B$ and metric $(V,d)$ where each vertex $v$ has a job with deterministic reward and random processing time (drawn from a known distribution). The processing times are…

Data Structures and Algorithms · Computer Science 2014-05-12 Nikhil Bansal , Viswanath Nagarajan

This paper presents a class of Dynamic Multi-Armed Bandit problems where the reward can be modeled as the noisy output of a time varying linear stochastic dynamic system that satisfies some boundedness constraints. The class allows many…

Machine Learning · Computer Science 2017-10-10 T. W. U. Madhushani , D. H. S. Maithripala , N. E. Leonard

We study an infinite horizon optimal stopping problem which arises naturally in the optimal timing of a firm/project sale or in the valuation of natural resources: the functional to be maximised is a sum of a discounted running reward and a…

Optimization and Control · Mathematics 2016-12-08 Jan Palczewski , Lukasz Stettner