中文
相关论文

相关论文: Some results on the Gittins index for a normal rew…

200 篇论文

The Chernoff bound is a well-known tool for obtaining a high probability bound on the expectation of a Bernoulli random variable in terms of its sample average. This bound is commonly used in statistical learning theory to upper bound the…

机器学习 · 统计学 2022-05-18 Andrew Y. K. Foong , Wessel P. Bruinsma , David R. Burt

A class of stochastic optimal control problems involving optimal stopping is considered. Methods of Krylov are adapted to investigate the numerical solutions of the corresponding normalized Bellman equations and to estimate the rate of…

最优化与控制 · 数学 2014-12-18 István Gyöngy , David Šiška

The aim of this paper is to present a result of discrete approximation of some class of stable self-similar stationary increments processes. The properties of such processes were intensively investigated, but little is known on the context…

概率论 · 数学 2008-01-18 Clément Dombry , Nadine Guillotin-Plantard

We introduce a consistent estimator of the extreme value index under random truncation based on a single sample fraction of top observations from truncated and truncation data. We establish the asymptotic normality of the proposed estimator…

统计理论 · 数学 2015-03-02 S. Benchaira , D. Meraghni , A. Necir

In this note, we introduce a general version of the well-known elliptical potential lemma that is a widely used technique in the analysis of algorithms in sequential learning and decision-making problems. We consider a stochastic linear…

机器学习 · 统计学 2022-01-20 Nima Hamidi , Mohsen Bayati

We consider optimal stopping problems for a Brownian motion and a geometric Brownian motion with a "disorder", assuming that the moment of a disorder is uniformly distributed on a finite interval. Optimal stopping rules are found as the…

统计理论 · 数学 2012-12-18 A. N. Shiryaev , M. V. Zhitlukhin

This paper investigates the problem of generalized linear bandits with heavy-tailed rewards, whose $(1+\epsilon)$-th moment is bounded for some $\epsilon\in (0,1]$. Although there exist methods for generalized linear bandits, most of them…

机器学习 · 计算机科学 2023-10-31 Bo Xue , Yimu Wang , Yuanyu Wan , Jinfeng Yi , Lijun Zhang

Regularization by the Shannon entropy enables us to efficiently and approximately solve optimal transport problems on a finite set. This paper is concerned with regularized optimal transport problems via Bregman divergence. We introduce the…

最优化与控制 · 数学 2025-04-10 Keiichi Morikuni , Koya Sakakibara , Asuka Takatsu

We consider kernel smoothed Grenander-type estimators for a monotone hazard rate and a monotone density in the presence of randomly right censored data. We show that they converge at rate $n^{2/5}$ and that the limit distribution at a fixed…

统计理论 · 数学 2018-05-18 Hendrik P. Lopuhaä , Eni Musta

The sub-Gaussian stable distribution is a heavy-tailed elliptically contoured law which has interesting applications in signal processing and financial mathematics. This work addresses the problem of feasible estimation of distributions. We…

统计理论 · 数学 2022-08-04 Taras Bodnar , Dmitry Otryakhin , Erik Thorsen

The article starts with generalizations of some classical results and new truncation error upper bounds in the sampling theorem for bandlimited stochastic processes. Then, it investigates $L_p([0,T])$ and uniform approximations of…

概率论 · 数学 2016-06-06 Yuriy Kozachenko , Andriy Olenko

For a wide class of monotonic functions $f$, we develop a Chernoff-style concentration inequality for quadratic forms $Q_f \sim \sum\limits_{i=1}^n f(\eta_i) (Z_i + \delta_i)^2$, where $Z_i \sim N(0,1)$. The inequality is expressed in terms…

统计理论 · 数学 2019-11-14 Robert E. Gallagher , Louis J. M. Aslett , David Steinsaltz , Ryan R. Christ

In mathematical finance, Levy processes are widely used for their ability to model both continuous variation and abrupt, discontinuous jumps. These jumps are practically relevant, so reliable inference on the feature that controls jump…

统计理论 · 数学 2021-09-21 Zhe Wang , Ryan Martin

We investigate the stability of equilibrium-induced optimal values with respect to (w.r.t.) reward functions $f$ and transition kernels $Q$ for time-inconsistent stopping problems under nonexponential discounting in discrete time. First,…

最优化与控制 · 数学 2022-05-19 Erhan Bayraktar , Zhenhua Wang , Zhou Zhou

Youden's index cutoff is a classifier mapping a patient's diagnostic test outcome and available covariate information to a diagnostic category. Typically the cutoff is estimated indirectly by first modeling the conditional distributions of…

统计方法学 · 统计学 2021-09-06 Nicholas Syring

We obtain the first probabilistic proof of continuous differentiability of time-dependent optimal boundaries in optimal stopping problems. The underlying stochastic dynamics is a one-dimensional, time-inhomogeneous diffusion. The gain…

概率论 · 数学 2024-05-28 Tiziano De Angelis , Damien Lamberton

In this paper, we develop a general theory of truncated inverse binomial sampling. In this theory, the fixed-size sampling and inverse binomial sampling are accommodated as special cases. In particular, the classical Chernoff-Hoeffding…

统计理论 · 数学 2019-08-20 Xinjia Chen

We consider optimal stopping problems with finite-time horizon and state-dependent discounting. The underlying process is a one-dimensional linear diffusion and the gain function is time-homogeneous and difference of two convex functions.…

概率论 · 数学 2022-01-19 Tiziano De Angelis

The trade-off between the cost of acquiring and processing data, and uncertainty due to a lack of data is fundamental in machine learning. A basic instance of this trade-off is the problem of deciding when to make noisy and costly…

机器学习 · 统计学 2017-03-30 Christopher R. Dance , Tomi Silander

In this paper we extend Stein's method to the distribution of the product of $n$ independent mean zero normal random variables. A Stein equation is obtained for this class of distributions, which reduces to the classical normal Stein…

概率论 · 数学 2017-05-30 Robert E. Gaunt