English
Related papers

Related papers: The Distributional Tail of Worst-Case Quickselect

200 papers

Models for extreme values are generally derived from limit results, which are meant to be good enough approximations when applied to finite samples. Depending on the speed of convergence of the process underlying the data, these…

Statistics Theory · Mathematics 2019-02-20 Thomas Lugrin , Anthony C. Davison , Jonathan A. Tawn

Using tail bounds, we introduce a new probabilistic condition for function estimation in stochastic derivative-free optimization which leads to a reduction in the number of samples and eases algorithmic analyses. Moreover, we develop simple…

Optimization and Control · Mathematics 2023-06-16 Francesco Rinaldi , Luis Nunes Vicente , Damiano Zeffiro

In this paper, we provide novel optimal (or near optimal) convergence rates for a clipped version of the stochastic subgradient method. We consider nonsmooth convex problems over possibly unbounded domains, under heavy-tailed noise that…

Optimization and Control · Mathematics 2025-04-21 Daniela Angela Parletta , Andrea Paudice , Saverio Salzo

We study the problem of heavy-tailed mean estimation in settings where the variance of the data-generating distribution does not exist. Concretely, given a sample $\mathbf{X} = \{X_i\}_{i = 1}^n$ from a distribution $\mathcal{D}$ over…

Statistics Theory · Mathematics 2020-12-10 Yeshwanth Cherapanamjeri , Nilesh Tripuraneni , Peter L. Bartlett , Michael I. Jordan

The multidimensional distributions with heavy tails attracted recently the attention of several papers on Applied Probability. However, the most of the works of the last decades are focused on multivariate regular variation, while the rest…

Probability · Mathematics 2026-03-10 Dimitrios G. Konstantinides , Charalampos D. Passalidis

We study the generalization properties of unregularized gradient methods applied to separable linear classification -- a setting that has received considerable attention since the pioneering work of Soudry et al. (2018). We establish tight…

Machine Learning · Computer Science 2023-03-03 Matan Schliserman , Tomer Koren

We suggest approximating the distribution of the sum of independent and identically distributed random variables with a Pareto-like tail by combining extreme value approximations for the largest summands with a normal approximation for the…

Probability · Mathematics 2018-02-05 Ulrich K. Mueller

Constant-specified and exponential concentration inequalities play an essential role in the finite-sample theory of machine learning and high-dimensional statistics area. We obtain sharper and constants-specified concentration inequalities…

Statistics Theory · Mathematics 2022-07-04 Huiming Zhang , Haoyu Wei

This work explores the bounds of the variance of unilaterally truncated Gaussian distributions (UTGDs) and scaled chi distributions (UTSCDs) with fixed means. For any arbitrary Gaussian distribution function, $f(x;\mu,\sigma)$, with a…

Statistics Theory · Mathematics 2025-11-17 Robert J. Petrella

A sharp, distribution free, non-asymptotic result is proved for the concentration of a random function around the mean function, when the randomization is generated by a finite sequence of independent data and the random functions satisfy…

Probability · Mathematics 2023-12-25 Thomas Anton , Sutanuka Roy , Rabee Tourky

A simple way of obtaining robust estimates of the "center" (or the "location") and of the "scatter" of a dataset is to use the maximum likelihood estimate with a class of heavy-tailed distributions, regardless of the "true" distribution…

Statistics Theory · Mathematics 2023-11-28 Pavol Ševera

In the "stochastic $\delta N$ formalism", the statistics of the inflationary density perturbation are obtained from the first passage distribution of a stochastic process. We develop a general framework in which to evaluate the rare tail of…

Cosmology and Nongalactic Astrophysics · Physics 2025-10-07 Jaime Calderón-Figueroa , David Seery

We consider a class of hypothesis testing problems where the null hypothesis postulates $M$ distributions for the observed data, and there is only one possible distribution under the alternative. We show that one can use a stochastic mirror…

We investigate a way of comparing and classifying tails of random variables. Our approach extends the notion of classical indices, such as exponential and moment indices, which are widely used measuring heaviness of tail functions. A…

Probability · Mathematics 2013-10-07 Jaakko Lehtomaa

We consider the \mnk{classical} problem of a controller activating (or sampling) sequentially from a finite number of $N \geq 2$ populations, specified by unknown distributions. Over some time horizon, at each time $n = 1, 2, \ldots$, the…

Machine Learning · Statistics 2015-12-18 Wesley Cowan , Michael N. Katehakis

While numerous works have focused on devising efficient algorithms for reinforcement learning (RL) with uniformly bounded rewards, it remains an open question whether sample or time-efficient algorithms for RL with large state-action space…

Machine Learning · Computer Science 2024-03-08 Jiayi Huang , Han Zhong , Liwei Wang , Lin F. Yang

We study the replacement paths problem in the $\mathsf{CONGEST}$ model of distributed computing. Given an $s$-$t$ shortest path $P$, the goal is to compute, for every edge $e$ in $P$, the shortest-path distance from $s$ to $t$ avoiding $e$.…

Data Structures and Algorithms · Computer Science 2025-08-28 Yi-Jun Chang , Yanyu Chen , Dipan Dey , Gopinath Mishra , Hung Thuan Nguyen , Bryce Sanchez

We study a new estimator for the tail index of a distribution in the Frechet domain of attraction that arises naturally by computing subsample maxima. This estimator is equivalent to taking a U-statistic over a Hill estimator with two order…

Methodology · Statistics 2015-03-20 Stefan Wager

We study the dynamics of a continuous-time model of the Stochastic Gradient Descent (SGD) for the least-square problem. Indeed, pursuing the work of Li et al. (2019), we analyze Stochastic Differential Equations (SDEs) that model SGD either…

Machine Learning · Computer Science 2024-07-03 Adrien Schertzer , Loucas Pillaud-Vivien

We establish maximal concentration bounds for the iterates generated by stochastic approximation algorithms with general step sizes, where the noise has a finite-state Markovian component plus a Martingale-difference component. When the…

Probability · Mathematics 2026-05-21 Shubhada Agrawal , Siva Theja Maguluri , Martin Zubeldia