English
Related papers

Related papers: On the missing log in upper tail estimates

200 papers

In high-dimensional statistics, the Lasso is a cornerstone method for simultaneous variable selection and parameter estimation. However, its reliance on the squared loss function renders it highly sensitive to outliers and heavy-tailed…

Machine Learning · Statistics 2025-11-20 The Tien Mai

In this paper, we develop a novel high-dimensional coefficient estimation procedure based on high-frequency data. Unlike usual high-dimensional regression procedures such as LASSO, we additionally handle the heavy-tailedness of…

Methodology · Statistics 2025-10-22 Minseok Shin , Donggyu Kim

In this paper, we propose a uniformly dithered 1-bit quantization scheme for high-dimensional statistical estimation. The scheme contains truncation, dithering, and quantization as typical steps. As canonical examples, the quantization…

Machine Learning · Statistics 2023-01-23 Junren Chen , Cheng-Long Wang , Michael K. Ng , Di Wang

We extend a result of Goldreich and Ron about estimating the collision probability of a hash function. Their estimate has a polynomial tail. We prove that when the load factor is greater than a certain constant, the estimator has a gaussian…

Data Structures and Algorithms · Computer Science 2007-05-23 Dawei Hong , Jean-Camille Birget , Shushuang Man

In this paper, we propose a new accelerated stochastic first-order method called clipped-SSTM for smooth convex stochastic optimization with heavy-tailed distributed noise in stochastic gradients and derive the first high-probability…

Optimization and Control · Mathematics 2020-10-26 Eduard Gorbunov , Marina Danilova , Alexander Gasnikov

It is not uncommon that real-world data are distributed with a long tail. For such data, the learning of deep neural networks becomes challenging because it is hard to classify tail classes correctly. In the literature, several existing…

Computer Vision and Pattern Recognition · Computer Science 2024-07-19 Mengke Li , Yiu-ming Cheung , Yang Lu , Zhikai Hu , Weichao Lan , Hui Huang

We study the fundamental task of outlier-robust mean estimation for heavy-tailed distributions in the presence of sparsity. Specifically, given a small number of corrupted samples from a high-dimensional heavy-tailed distribution whose mean…

Data Structures and Algorithms · Computer Science 2022-11-30 Ilias Diakonikolas , Daniel M. Kane , Jasper C. H. Lee , Ankit Pensia

This paper develops robust inference methods for predictive regressions that address key challenges posed by endogenously persistent or heavy-tailed regressors, as well as persistent volatility in errors. Building on the Cauchy estimation…

Econometrics · Economics 2026-04-21 Rustam Ibragimov , Jihyun Kim , Anton Skrobotov

In a number of applications, particularly in financial and actuarial mathematics, it is of interest to characterize the tail distribution of a random variable $V$ satisfying the distributional equation $V\stackrel{\mathcal{D}}{=}f(V)$,…

Probability · Mathematics 2014-07-04 Jeffrey F. Collamore , Guoqing Diao , Anand N. Vidyashankar

We introduce a new type of estimator for the spectral tail process of a regularly varying time series. The approach is based on a characterizing invariance property of the spectral tail process, which is incorporated into the new estimator…

Statistics Theory · Mathematics 2021-03-16 Holger Drees , Anja Janßen , Sebastian Neblung

The autoregressive (AR) model is a widely used model to understand time series data. Traditionally, the innovation noise of the AR is modeled as Gaussian. However, many time series applications, for example, financial time series data, are…

Applications · Statistics 2019-03-27 Junyan Liu , Sandeep Kumar , Daniel P. Palomar

Mixup is a popular data augmentation method, with many variants subsequently proposed. These methods mainly create new examples via convex combination of random data pairs and their corresponding one-hot labels. However, most of them adhere…

Computer Vision and Pattern Recognition · Computer Science 2021-10-12 Shaoyu Zhang , Chen Chen , Xiujuan Zhang , Silong Peng

This paper presents an adaptive version of the Hill estimator based on Lespki's model selection method. This simple data-driven index selection method is shown to satisfy an oracle inequality and is checked to achieve the lower bound…

Statistics Theory · Mathematics 2015-12-16 Stéphane Boucheron , Maud Thomas

Bucket Sort is known to run in expected linear time when the input keys are distributed independently and uniformly at random in the interval $[0,1)$. The analysis holds even when a quadratic time algorithm is used to sort the keys in each…

Data Structures and Algorithms · Computer Science 2020-02-26 Ioana O. Bercea , Guy Even

We introduce a method to estimate simultaneously the tail and the threshold parameters of an extreme value regression model. This standard model finds its use in finance to assess the effect of market variables on extreme loss distributions…

Methodology · Statistics 2023-04-17 Julien Hambuckers , Marie Kratz , Antoine Usseglio-Carleve

We consider solutions to so-called stochastic fixed point equation $R \stackrel{d}{=} \Psi(R)$, where $\Psi $ is a random Lipschitz function and $R$ is a random variable independent of $\Psi$. Under the assumption that $\Psi$ can be…

Probability · Mathematics 2017-06-14 Ewa Damek , Piotr Dyszewski

Both parametric distribution functions appearing in extreme value theory - the generalized extreme value distribution and the generalized Pareto distribution - have log-concave densities if the extreme value index gamma is in [-1,0].…

Statistics Theory · Mathematics 2023-04-17 Samuel Müller , Kaspar Rufibach

The approach used by Kalashnikov and Tsitsiashvili for constructing upper bounds for the tail distribution of a geometric sum with subexponential summands is reconsidered. By expressing the problem in a more probabilistic light, several…

Probability · Mathematics 2009-03-18 Andrew Richards

A novel statistical method is proposed and investigated for estimating a heavy tailed density under mild smoothness assumptions. Statistical analyses of heavy-tailed distributions are susceptible to the problem of sparse information in the…

Methodology · Statistics 2022-11-18 Surya T Tokdar , Sheng Jiang , Erika L Cunningham

Variable selection is a classic problem in statistics. In this paper, we consider a Bayes variable selection problem based on spike-and-slab prior with mixed normal distribution proposed by Ro\v{c}kov\'a and George (2014). Motivated by…

Methodology · Statistics 2023-03-08 Lin Guoqiang