English
Related papers

Related papers: On the missing log in upper tail estimates

200 papers

In this paper, we develop connections between two seemingly disparate, but central, models in robust statistics: Huber's epsilon-contamination model and the heavy-tailed noise model. We provide conditions under which this connection…

Machine Learning · Statistics 2019-07-03 Adarsh Prasad , Sivaraman Balakrishnan , Pradeep Ravikumar

In this paper, we address the problem of embedded feature selection for ranking on top of the list problems. We pose this problem as a regularized empirical risk minimization with $p$-norm push loss function ($p=\infty$) and sparsity…

Machine Learning · Computer Science 2012-07-03 Alain Rakotomamonjy

Recently there has been much interest in "sparsifying" sums of rank one matrices: modifying the coefficients such that only a few are nonzero, while approximately preserving the matrix that results from the sum. Results of this sort have…

Discrete Mathematics · Computer Science 2018-01-30 Marcel K. de Carli Silva , Nicholas J. A. Harvey , Cristiane M. Sato

We study high-dimensional regression in principal components space when the predictors are observed with additive measurement error and the response errors may be heavy-tailed. The starting point is the $\ell_1$-penalized…

Methodology · Statistics 2026-04-07 Long Feng , Xiaoyi Wang , Le Zhou

This work studies applications and generalizations of a simple estimation technique that provides exponential concentration under heavy-tailed distributions, assuming only bounded low-order moments. We show that the technique can be used…

Machine Learning · Computer Science 2016-04-19 Daniel Hsu , Sivan Sabato

We present a method for upper and lower bounding the right and the left tail probabilities of continuous random variables (RVs). For the right tail probability of RV $X$ with probability density function $f (x)$, this method requires first…

Probability · Mathematics 2026-01-07 Nikola Zlatanov

This study proposes a reversible jump Markov chain Monte Carlo method for estimating parameters of lognormal distribution mixtures for income. Using simulated data examples, we examined the proposed algorithm's performance and the accuracy…

Econometrics · Economics 2025-10-28 Kazuhiko Kakamu

We extend the exponential formula by Bender and Canfield (1996), which relates log-concavity and the cycle index polynomials. The extension clarifies the log-convexity relation. The proof is by noticing the property of a compound Poisson…

Combinatorics · Mathematics 2017-07-31 Muneya Matsui

Bias reduction in tail estimation has received considerable interest in extreme value analysis. Estimation methods that minimize the bias while keeping the mean squared error (MSE) under control, are especially useful when applying…

Statistics Theory · Mathematics 2016-06-21 Gaonyalelwe Maribe , Andréhette Verster , Jan Beirlant

A notoriously difficult challenge in extreme value theory is the choice of the number $k\ll n$, where $n$ is the total sample size, of extreme data points to consider for inference of tail quantities. Existing theoretical guarantees for…

Other Statistics · Statistics 2025-05-30 Johannes Lederer , Anne Sabourin , Mahsa Taheri

This paper studies the quantization of heavy-tailed data in some fundamental statistical estimation problems, where the underlying distributions have bounded moments of some order. We propose to truncate and properly dither the data prior…

Statistics Theory · Mathematics 2023-07-27 Junren Chen , Michael K. Ng , Di Wang

Reduced-rank regression recognises the possibility of a rank-deficient matrix of coefficients. We propose a novel Bayesian model for estimating the rank of the coefficient matrix, which obviates the need for post-processing steps and allows…

Methodology · Statistics 2024-02-14 Maria F. Pintado , Matteo Iacopini , Luca Rossini , Alexander Y. Shestopaloff

A fundamental problem in analysis of complex systems is getting a reliable estimate of entropy of their probability distributions over the state space. This is difficult because unsampled states can contribute substantially to the entropy,…

Data Analysis, Statistics and Probability · Physics 2023-07-19 Damián G. Hernández , Ahmed Roman , Ilya Nemenman

In this paper we consider the problem of estimating the joint upper and lower tail large deviations of the edge eigenvalues of an Erd\H{o}s-R\'enyi random graph $\mathcal{G}_{n,p}$, in the regime of $p$ where the edge of the spectrum is no…

Probability · Mathematics 2020-04-02 Bhaswar B. Bhattacharya , Sohom Bhattacharya , Shirshendu Ganguly

We develop estimation for potentially high-dimensional additive structural equation models. A key component of our approach is to decouple order search among the variables from feature or edge selection in a directed acyclic graph encoding…

Methodology · Statistics 2014-12-02 Peter Bühlmann , Jonas Peters , Jan Ernest

Low-rank matrix estimation under heavy-tailed noise is challenging, both computationally and statistically. Convex approaches have been proven statistically optimal but suffer from high computational costs, especially since robust loss…

Statistics Theory · Mathematics 2023-05-12 Yinan Shen , Jingyang Li , Jian-Feng Cai , Dong Xia

Following a strategy recently developed by Ivan Nourdin and Giovanni Peccati, we provide a general technique to compare the tail of a given random variable to that of a reference distribution. This enables us to give concrete conditions to…

Probability · Mathematics 2010-07-06 Richard Eden , Frederi Viens

Let $\{\xi_n\}$ be a sequence of independent and identically distributed random variables. In this paper we study the comparison for two upper tail probabilities $\mathbb{P}\{\sum_{n=1}^{\infty}a_n|\xi_n|^p\geq r\}$ and…

Probability · Mathematics 2013-02-12 Fuchang Gao , Zhenxia Liu , Xiangfeng Yang

We develop a general embedding method based on the Friedman-Pippenger tree embedding technique (1987) and its algorithmic version, essentially due to Aggarwal et al. (1996), enhanced with a roll-back idea allowing to sequentially retrace…

Combinatorics · Mathematics 2021-03-22 Nemanja Draganić , Michael Krivelevich , Rajko Nenadov

In the particular case we have insertions/deletions at the tail of a given set S of $n$ one-dimensional elements, we present a simpler and more concrete algorithm than that presented in [Anderson, 2007] achieving the same (but also…

Data Structures and Algorithms · Computer Science 2008-12-18 Spyros Sioutas