English
Related papers

Related papers: Lower tail large deviations of the stochastic six …

200 papers

Heavy tails are often found in practice, and yet they are an Achilles heel of a variety of mainstream random probability measures such as the Dirichlet process (DP). The first contribution of this paper focuses on characterizing the tails…

Statistics Theory · Mathematics 2025-04-04 Vianey Palacios Ramirez , Miguel de Carvalho , Luis Gutierrez Inostroza

We study the problem of differentially private stochastic convex optimization (DP-SCO) with heavy-tailed gradients, where we assume a $k^{\text{th}}$-moment bound on the Lipschitz constants of sample functions rather than a uniform bound.…

Data Structures and Algorithms · Computer Science 2024-06-06 Hilal Asi , Daogao Liu , Kevin Tian

As a fundamental problem in machine learning and differential privacy (DP), DP linear regression has been extensively studied. However, most existing methods focus primarily on either regular data distributions or low-dimensional cases with…

Machine Learning · Computer Science 2025-06-10 Xizhi Tian , Meng Ding , Touming Tao , Zihang Xiang , Di Wang

Likelihood-based procedures are a common way to estimate tail dependence parameters. They are not applicable, however, in non-differentiable models such as those arising from recent max-linear structural equation models. Moreover, they can…

Methodology · Statistics 2016-01-20 John H. J. Einmahl , Anna Kiriliouk , Johan Segers

Recently, several studies consider the stochastic optimization problem but in a heavy-tailed noise regime, i.e., the difference between the stochastic gradient and the true gradient is assumed to have a finite $p$-th moment (say being upper…

Optimization and Control · Mathematics 2023-05-23 Zijian Liu , Zhengyuan Zhou

Characterizing the differential privacy (DP) of learning algorithms has become a major challenge in recent years. In parallel, many studies suggested investigating the behavior of stochastic gradient descent (SGD) with heavy-tailed noise,…

Machine Learning · Statistics 2025-11-20 Benjamin Dupuis , Mert Gürbüzbalaban , Umut Şimşekli , Jian Wang , Sinan Yildirim , Lingjiong Zhu

We study the probability that the random graph $G(n,p)$ is triangle-free. When $p =o(n^{-1/2})$ or $p = \omega(n^{-1/2})$ the asymptotics of the logarithm of this probability are known via Janson's inequality in the former case and via…

Probability · Mathematics 2024-11-28 Matthew Jenssen , Will Perkins , Aditya Potukuchi , Michael Simkin

This article introduces a non-parametric information-theoretic approach to inference about the tail of a continuous or a discrete distribution. Leveraging a new concept named tail profile -- a set of information-theoretic quantities…

Applications · Statistics 2025-03-19 Jialin Zhang , Zhiyi Zhang

Standard definition of the stochastic Risk-Sensitive Linear-Quadratic (RS-LQ) control depends on the risk parameter, which is normally left to be set exogenously. We reconsider the classical approach and suggest two alternatives resolving…

Statistical Mechanics · Physics 2015-06-04 Michael Chertkov , Igor Kolokolov , Vladimir Lebedev

It has repeatedly been observed that loss minimization by stochastic gradient descent (SGD) leads to heavy-tailed distributions of neural network parameters. Here, we analyze a continuous diffusion approximation of SGD, called homogenized…

Machine Learning · Statistics 2024-02-05 Zhe Jiao , Martin Keller-Ressel

Gradient clipping is a commonly used technique to stabilize the training process of neural networks. A growing body of studies has shown that gradient clipping is a promising technique for dealing with the heavy-tailed behavior that emerged…

Machine Learning · Computer Science 2023-07-26 Shaojie Li , Yong Liu

The gradient noise (GN) in the stochastic gradient descent (SGD) algorithm is often considered to be Gaussian in the large data regime by assuming that the \emph{classical} central limit theorem (CLT) kicks in. This assumption is often made…

Machine Learning · Statistics 2019-12-03 Umut Şimşekli , Mert Gürbüzbalaban , Thanh Huy Nguyen , Gaël Richard , Levent Sagun

We study the problem of differentially-private (DP) stochastic (convex-concave) saddle-points in the $\ell_1$ setting. We propose $(\varepsilon, \delta)$-DP algorithms based on stochastic mirror descent that attain nearly…

Optimization and Control · Mathematics 2025-11-17 Tomás González , Cristóbal Guzmán , Courtney Paquette

Modern risk modelling approaches deal with vectors of multiple components. The components could be, for example, returns of financial instruments or losses within an insurance portfolio concerning different lines of business. One of the…

Probability · Mathematics 2021-05-12 Miriam Hägele , Jaakko Lehtomaa

We present a method for upper and lower bounding the right and the left tail probabilities of continuous random variables (RVs). For the right tail probability of RV $X$ with probability density function $f (x)$, this method requires first…

Probability · Mathematics 2026-01-07 Nikola Zlatanov

We propose a new and interpretable class of high-dimensional tail dependence models based on latent linear factor structures. Specifically, extremal dependence of an observable vector is assumed to be driven by a lower-dimensional latent…

Methodology · Statistics 2026-02-27 Alexis Boulin , Axel Bücher

Many real-world prediction tasks have outcome variables that have characteristic heavy-tail distributions. Examples include copies of books sold, auction prices of art pieces, demand for commodities in warehouses, etc. By learning…

Machine Learning · Computer Science 2023-10-13 Xindi Wang , Onur Varol , Tina Eliassi-Rad

The stochastic gradient noise (SGN) is a significant factor in the success of stochastic gradient descent (SGD). Following the central limit theorem, SGN was initially modeled as Gaussian, and lately, it has been suggested that stochastic…

Machine Learning · Computer Science 2023-03-07 Barak Battash , Ofir Lindenbaum

The non-asymptotic tail bounds of random variables play crucial roles in probability, statistics, and machine learning. Despite much success in developing upper bounds on tail probability in literature, the lower bounds on tail…

Probability · Mathematics 2020-09-08 Anru R. Zhang , Yuchen Zhou

Large deviations for sums of i.i.d.\ random variables with stretched-exponential tails (also called Weibull or semi-exponential tails) have been well understood since the 60's, going back to Nagaev's seminal work. Many extensions in the…

Probability · Mathematics 2026-02-04 Nina Gantert , Joscha Prochno , Philipp Tuchel