English
Related papers

Related papers: Tail Annealing for Heavy-Tailed Flow Matching

200 papers

Heavy-tailed distributions have been studied in statistics, random matrix theory, physics, and econometrics as models of correlated systems, among other domains. Further, heavy-tail distributed eigenvalues of the covariance matrix of the…

Machine Learning · Computer Science 2021-05-25 John Y. Shin

In the real open world, data tends to follow long-tailed class distributions, motivating the well-studied long-tailed recognition (LTR) problem. Naive training produces models that are biased toward common classes in terms of higher…

Computer Vision and Pattern Recognition · Computer Science 2022-03-29 Shaden Alshammari , Yu-Xiong Wang , Deva Ramanan , Shu Kong

For long-tailed classification, most works often pretrain a big model on a large-scale dataset, and then fine-tune the whole model for adapting to long-tailed data. Though promising, fine-tuning the whole pretrained model tends to suffer…

Computer Vision and Pattern Recognition · Computer Science 2023-03-29 Bowen Dong , Pan Zhou , Shuicheng Yan , Wangmeng Zuo

Flexible spatial models that allow transitions between tail dependence classes have recently appeared in the literature. However, inference for these models is computationally prohibitive, even in moderate dimensions, due to the necessity…

Statistics Theory · Mathematics 2020-12-03 Likun Zhang , Benjamin A. Shaby , Jennifer L. Wadsworth

The imbalanced distribution of long-tailed data presents a significant challenge for deep learning models, causing them to prioritize head classes while neglecting tail classes. Two key factors contributing to low recognition accuracy are…

Computer Vision and Pattern Recognition · Computer Science 2025-06-03 Mengke Li , Zhikai Hu , Yang Lu , Weichao Lan , Yiu-ming Cheung , Hui Huang

This paper studies the quantization of heavy-tailed data in some fundamental statistical estimation problems, where the underlying distributions have bounded moments of some order. We propose to truncate and properly dither the data prior…

Statistics Theory · Mathematics 2023-07-27 Junren Chen , Michael K. Ng , Di Wang

We consider a fluid queue fed by multiple On-Off flows with heavy-tailed (regularly varying) On periods. Under fairly mild assumptions, we prove that the workload distribution is asymptotically equivalent to that in a reduced system. The…

Probability · Mathematics 2016-09-07 Bert Zwart , Sem Borst , Michel Mandjes

This paper discovers fundamental principles of the backoff process that governs the performance of IEEE 802.11. A simplistic principle founded upon regular variation theory is that the backoff time has a truncated Pareto-type tail…

Networking and Internet Architecture · Computer Science 2010-08-23 Jeong-woo Cho , Yuming Jiang

Let $X$ be lognormal$(\mu,\sigma^2)$ with density $f(x)$, let $\theta>0$ and define ${L}(\theta)=E e^{-\theta X}$. We study properties of the exponentially tilted density (Esscher transform) $f_\theta(x) =e^{-\theta x}f(x)/{L}(\theta)$, in…

Probability · Mathematics 2014-03-20 Soren Asmussen , Jens Ledet Jensen , Leonardo Rojas-Nandayapa

In the liquefied natural gas (LNG) shipping industry, the phenomenon of sloshing can lead to the occurrence of very high pressures in the tanks of the vessel. The issue of modelling or estimating the probability of the simultaneous…

Statistics Theory · Mathematics 2013-12-03 Antoine Dematteo , Stéphan CLEMENCON , Nicolas Vayatis , Mathilde Mougeot

Modern machine learning is transforming jet tagging at the LHC, but the leading transformer architectures are large, not particularly fast, and training-intensive. We present a slim version of the L-GATr tagger, reduce the number of…

High Energy Physics - Phenomenology · Physics 2026-01-29 Antoine Petitjean , Tilman Plehn , Jonas Spinner , Ullrich Köthe

Real-world datasets typically exhibit long-tailed (LT) distributions, where a few head classes dominate and many tail classes are severely underrepresented. While recent work shows that parameter-efficient fine-tuning (PEFT) methods like…

Machine Learning · Computer Science 2026-01-27 Masih Aminbeidokhti , Subhankar Roy , Eric Granger , Elisa Ricci , Marco Pedersoli

State-space models are pivotal for dynamic system analysis but often struggle with outlier data that deviates from Gaussian distributions, frequently exhibiting skewness and heavy tails. This paper introduces a robust extension utilizing…

Signal Processing · Electrical Eng. & Systems 2025-07-31 Yifan Yu , Shengjie Xiu , Daniel P. Palomar

This paper studies low-rank matrix completion in the presence of heavy-tailed and possibly asymmetric noise, where we aim to estimate an underlying low-rank matrix given a set of highly incomplete noisy entries. Though the matrix completion…

Statistics Theory · Mathematics 2022-06-10 Bingyan Wang , Jianqing Fan

High-dimensional data arise routinely in modern statistics, econometrics, finance, genomics, and machine learning. While a large body of existing methodology is developed under Gaussian or light-tailed assumptions, many real data sets…

Methodology · Statistics 2026-04-16 Long Feng

Learning rate configuration is a fundamental aspect of modern deep learning. The prevailing practice of applying a uniform learning rate across all layers overlooks the structural heterogeneity of Transformers, potentially limiting their…

Machine Learning · Computer Science 2026-05-28 Di He , Songjun Tu , Keyu Wang , Lu Yin , Shiwei Liu

We consider removing lower order statistics from the classical Hill estimator in extreme value statistics, and compensating for it by rescaling the remaining terms. Trajectories of these trimmed statistics as a function of the extent of…

Methodology · Statistics 2020-06-30 Martin Bladt , Hansjoerg Albrecher , Jan Beirlant

Recognizing driving behaviors is important for downstream tasks such as reasoning, planning, and navigation. Existing video recognition approaches work well for common behaviors (e.g. "drive straight", "brake", "turn left/right"). However,…

Computer Vision and Pattern Recognition · Computer Science 2024-05-10 Chirag Parikh , Ravi Shankar Mishra , Rohan Chandra , Ravi Kiran Sarvadevabhatla

The generalization gap on the long-tailed data sets is largely owing to most categories only occupying a few training samples. Decoupled training achieves better performance by training backbone and classifier separately. What causes the…

Computer Vision and Pattern Recognition · Computer Science 2023-03-13 Zhiwei Zhang

Flow models transform data gradually from one modality (e.g. noise) onto another (e.g. images). Such models are parameterized by a time-dependent velocity field, trained to fit segments connecting pairs of source and target points. When the…

Machine Learning · Computer Science 2025-10-01 Stephen Zhang , Alireza Mousavi-Hosseini , Michal Klein , Marco Cuturi