English
Related papers

Related papers: Bias correction in multivariate extremes

200 papers

We develop an efficient simulation algorithm for computing the tail probabilities of the infinite series $S = \sum_{n \geq 1} a_n X_n$ when random variables $X_n$ are heavy-tailed. As $S$ is the sum of infinitely many random variables, any…

Probability · Mathematics 2016-09-08 Henrik Hult , Sandeep Juneja , Karthyek Murthy

In an influential critique of empirical practice, Freedman (2008) showed that the linear regression estimator was biased for the analysis of randomized controlled trials under the randomization model. Under Freedman's assumptions, we derive…

Methodology · Statistics 2021-10-26 Haoge Chang , Joel Middleton , P. M. Aronow

Many random phenomena, including life-testing and environmental data, show positive values and excess zeros, which pose modeling challenges. In life testing, immediate failures result in zero lifetimes, often due to defects or poor quality,…

Methodology · Statistics 2026-02-06 Shivshankar Nila , Ishapathik Das , N. Balakrishna

This article proposes a new method of truncated estimation to estimate the tail index $\alpha$ of the extremely heavy-tailed distribution with infinite mean or variance. We not only present two truncated estimators $\hat{\alpha}$ and…

Statistics Theory · Mathematics 2022-09-13 F. Q. Tang , D. Han

Heavy-tailed errors impair the accuracy of the least squares estimate, which can be spoiled by a single grossly outlying observation. As argued in the seminal work of Peter Huber in 1973 [{\it Ann. Statist.} {\bf 1} (1973) 799--821], robust…

Statistics Theory · Mathematics 2017-11-16 Wen-Xin Zhou , Koushiki Bose , Jianqing Fan , Han Liu

The distributed Hill estimator is a divide-and-conquer algorithm for estimating the extreme value index when data are stored in multiple machines. In applications, estimates based on the distributed Hill estimator can be sensitive to the…

Methodology · Statistics 2021-12-21 Liujun Chen , Deyuan Li , Chen Zhou

Variance reduction is a family of powerful mechanisms for stochastic optimization that appears to be helpful in many machine learning tasks. It is based on estimating the exact gradient with some recursive sequences. Previously, many papers…

Optimization and Control · Mathematics 2025-11-07 Aleksandr Shestakov , Valery Parfenov , Aleksandr Beznosikov

In this paper, we consider the beta prime regression model recently proposed by \cite{bour18}, which is tailored to situations where the response is continuous and restricted to the positive real line with skewed and long tails and the…

Methodology · Statistics 2020-08-28 Francisco M. C. Medeiros , Mariana C. Araújo , Marcelo Bourguignon

While the estimation of risk is an important question in the daily business of banking and insurance, many existing plug-in estimation procedures suffer from an unnecessary bias. This often leads to the underestimation of risk and…

Risk Management · Quantitative Finance 2022-01-28 Marcin Pitera , Thorsten Schmidt

Model selection aims to identify a sufficiently well performing model that is possibly simpler than the most complex model among a pool of candidates. However, the decision-making process itself can inadvertently introduce non-negligible…

Methodology · Statistics 2024-08-08 Yann McLatchie , Aki Vehtari

The subject of tail estimation for randomly censored data from a heavy tailed distribution receives growing attention, motivated by applications for instance in actuarial statistics. The bias of the available estimators of the extreme value…

Methodology · Statistics 2017-05-19 Jan Beirlant , Gaonyalelwe Maribe , Andrehette Verster

We study the bias of classical quantile regression and instrumental variable quantile regression estimators. While being asymptotically first-order unbiased, these estimators can have non-negligible second-order biases. We derive a…

Econometrics · Economics 2025-12-17 Grigory Franguridi , Bulat Gafarov , Kaspar Wuthrich

Observational studies are a key resource for causal inference but are often affected by systematic biases. Prior work has focused mainly on detecting these biases, via sensitivity analyses and comparisons with randomized controlled trials,…

Methodology · Statistics 2025-06-03 Ilker Demirel , Zeshan Hussain , Piersilvio De Bartolomeis , David Sontag

Compared to nonparametric estimators in the multivariate setting, kernel estimators for functional data models have a larger order of bias. This is problematic for constructing confidence regions or statistical tests since the bias might…

Statistics Theory · Mathematics 2025-11-21 Melanie Birke , Tim Greger

Randomized experiments are the gold standard for investigating causal relationships, with comparisons of potential outcomes under different treatment groups used to estimate treatment effects. However, outcomes with heavy-tailed…

Methodology · Statistics 2024-07-09 Hongzi Li , Wei Ma , Yingying Ma , Hanzhong Liu

Regular variation is often used as the starting point for modeling multivariate heavy-tailed data. A random vector is regularly varying if and only if its radial part $R$ is regularly varying and is asymptotically independent of the angular…

Statistics Theory · Mathematics 2018-03-28 Phyllis Wan , Richard A. Davis

Consider a random sample from a bivariate distribution function $F$ in the max-domain of attraction of an extreme-value distribution function $G$. This $G$ is characterized by two extreme-value indices and a spectral measure, the latter…

Statistics Theory · Mathematics 2009-09-01 John H. J. Einmahl , Johan Segers

While the {estimation} of risk is an important question in the daily business of banking and insurance, many existing plug-in estimation procedures suffer from an unnecessary bias. This often leads to the underestimation of risk and…

Risk Management · Quantitative Finance 2022-02-04 Marcin Pitera , Thorsten Schmidt

We study the asymptotic behaviour of widely used tests for evaluating and comparing predictive accuracy when forecast errors exhibit heavy tails. In particular, when loss differentials have infinite variance, the Diebold-Mariano test…

Methodology · Statistics 2026-05-20 Jonas F. Frederiksen , Muneya Matsui , Rasmus S. Pedersen

This paper presents a theoretical analysis of sample selection bias correction. The sample bias correction technique commonly used in machine learning consists of reweighting the cost of an error on each training point of a biased sample to…

Machine Learning · Computer Science 2008-12-18 Corinna Cortes , Mehryar Mohri , Michael Riley , Afshin Rostamizadeh