English
Related papers

Related papers: Variance Estimation with Dependence and Heterogene…

200 papers

A classical statistical inequality is used to show that the distance covariance of two bounded random vectors is bounded from above by a simple function of the dimensionality and the bounds of the random vectors. Two special cases that…

Probability · Mathematics 2023-06-30 John Çamkıran

We present a general central limit theorem with simple, easy-to-check covariance-based sufficient conditions for triangular arrays of random vectors when all variables could be interdependent. The result is constructed from Stein's method,…

Heteroskedastic errors can lead to inaccurate statistical conclusions if they are not properly handled. We introduce a test for heteroskedasticity for the nonparametric regression model with multiple covariates. It is based on a suitable…

Methodology · Statistics 2018-02-21 Justin Chown , Ursula U. Müller

For linear regression models with cross-section or panel data, it is natural to assume that the disturbances are clustered in two dimensions. However, the finite-sample properties of two-way cluster-robust tests and confidence intervals are…

Econometrics · Economics 2026-03-13 James G. MacKinnon , Morten Ørregaard Nielsen , Matthew D. Webb

In this paper, we propose a novel approach to detect heteroskedasticity in regression models with regressors contaminated by measurement error. Specifically, inspired by the integrated conditional moment (ICM) approach, we construct test…

Econometrics · Economics 2026-05-20 Xiaojun Song , Jichao Yuan

We study a linear high-dimensional regression model in a semi-supervised setting, where for many observations only the vector of covariates $X$ is given with no response $Y$. We do not make any sparsity assumptions on the vector of…

Statistics Theory · Mathematics 2021-09-03 Ilan Livne , David Azriel , Yair Goldberg

We give necessary and sufficient conditions for two sub-vectors of a random vector with a multivariate extreme value distribution, corresponding to the limit distribution of the maximum of a multidimensional stationary sequence with…

Probability · Mathematics 2010-06-09 Clara Viseu , Luísa Pereira , Ana Paula Martins , Helena Ferreira

Inference for statistics of a stationary time series often involve nuisance parameters and sampling distributions that are difficult to estimate. In this paper, we propose the method of orthogonal samples, which can be used to address some…

Methodology · Statistics 2016-11-03 Suhasini Subba Rao

We study the problem of estimating the mean of a random vector in $\mathbb{R}^d$ based on an i.i.d.\ sample, when the accuracy of the estimator is measured by a general norm on $\mathbb{R}^d$. We construct an estimator (that depends on the…

Statistics Theory · Mathematics 2018-06-19 Gábor Lugosi , Shahar Mendelson

We propose leave-out estimators of quadratic forms designed for the study of linear models with unrestricted heteroscedasticity. Applications include analysis of variance and tests of linear restrictions in models with many regressors. An…

Econometrics · Economics 2019-08-28 Patrick Kline , Raffaele Saggio , Mikkel Sølvsten

A weighted likelihood technique for robust estimation of a multivariate Wrapped Normal distribution for data points scattered on a p-dimensional torus is proposed. The occurrence of outliers in the sample at hand can badly compromise…

Methodology · Statistics 2021-07-01 Giovanni Saraceno , Claudio Agostinelli , Luca Greco

Prediction intervals are commonly used in meta-analysis with random-effects models. One widely used method, the Higgins-Thompson-Spiegelhalter prediction interval, replaces the heterogeneity parameter with its point estimate, but its…

Methodology · Statistics 2025-11-14 Kengo Nagashima , Hisashi Noma , Toshi A. Furukawa

We study the problem of bivariate discrete or continuous probability density estimation under low-rank constraints.For discrete distributions, we assume that the two-dimensional array to estimate is a low-rank probability matrix. In the…

Statistics Theory · Mathematics 2024-10-23 Julien Chhor , Olga Klopp , Alexandre Tsybakov

Collaboration between different data centers is often challenged by heterogeneity across sites. To account for the heterogeneity, the state-of-the-art method is to re-weight the covariate distributions in each site to match the distribution…

Machine Learning · Statistics 2024-04-25 Tianyu Guo , Sai Praneeth Karimireddy , Michael I. Jordan

Most linear experimental design problems assume homogeneous variance although heteroskedastic noise is present in many realistic settings. Let a learner have access to a finite set of measurement vectors $\mathcal{X}\subset \mathbb{R}^d$…

Statistics Theory · Mathematics 2024-09-19 Justin Weltz , Tanner Fiez , Alexander Volfovsky , Eric Laber , Blake Mason , Houssam Nassif , Lalit Jain

The control variates method is a classical variance reduction technique for Monte Carlo estimators that exploits correlated auxiliary variables without introducing bias. In many applications, the quantity of interest can be expressed as a…

Statistics Theory · Mathematics 2025-11-10 Louison Bocquet-Nouaille , Jérôme Morio , Benjamin Bobbia

We establish a Cram\'er-type moderate deviation result for self-normalized sums of weakly dependent random variables, where the moment requirement is much weaker than the non-self-normalized counterpart. The range of the moderate deviation…

Statistics Theory · Mathematics 2014-09-15 Xiaohong Chen , Qi-Man Shao , Wei Biao Wu

The uncertainty or the variability of the data may be treated by considering, rather than a single value for each data, the interval of values in which it may fall. This paper studies the derivation of basic description statistics for…

Computation · Statistics 2008-12-18 Marie Chavent , Jérôme Saracco

Standard inference about a scalar parameter estimated via GMM amounts to applying a t-test to a particular set of observations. If the number of observations is not very large, then moderately heavy tails can lead to poor behavior of the…

Econometrics · Economics 2020-07-15 Ulrich K. Mueller

We want to reconstruct a signal based on inhomogeneous data (the amount of data can vary strongly), using the model of regression with a random design. Our aim is to understand the consequences of inhomogeneity on the accuracy of estimation…

Statistics Theory · Mathematics 2016-08-16 Stéphane Gaiffas