English
Related papers

Related papers: Creating Jackknife and Bootstrap estimates of the …

200 papers

The quality of a summarization evaluation metric is quantified by calculating the correlation between its scores and human annotations across a large number of summaries. Currently, it is unclear how precise these correlation estimates are,…

Computation and Language · Computer Science 2021-07-28 Daniel Deutsch , Rotem Dror , Dan Roth

We introduce an estimation method of covariance matrices in a high-dimensional setting, i.e., when the dimension of the matrix, , is larger than the sample size . Specifically, we propose an orthogonally equivariant estimator. The…

Statistics Theory · Mathematics 2020-12-04 Samprit Banerjee , Stefano Monni

Jackknife instrumental variable estimation (JIVE) is a classic method to leverage many weak instrumental variables (IVs) to estimate linear structural models, overcoming the bias of standard methods like two-stage least squares. In this…

Statistics Theory · Mathematics 2024-10-08 Aurélien Bibaut , Nathan Kallus , Apoorva Lal

For studying or reducing the bias of functionals of the Kaplan-Meier survival estimator, the jackknifing approach of Stute and Wang (1994) is natural. We have studied the behavior of the jackknife estimate of bias under different…

Methodology · Statistics 2013-12-17 Md Hasinur Rahaman Khan , J. Ewart H. Shaw

The maximum likelihood estimator in nonlinear panel data models with interactive fixed effects is biased. Several bias correction methods, such as analytical and jackknife approaches, have been proposed to enable valid inference. This paper…

Econometrics · Economics 2026-04-30 Haoyuan Xu , Wei Miao , Geert Dhaene , Jad Beyhum

This paper develops distribution theory and bootstrap-based inference methods for a broad class of convex pairwise difference estimators. These estimators minimize a kernel-weighted convex-in-parameter function over observation pairs with…

Econometrics · Economics 2026-05-29 Matias D. Cattaneo , Michael Jansson , Kenichi Nagasawa

We introduce an efficient implementation of sparse recovery methods for the problem of harmonic estimation with 2D sparse arrays using a single snapshot. By imposing a uniformity constraint on the harmonic grids of the subdictionaries used…

Signal Processing · Electrical Eng. & Systems 2025-09-18 Youval Klioui

Subsampling algorithms for various parametric regression models with massive data have been extensively investigated in recent years. However, all existing studies on subsampling heavily rely on clean massive data. In practical…

Statistics Theory · Mathematics 2025-06-11 Jiangshan Ju , Mingqiu Wang , Shengli Zhao

One of the goals in scaling sequential machine learning methods pertains to dealing with high-dimensional data spaces. A key related challenge is that many methods heavily depend on obtaining the inverse covariance matrix of the data. It is…

Computation · Statistics 2017-07-28 Tomer Lancewicki

Although the operator (spectral) norm is one of the most widely used metrics for covariance estimation, comparatively little is known about the fluctuations of error in this norm. To be specific, let $\hat\Sigma$ denote the sample…

Statistics Theory · Mathematics 2019-09-16 Miles E. Lopes , N. Benjamin Erichson , Michael W. Mahoney

Analyses of the galaxy N-Point Correlation Functions (NPCFs) have a large number of degrees of freedom, meaning one cannot directly estimate an invertible covariance matrix purely from mock catalogs, as has been the standard approach for…

Cosmology and Nongalactic Astrophysics · Physics 2025-07-02 Jessica Chellino , Alessandro Greco , Simon May , Zachary Slepian

Let $\hat\Sigma=\frac{1}{n}\sum_{i=1}^n X_i\otimes X_i$ denote the sample covariance operator of centered i.i.d.~observations $X_1,\dots,X_n$ in a real separable Hilbert space, and let $\Sigma=\mathbb{E}(X_1\otimes X_1)$. The focus of this…

Statistics Theory · Mathematics 2024-01-25 Miles E. Lopes

There is a great need for robust techniques in data mining and machine learning contexts where many standard techniques such as principal component analysis and linear discriminant analysis are inherently susceptible to outliers.…

Methodology · Statistics 2015-09-28 Garth Tarr , Samuel Müller , Neville C. Weber

In the context of principal components analysis (PCA), the bootstrap is commonly applied to solve a variety of inference problems, such as constructing confidence intervals for the eigenvalues of the population covariance matrix $\Sigma$.…

Statistics Theory · Mathematics 2022-02-17 Junwen Yao , Miles E. Lopes

This paper proposes methods for likelihood-based inference in multivariate linear regressions when the correlation matrix of the responses is separable; that is, it has a Kronecker product structure, but the variances are unrestricted. The…

Computation · Statistics 2026-04-16 Karl Oskar Ekvall

Parameter inference with an estimated covariance matrix systematically loses information due to the remaining uncertainty of the covariance matrix. Here, we quantify this loss of precision and develop a framework to hypothetically restore…

Cosmology and Nongalactic Astrophysics · Physics 2017-03-16 Elena Sellentin , Alan F. Heavens

General positivity constraints linking various powers of observables in energy eigenstates can be used to sharply locate acceptable regions for the energy eigenvalues, provided that efficient recursive methods are available to calculate the…

High Energy Physics - Theory · Physics 2023-12-22 Zane Ozzello , Yannick Meurice

We explore the ability of normalizing flow (NF) generative models to reproduce weak-lensing summary statistics when trained on a set of cosmological simulations. Our analysis focuses on how accurately NF models recover the mean, standard…

Cosmology and Nongalactic Astrophysics · Physics 2026-01-29 Joaquin Armijo , Leander Thiele , Jia Liu

Little attention has been given to the correlation coefficient when data come from discrete or continuous non-normal populations. In this article, we consider the efficiency of two correlation coefficients which are from the same family,…

Methodology · Statistics 2015-11-06 Michael Tsagris , Ioannis Elmatzoglou , Christos C. Frangos

In critical lattice models, distance ($r$) dependent correlation functions contain power laws $r^{-2\Delta}$ governed by scaling dimensions $\Delta$ of an underlying continuum field theory. In Monte Carlo simulations, the leading dimensions…

Statistical Mechanics · Physics 2025-04-15 Anders W. Sandvik