English
Related papers

Related papers: WeSpeR: Computing non-linear shrinkage formulas fo…

200 papers

We describe an approximate statistical model for the sample variance distribution of the non-linear matter power spectrum that can be calibrated from limited numbers of simulations. Our model retains the common assumption of a multivariate…

A significant hurdle for analyzing large sample data is the lack of effective statistical computing and inference methods. An emerging powerful approach for analyzing large sample data is subsampling, by which one takes a random subsample…

Methodology · Statistics 2015-11-24 Rong Zhu , Ping Ma , Michael W. Mahoney , Bin Yu

We present a unified treatment of the Fourier spectra of spherically symmetric nonlocal diffusion operators. We develop numerical and analytical results for the class of kernels with weak algebraic singularity as the distance between source…

Numerical Analysis · Mathematics 2019-09-04 Yu Li , Richard Mikael Slevinsky

Weighted sampling is a fundamental tool in data analysis and machine learning pipelines. Samples are used for efficient estimation of statistics or as sparse representations of the data. When weight distributions are skewed, as is often the…

Machine Learning · Computer Science 2020-08-18 Edith Cohen , Rasmus Pagh , David P. Woodruff

Analyzing large samples of high-dimensional data under dependence is a challenging statistical problem as long time series may have change points, most importantly in the mean and the marginal covariances, for which one needs valid tests.…

Methodology · Statistics 2022-11-07 Fabian Mies , Ansgar Steland

We establish a general form of explicit, input-dependent, measure-valued warpings for learning nonstationary kernels. While stationary kernels are ubiquitous and simple to use, they struggle to adapt to functions that vary in smoothness…

Machine Learning · Computer Science 2020-10-12 Anthony Tompkins , Rafael Oliveira , Fabio Ramos

A new computational scheme for the nonlinear cosmological matter power spectrum (PS) is presented. Our method is based on evolution equations in time, which can be cast in a form extremely convenient for fast numerical evaluations. A…

Cosmology and Nongalactic Astrophysics · Physics 2015-06-05 Stefano Anselmi , Massimo Pietroni

Building spatial process models that capture nonstationary behavior while delivering computationally efficient inference is challenging. Nonstationary spatially varying kernels (see, e.g., Paciorek, 2003) offer flexibility and richness, but…

Methodology · Statistics 2025-07-01 Sébastien Coube-Sisqueille , Sudipto Banerjee , Benoît Liquet

We derive an efficient stochastic algorithm for inverse problems that present an unknown linear forcing term and a set of nonlinear parameters to be recovered. It is assumed that the data is noisy and that the linear part of the problem is…

Numerical Analysis · Mathematics 2019-09-17 Darko Volkov

Quantile regression is fundamental to distributional modeling, yet independent estimation of multiple quantiles frequently produces crossing -- where estimated quantile functions violate monotonicity, implying impossible negative…

Machine Learning · Statistics 2025-12-16 Kaihua Chang

Asymptotic concentration behaviors of linear combinations of weight distributions on the random linear code ensemble are presented. Many important properties of a binary linear code can be expressed as the form of a linear combination of…

Information Theory · Computer Science 2008-03-10 Tadashi Wadayama

Approximation of non-linear kernels using random feature maps has become a powerful technique for scaling kernel methods to large datasets. We propose $\textit{Tensor Sketch}$, an efficient random feature map for approximating polynomial…

Data Structures and Algorithms · Computer Science 2025-05-20 Ninh Pham , Rasmus Pagh

We study parameter estimation and asymptotic inference for sparse nonlinear regression. More specifically, we assume the data are given by $y = f( x^\top \beta^* ) + \epsilon$, where $f$ is nonlinear. To recover $\beta^*$, we propose an…

Machine Learning · Statistics 2015-11-17 Zhuoran Yang , Zhaoran Wang , Han Liu , Yonina C. Eldar , Tong Zhang

The factor modeling for high-dimensional time series is powerful in discovering latent common components for dimension reduction and information extraction. Most available estimation methods can be divided into two categories: the…

Methodology · Statistics 2026-05-26 Xinghao Qiao , Zihan Wang , Qiwei Yao , Bo Zhang

Slowly convergent or divergent sequences and series occur abundantly in quantum physics and quantum chemistry. These convergence problems can be overcome with the help of nonlinear sequence transformations (Wynn's epsilon or rho algorithm,…

Mathematical Physics · Physics 2007-05-23 Ernst Joachim Weniger

We give a mathematical framework for weighted ensemble (WE) sampling, a binning and resampling technique for efficiently computing probabilities in molecular dynamics. We prove that WE sampling is unbiased in a very general setting that…

Numerical Analysis · Mathematics 2018-10-17 David Aristoff

We combine Tyler's robust estimator of the dispersion matrix with nonlinear shrinkage. This approach delivers a simple and fast estimator of the dispersion matrix in elliptical models that is robust against both heavy tails and high…

Methodology · Statistics 2023-05-31 Simon Hediger , Jeffrey Näf , Michael Wolf

We investigate optimal subsampling for quantile regression. We derive the asymptotic distribution of a general subsampling estimator and then derive two versions of optimal subsampling probabilities. One version minimizes the trace of the…

Computation · Statistics 2020-01-29 HaiYing Wang , Yanyuan Ma

We introduce a nonparametric spectral density estimator for continuous-time and continuous-space processes measured at fully irregular locations. Our estimator is constructed using a weighted nonuniform Fourier sum whose weights yield a…

Methodology · Statistics 2025-10-07 Christopher J. Geoga , Paul G. Beckman

This paper considers the problem of estimating a high-dimensional (HD) covariance matrix when the sample size is smaller, or not much larger, than the dimensionality of the data, which could potentially be very large. We develop a…

Methodology · Statistics 2019-05-22 Esa Ollila , Elias Raninen