English
Related papers

Related papers: High Dimensional Gaussian and Bootstrap Approximat…

200 papers

A major effort in modern high-dimensional statistics has been devoted to the analysis of linear predictors trained on nonlinear feature embeddings via empirical risk minimization (ERM). Gaussian equivalence theory (GET) has emerged as a…

Statistics Theory · Mathematics 2025-12-04 Garrett G. Wen , Hong Hu , Yue M. Lu , Zhou Fan , Theodor Misiakiewicz

In many applications, linear models fit the data poorly. This article studies an appealing alternative, the generalized regression model. This model only assumes that there exists an unknown monotonically increasing link function connecting…

Methodology · Statistics 2017-07-24 Fang Han , Hongkai Ji , Zhicheng Ji , Honglang Wang

Gaussian Processes (GPs) provide powerful probabilistic frameworks for interpolation, forecasting, and smoothing, but have been hampered by computational scaling issues. Here we investigate data sampled on one dimension (e.g., a scalar or…

Machine Learning · Statistics 2022-08-04 Jackson Loper , David Blei , John P. Cunningham , Liam Paninski

Nearly all statistical inference methods were developed for the regime where the number $N$ of data samples is much larger than the data dimension $p$. Inference protocols such as maximum likelihood (ML) or maximum a posteriori probability…

Disordered Systems and Neural Networks · Physics 2020-07-09 ACC Coolen , M Sheikh , A Mozeika , F Aguirre-Lopez , F Antenucci

The Bonferroni adjustment, or the union bound, is commonly used to study rate optimality properties of statistical methods in high-dimensional problems. However, in practice, the Bonferroni adjustment is overly conservative. The extreme…

Methodology · Statistics 2020-01-13 Hang Deng , Cun-Hui Zhang

High-dimensional statistical learning (HDSL) has wide applications in data analysis, operations research, and decision-making. Despite the availability of multiple theoretical frameworks, most existing HDSL schemes stipulate the following…

Statistics Theory · Mathematics 2021-10-25 Hongcheng Liu , Yinyu Ye , Hung Yi Lee

We establish a central limit theorem for (a sequence of) multivariate martingales which dimension potentially grows with the length $n$ of the martingale. A consequence of the results are Gaussian couplings and a multiplier bootstrap for…

Statistics Theory · Mathematics 2018-09-11 Alexandre Belloni , Roberto I. Oliveira

Multidimensional scaling (MDS) is widely used to reconstruct a low-dimensional representation of high-dimensional data while preserving pairwise distances. However, Bayesian MDS approaches based on Markov chain Monte Carlo (MCMC) face…

Methodology · Statistics 2026-02-26 Jiarui Zhang , Jiguo Cao , Liangliang Wang

This is Part II of a two-part work on the estimation for a multi-layer generalized linear model (ML-GLM) in large system limits. In Part I, we had analyzed the asymptotic performance of an exact MMSE estimator, and obtained a set of coupled…

Information Theory · Computer Science 2020-07-21 Qiuyun Zou , Haochuan Zhang , Hongwen Yang

This paper presents the hierarchical generalized linear model (HGLM) for loss reserving in a non-life insurance company. Because in this case the error of prediction is expressed by a complex analytical formula, the error bootstrap…

Risk Management · Quantitative Finance 2016-12-14 Alicja Wolny-Dominiak

One reason why standard formulations of the central limit theorems are not applicable in high-dimensional and non-stationary regimes is the lack of a suitable limit object. Instead, suitable distributional approximations can be used, where…

Statistics Theory · Mathematics 2024-12-20 Fabian Mies

Generalized compressed sensing (GCS) is a paradigm in which a structured high-dimensional signal may be recovered from random, under-determined, and corrupted linear measurements. Generalized Lasso (GL) programs are effective for solving…

Information Theory · Computer Science 2022-08-25 Aaron Berk , Yaniv Plan , Özgür Yilmaz

Phase-averaged dilute bubbly flow models require high-order statistical moments of the bubble population. The method of classes, which directly evolve bins of bubbles in the probability space, are accurate but computationally expensive.…

Gaussian universality results assert that the properties of many estimators remain unchanged when the input data are replaced by Gaussians. Such results have gained popularity in high-dimensional statistics and machine learning, as…

Probability · Mathematics 2025-12-03 Kevin Han Huang , Morgane Austern , Peter Orbanz

Gaussian and discrete non-Gaussian spatial datasets are common across fields like public health, ecology, geosciences, and social sciences. Bayesian spatial generalized linear mixed models (SGLMMs) are a flexible class of models for…

Methodology · Statistics 2025-01-27 Jin Hyung Lee , Ben Seiyon Lee

The Collective Graphical Model (CGM) models a population of independent and identically distributed individuals when only collective statistics (i.e., counts of individuals) are observed. Exact inference in CGMs is intractable, and previous…

Machine Learning · Computer Science 2014-05-21 Li-Ping Liu , Daniel Sheldon , Thomas G. Dietterich

The statistical framework of Generalized Linear Models (GLM) can be applied to sequential problems involving categorical or ordinal rewards associated, for instance, with clicks, likes or ratings. In the example of binary rewards, logistic…

Machine Learning · Computer Science 2020-03-24 Yoan Russac , Olivier Cappé , Aurélien Garivier

Gaussian Process (GP) kernels are central to Bayesian optimization (BO), yet designing effective kernels for high-dimensional problems still relies on extensive manual engineering. Existing automated approaches struggle in high dimensions…

Machine Learning · Computer Science 2026-05-21 Taeyoung Yun , Woocheol Shin , Inhyuck Song , Jaewoo Lee , Jinkyoo Park

The Gaussian Process Latent Variable Model (GP-LVM) is a non-linear probabilistic method of embedding a high dimensional dataset in terms low dimensional `latent' variables. In this paper we illustrate that maximum a posteriori (MAP)…

Machine Learning · Statistics 2013-07-02 James Barrett , Anthony C. C. Coolen

Inference for functional linear models in the presence of heteroscedastic errors has received insufficient attention given its practical importance; in fact, even a central limit theorem has not been studied in this case. At issue,…

Statistics Theory · Mathematics 2024-05-27 Hyemin Yeon , Xiongtao Dai , Daniel John Nordman