English
Related papers

Related papers: Finite sample expansions and risk bounds in high-d…

200 papers

We explore the Wilks phenomena in two random graph models: the $\beta$-model and the Bradley-Terry model. For two increasing dimensional null hypotheses, including a specified null $H_0: \beta_i=\beta_i^0$ for $i=1,\ldots, r$ and a…

Statistics Theory · Mathematics 2025-06-27 Ting Yan , Yuanzhang Li , Jinfeng Xu , Yaning Yang , Ji Zhu

We investigate statistical properties of the optimal value of the Sample Average Approximation of stochastic programs, continuing the study in Kr\"atschmer (2023). Central Limit Theorem type results are derived for the optimal value. As a…

Optimization and Control · Mathematics 2023-12-12 Volker Krätschmer

We propose a new sparsity-smoothness penalty for high-dimensional generalized additive models. The combination of sparsity and smoothness is crucial for mathematical theory as well as performance for finite-sample data. We present a…

Machine Learning · Statistics 2009-11-18 Lukas Meier , Sara van de Geer , Peter Bühlmann

Maximum likelihood estimation in nonlinear models can exhibit substantial instability in finite samples when the data provide limited information about certain parameters. Such instability is driven by rare but extreme realizations of the…

Methodology · Statistics 2026-04-15 Masamune Iwasawa

Large samples have been generated routinely from various sources. Classic statistical models, such as smoothing spline ANOVA models, are not well equipped to analyze such large samples due to expensive computational costs. In particular,…

Methodology · Statistics 2020-04-23 Xiaoxiao Sun , Wenxuan Zhong , Ping Ma

In extreme value theory and other related risk analysis fields, probability weighted moments (PWM) have been frequently used to estimate the parameters of classical extreme value distributions. This method-of-moment technique can be applied…

Statistics Theory · Mathematics 2023-06-21 Anna Ben-Hamou , Philippe Naveau , Maud Thomas

In statistical mechanics, evaluating finite-size macroscopic fluctuations typically relies on Edgeworth expansions. However, these perturbative methods append additive polynomial corrections that inevitably break down in the large deviation…

Information Theory · Computer Science 2026-04-15 Hiroki Suyari

We consider the problem of simultaneous variable selection and constant coefficient identification in high-dimensional varying coefficient models based on B-spline basis expansion. Both objectives can be considered as some type of model…

Methodology · Statistics 2010-08-16 Heng Lian

This article studies local and global inference for smoothing spline estimation in a unified asymptotic framework. We first introduce a new technical tool called functional Bahadur representation, which significantly generalizes the…

Statistics Theory · Mathematics 2013-11-27 Zuofeng Shang , Guang Cheng

We solve the problem of estimating the distribution of presumed i.i.d. observations for the total variation loss. Our approach is based on density models and is versatile enough to cope with many different ones, including some density…

Statistics Theory · Mathematics 2024-01-05 Y. Baraud , H. Halconruy , G. Maillard

This paper derives central limit theorems (CLTs) for general linear spectral statistics (LSS) of three important multi-spiked Hermitian random matrix ensembles. The first is the most common spiked scenario, proposed by Johnstone, which is a…

Statistics Theory · Mathematics 2014-06-05 Damien Passemier , Matthew R. Mckay , Yang Chen

Feature selection has been widely used to alleviate compute requirements during training, elucidate model interpretability, and improve model generalizability. We propose SLM -- Sparse Learnable Masks -- a canonical approach for end-to-end…

Machine Learning · Computer Science 2023-04-07 Yihe Dong , Sercan O. Arik

Mixed outcome endpoints that combine multiple continuous and discrete components to form co-primary, multiple primary or composite endpoints are often employed as primary outcome measures in clinical trials. There are many advantages to…

Methodology · Statistics 2019-12-12 Martina McMenamin , Jessica K. Barrett , Anna Berglind , James M. S. Wason

Large language models (LLMs) are increasingly used as decision-support tools in data-constrained scientific workflows, where correctness and validity are critical. However, evaluation practices often emphasize stability or reproducibility…

Machine Learning · Computer Science 2026-03-18 Nazia Riasat

The Highly-Adaptive-Lasso(HAL)-TMLE is an efficient estimator of a pathwise differentiable parameter in a statistical model that at minimal (and possibly only) assumes that the sectional variation norm of the true nuisance parameters are…

Statistics Theory · Mathematics 2017-09-01 Mark van der Laan

Imbalanced classification and spurious correlation are common challenges in data science and machine learning. Both issues are linked to data imbalance, with certain groups of data samples significantly underrepresented, which in turn would…

Machine Learning · Statistics 2026-02-10 Ryumei Nakada , Yichen Xu , Lexin Li , Linjun Zhang

We establish finite sample bounds for the error of standard and waste-free SMC samplers. Our results cover estimates of both expectations and normalising constants of the target distributions. We consider first an arbitrary sequence of…

Computation · Statistics 2026-04-15 Yvann Le Fay , Nicolas Chopin , Matti Vihola

Many practical optimization problems lack strong convexity. Fortunately, recent studies have revealed that first-order algorithms also enjoy linear convergences under various weaker regularity conditions. While the relationship among…

Optimization and Control · Mathematics 2026-02-05 Feng-Yi Liao , Lijun Ding , Yang Zheng

We consider optimization algorithms that successively minimize simple Taylor-like models of the objective function. Methods of Gauss-Newton type for minimizing the composition of a convex function and a smooth map are common examples. Our…

Optimization and Control · Mathematics 2016-10-12 Dmitriy Drusvyatskiy , Alexander D. Ioffe , Adrian S. Lewis

We consider the estimation of smoothing parameters and variance components in models with a regular log likelihood subject to quadratic penalization of the model coefficients, via a generalization of the method of Fellner (1986) and Schall…

Methodology · Statistics 2020-08-11 Simon N. Wood , Matteo Fasiolo