English
Related papers

Related papers: Coverage of Credible Sets for Regression under Var…

200 papers

We consider the confidence interval centered on a frequentist model averaged estimator that was proposed by Buckland, Burnham & Augustin (1997). In the context of a simple testbed situation involving two linear regression models, we derive…

Methodology · Statistics 2023-06-29 Paul Kabaila , Alan H. Welsh , Christeen Wijethunga

Credible intervals and credible sets, such as highest posterior density (HPD) intervals, form an integral statistical tool in Bayesian phylogenetics, both for phylogenetic analyses and for development. Readily available for continuous…

Data Structures and Algorithms · Computer Science 2026-05-05 Jonathan Klawitter , Alexei J. Drummond

Reacting against the limitation of statistics to decision procedures, R. A. Fisher proposed for inductive reasoning the use of the fiducial distribution, a parameter-space distribution of epistemological probability transferred directly…

Statistics Theory · Mathematics 2013-03-01 David R. Bickel

High-dimensional linear models have been widely studied, but the developments in high-dimensional generalized linear models, or GLMs, have been slower. In this paper, we propose an empirical or data-driven prior leading to an empirical…

Statistics Theory · Mathematics 2025-07-09 Yiqi Tang , Ryan Martin

Gaussian process (GP) regression is a powerful interpolation technique due to its flexibility in capturing non-linearity. In this paper, we provide a general framework for understanding the frequentist coverage of point-wise and…

Statistics Theory · Mathematics 2017-08-17 Yun Yang , Anirban Bhattacharya , Debdeep Pati

Inference methods for computing confidence intervals in parametric settings usually rely on consistent estimators of the parameter of interest. However, it may be computationally and/or analytically burdensome to obtain such estimators in…

Methodology · Statistics 2024-09-20 Samuel Orso , Mucyo Karemera , Maria-Pia Victoria-Feser , Stéphane Guerrier

Randomized Controlled Trials (RCTs) represent a gold standard when developing policy guidelines. However, RCTs are often narrow, and lack data on broader populations of interest. Causal effects in these populations are often estimated using…

Machine Learning · Computer Science 2023-03-07 Zeshan Hussain , Michael Oberst , Ming-Chieh Shih , David Sontag

We study variable selection (also called support recovery) in high-dimensional sparse linear regression when one has external information on which variables are likely to be associated with the response. Consistent recovery is only possible…

Statistics Theory · Mathematics 2026-02-16 Paul Rognon-Vael , David Rossell , Piotr Zwiernik

Selection bias arises when the probability that an observation enters a dataset depends on variables related to the quantities of interest, leading to systematic distortions in estimation and uncertainty quantification. For example, in…

Given a sequence of observable variables $\{(x_1, y_1), \ldots, (x_n, y_n)\}$, the conformal prediction method estimates a confidence set for $y_{n+1}$ given $x_{n+1}$ that is valid for any finite sample size by merely assuming that the…

Machine Learning · Computer Science 2023-07-12 Etash Kumar Guha , Eugene Ndiaye , Xiaoming Huo

An informative sampling design leads to unit inclusion probabilities that are correlated with the response variable of interest. However, multistage sampling designs may also induce higher order dependencies, which are typically ignored in…

Methodology · Statistics 2019-01-23 Matthew R. Williams , Terrance D. Savitsky

Amortized variational inference is an often employed framework in simulation-based inference that produces a posterior approximation that can be rapidly computed given any new observation. Unfortunately, there are few guarantees about the…

Methodology · Statistics 2024-07-26 Yash Patel , Declan McNamara , Jackson Loper , Jeffrey Regier , Ambuj Tewari

Many standard estimators, when applied to adaptively collected data, fail to be asymptotically normal, thereby complicating the construction of confidence intervals. We address this challenge in a semi-parametric context: estimating the…

Statistics Theory · Mathematics 2025-03-04 Licong Lin , Koulik Khamaru , Martin J. Wainwright

In the setting of nonparametric multivariate regression with unknown error variance, we study asymptotic properties of a Bayesian method for estimating a regression function f and its mixed partial derivatives. We use a random series of…

Statistics Theory · Mathematics 2016-04-13 William Weimin Yoo , Subhashis Ghosal

A subvector of predictor that satisfies the ignorability assumption, whose index set is called a sufficient adjustment set, is crucial for conducting reliable causal inference based on observational data. In this paper, we propose a general…

Methodology · Statistics 2024-08-20 Wei Luo , Fei Qin , Lixing Zhu

We consider inference for M-estimators after model selection using a sparsity-inducing penalty. While existing methods for this task require bespoke inference procedures, we propose a simpler approach, which relies on two insights: (i)…

Methodology · Statistics 2026-01-21 Ronan Perry , Snigdha Panigrahi , Daniela Witten

Penalized B-splines are routinely used in additive models to describe smooth changes in a response with quantitative covariates. It is typically done through the conditional mean in the exponential family using generalized additive models…

Methodology · Statistics 2020-05-12 Philippe Lambert

We propose principled prediction intervals to quantify the uncertainty of a large class of synthetic control predictions (or estimators) in settings with staggered treatment adoption, offering precise non-asymptotic coverage probability…

Econometrics · Economics 2025-02-04 Matias D. Cattaneo , Yingjie Feng , Filippo Palomba , Rocio Titiunik

The challenges posed by complex stochastic models used in computational ecology, biology and genetics have stimulated the development of approximate approaches to statistical inference. Here we focus on Synthetic Likelihood (SL), a…

Methodology · Statistics 2017-06-09 Matteo Fasiolo , Simon N. Wood , Florian Hartig , Mark V. Bravington

We consider the problem of parametric statistical inference when likelihood computations are prohibitively expensive but sampling from the model is possible. Several so-called likelihood-free methods have been developed to perform inference…

Machine Learning · Statistics 2020-09-14 Owen Thomas , Ritabrata Dutta , Jukka Corander , Samuel Kaski , Michael U. Gutmann