中文
相关论文

相关论文: Asymptotically well-calibrated Bayesian $p$-value …

200 篇论文

This paper provides conditions under which subsampling and the bootstrap can be used to construct estimators of the quantiles of the distribution of a root that behave well uniformly over a large class of distributions $\mathbf{P}$. These…

统计理论 · 数学 2013-02-19 Joseph P. Romano , Azeem M. Shaikh

When comparing two distributions, it is often helpful to learn at which quantiles or values there is a statistically significant difference. This provides more information than the binary "reject" or "do not reject" decision of a global…

统计理论 · 数学 2018-08-16 Matt Goldman , David M. Kaplan

We present improved methods for calculating confidence intervals and $p$-values in situations where standard asymptotic approaches fail due to small sample sizes. We apply these techniques to a specific class of statistical model that can…

数据分析、统计与概率 · 物理学 2024-01-11 Enzo Canonero , Alessandra Rosalba Brazzale , Glen Cowan

This paper develops tests for inequality constraints of nonparametric regression functions. The test statistics involve a one-sided version of $L_p$-type functionals of kernel estimators $(1 \leq p < \infty)$. Drawing on the approach of…

统计理论 · 数学 2023-08-28 Sokbae Lee , Kyungchul Song , Yoon-Jae Whang

In Bayesian statistics, one's prior beliefs about underlying model parameters are revised with the information content of observed data from which, using Bayes' rule, a posterior belief is obtained. A non-trivial example taken from the…

高能物理 - 唯象学 · 物理学 2007-05-23 J. Charles , A. Hocker , H. Lacker , F. R. Le Diberder , S. T'Jampens

Machine learning (ML) is transforming healthcare, but safe clinical decisions demand reliable uncertainty estimates that standard ML models fail to provide. Conformal prediction (CP) is a popular tool that allows users to turn heuristic…

Motivated by real-world machine learning applications, we analyze approximations to the non-asymptotic fundamental limits of statistical classification. In the binary version of this problem, given two training sequences generated according…

信息论 · 计算机科学 2018-12-07 Lin Zhou , Vincent Y. F. Tan , Mehul Motani

Posterior sampling with the spike-and-slab prior [MB88], a popular multimodal distribution used to model uncertainty in variable selection, is considered the theoretical gold standard method for Bayesian sparse linear regression [CPS09,…

机器学习 · 统计学 2025-03-05 Syamantak Kumar , Purnamrita Sarkar , Kevin Tian , Yusong Zhu

Multi-class classification methods that produce sets of probabilistic classifiers, such as ensemble learning methods, are able to model aleatoric and epistemic uncertainty. Aleatoric uncertainty is then typically quantified via the Bayes…

机器学习 · 统计学 2023-04-20 Thomas Mortier , Viktor Bengs , Eyke Hüllermeier , Stijn Luca , Willem Waegeman

Proper scoring rules evaluate the quality of probabilistic predictions, playing an essential role in the pursuit of accurate and well-calibrated models. Every proper score decomposes into two fundamental components -- proper calibration…

Hypothesis testing in contingency tables is usually based on asymptotic results, thereby restricting its proper use to large samples. To study these tests in small samples, we consider the likelihood ratio test and define an accurate index,…

统计方法学 · 统计学 2018-10-04 Natalia L. Oliveira , Carlos A. de B. Pereira , Marcio A. Diniz , Adriano Polpo

Gaussian processes are notorious for scaling cubically with the size of the training set, preventing application to very large regression problems. Computation-aware Gaussian processes (CAGPs) tackle this scaling issue by exploiting…

机器学习 · 统计学 2025-03-24 Disha Hegde , Mohamed Adil , Jon Cockayne

This paper aims to examine the characteristics of the posterior distribution of covariance/precision matrices in a "large $p$, large $n$" scenario, where $p$ represents the number of variables and $n$ is the sample size. Our analysis…

统计理论 · 数学 2026-02-02 Partha Sarkar , Kshitij Khare , Malay Ghosh , Matt P. Wand

We introduce a Bayesian framework for mixed-type multivariate regression using continuous shrinkage priors. Our framework enables joint analysis of mixed continuous and discrete outcomes and facilitates variable selection from the $p$…

统计理论 · 数学 2024-12-20 Shao-Hsuan Wang , Ray Bai , Hsin-Hsiung Huang

We study a large-scale one-sided multiple testing problem in which test statistics follow normal distributions with unit variance, and the goal is to identify signals with positive mean effects. A conventional approach is to compute…

统计方法学 · 统计学 2026-05-15 Kwangok Seo , Johan Lim , Hyungwon Choi , Jaesik Jeong

A Bayesian non-parametric framework for studying time-to-event data is proposed, where the prior distribution is allowed to depend on an additional random source, and may update with the sample size. Such scenarios are natural, for…

统计方法学 · 统计学 2025-05-06 Martin Bladt , Jorge González Cázares

Most supervised machine learning tasks are subject to irreducible prediction errors. Probabilistic predictive models address this limitation by providing probability distributions that represent a belief over plausible targets, rather than…

机器学习 · 统计学 2022-10-25 David Widmann , Fredrik Lindsten , Dave Zachariah

Motivated by the weak limit of the Kolmogorov-Smirnov test statistics, in this contribution, we concern the asymptotics of \begin{align*} \mathbb{P}\left\{\sup_{\boldsymbol{x}\in [0,1]^n}\left(W(\boldsymbol{x})\Big|…

概率论 · 数学 2018-02-27 Long Bai , David Kalaj

Combining dependent p-values poses a long-standing challenge in statistical inference, particularly when aggregating findings from multiple methods to enhance signal detection. Recently, p-value combination tests based on regularly…

统计方法学 · 统计学 2025-04-22 Lin Gui , Yuchao Jiang , Jingshu Wang

We study the calibration of Gaussian process (GP) predictive distributions in the interpolation setting from a design-marginal perspective. Conditioning on the data and averaging over a design measure \mu, we formalize \mu-coverage for…

机器学习 · 统计学 2025-12-08 Aurélien Pion , Emmanuel Vazquez