English
Related papers

Related papers: Data-driven Smooth Tests for Normality in ANOVA Wh…

200 papers

We study distribution-free nonparametric regression following a notion of average smoothness initiated by Ashlagi et al. (2021), which measures the "effective" smoothness of a function with respect to an arbitrary unknown underlying…

Machine Learning · Computer Science 2024-02-14 Steve Hanneke , Aryeh Kontorovich , Guy Kornowski

Given a random sample of observations, mixtures of normal densities are often used to estimate the unknown continuous distribution from which the data come. Here we propose the use of this semiparametric framework for testing symmetry about…

Methodology · Statistics 2012-04-23 Silvia Bacci , Francesco Bartolucci

Several anomaly detection and classification methods rely on large amounts of non-anomalous or "normal" samples under the assump- tion that anomalous data is typically harder to acquire. This hypothesis becomes questionable in Few-Shot…

Machine Learning · Computer Science 2025-08-01 Aymane Abdali , Bartosz Boguslawski , Lucas Drumetz , Vincent Gripon

This paper proposes self-normalized tests for multistep conditional predictive ability in forecast comparison. By normalizing the sample mean of the transformed loss differential using functionals of its cumulative sum (CUSUM) process,…

Statistics Theory · Mathematics 2026-05-11 Qitong Chen , Shuwen Lai

Quantifying the uncertainty of predictions is a core problem in modern statistics. Methods for predictive inference have been developed under a variety of assumptions, often -- for instance, in standard conformal prediction -- relying on…

Methodology · Statistics 2024-09-13 Edgar Dobriban , Mengxin Yu

The homogeneity problem for testing if more than two different samples come from the same population is considered for the case of functional data. The methodological results are motivated by the study of homogeneity of electronic devices…

We consider three problems in high-dimensional Gaussian linear mixed models. Without any assumptions on the design for the fixed effects, we construct an asymptotic $F$-statistic for testing whether a collection of random effects is zero,…

Statistics Theory · Mathematics 2019-07-30 Michael Law , Ya'acov Ritov

In modern data analysis, it is common to select a model before performing statistical inference. Selective inference tools make adjustments for the model selection process in order to ensure reliable inference post selection. In this paper,…

Methodology · Statistics 2025-02-24 Yumeng Wang , Snigdha Panigrahi , Xuming He

Randomization tests deliver exact finite-sample Type 1 error control when the null satisfies the randomization hypothesis. In practice, achieving these guarantees often requires stronger conditions than the null hypothesis of primary…

Econometrics · Economics 2026-04-03 Deniz Dutz , Xinyi Zhang

We study average treatment effect (ATE) estimation under complete randomization with many covariates in a design-based, finite-population framework. In randomized experiments, regression adjustment can improve precision of estimators using…

Statistics Theory · Mathematics 2025-11-12 Dogyoon Song

When we use the normal mixture model, the optimal number of the components describing the data should be determined. Testing homogeneity is good for this purpose; however, to construct its theory is challenging, since the test statistic…

Statistics Theory · Mathematics 2019-12-24 Natsuki Kariya , Sumio Watanabe

Data-driven risk analysis involves the inference of probability distributions from measured or simulated data. In the case of a highly reliable system, such as the electricity grid, the amount of relevant data is often exceedingly limited,…

Methodology · Statistics 2017-07-11 Simon H. Tindemans , Goran Strbac

In applications of group testing in networks, e.g. identifying individuals who are infected by a disease spread over a network, exploiting correlation among network nodes provides fundamental opportunities in reducing the number of tests…

Information Theory · Computer Science 2023-03-21 Hesam Nikpey , Jungyeol Kim , Xingran Chen , Saswati Sarkar , Shirin Saeedi Bidokhti

In this paper, we focus on the problem of stable prediction across unknown test data, where the test distribution is agnostic and might be totally different from the training one. In such a case, previous machine learning methods might…

Machine Learning · Computer Science 2020-06-11 Kun Kuang , Bo Li , Peng Cui , Yue Liu , Jianrong Tao , Yueting Zhuang , Fei Wu

The paper introduces a novel approach to global sensitivity analysis, grounded in the variance-covariance structure of random variables derived from random measures. The proposed methodology facilitates the application of…

Methodology · Statistics 2025-10-20 Caleb Deen Bastian , Herschel Rabitz , Grzegorz A Rempala

Hypothesis tests based on linear models are widely accepted by organizations that regulate clinical trials. These tests are derived using strong assumptions about the data-generating process so that the resulting inference can be based on…

Applications · Statistics 2018-09-13 Kellie Ottoboni , Fraser Lewis , Luigi Salmaso

Nested error regression models are useful tools for analysis of grouped data, especially in the case of small area estimation. This paper suggests a nested error regression model using uncertain random effects in which the random effect in…

Methodology · Statistics 2017-02-28 Shonosuke Sugasawa , Tatsuya Kubokawa

We propose a two-sample testing procedure based on learned deep neural network representations. To this end, we define two test statistics that perform an asymptotic location test on data samples mapped onto a hidden layer. The tests are…

Machine Learning · Statistics 2020-03-11 Matthias Kirchler , Shahryar Khorasani , Marius Kloft , Christoph Lippert

We study goodness-of-fit testing for non-causal autoregressive time series with non-Gaussian stable noise. To model time series exhibiting sharp spikes or occasional bursts of outlying observations, the exponent of the non-Gaussian stable…

Statistics Theory · Mathematics 2012-09-19 Yunwei Cui , Rongning Wu , Thomas J. Fisher

This paper deals with a general class of transformation models that contains many important semiparametric regression models as special cases. It develops a self-induced smoothing for the maximum rank correlation estimator, resulting in…

Methodology · Statistics 2013-02-28 Junyi Zhang , Zhezhen Jin , Yongzhao Shao , Zhiliang Ying