English
Related papers

Related papers: Composite Goodness-of-fit Tests with Kernels

200 papers

Interpolation and approximation of functionals with conditionally positive definite kernels is considered on sets of centers that are not determining for polynomials. It is shown that polynomial consistency is sufficient in order to define…

Numerical Analysis · Mathematics 2025-08-26 Oleg Davydov

Kernel discrepancies are a powerful tool for analyzing worst-case errors in quasi-Monte Carlo (QMC) methods. Building on recent advances in optimizing such discrepancy measures, we extend the subset selection problem to the setting of…

Machine Learning · Statistics 2025-11-05 Deyao Chen , François Clément , Carola Doerr , Nathan Kirk

This work builds a unified framework for the study of quadratic form distance measures as they are used in assessing the goodness of fit of models. Many important procedures have this structure, but the theory for these methods is dispersed…

Statistics Theory · Mathematics 2008-12-18 Bruce G. Lindsay , Marianthi Markatou , Surajit Ray , Ke Yang , Shu-Chuan Chen

Kernel-based modal statistical methods include mode estimation, regression, and clustering. Estimation accuracy of these methods depends on the kernel used as well as the bandwidth. We study effect of the selection of the kernel function to…

Machine Learning · Statistics 2023-04-21 Ryoya Yamasaki , Toshiyuki Tanaka

In this paper, we propose a test for the equality of multiple distributions based on kernel mean embeddings. Our framework provides a flexible way to handle multivariate or even high-dimensional data by virtue of kernel methods and allows…

Statistics Theory · Mathematics 2020-06-08 Ilmun Kim

We propose kernel-type smoothed Kolmogorov-Smirnov and Cram\'{e}r-von Mises tests for data on general interval, using bijective transformations. Though not as severe as in the kernel density estimation, utilizing naive kernel method…

Methodology · Statistics 2020-05-29 Rizky Reza Fauzi , Yoshihiko Maesono

The characterization of covariate effects on model parameters is a crucial step during pharmacokinetic/pharmacodynamic analyses. While covariate selection criteria have been studied extensively, the choice of the functional relationship…

Methodology · Statistics 2024-04-09 Niklas Hartung , Martin Wahl , Abhishake Rastogi , Wilhelm Huisinga

A common problem in physics is to fit regression data by a parametric class of functions, and to decide whether a certain functional form allows for a good fit of the data. Common goodness of fit methods are based on the calculation of the…

Astrophysics · Physics 2009-11-07 N. Bissantz , A. Munk

We consider the variable selection problem for two-sample tests, aiming to select the most informative variables to determine whether two collections of samples follow the same distribution. To address this, we propose a novel framework…

Machine Learning · Statistics 2024-12-23 Jie Wang , Santanu S. Dey , Yao Xie

Imbalanced response variable distribution is a common occurrence in data science. In fields such as fraud detection, medical diagnostics, system intrusion detection and many others where abnormal behavior is rarely observed the data under…

Machine Learning · Computer Science 2019-11-21 Firuz Kamalov

Kernel density estimation is a widely used nonparametric approach to estimate an unknown distribution. Recent work in Bayesian predictive inference has considered stochastic processes formed by specifying the predictive distribution for the…

Methodology · Statistics 2026-05-15 Torey Hilbert

A goodness-of-fit test for the fitting of a parametric model to data obtained from a detector with finite resolution and limited acceptance is proposed. The parameters of the model are found by minimization of a statistic that is used for…

Data Analysis, Statistics and Probability · Physics 2015-03-17 N. D. Gagunashvili

We use a Stein identity to define a new class of parametric distributions which we call ``independent additive weighted bias distributions.'' We investigate related $L^2$-type discrepancy measures, empirical versions of which not only…

Methodology · Statistics 2023-04-27 Bruno Ebner , Yvik Swan

It is well known that nonparametric regression estimation and inference procedures are subject to the curse of dimensionality. Moreover, model interpretability usually decreases with the data dimension. Therefore, model-free variable…

Methodology · Statistics 2025-05-22 Daniel Diz-Castro , Manuel Febrero-Bande , Wenceslao González-Manteiga

An accurate assessment of a model's complexity is crucial for topics such as interpretation, generalization, and model selection. However, most existing complexity measures either rely on heuristic assumptions or are computationally…

Machine Learning · Statistics 2026-05-21 Oskar Allerbo , Thomas B. Schön

When modeling a probability distribution with a Bayesian network, we are faced with the problem of how to handle continuous variables. Most previous work has either solved the problem by discretizing, or assumed that the data are generated…

Machine Learning · Computer Science 2013-02-21 George H. John , Pat Langley

We consider two division models for structured cell populations, where cells can grow, age and divide. These models have been introduced in the literature under the denomination of `mitosis' and `adder' models. In the recent years, there…

We survey classical kernel methods for providing nonparametric solutions to problems involving measurement error. In particular we outline kernel-based methodology in this setting, and discuss its basic properties. Then we point to close…

Methodology · Statistics 2010-03-02 Aurore Delaigle , Peter Hall

Consider a random sample from a continuous multivariate distribution function $F$ with copula $C$. In order to test the null hypothesis that $C$ belongs to a certain parametric family, we construct an empirical process on the unit hypercube…

Statistics Theory · Mathematics 2018-12-20 Sami Umut Can , John H. J. Einmahl , Roger J. A. Laeven

There has been great interest recently in applying nonparametric kernel mixtures in a hierarchical manner to model multiple related data samples jointly. In such settings several data features are commonly present: (i) the related samples…

Methodology · Statistics 2017-04-18 Jacopo Soriano , Li Ma