English
Related papers

Related papers: Estimation of Goodness-of-Fit in Multidimensional …

200 papers

In the statistical literature, as well as in artificial intelligence and machine learning, measures of discrepancy between two probability distributions are largely used to develop measures of goodness-of-fit. We concentrate on quadratic…

Methodology · Statistics 2025-10-01 Marianthi Markatou , Giovanni Saraceno

Different types of two- and three-dimensional representations of a finite metric space are studied that focus on the accurate representation of the linear order among the distances rather than their actual values. Lower and upper bounds for…

Combinatorics · Mathematics 2007-05-23 Jobst Heitzig

The goal of clustering is to group similar objects into meaningful partitions. This process is well understood when an explicit similarity measure between the objects is given. However, far less is known when this information is not readily…

Machine Learning · Computer Science 2020-10-12 Michaël Perrot , Pascal Mattia Esser , Debarghya Ghoshdastidar

The problem of making practical, useful goodness of fit tests in the Bayesian paradigm is largely open. We introduce a class of special cases (testing for uniformity: have the cards been shuffled enough; does my random generator work) and a…

Methodology · Statistics 2018-04-11 Persi Diaconis , Guanyang Wang

Suppose we have an observed path from a point process counting event occurrences in a large population. Based on the observed path, we would like to test the null hypothesis that the conditional intensity of the point process belongs to a…

Statistics Theory · Mathematics 2026-05-18 Sami Umut Can , Estate V. Khmaladze , Roger J. A. Laeven

We develop here several goodness-of-fit tests for testing the k-monotonicity of a discrete density, based on the empirical distribution of the observations. Our tests are non-parametric, easy to implement and are proved to be asymptotically…

Methodology · Statistics 2017-08-30 Jade Giguelay , Sylvie Huet

Log-linear models are widely used to express the association in multivariate frequency data on contingency tables. The paper focuses on the power analysis for testing the goodness-of-fit hypothesis for this model type. Conventionally, for…

Methodology · Statistics 2024-05-06 Anna Klimova

Performance of classifiers is often measured in terms of average accuracy on test data. Despite being a standard measure, average accuracy fails in characterizing the fit of the model to the underlying conditional law of labels given the…

Methodology · Statistics 2023-09-01 Adel Javanmard , Mohammad Mehrabi

A survey of goodness-of-fit and symmetry tests based on the characterization properties of distributions is presented. This approach became popular in recent years. In most cases the test statistics are functionals of $U$-empirical…

Statistics Theory · Mathematics 2017-07-07 Ya. Yu. Nikitin

We present a maximum likelihood method for fitting two-dimensional model distributions to stellar data in colour-magnitude space. This allows one to include (for example) binary stars in an isochronal population. The method also allows one…

Astrophysics · Physics 2009-11-11 Tim Naylor , Rob Jeffries

Multi-model mimicry (MMM) is a flexible model selection technique for comparison of multiple, non-nested models on any desired goodness-of-fit criteria. Applicable to any set of candidate models that are 1) able to be fit to observed data,…

Methodology · Statistics 2019-12-17 Lachlann McArthur , Melissa A. Humphries

Clinical trials involving paired organs often yield a mixture of unilateral and bilateral data, where each subject may contribute either one or two responses under certain circumstances. While unilateral responses from different individuals…

Methodology · Statistics 2025-07-01 Jia Zhou , Chang-Xing Ma

Despite compelling theoretical arguments, the use of clusters as cosmological probes is, in practice, frequently questioned because of the many uncertainties impinging on cluster mass estimates. Our aim is to develop a fully self-consistent…

Cosmology and Nongalactic Astrophysics · Physics 2017-11-29 M. Pierre , A. Valotti , L. Faccioli , N. Clerc , R. Gastaud , E. Koulouridis , F. Pacaud

We investigate sample-based learning of conditional distributions on multi-dimensional unit boxes, allowing for different dimensions of the feature and target spaces. Our approach involves clustering data near varying query points in the…

Machine Learning · Statistics 2024-06-14 Cyril Bénézet , Ziteng Cheng , Sebastian Jaimungal

Clusters of sizes ranging from two to five are studied by variational quantum Monte Carlo techniques. The clusters consist of Ar, Ne and hypothetical lighter (``$1 \over 2$-Ne") atoms. A general form of trial function is developed for which…

chem-ph · Physics 2009-10-22 Andrei Mushinski , M. P. Nightingale

As the largest gravitationally bound objects in the universe, clusters of galaxies may contain a fair sample of the baryonic mass fraction of the universe. Since the gas mass fraction from the hot ICM is believed to be constant in time, the…

Astrophysics · Physics 2009-10-30 K. Rines , W. Forman , U. Pen , C. Jones , R. Burg

Various distribution free goodness-of-fit test procedures have been extracted from literature. We present two new binning free tests, the univariate three-region-test and the multivariate energy test. The power of the selected tests with…

Probability · Mathematics 2007-05-23 B. Aslan , G. Zech

There is no, nor will there ever be, single best clustering algorithm. Nevertheless, we would still like to be able to distinguish between methods that work well on certain task types and those that systematically underperform. Clustering…

Machine Learning · Computer Science 2025-10-16 Marek Gagolewski

[...] This paper presents a rigorous derivation of a goodness-of-fit statistics for colour-magnitude diagrams (CMD). We discuss the reliability of the underlying assumptions and their validity. We derived the distribution of the sum of…

Solar and Stellar Astrophysics · Physics 2021-05-26 G. Valle , M. Dell'Omodarme , E. Tognelli

Bayesian model-based clustering is a widely applied procedure for discovering groups of related observations in a dataset. These approaches use Bayesian mixture models, estimated with MCMC, which provide posterior samples of the model…

Methodology · Statistics 2018-09-24 Ketong Wang , Michael D. Porter