English
Related papers

Related papers: Improving Pearson's chi-squared test: hypothesis t…

200 papers

Applying the standard weighted mean formula, [sum_i {n_i sigma^{-2}_i}] / [sum_i {sigma^{-2}_i}], to determine the weighted mean of data, n_i, drawn from a Poisson distribution, will, on average, underestimate the true mean by ~1 for all…

Astrophysics · Physics 2009-10-31 Kenneth J. Mighell

We introduce a novel probabilistic group testing framework, termed Poisson group testing, in which the number of defectives follows a right-truncated Poisson distribution. The Poisson model has a number of new applications, including…

Information Theory · Computer Science 2023-07-19 Amin Emad , Olgica Milenkovic

Least-squares fits are an important tool in many data analysis applications. In this paper, we review theoretical results, which are relevant for their application to data from counting experiments. Using a simple example, we illustrate the…

Data Analysis, Statistics and Probability · Physics 2019-06-07 Hans Dembinski , Michael Schmelling , Roland Waldi

Given samples from an unknown distribution $p$, is it possible to distinguish whether $p$ belongs to some class of distributions $\mathcal{C}$ versus $p$ being far from every distribution in $\mathcal{C}$? This fundamental question has…

Data Structures and Algorithms · Computer Science 2015-12-09 Jayadev Acharya , Constantinos Daskalakis , Gautam Kamath

This paper discusses estimation and limited information goodness-of-fit test statistics in factor models for binary data using pairwise likelihood estimation and sampling weights. The paper extends the applicability of pairwise likelihood…

Methodology · Statistics 2026-03-30 Haziq Jamil , Irini Moustaki , Chris Skinner

The ultimate limits of quantum state discrimination are often thought to be captured by asymptotic bounds that restrict the achievable error probabilities, notably the quantum Chernoff and Hoeffding bounds. Here we study hypothesis testing…

Quantum Physics · Physics 2025-12-10 Kaiyuan Ji , Bartosz Regula

Statistical data is often analyzed as a contingency table, sometimes with empty cells called zeros. Such sparse tables can be due to scarse observations classified in numerous categories, as for example in genetic association studies. Thus,…

Statistics Theory · Mathematics 2010-07-28 Audrey Finkler

Statistical data is often analyzed as a contingency table, sometimes with empty cells called zeros. Such sparse tables can be due to scarse observations classified in numerous categories, as for example in genetic association studies. Thus,…

Statistics Theory · Mathematics 2010-07-28 Audrey Finkler

We give a general unified method that can be used for $L_1$ {\em closeness testing} of a wide range of univariate structured distribution families. More specifically, we design a sample optimal and computationally efficient algorithm for…

Data Structures and Algorithms · Computer Science 2015-08-25 Ilias Diakonikolas , Daniel M. Kane , Vladimir Nikishkin

In this paper, we develop new test statistics for private hypothesis testing. These statistics are designed specifically so that their asymptotic distributions, after accounting for noise added for privacy concerns, match the asymptotics of…

Statistics Theory · Mathematics 2016-10-26 Daniel Kifer , Ryan Rogers

Active learning can reduce the number of samples needed to perform a hypothesis test and to estimate the parameters of a model. In this paper, we revisit the work of Chernoff that described an asymptotically optimal algorithm for performing…

Machine Learning · Statistics 2022-03-14 Subhojyoti Mukherjee , Ardhendu Tripathy , Robert Nowak

Consider two problems about an unknown probability distribution $p$: 1. How many samples from $p$ are required to test if $p$ is supported on $n$ elements or not? Specifically, given samples from $p$, determine whether it is supported on at…

Data Structures and Algorithms · Computer Science 2026-05-27 Renato Ferreira Pinto , Nathaniel Harms

We describe an approximation to the widely-used Poisson-likelihood chi-square using a linear combination of Neyman's and Pearson's chi-squares, namely "combined Neyman-Pearson chi-square" ($\chi^2_{\mathrm{CNP}}$). Through analytical…

Data Analysis, Statistics and Probability · Physics 2020-02-26 Xiangpan Ji , Wenqiang Gu , Xin Qian , Hanyu Wei , Chao Zhang

This paper revisits a meta-analysis method proposed by Pearson [Biometrika 26 (1934) 425--442] and first used by David [Biometrika 26 (1934) 1--11]. It was thought to be inadmissible for over fifty years, dating back to a paper of Birnbaum…

Statistics Theory · Mathematics 2009-11-19 Art B. Owen

The problem of robust hypothesis testing is studied, where under the null and the alternative hypotheses, the data-generating distributions are assumed to be in some uncertainty sets, and the goal is to design a test that performs well…

Signal Processing · Electrical Eng. & Systems 2023-08-08 Zhongchang Sun , Shaofeng Zou

Meta-analysis seeks to combine the results of several experiments in order to improve the accuracy of decisions. It is common to use a test for homogeneity to determine if the results of the several experiments are sufficiently similar to…

Methodology · Statistics 2009-08-01 Elena Kulinskaya , Michael B. Dollinger , Kirsten Bjørkestøl

An active hypothesis testing problem is formulated. In this problem, the agent can perform a fixed number of experiments and then decide on one of the hypotheses. The agent is also allowed to declare its experiments inconclusive if needed.…

Information Theory · Computer Science 2019-01-23 Dhruva Kartik , Ashutosh Nayyar , Urbashi Mitra

Hypothesis testing is a fundamental issue in statistical inference and has been a crucial element in the development of information sciences. The Chernoff bound gives the minimal Bayesian error probability when discriminating two hypotheses…

Quantum Physics · Physics 2009-11-13 J. Calsamiglia , R. Munoz-Tapia , Ll. Masanes , A. Acin , E. Bagan

In this work, we give a novel general approach for distribution testing. We describe two techniques: our first technique gives sample-optimal testers, while our second technique gives matching sample lower bounds. As a consequence, we…

Data Structures and Algorithms · Computer Science 2016-05-10 Ilias Diakonikolas , Daniel M. Kane

The paper proposes one-to-one transformation of the vector of components $\{Y_{in}\}_{i=1}^m$ of Pearson's chi-square statistic, \[Y_{in}=\frac{\nu_{in}-np_i}{\sqrt{np_i}},\qquad i=1,\ldots,m,\] into another vector $\{Z_{in}\}_{i=1}^m$,…

Statistics Theory · Mathematics 2014-01-06 Estate Khmaladze