English
Related papers

Related papers: A Robust Wald-type Test for Testing the Equality o…

200 papers

We introduce a generalized formulation of mutual information (MI) based on the extended Bregman divergence, a framework that subsumes the generalized S-Bregman (GSB) divergence family. The GSB divergence unifies two important classes of…

Methodology · Statistics 2026-02-05 Arijit Pyne

Linear regression with the classical normality assumption for the error distribution may lead to an undesirable posterior inference of regression coefficients due to the potential outliers. This paper considers the finite mixture of two…

Methodology · Statistics 2021-01-12 Yasuyuki Hamura , Kaoru Irie , Shonosuke Sugasawa

We study a novel class of affine invariant and consistent tests for multivariate normality. The tests are based on a characterization of the standard $d$-variate normal distribution by means of the unique solution of an initial value…

Statistics Theory · Mathematics 2020-07-07 Bruno Ebner , Norbert Henze , David Strieder

A large number of models of the species abundance distribution (SAD) have been proposed, many of which are generically similar to the log-normal distribution, from which they are often indistinguishable when describing a given data set.…

Populations and Evolution · Quantitative Biology 2008-05-02 Craig R. Powell , Alan J. McKane

The robust rank-order test (Fligner and Policello, 1981) was designed as an improvement of the non-parametric Wilcoxon-Mann-Whitney U-test to be more appropriate when the samples being compared have unequal variance. However, it tends to be…

Methodology · Statistics 2020-09-08 Nirvik Sinha

In Econometrics, the Breusch-Pagan test-statistic has become an iconic application of the Lagrange multipliers (LM) test. We shall introduce beta-score LM tests for heteroscedasticity in linear regression models, which trades-off the degree…

Statistics Theory · Mathematics 2023-01-19 Nirian Martin

In this paper, we propose a novel approach to test the equality of high-dimensional mean vectors of several populations via the weighted $L_2$-norm. We establish the asymptotic normality of the test statistics under the null hypothesis. We…

Statistics Theory · Mathematics 2024-02-01 Jianghao Li , Zhenzhen Niu , Shizhe Hong , Zhidong Bai

Normalizing flows are a popular class of models for approximating probability distributions. However, their invertible nature limits their ability to model target distributions whose support have a complex topological structure, such as…

Machine Learning · Statistics 2022-02-25 Vincent Stimper , Bernhard Schölkopf , José Miguel Hernández-Lobato

Motivated by applications in text mining and discrete distribution inference, we investigate the testing for equality of probability mass functions of $K$ groups of high-dimensional multinomial distributions. A test statistic, which is…

Methodology · Statistics 2023-11-28 T. Tony Cai , Zheng Tracy Ke , Paxton Turner

We study a likelihood ratio test for the location of the mode of a log-concave density. Our test is based on comparison of the log-likelihoods corresponding to the unconstrained maximum likelihood estimator of a log-concave density and the…

Statistics Theory · Mathematics 2018-06-05 Charles R. Doss , Jon A. Wellner

In this work, we revisit the problem of uniformity testing of discrete probability distributions. A fundamental problem in distribution testing, testing uniformity over a known domain has been addressed over a significant line of works, and…

Data Structures and Algorithms · Computer Science 2017-08-17 Tuğkan Batu , Clément L. Canonne

Zhang (2019) presented a general estimation approach based on the Gaussian distribution for general parametric models where the likelihood of the data is difficult to obtain or unknown, but the mean and variance-covariance matrix are known.…

Statistics Theory · Mathematics 2023-02-15 Ángel Felipe , María Jaenada , Pedro Miranda , Leandro Pardo

We study the equivalence testing problem where the goal is to determine if the given two unknown distributions on $[n]$ are equal or $\epsilon$-far in the total variation distance in the conditional sampling model (CFGM, SICOMP16; CRS,…

Data Structures and Algorithms · Computer Science 2023-08-23 Diptarka Chakraborty , Sourav Chakraborty , Gunjan Kumar

We propose a likelihood ratio test framework for testing normal mean vectors in high-dimensional data under two common scenarios: the one-sample test and the two-sample test with equal covariance matrices. We derive the test statistics…

Methodology · Statistics 2018-09-25 Zongliang Hu , Tiejun Tong , Marc G. Genton

We give a general unified method that can be used for $L_1$ {\em closeness testing} of a wide range of univariate structured distribution families. More specifically, we design a sample optimal and computationally efficient algorithm for…

Data Structures and Algorithms · Computer Science 2015-08-25 Ilias Diakonikolas , Daniel M. Kane , Vladimir Nikishkin

Analyzing polytomous response from a complex survey scheme, like stratified or cluster sampling is very crucial in several socio-economics applications. We present a class of minimum quasi weighted density power divergence estimators for…

Methodology · Statistics 2019-04-05 Elena Castilla , Abhik Ghosh , Nirian Martin , Leandro Pardo

Two-sample tests for multivariate data and especially for non-Euclidean data are not well explored. This paper presents a novel test statistic based on a similarity graph constructed on the pooled observations from the two samples. It can…

Methodology · Statistics 2024-08-12 Hao Chen , Jerome H. Friedman

We propose a two-sample test for high-dimensional means that requires neither distributional nor correlational assumptions, besides some weak conditions on the moments and tail properties of the elements in the random vectors. This…

Methodology · Statistics 2019-04-17 Kaijie Xue , Fang Yao

From the distributional characterizations that lie at the heart of Stein's method we derive explicit formulae for the mass functions of discrete probability laws that identify those distributions. These identities are applied to develop…

Methodology · Statistics 2022-02-16 Steffen Betsch , Bruno Ebner , Franz Nestmann

There is no agreement over which statistical distribution is most appropriate for modelling citation count data. This is important because if one distribution is accepted then the relative merits of different citation-based indicators, such…

Digital Libraries · Computer Science 2016-03-17 Mike Thelwall