English
Related papers

Related papers: Two-Sample Test for High-Dimensional Covariance Ma…

200 papers

Assume that X is a set of sample statistics which follow a special case Central Limit Theorem, namely: as the sample size n increases the corresponding distribution becomes multivariate Normal with the mean (of each X) equal to zero and…

Statistics Theory · Mathematics 2014-11-21 Hao Yuan Zhang , Jan Vrbik

Two-sample inference for the difference of population means typically relies upon a Central Limit Theorem approximation. When data are drawn from a Negative Binomial distribution, previous work of Shilane et al. (2010) showed that a Normal…

Methodology · Statistics 2012-03-06 David Shilane , Derek Bean

The problem of detecting changes in covariance for a single pair of features has been studied in some detail, but may be limited in importance or general applicability. In contrast, testing equality of covariance matrices of a {\it set} of…

Methodology · Statistics 2017-12-12 Yi-Hui Zhou

We address the issue of performing testing inference in generalized linear models when the sample size is small. This class of models provides a straightforward way of modeling normal and non-normal data and has been widely used in several…

Methodology · Statistics 2013-08-16 Tiago M. Vargas , Silvia L. P. Ferrari , Artur J. Lemonte

The likelihood ratio test is widely used in exploratory factor analysis to assess the model fit and determine the number of latent factors. Despite its popularity and clear statistical rationale, researchers have found that when the…

Statistics Theory · Mathematics 2025-01-08 Yinqiu He , Zi Wang , Gongjun Xu

Testing the homogeneity of two distributions is fundamental in statistics, but classical procedures may fail under nonignorable nonresponse. In many surveys, callback data record repeated contact attempts and provide auxiliary information…

Methodology · Statistics 2026-04-24 Xinyu Wang , Tao Yu , Chunlin Wang , Pengfei Li

We present a novel method for testing the hypothesis of equality of two correlation matrices using paired high-dimensional datasets. We consider test statistics based on the average of squares, maximum and sum of exceedances of Fisher…

Methodology · Statistics 2018-04-10 Adria Caballe , Natalia Bochkina , Claus Mayer , Ioannis Papastathopoulos

Testing covariance structure is of importance in many areas of statistical analysis, such as microarray analysis and signal processing. Conventional tests for finite-dimensional covariance cannot be applied to high-dimensional data in…

Statistics Theory · Mathematics 2013-10-31 Rongmao Zhang , Liang Peng , Ruodu Wang

This paper investigates a statistical procedure for testing the equality of two independent estimated covariance matrices when the number of potentially dependent data vectors is large and proportional to the size of the vectors, that is,…

Statistics Theory · Mathematics 2020-06-01 Rémy Mariétan , Stephan Morgenthaler

Wilk's theorem, which offers universal chi-squared approximations for likelihood ratio tests, is widely used in many scientific hypothesis testing problems. For modern datasets with increasing dimension, researchers have found that the…

Statistics Theory · Mathematics 2020-08-14 Yinqiu He , Bo Meng , Zhenghao Zeng , Gongjun Xu

For random samples of size n obtained from p-variate normal distributions, we consider the classical likelihood ratio tests (LRT) for their means and covariance matrices in the high-dimensional setting. These test statistics have been…

Statistics Theory · Mathematics 2013-06-04 Tiefeng Jiang , Fan Yang

Hotelling's T-squared test is a classical tool to test if the normal mean of a multivariate normal distribution is a specified one or the means of two multivariate normal means are equal. When the population dimension is higher than the…

Statistics Theory · Mathematics 2021-08-17 Tiefeng Jiang , Ping Li

We propose novel methodology for testing equality of model parameters between two high-dimensional populations. The technique is very general and applicable to a wide range of models. The method is based on sample splitting: the data is…

Methodology · Statistics 2013-01-17 Nicolas Städler , Sach Mukherjee

We investigate a generalized empirical likelihood approach in a two-group setting where the constraints on parameters have a form of U-statistics. In this situation, the summands that consist of the constraints for the empirical likelihood…

Methodology · Statistics 2015-05-04 Jihnhee Yu , Luge Yang , Albert Vexler , Alan D. Hutson

Pearson's Chi-square test is a widely used tool for analyzing categorical data, yet its statistical power has remained theoretically underexplored. Due to the difficulties in obtaining its power function in the usual manner, Cochran (1952)…

Methodology · Statistics 2024-09-24 Qingyang Zhang

We introduce a unified approach to testing a variety of rather general null hypotheses that can be formulated in terms of covariances matrices. These include as special cases, for example, testing for equal variances, equal traces, or for…

Statistics Theory · Mathematics 2020-12-23 Paavo Sattler , Arne C. Bathke , Markus Pauly

For a multivariate linear model, Wilk's likelihood ratio test (LRT) constitutes one of the cornerstone tools. However, the computation of its quantiles under the null or the alternative requires complex analytic approximations and more…

Methodology · Statistics 2018-01-23 Z. Bai , D. Jiang , J. Yao , S. Zheng

The problem of testing changes in covariance has received increasing attention in recent years, especially in the context of high-dimensional testing. A number of approaches have been proposed, all limited to the two-sample problem and…

Methodology · Statistics 2016-09-06 Yi-Hui Zhou

In this research, inferential theory for hypothesis testing under general convex cone alternatives for correlated data is developed. While there exists extensive theory for hypothesis testing under smooth cone alternatives with independent…

Statistics Theory · Mathematics 2007-06-13 Ramani S. Pilla

A common problem in genetics is that of testing whether a set of highly dependent gene expressions differ between two populations, typically in a high-dimensional setting where the data dimension is larger than the sample size. Most…

Methodology · Statistics 2015-03-11 Måns Thulin