中文
相关论文

相关论文: Two-sample Behrens--Fisher problems for high-dimen…

200 篇论文

Since model selection is ubiquitous in data analysis, reproducibility of statistical results demands a serious evaluation of reliability of the employed model selection method, no matter what label it may have in terms of good properties.…

统计方法学 · 统计学 2017-05-01 Yanjia Yu , Yi Yang , Yuhong Yang

The test of independence is a crucial component of modern data analysis. However, traditional methods often struggle with the complex dependency structures found in high-dimensional data. To overcome this challenge, we introduce a novel…

统计方法学 · 统计学 2024-09-13 Mingshuo Liu , Doudou Zhou , Hao Chen

This study proposes a simple, trustworthy Chow test in the presence of heteroscedasticity and autocorrelation. The test is based on a series heteroscedasticity and autocorrelation robust variance estimator with judiciously crafted basis…

计量经济学 · 经济学 2019-11-12 Yixiao Sun , Xuexin Wang

Fisher randomization tests for Neyman's null hypothesis of no average treatment effects are considered in a finite population setting associated with completely randomized experiments with more than two treatments. The consequences of using…

统计理论 · 数学 2017-07-26 Peng Ding , Tirthankar Dasgupta

We consider the problem of testing the equality of conditional distributions of a response variable given a vector of covariates between two populations. Such a hypothesis testing problem can be motivated from various machine learning and…

统计方法学 · 统计学 2023-02-24 Xiaoyu Hu , Jing Lei

This paper is concerned with estimation and inference for ultrahigh dimensional partially linear single-index models. The presence of high dimensional nuisance parameter and nuisance unknown function makes the estimation and inference…

统计方法学 · 统计学 2024-04-09 Shijie Cui , Xu Guo , Zhe Zhang

Let $(X\_1,\ldots,X\_n)$ be a $d$-dimensional i.i.d sample from a distribution with density $f$. The problem of detection of a two-component mixture is considered. Our aim is to decide whether $f$ is the density of a standard Gaussian…

统计理论 · 数学 2015-10-01 Béatrice Laurent , Clément Marteau , Cathy Maugis-Rabusseau

Testing equality of two multivariate distributions is a classical problem for which many non-parametric tests have been proposed over the years. Most of the popular two-sample tests, which are asymptotically distribution-free, are based…

统计理论 · 数学 2019-04-17 Bhaswar B. Bhattacharya

Suppose that a data analyst wishes to report the results of a least squares linear regression only if the overall null hypothesis, $H_0^{1:p}: \beta_1= \beta_2 = \ldots = \beta_p=0$, is rejected. This practice, which we refer to as…

统计方法学 · 统计学 2026-05-12 Olivia McGough , Daniela Witten , Daniel Kessler

The Fisher-Snedecor $\mathcal{F}$ distribution has been recently proposed as a more accurate and mathematically tractable composite fading model than traditional established models in some practical cases. In this paper, we firstly derive…

信息论 · 计算机科学 2019-11-27 Hongyang Du , Jiayi Zhang , Kostas P. Peppas , Hui Zhao , Bo Ai , Xiaodan Zhang

In small sample studies with binary outcome data, use of a normal approximation for hypothesis testing can lead to substantial inflation of the type-I error-rate. Consequently, exact statistical methods are necessitated, and accordingly,…

统计方法学 · 统计学 2017-11-29 Michael Grayling , Adrian Mander , James Wason

We present a novel method for testing the hypothesis of equality of two correlation matrices using paired high-dimensional datasets. We consider test statistics based on the average of squares, maximum and sum of exceedances of Fisher…

统计方法学 · 统计学 2018-04-10 Adria Caballe , Natalia Bochkina , Claus Mayer , Ioannis Papastathopoulos

Combining dependent tests of significance has broad applications but the $p$-value calculation is challenging. Current moment-matching methods (e.g., Brown's approximation) for Fisher's combination test tend to significantly inflate the…

统计方法学 · 统计学 2020-03-04 Hong Zhang , Zheyang Wu

The F-measure, which has originally been introduced in information retrieval, is nowadays routinely used as a performance metric for problems such as binary classification, multi-label classification, and structured output prediction.…

Testing hypothesis of independence between two random elements on a joint alphabet is a fundamental exercise in statistics. Pearson's chi-squared test is an effective test for such a situation when the contingency table is relatively small.…

统计理论 · 数学 2025-03-19 Jialin Zhang , Zhiyi Zhang

We propose a non-parametric, two-sample Bayesian test for checking whether or not two data sets share a common distribution. The test makes use of data splitting ideas and does not require priors for high-dimensional parameter vectors as do…

统计方法学 · 统计学 2020-03-16 Jeffery Hart , Taeryon Choi , Naveed Merchant

The problem of detecting changes in covariance for a single pair of features has been studied in some detail, but may be limited in importance or general applicability. In contrast, testing equality of covariance matrices of a {\it set} of…

统计方法学 · 统计学 2017-12-12 Yi-Hui Zhou

We consider the problem of testing the mean of high-dimensional data when the dimension may grow without explicit rate restrictions relative to the sample size. The proposed procedure is based on the statistic V_n = n||Xn||^2, which avoids…

统计理论 · 数学 2026-05-18 Dietmar Ferger

After rejecting the null hypothesis in the analysis of variance, the next step is to make the pairwise comparisons to find out differences in means. The purpose of this paper is threefold. The foremost aim is to suggest expression for…

统计方法学 · 统计学 2023-06-22 Elsayed A. H. Elamir

A number of applications require two-sample testing on ranked preference data. For instance, in crowdsourcing, there is a long-standing question of whether pairwise comparison data provided by people is distributed similar to…

机器学习 · 统计学 2020-11-20 Charvi Rastogi , Sivaraman Balakrishnan , Nihar B. Shah , Aarti Singh