中文
相关论文

相关论文: A test for k sample Behrens-Fisher problem in high…

200 篇论文

For a set of dependent random variables, without stationary or the strong mixing assumptions, we derive the asymptotic independence between their sums and maxima. Then we apply this result to high-dimensional testing problems, where we…

统计方法学 · 统计学 2022-05-12 Long Feng , Tiefeng Jiang , Xiaoyun Li , Binghui Liu

The $k$-of-$n$ testing problem involves performing $n$ independent tests sequentially, in order to determine whether/not at least $k$ tests pass. The objective is to minimize the expected cost of testing. This is a fundamental and…

数据结构与算法 · 计算机科学 2026-03-26 Rayen Tan , Viswanath Nagarajan

A common problem in genetics is that of testing whether a set of highly dependent gene expressions differ between two populations, typically in a high-dimensional setting where the data dimension is larger than the sample size. Most…

统计方法学 · 统计学 2015-03-11 Måns Thulin

Along the lines of Janssen's and Pfanzagl's work the testing theory for statistical functionals is further developed for non-parametric one-sample problems. Efficient tests for the one-sided and two-sided problems are derived for…

统计理论 · 数学 2013-11-11 Vladimir Ostrovski

We propose a two-sample test for covariance matrices in the high-dimensional regime, where the dimension diverges proportionally to the sample size. Our hybrid test combines a Frobenius-norm-based statistic as considered in Li and Chen…

统计理论 · 数学 2025-06-10 Thomas Lam , Nina Dörnemann , Holger Dette

Two semimetrics on probability distributions are proposed, given as the sum of differences of expectations of analytic functions evaluated at spatial or frequency locations (i.e, features). The features are chosen so as to maximize the…

机器学习 · 统计学 2016-10-31 Wittawat Jitkrittum , Zoltan Szabo , Kacper Chwialkowski , Arthur Gretton

We discuss a one-sample location test that can be used in the case of high-dimensional data. For high-dimensional data, the power of Hotelling's test decrises when the dimension is close to the sample size. To address this loss of power,…

统计理论 · 数学 2014-05-13 Masashi Hyodo , Takahiro Nishiyama

This paper studies new tests for the number of latent factors in a large cross-sectional factor model with small time dimension. These tests are based on the eigenvalues of variance-covariance matrices of (possibly weighted) asset returns,…

计量经济学 · 经济学 2022-10-31 Alain-Philippe Fortin , Patrick Gagliardini , Olivier Scaillet

Experimental comparisons of performance represent an important aspect of research on optimization algorithms. In this work we present a methodology for defining the required sample sizes for designing experiments with desired statistical…

神经与进化计算 · 计算机科学 2018-10-16 Felipe Campelo , Fernanda Takahashi

In the problem of composite hypothesis testing, identifying the potential uniformly most powerful (UMP) unbiased test is of great interest. Beyond typical hypothesis settings with exponential family, it is usually challenging to prove the…

统计方法学 · 统计学 2022-08-03 Tianyu Zhan , Jian Kang

Combining individual p-values to aggregate multiple small effects has a long-standing interest in statistics, dating back to the classic Fisher's combination test. In modern large-scale data analysis, correlation and sparsity are common…

统计方法学 · 统计学 2018-11-30 Yaowu Liu , Jun Xie

Missing data is a common issue in many biomedical studies. Under a paired design, some subjects may have missing values in either one or both of the conditions due to loss of follow-up, insufficient biological samples, etc. Such partially…

In typical high dimensional statistical inference problems, confidence intervals and hypothesis tests are performed for a low dimensional subset of model parameters under the assumption that the parameters of interest are unconstrained.…

统计方法学 · 统计学 2019-11-19 Ming Yu , Varun Gupta , Mladen Kolar

In this paper, we consider the problem of testing independence in high-dimensional settings with missing data. Building upon a recently proposed Kendall-based statistic, we introduce two new modifications specifically designed to…

统计方法学 · 统计学 2026-04-28 Marija Cuparić , Bojana Milošević , Jelena Radojević

Bayesian models offer great flexibility for clustering applications---Bayesian nonparametrics can be used for modeling infinite mixtures, and hierarchical Bayesian models can be utilized for sharing clusters across multiple data sets. For…

机器学习 · 计算机科学 2012-06-15 Brian Kulis , Michael I. Jordan

We consider three problems in high-dimensional Gaussian linear mixed models. Without any assumptions on the design for the fixed effects, we construct an asymptotic $F$-statistic for testing whether a collection of random effects is zero,…

统计理论 · 数学 2019-07-30 Michael Law , Ya'acov Ritov

This paper establishes the asymptotic independence between the quadratic form and maximum of a sequence of independent random variables. Based on this theoretical result, we find the asymptotic joint distribution for the quadratic form and…

统计方法学 · 统计学 2023-08-03 Dachuan Chen , Decai Liang , Long Feng

In this paper, we consider testing the homogeneity of risk differences in independent binomial distributions especially when data are sparse. We point out some drawback of existing tests in either controlling a nominal size or obtaining…

统计方法学 · 统计学 2018-05-31 Junyong Park , Iris Ivy Gauran

In this paper, we consider procedures for testing hypotheses on the dimension of the linear span generated by a growing number of $p\times p$ covariance matrices from independent $q$ populations. Under a proper limiting scheme where all the…

统计理论 · 数学 2026-02-16 Tianxing Mei , Chen Wang , Jianfeng Yao

Statistical techniques are used in all branches of science to determine the feasibility of quantitative hypotheses. One of the most basic applications of statistical techniques in comparative analysis is the test of equality of two…

统计方法学 · 统计学 2018-05-01 Ayanendranath Basu , Abhijit Mandal , Nirian Martin , Leandro Pardo