中文
相关论文

相关论文: Two-sample tests for high-dimension, strongly spik…

200 篇论文

Estimating the number of spikes in a spiked model is an important problem in many areas such as signal processing. Most of the classical approaches assume a large sample size $n$ whereas the dimension $p$ of the observations is kept small.…

统计理论 · 数学 2014-06-04 Damien Passemier , Jian-Feng Yao

High-dimensional data, where the dimension of the feature space is much larger than sample size, arise in a number of statistical applications. In this context, we construct the generalized multivariate sign transformation, defined as a…

统计方法学 · 统计学 2021-07-05 Subhabrata Majumdar , Snigdhansu Chatterjee

Non-uniform sampling arises when an experimenter does not have full control over the sampling characteristics of the process under investigation. Moreover, it is introduced intentionally in algorithms such as Bayesian optimization and…

机器学习 · 统计学 2020-07-03 Stijn de Waele

This paper proposes a new mutual independence test for a large number of high dimensional random vectors. The test statistic is based on the characteristic function of the empirical spectral distribution of the sample covariance matrix. The…

统计理论 · 数学 2012-05-31 G. M. Pan , J. Gao , Y. Yang , M. Guo

This paper deals with a class of nonparametric two-sample tests for ordered alternatives. The test statistics proposed are based on the number of observations from one sample that precede or exceed a threshold specified by the other sample,…

统计理论 · 数学 2016-11-01 Eugenia Stoimenova , N. Balakrishnan

In this paper, a robust non-parametric measure of statistical dependence, or correlation, between two random variables is presented. The proposed coefficient is a permutation-like statistic that quantifies how much the observed sample S_n :…

统计方法学 · 统计学 2020-07-27 Rami Mahdi

We consider the linear regression problem under semi-supervised settings wherein the available data typically consists of: (i) a small or moderate sized 'labeled' data, and (ii) a much larger sized 'unlabeled' data. Such data arises…

统计方法学 · 统计学 2018-07-02 Abhishek Chakrabortty , Tianxi Cai

Several approaches to testing the hypothesis that two histograms are drawn from the same distribution are investigated. We note that single-sample continuous distribution tests may be adapted to this two-sample grouped data situation. The…

数据分析、统计与概率 · 物理学 2008-04-03 Frank C. Porter

Two-sample tests for multivariate data and especially for non-Euclidean data are not well explored. This paper presents a novel test statistic based on a similarity graph constructed on the pooled observations from the two samples. It can…

统计方法学 · 统计学 2024-08-12 Hao Chen , Jerome H. Friedman

Factor analysis for high-dimensional data is a canonical problem in statistics and has a wide range of applications. However, there is currently no factor model tailored to effectively analyze high-dimensional count responses with…

统计方法学 · 统计学 2024-08-21 Wei Liu , Qingzhi Zhong

In this paper, we propose a novel approach to test the equality of high-dimensional mean vectors of several populations via the weighted $L_2$-norm. We establish the asymptotic normality of the test statistics under the null hypothesis. We…

统计理论 · 数学 2024-02-01 Jianghao Li , Zhenzhen Niu , Shizhe Hong , Zhidong Bai

Statistical techniques are used in all branches of science to determine the feasibility of quantitative hypotheses. One of the most basic applications of statistical techniques in comparative analysis is the test of equality of two…

统计方法学 · 统计学 2018-05-01 Ayanendranath Basu , Abhijit Mandal , Nirian Martin , Leandro Pardo

Targeted syntactic evaluation of subject-verb number agreement in English (TSE) evaluates language models' syntactic knowledge using hand-crafted minimal pairs of sentences that differ only in the main verb's conjugation. The method…

计算与语言 · 计算机科学 2021-04-21 Benjamin Newman , Kai-Siang Ang , Julia Gong , John Hewitt

Motivated by dimension reduction in regression analysis and signal detection, we investigate the order determination for large dimension matrices including spiked models of which the numbers of covariates are proportional to the sample…

统计方法学 · 统计学 2019-11-01 Yicheng Zeng , Lixing Zhu

In this paper, we study limiting laws and consistent estimation criteria for the extreme eigenvalues in a spiked covariance model of dimension $p$. Firstly, for fixed $p$, we propose a generalized estimation criterion that can consistently…

统计理论 · 数学 2026-03-26 Jianwei Hu , Jingfei Zhang , Jianhua Guo , Ji Zhu

In this article, we consider the complete independence test of high-dimensional data. Based on Chatterjee coefficient, we pioneer the development of quadratic test and extreme value test which possess good testing performance for…

统计理论 · 数学 2024-09-17 Liqi Xia , Ruiyuan Cao , Jiang Du , Jun Dai

We propose a test of many zero parameter restrictions in a high dimensional linear iid regression model with $k$ $>>$ $n$ regressors. The test statistic is formed by estimating key parameters one at a time based on many low dimension…

统计理论 · 数学 2023-12-12 Jonathan B. Hill

For a high-dimensional linear model with a finite number of covariates measured with error, we study statistical inference on the parameters associated with the error-prone covariates, and propose a new corrected decorrelated score test and…

统计方法学 · 统计学 2020-01-29 Mengyan Li , Runze Li , Yanyuan Ma

We consider multivariate two-sample tests of means, where the location shift between the two populations is expected to be related to a known graph structure. An important application of such tests is the detection of differentially…

定量方法 · 定量生物学 2014-05-16 Laurent Jacob , Pierre Neuvial , Sandrine Dudoit

We propose and analyse a new Milstein type scheme for simulating stochastic differential equations (SDEs) with highly nonlinear coefficients. Our work is motivated by the need to justify multi-level Monte Carlo simulations for…

数值分析 · 数学 2012-04-10 Desmond J. Higham , Xuerong Mao , Lukasz Szpruch