English
Related papers

Related papers: Two-Stage Testing in a high dimensional setting

200 papers

A/B tests, also known as randomized controlled experiments (RCTs), are the gold standard for evaluating the impact of new policies, products, or decisions. However, these tests can be costly in terms of time and resources, potentially…

Machine Learning · Statistics 2025-01-03 Shima Nassiri , Mohsen Bayati , Joe Cooprider

Score-based statistical models play an important role in modern machine learning, statistics, and signal processing. For hypothesis testing, a score-based hypothesis test is proposed in \cite{wu2022score}. We analyze the performance of this…

Signal Processing · Electrical Eng. & Systems 2024-02-06 Enmao Diao , Taposh Banerjee , Vahid Tarokh

This paper analyses the use of bootstrap methods to test for parameter change in linear models estimated via Two Stage Least Squares (2SLS). Two types of test are considered: one where the null hypothesis is of no change and the alternative…

Econometrics · Economics 2020-02-03 Otilia Boldea , Adriana Cornea-Madeira , Alastair R. Hall

While the problem of testing multivariate normality has received considerable attention in the classical low-dimensional setting where the sample size $n$ is much larger than the feature dimension $d$ of the data, there is presently a…

Methodology · Statistics 2025-12-23 Xin Bing , Derek Latremouille

The $\gamma$-FDP and $k$-FWER multiple testing error metrics, which are tail probabilities of the respective error statistics, have become popular recently as less-stringent alternatives to the FDR and FWER. We propose general and flexible…

Methodology · Statistics 2016-12-20 Jay Bartroff

We investigate the problem of statistical inference for logistic regression with high-dimensional covariates in settings where dependence among individuals is induced by an underlying Markov random field. Going beyond the pairwise…

Statistics Theory · Mathematics 2026-03-23 Josh Miles , Sohom Bhattacharya

We address a common problem in large-scale data analysis, and especially the field of genetics, the huge-scale testing problem, where millions to billions of hypotheses are tested together creating a computational challenge to perform…

Methodology · Statistics 2015-01-22 Vered Madar , Sandra Batista

We investigate the significance of change-points within fully nonparametric regression contexts, with a particular focus on panel data where data generation processes vary across units, and error terms may display complex dependency…

Econometrics · Economics 2025-01-07 Likai Chen , Georg Keilbar , Liangjun Su , Weining Wang

In preliminary analysis of control charts, one may encounter multiple shifts and/or outliers especially with a large number of observations. The following paper addresses this problem. A statistical model for detecting and estimating…

Applications · Statistics 2014-03-05 Issac Shams , Saeede Ajorlou , Kai Yang

Consider a case-control study in which the aim is to assess the effect of a factor on disease occurrence. We suppose that this factor is dichotomous. Also suppose that the data consists of two strata, each stratum summarized by a two-by-two…

Methodology · Statistics 2017-10-18 Paul Kabaila , Dilshani Tissera

The analysis of screening experiments is often done in two stages, starting with factor selection via an analysis under a main effects model. The success of this first stage is influenced by three components: (1) main effect estimators'…

Methodology · Statistics 2024-03-19 Jonathan W. Stallrich , Michael McKibben

A daunting challenge faced by modern biological sciences is finding an efficient and computationally feasible approach to deal with the curse of high dimensionality. The problem becomes even more severe when the research focus is on…

Methodology · Statistics 2013-04-16 Hung Hung , Yu-Tin Lin , Pengwen Chen , Chen-Chien Wang , Su-Yun Huang , Jung-Ying Tzeng

This paper studies inference in two-stage randomized experiments under covariate-adaptive randomization. In the initial stage of this experimental design, clusters (e.g., households, schools, or graph partitions) are stratified and randomly…

Econometrics · Economics 2026-01-16 Jizhou Liu

Testing mutual independence among multiple random variables is a fundamental problem in statistics, with wide applications in genomics, finance, and neuroscience. In this paper, we propose a new class of tests for high-dimensional mutual…

Applications · Statistics 2026-01-28 Ping Zhao , Huifang Ma

The paper considers variable selection in linear regression models where the number of covariates is possibly much larger than the number of observations. High dimensionality of the data brings in many complications, such as (possibly…

Methodology · Statistics 2016-11-29 Haeran Cho , Piotr Fryzlewicz

Classification is a fundamental problem in machine learning and data mining. During the past decades, numerous classification methods have been presented based on different principles. However, most existing classifiers cast the…

Machine Learning · Computer Science 2019-04-23 Zengyou He , Chaohua Sheng , Yan Liu , Quan Zou

Particularly in genomics, but also in other fields, it has become commonplace to undertake highly multiple Student's $t$-tests based on relatively small sample sizes. The literature on this topic is continually expanding, but the main…

Statistics Theory · Mathematics 2010-10-11 Peter Hall , Qiying Wang

Two-stage randomized experiments are becoming an increasingly popular experimental design for causal inference when the outcome of one unit may be affected by the treatment assignments of other units in the same cluster. In this paper, we…

Methodology · Statistics 2022-10-21 Zhichao Jiang , Kosuke Imai , Anup Malani

Gene expression levels, hormone secretion, and internal body temperature each oscillate over an approximately 24-hour cycle, or display circadian rhythms. Many circadian biology studies have investigated how these rhythms vary across…

We consider testing the equality of two high-dimensional covariance matrices by carrying out a multi-level thresholding procedure, which is designed to detect sparse and faint differences between the covariances. A novel U-statistic…

Statistics Theory · Mathematics 2019-10-30 Song Xi Chen , Bin Guo , Yumou Qiu
‹ Prev 1 8 9 10 Next ›