English
Related papers

Related papers: Multiple Testing of Partial Conjunction Hypotheses…

200 papers

Genome-wide association studies (GWAS) have identified hundreds of loci at very stringent levels of statistical significance across many different human traits. However, it is now clear that very large samples (n~10^4-10^5) are needed to…

Genomics · Quantitative Biology 2013-08-20 Inti Pedroso

There is a significant literature on methods for incorporating knowledge into multiple testing procedures so as to improve their power and precision. Some common forms of prior knowledge include (a) beliefs about which hypotheses are null,…

Methodology · Statistics 2019-08-07 Aaditya Ramdas , Rina Foygel Barber , Martin J. Wainwright , Michael I. Jordan

In this study, we propose a two-stage procedure for hypothesis testing, where the first stage is conventional hypothesis testing and the second is an equivalence testing procedure using an introduced Empirical Equivalence Bound. In 2016,…

Methodology · Statistics 2020-01-01 Yi Zhao , Brian S. Caffo , Joshua B. Ewen

Identifying phenotypes plays an important role in furthering our understanding of disease biology through practical applications within healthcare and the life sciences. The challenge of dealing with the complexities and noise within…

Applications · Statistics 2023-04-28 Andre Vauvelle , Hamish Tomlinson , Aaron Sim , Spiros Denaxas

Identifying multivariate dependencies in high-dimensional data is an important problem in large-scale inference. This problem has motivated recent advances in mining (partial) correlations, which focus on the challenging ultra-high…

Methodology · Statistics 2025-09-23 Emily Neo , Peter Radchenko , Bala Rajaratnam

This paper builds upon the fundamental work of Niwa et al. [34], which provides the unique possibility to analyze the relative aggregation/folding propensity of the elements of the entire Escherichia coli (E. coli) proteome in a cell-free…

Computational Engineering, Finance, and Science · Computer Science 2015-07-22 Lorenzo Livi , Alessandro Giuliani , Antonello Rizzi

Integrative analyses based on statistically relevant associations between genomics and a wealth of intermediary phenotypes (such as imaging) provide vital insights into their clinical relevance in terms of the disease mechanisms. Estimates…

Applications · Statistics 2022-08-16 Snigdha Panigrahi , Shariq Mohammed , Arvind Rao , Veerabhadran Baladandayuthapani

A standard practice in statistical hypothesis testing is to mention the p-value alongside the accept/reject decision. We show the advantages of mentioning an e-value instead. With p-values, it is not clear how to use an extreme observation…

Methodology · Statistics 2024-04-04 Peter Grünwald

Recombinant Inbred Lines derived from divergent parental lines can display extensive segregation distortion and long-range linkage disequilibrium (LD) between distant loci. These genomic signatures are consistent with epistatic selection…

Applications · Statistics 2018-01-03 P. Behrouzi , E. C. Wit

In an empirical Bayesian setting, we provide a new multiple testing method, useful when an additional covariate is available, that influences the probability of each null hypothesis being true. We measure the posterior significance of each…

Applications · Statistics 2008-07-30 Egil Ferkingstad , Arnoldo Frigessi , Håvard Rue , Gudmar Thorleifsson , Augustine Kong

A number of biomedical problems require performing many hypothesis tests, with an attendant need to apply stringent thresholds. Often the data take the form of a series of predictor vectors, each of which must be compared with a single…

Methodology · Statistics 2014-05-13 Yi-Hui Zhou , Fred Wright

Many testing problems are readily amenable to randomised tests such as those employing data splitting. However despite their usefulness in principle, randomised tests have obvious drawbacks. Firstly, two analyses of the same dataset may…

Methodology · Statistics 2024-09-05 F. Richard Guo , Rajen D. Shah

For a property $P$ and a sub-property $P'$, we say that $P$ is $P'$-partially testable with $q$ queries if there exists an algorithm that distinguishes, with high probability, inputs in $P'$ from inputs $\epsilon$-far from $P$ by using $q$…

Computational Complexity · Computer Science 2013-06-07 Eldar Fischer , Yonatan Goldhirsh , Oded Lachish

Transitioning from Phase 2 to Phase 3 in drug development, at a rate of $\approx$40%, is the most stringent among phase transitions (Hay et al. (2014)). Yet, success rate at Phase 3 leading to approval is only $\approx$50% (Arrowsmith…

Methodology · Statistics 2025-10-29 Yujia Sun , Yang Han , Xingya Wang , Szu-Yu Tang , Yushi Liu , Jason C. Hsu

Correlation analysis is a fundamental problem in statistics. In this paper, we consider the correlation detection problem between a pair of Erdos-Renyi graphs. Specifically, the problem is formulated as a hypothesis testing problem: under…

Statistics Theory · Mathematics 2026-01-21 Dong Huang , Pengkun Yang

We propose a resampling-based fast variable selection technique for detecting relevant single nucleotide polymorphisms (SNP) in a multi-marker mixed effect model. Due to computational complexity, current practice primarily involves testing…

Applications · Statistics 2025-04-30 Subhabrata Majumdar , Saonli Basu , Matt McGue , Snigdhansu Chatterjee

Given a family of null hypotheses $H_{1},\ldots,H_{s}$, we are interested in the hypothesis $H_{s}^{\gamma}$ that at most $\gamma-1$ of these null hypotheses are false. Assuming that the corresponding $p$-values are independent, we are…

Applications · Statistics 2021-04-28 Anh-Tuan Hoang , Thorsten Dickhaus

Fitting models to data is an important part of the practice of science. Advances in machine learning have made it possible to fit more -- and more complex -- models, but have also exacerbated a problem: when multiple models fit the data…

Methodology · Statistics 2025-10-27 Alexandre René , André Longtin

Meta-analysis of multiple genome-wide association studies (GWAS) is effective for detecting single or multi marker associations with complex traits. We develop a flexible procedure ("STAMP") based on mixture models to perform region based…

Methodology · Statistics 2018-01-01 Andriy Derkach , Ruth M. Pfeiffer

We consider the problem of testing the significance of features in high-dimensional settings. In particular, we test for differentially-expressed genes in a microarray experiment. We wish to identify genes that are associated with some type…

Applications · Statistics 2008-11-12 Daniela M. Witten , Robert Tibshirani