中文
相关论文

相关论文: Estimating the proportion of false null hypotheses…

200 篇论文

The problem of multiple hypothesis testing arises when there are more than one hypothesis to be tested simultaneously for statistical significance. This is a very common situation in many data mining applications. For instance, assessing…

机器学习 · 统计学 2009-06-30 Sami Hanhijärvi , Kai Puolamäki , Gemma C. Garriga

A central problem in Binary Hypothesis Testing (BHT) is to determine the optimal tradeoff between the Type I error (referred to as false alarm) and Type II (referred to as miss) error. In this context, the exponential rate of convergence of…

信息论 · 计算机科学 2021-11-29 Sebastian Espinosa , Jorge F. Silva , Pablo Piantanida

We consider interval estimation of the difference between two binomial proportions. Several methods of constructing such an interval are known. Unfortunately those confidence intervals have poor coverage probability: it is significantly…

统计方法学 · 统计学 2019-03-11 Wojciech Zieliński

We consider a hypothesis testing problem where a part of data cannot be observed. Our helper observes the missed data and can send us a limited amount of information about them. What kind of this limited information will allow us to make…

信息论 · 计算机科学 2020-09-08 Marat V. Burnashev

We propose a novel finite-sample procedure for testing composite null hypotheses. Traditional likelihood ratio tests based on asymptotic $\chi^2$ approximations often exhibit substantial bias in small samples. Our procedure rejects the…

统计方法学 · 统计学 2026-01-07 Joonha Park , Ming Wang

As increasingly complex hypothesis-testing scenarios are considered in many scientific fields, analytic derivation of null distributions is often out of reach. To the rescue comes Monte Carlo testing, which may appear deceptively simple: as…

统计方法学 · 统计学 2015-04-13 Egil Ferkingstad , Lars Holden , Geir Kjetil Sandve

We consider the problem of inference on the signs of $n>1$ parameters. We aim to provide $1-\alpha$ post-hoc confidence bounds on the number of positive and negative (or non-positive) parameters. The guarantee is simultaneous, for all…

统计方法学 · 统计学 2024-03-05 Ruth Heller , Aldo Solari

We propose an improved method to study recent and near-future dark matter direct detection experiments with small numbers of observed events. Our method determines in a quantitative and halo-independent way whether the experiments point…

高能物理 - 唯象学 · 物理学 2015-03-24 Brian Feldstein , Felix Kahlhoefer

Bayesian hypothesis testing is re-examined from the perspective of an a priori assessment of the test statistic distribution under the alternative. By assessing the distribution of an observable test statistic, rather than prior parameter…

统计理论 · 数学 2018-08-28 Hedibert F. Lopes , Nicholas G. Polson

We consider the problem of estimating the proportion $\theta$ of true null hypotheses in a multiple testing context. The setup is classically modeled through a semiparametric mixture with two components: a uniform distribution on interval…

应用统计 · 统计学 2013-01-09 Van Hanh Nguyen , Catherine Matias

The search for new significant peaks over a energy spectrum often involves a statistical multiple hypothesis testing problem. Separate tests of hypothesis are conducted at different locations producing an ensemble of local p-values, the…

数据分析、统计与概率 · 物理学 2016-12-16 Sara Algeri , David A. van Dyk , Jan Conrad , Brandon Anderson

As a convention, p-value is often computed in frequentist hypothesis testing and compared with the nominal significance level of 0.05 to determine whether or not to reject the null hypothesis. The smaller the p-value, the more significant…

统计方法学 · 统计学 2020-02-25 Haolun Shi , Guosheng Yin

We consider the hypothesis testing problem of deciding whether an observed high-dimensional vector has independent normal components or, alternatively, if it has a small subset of correlated components. The correlated components may have a…

统计理论 · 数学 2012-06-04 Ery Arias-Castro , Sébastien Bubeck , Gábor Lugosi

In many large multiple testing problems the hypotheses are divided into families. Given the data, families with evidence for true discoveries are selected, and hypotheses within them are tested. Neither controlling the error-rate in each…

统计理论 · 数学 2011-06-21 Yoav Benjamini , Marina Bogomolov

We consider the problem of interval estimation of the odds ratio. An asymptotic confidence interval is widely applied in medical research. Unfortunately that confidence interval has a poor coverage probability: it is significantly smaller…

统计方法学 · 统计学 2020-11-19 Zofia Zielińska-Kolasińska , Wojciech Zieliński

Intuitively, unfamiliarity should lead to lack of confidence. In reality, current algorithms often make highly confident yet wrong predictions when faced with relevant but unfamiliar examples. A classifier we trained to recognize gender is…

计算机视觉与模式识别 · 计算机科学 2020-09-09 Zhizhong Li , Derek Hoiem

Multiple hypothesis testing is a fundamental problem in high dimensional inference, with wide applications in many scientific fields. In genome-wide association studies, tens of thousands of tests are performed simultaneously to find if any…

统计方法学 · 统计学 2011-11-16 Jianqing Fan , Xu Han , Weijie Gu

Most scientific disciplines use significance testing to draw conclusions about experimental or observational data. This classical approach provides a theoretical guarantee for controlling the number of false positives across a set of…

应用统计 · 统计学 2023-03-06 Stanley E. Lazic

It is quite common in modern research, for a researcher to test many hypotheses. The statistical (frequentist) hypothesis testing framework, does not scale with the number of hypotheses in the sense that naively performing many hypothesis…

统计方法学 · 统计学 2013-06-26 Jonathan Rosenblatt

Correlated proportions arise in longitudinal (panel) studies. A typical example is the ``opinion swing'' problem: ``Has the proportion of people favoring a politician changed after his recent speech to the nation on TV?''. Since the same…

统计理论 · 数学 2007-07-27 Guido Consonni , Luca La Rocca