中文
相关论文

相关论文: Assessing replicability with the sceptical p-value…

200 篇论文

Assessment of replicability is critical to ensure the quality and rigor of scientific research. In this paper, we discuss inference and modeling principles for replicability assessment. Targeting distinct application scenarios, we propose…

统计方法学 · 统计学 2021-05-11 Yi Zhao , Xiaoquan Wen

Empirical science needs to be based on facts and claims that can be reproduced. This calls for replicating the studies that proclaim the claims, but practice in most fields still fails to implement this idea. When such studies emerged in…

其他统计学 · 统计学 2025-08-27 Werner A. Stahel

The two-trials rule in drug regulation requires statistically significant results from two pivotal trials to demonstrate efficacy. However, it is unclear how the effect estimates from both trials should be combined to quantify the drug…

统计方法学 · 统计学 2025-07-08 Samuel Pawel , Małgorzata Roos , Leonhard Held

$P$-values that are derived from continuously distributed test statistics are typically uniformly distributed on $(0,1)$ under least favorable parameter configurations (LFCs) in the null hypothesis. Conservativeness of a $p$-value $P$…

统计方法学 · 统计学 2023-03-13 Daniel Ochieng , Anh-Tuan Hoang , Thorsten Dickhaus

High dimensional case control studies are ubiquitous in the biological sciences, particularly genomics. To maximise power while constraining cost and to minimise type-1 error rates, researchers typically seek to replicate findings in a…

统计方法学 · 统计学 2017-07-11 James Liley

The logical and practical difficulties associated with research interpretation using P values and null hypothesis significance testing have been extensively documented. This paper describes an alternative, likelihood-based approach to…

统计方法学 · 统计学 2021-09-21 Nicholas Adams , Gerard O'Reilly

Adaptive experiments use preliminary analyses of the data to inform further course of action and are commonly used in many disciplines including medical and social sciences. Because the null hypothesis and experimental design are…

统计方法学 · 统计学 2026-05-26 Tobias Freidling , Qingyuan Zhao , Zijun Gao

We study a large-scale one-sided multiple testing problem in which test statistics follow normal distributions with unit variance, and the goal is to identify signals with positive mean effects. A conventional approach is to compute…

统计方法学 · 统计学 2026-05-15 Kwangok Seo , Johan Lim , Hyungwon Choi , Jaesik Jeong

The reproducibility crisis has led to an increasing number of replication studies being conducted. Sample sizes for replication studies are often calculated using conditional power based on the effect estimate from the original study.…

统计方法学 · 统计学 2022-10-19 Charlotte Micheloud , Leonhard Held

Several scientific fields including psychology are undergoing a replication crisis. There are many reasons for this problem, one of which is a misuse of p-values. There are several alternatives to p-values, and in this paper we describe a…

统计方法学 · 统计学 2020-10-05 Brian D. Segal

We propose novel methodology for testing equality of model parameters between two high-dimensional populations. The technique is very general and applicable to a wide range of models. The method is based on sample splitting: the data is…

统计方法学 · 统计学 2013-01-17 Nicolas Städler , Sach Mukherjee

Replicability is a lynchpin for credible discoveries. The partial conjunction (PC) p-value, which combines individual base p-values from multiple similar studies, can gauge whether a feature of interest exhibits replicated signals across…

统计方法学 · 统计学 2025-07-29 Ninh Tran , Dennis Leung

For randomized controlled trials to be conclusive, it is important to set the target sample size accurately at the design stage. Comparing two normal populations, the sample size calculation requires specification of the variance other than…

统计方法学 · 统计学 2026-02-04 Hirotada Maeda , Satoshi Hattori , Tim Friede

Replicability analysis aims to identify the findings that replicated across independent studies that examine the same features. We provide powerful novel replicability analysis procedures for two studies for FWER and for FDR control on the…

统计方法学 · 统计学 2019-03-01 Marina Bogomolov , Ruth Heller

We present the expected values from p-value hacking as a choice of the minimum p-value among $m$ independents tests, which can be considerably lower than the "true" p-value, even with a single trial, owing to the extreme skewness of the…

应用统计 · 统计学 2018-01-29 Nassim Nicholas Taleb

Many testing problems are readily amenable to randomised tests such as those employing data splitting. However despite their usefulness in principle, randomised tests have obvious drawbacks. Firstly, two analyses of the same dataset may…

统计方法学 · 统计学 2024-09-05 F. Richard Guo , Rajen D. Shah

We consider one of the most basic multiple testing problems that compares expectations of multivariate data among several groups. As a test statistic, a conventional (approximate) $t$-statistic is considered, and we determine its rejection…

统计方法学 · 统计学 2016-12-20 Yoshiyuki Ninomiya , Satoshi Kuriki , Toshihiko Shiroishi , Toyoyuki Takada

We theoretically analyze the problem of testing for $p$-hacking based on distributions of $p$-values across multiple studies. We provide general results for when such distributions have testable restrictions (are non-increasing) under the…

计量经济学 · 经济学 2022-05-13 Graham Elliott , Nikolay Kudrin , Kaspar Wuthrich

In clinical studies upon which decisions are based there are two types of errors that can be made: a type I error arises when the decision is taken to declare a positive outcome when the truth is in fact negative, and a type II error arises…

统计方法学 · 统计学 2024-09-19 Andrew P Grieve

Observational healthcare data offer the potential to estimate causal effects of medical products on a large scale. However, the confidence intervals and p-values produced by observational studies only account for random error and fail to…

应用统计 · 统计学 2024-05-02 Jami J. Mulgrave , David Madigan , George Hripcsak