English
Related papers

Related papers: Misplaced Confidence in Observed Power

200 papers

Background: Screening trials require large sample sizes and long time-horizons to demonstrate mortality reductions. We recently proposed increasing statistical power by testing stored control-arm specimens, called the Intended Effect (IE)…

Applications · Statistics 2024-11-11 Hormuzd A. Katki , Li C. Cheung

Consistently checking the statistical significance of experimental results is one of the mandatory methodological steps to address the so-called "reproducibility crisis" in deep reinforcement learning. In this tutorial paper, we explain how…

Machine Learning · Computer Science 2018-07-06 Cédric Colas , Olivier Sigaud , Pierre-Yves Oudeyer

For randomized controlled trials to be conclusive, it is important to set the target sample size accurately at the design stage. Comparing two normal populations, the sample size calculation requires specification of the variance other than…

Methodology · Statistics 2026-02-04 Hirotada Maeda , Satoshi Hattori , Tim Friede

Interval-censored competing risks data arise when each study subject may experience an event or failure from one of several causes and the failure time is not observed exactly but rather known to lie in an interval between two successive…

Methodology · Statistics 2016-03-02 Lu Mao , D. Y. Lin , Donglin Zeng

Hazard ratios are ubiquitously used in time to event analysis to quantify treatment effects. Although hazard ratios are invaluable for hypothesis testing, other measures of association, both relative and absolute, may be used to fully…

Methodology · Statistics 2020-11-02 Federico Ambrogi , Simona Iacobelli , Per Kragh Andersen

In confirmatory cancer clinical trials, overall survival (OS) is normally a primary endpoint in the intention-to-treat (ITT) analysis under regulatory standards. After the tumor progresses, it is common that patients allocated to the…

Methodology · Statistics 2022-01-19 José L. Jiménez , Julia Niewczas , Alexander Bore , Carl-Fredrik Burman

From biotechnology to cyber-risks, most extreme technological risks cannot be reliably estimated from historical statistics. Therefore, engineers resort to predictive methods, such as fault/event trees in the framework of probabilistic…

Physics and Society · Physics 2014-08-26 D. Sornette , T. Maillart , W. Kroeger

Competing risks data are common in medical studies, and the sub-distribution hazard (SDH) ratio is considered an appropriate measure. However, because the limitations of hazard itself are not easy to interpret clinically and because the SDH…

Applications · Statistics 2021-10-19 Jingjing Lyu , Yawen Hou , Zheng Chen

Safety analyses in terms of adverse events (AEs) are an important aspect of benefit-risk assessments of therapies. Compared to efficacy analyses AE analyses are often rather simplistic. The probability of an AE of a specific type is…

Applications · Statistics 2021-05-20 Regina Stegherr , Claudia Schmoor , Michael Lübbert , Tim Friede , Jan Beyersmann

While running any experiment, we often have to consider the statistical power to ensure an effective study. Statistical power or power ensures that we can observe an effect with high probability if such a true effect exists. However,…

Methodology · Statistics 2023-06-21 Ajinkya K Mulay , Sean Lane , Erin Hennes

In this paper, we introduce a probabilistic approach to risk assessment of robot systems by focusing on the impact of uncertainties. While various approaches to identifying systematic hazards (e.g., bugs, design flaws, etc.) can be found in…

Robotics · Computer Science 2024-10-28 Woo-Jeong Baek , Tom P. Huck , Joschka Haas , Jonas Lewandrowski , Tamim Asfour , Torsten Kröger

Maximizing statistical power in experimental design often involves imbalanced treatment allocation, but several challenges hinder its practical adoption: (1) the misconception that equal allocation always maximizes power, (2) when only…

Methodology · Statistics 2025-09-17 Stef Baas , Lukas Pin , Sofía S. Villar , William F. Rosenberger

Prescription opioids relieve moderate-to-severe pain after surgery, but overprescription can lead to misuse and overdose. Understanding factors associated with post-surgical opioid refills is crucial for improving pain management and…

Methodology · Statistics 2025-09-16 Eileen Yang , Donglin Zeng , Mark Bicket , Yi Li

The use of the non-parametric Restricted Mean Survival Time endpoint (RMST) has grown in popularity as trialists look to analyse time-to-event outcomes without the restrictions of the proportional hazards assumption. In this paper, we…

Methodology · Statistics 2023-11-06 Emily Alger , David S. Robertson , Abigail J. Burdon

Background: Often when undertaking meta-analyses of time-to-event (TTE) outcomes, especially in a Health Technology Assessment context, a hazard ratio (HR) scale is used. However, issues arise when there is evidence of non-proportional…

Methodology · Statistics 2026-05-21 Rhiannon K Owen , Keith R Abrams

It is generally believed that more observations provide more information. However, we observe that in the independence test for rare events, the power of the test is, surprisingly, determined by the number of rare events rather than the…

Methodology · Statistics 2025-06-17 Danyang Huang , Liyuan Wang , Liping Zhu

The most dangerous error in clinical trial interpretation is equating p > 0.05 with no effect. This review provides a practical, algorithm-based framework for classifying randomized controlled trial (RCT) results into six distinct…

Methodology · Statistics 2026-04-13 Ibrahim Halil Tanboga

Much research on Machine Learning testing relies on empirical studies that evaluate and show their potential. However, in this context empirical results are sensitive to a number of parameters that can adversely impact the results of the…

Software Engineering · Computer Science 2023-09-12 Salah Ghamizi , Maxime Cordy , Yuejun Guo , Mike Papadakis , And Yves Le Traon

Objective: Randomised controlled trials (RCTs) are widely considered as gold standard for assessing the effectiveness of new health interventions. When treatment non-compliance is present in RCTs, the treatment effect in the subgroup of…

Applications · Statistics 2025-03-25 Theodosios Papazoglou , Ed Waddingham , Alastair Young

Propensity Score Matching (PSM) is an useful method to reduce the impact ofTreatment - Selection Bias in the estimation of causal effects in observational studies. After matching, the PSM significantly reduces the sample under…

Methodology · Statistics 2019-02-01 Daniel García Iglesias