English
Related papers

Related papers: The evidence contained in the P-value is context d…

200 papers

Analysis of credibility is a reverse-Bayes technique that has been proposed by Matthews (2001) to overcome some of the shortcomings of significance tests. A significant result is deemed credible if current knowledge about the effect size is…

Methodology · Statistics 2017-12-11 Leonhard Held

There are two distinct definitions of 'P-value' for evaluating a proposed hypothesis or model for the process generating an observed dataset. The original definition starts with a measure of the divergence of the dataset from what was…

Other Statistics · Statistics 2023-09-25 Sander Greenland

I present a critique of the methods used in a typical paper. This leads to three broad conclusions about the conventional use of statistical methods. First, results are often reported in an unnecessarily obscure manner. Second, the null…

Applications · Statistics 2013-03-05 Michael Wood

A recurring debate in the philosophy of statistics concerns what, exactly, should count as a measure of evidence for or against a given hypothesis. P-values, likelihood ratios, and Bayes factors all have their defenders. In this paper we…

Methodology · Statistics 2026-03-26 Ben Chugg , Aaditya Ramdas , Peter Grünwald

Hypothesis testing in contingency tables is usually based on asymptotic results, thereby restricting its proper use to large samples. To study these tests in small samples, we consider the likelihood ratio test and define an accurate index,…

Methodology · Statistics 2018-10-04 Natalia L. Oliveira , Carlos A. de B. Pereira , Marcio A. Diniz , Adriano Polpo

Null hypothesis statistical significance tests (NHST) are widely used in quantitative research in the empirical sciences including scientometrics. Nevertheless, since their introduction nearly a century ago significance tests have been…

Other Statistics · Statistics 2014-02-06 Jesper W. Schneider

Despite its importance to experimental design, statistical power (the probability that, given a real effect, an experiment will reject the null hypothesis) has largely been ignored by the NLP community. Underpowered experiments make it more…

Computation and Language · Computer Science 2020-10-15 Dallas Card , Peter Henderson , Urvashi Khandelwal , Robin Jia , Kyle Mahowald , Dan Jurafsky

Incorrect usage of $p$-values, particularly within the context of significance testing using the arbitrary .05 threshold, has become a major problem in modern statistical practice. The prevalence of this problem can be traced back to the…

Methodology · Statistics 2021-07-19 MaryLena Bleile

We introduce a method for calculating \(p\)-values to test causal hypotheses in qualitative research \emph{a la} process tracing. As in an experiment, our \(p\)-value tells us how often one would make the same or more compelling…

Methodology · Statistics 2025-08-04 Matias Lopez , Jake Bowers

P-values and null-hypothesis significance testing are popular data-analytical tools in functional neuroimaging. Sparked by the analysis of resting-state fMRI data, there has been a resurgence of interest in the validity of some of the…

Quantitative Methods · Quantitative Biology 2021-08-10 Dirk Ostwald , Sebastian Schneider , Rasmus Bruckner , Lilla Horvath

In this paper, the authors first provide an overview of two major developments on complex survey data analysis: the empirical likelihood methods and statistical inference with non-probability survey samples, and highlight the important…

Methodology · Statistics 2025-08-14 Yilin Chen , Pengfei Li , J. N. K. Rao , Changbao Wu

In the recent advances of natural language processing, the scale of the state-of-the-art models and datasets is usually extensive, which challenges the application of sample-based explanation methods in many aspects, such as explanation…

Computation and Language · Computer Science 2021-06-10 Wei Zhang , Ziming Huang , Yada Zhu , Guangnan Ye , Xiaodong Cui , Fan Zhang

The American Statistical Association (ASA) statement on statistical significance and P-values \cite{wasserstein2016asa} cautioned statisticians against making scientific decisions solely on the basis of traditional P-values. The statement…

Methodology · Statistics 2024-02-22 Abhisek Chakraborty , Megan H. Murray , Ilya Lipkovich , Yu Du

Following discussions in 2010 and 2011, scientometric evaluators have increasingly abandoned relative indicators in favor of comparing observed with expected citation ratios. The latter method provides parameters with error values allowing…

Digital Libraries · Computer Science 2018-08-30 Loet Leydesdorff , Tobias Opthof

Research in Sports Sciences is supported often by inferences based on the declaration of the value of the statistic statistically significant or nonsignificant on the bases of a P value derived from a null-hypothesis test. Taking into…

Applications · Statistics 2018-02-09 Rui Marcelino , Bruno Natale Pasquarelli , Jaime Sampaio

The internal validity of observational study is often subject to debate. In this study, we define the unobserved sample based on the counterfactuals and formalize its relationship with the null hypothesis statistical testing (NHST) for…

Methodology · Statistics 2020-05-27 Tenglong Li , Kenneth A. Frank

The mid-p-value is a proposed improvement on the ordinary p-value for the case where the test statistic is partially or completely discrete. In this case, the ordinary p-value is conservative, meaning that its null distribution is larger…

Statistics Theory · Mathematics 2017-06-02 Patrick Rubin-Delanchy , Nicholas A. Heard , Daniel John Lawson

This document presents the statistical methods used to process low-level measurements in the presence of noise. These methods can be classical or Bayesian. The question is placed in the general framework of the problem of nuisance…

Instrumentation and Detectors · Physics 2024-03-20 Guillaume Manificat , Salima Helali , Patrick Bouisset

It has long been a puzzle why, despite sustained reform efforts, many applied scientific fields remain dominated by Null Hypothesis Significance Testing (NHST), a framework that dichotomizes study results and privileges "statistically…

Other Statistics · Statistics 2026-03-24 Carol Ting

It is well known that there is no direct one-to-one relation between $p$-values and likelihood ratios or Bayes factors, since their relation crucially involves the sample size $n$. We investigate their (asymptotic) relation in a…

Statistics Theory · Mathematics 2026-03-23 Wouter Kager , Ronald Meester