English
Related papers

Related papers: Statistics against irritations: a response to Dick…

200 papers

One way of evaluating individual scientists is the determination of the number of highly cited publications, where the threshold is given by a large reference set. It is shown that this indicator behaves in a counterintuitive way, leading…

Digital Libraries · Computer Science 2013-05-09 Michael Schreiber

The existence of "Hot Hands" and "Streaks" in sports and gambling is hotly debated, but there is no uncertainty about the recent batting-average of the New York Times: it is now two-for-two in mangling and misunderstanding elementary…

History and Overview · Mathematics 2015-12-31 Dan Gusfield

Many writers have observed that default logics appear to contain the "lottery paradox" of probability theory. This arises when a default "proof by contradiction" lets us conclude that a typical X is not a Y where Y is an unusual subclass of…

Artificial Intelligence · Computer Science 2013-04-08 Eric Neufeld , J. D. Horton

We present a formal measure of argument strength, which combines the ideas that conclusions of strong arguments are (i) highly probable and (ii) their uncertainty is relatively precise. Likewise, arguments are weak when their conclusion…

Artificial Intelligence · Computer Science 2017-03-10 Niki Pfeifer , Hanna Pankka

We review the well known Bertrand paradoxes, and we first maintain that they do not point to any probabilistic inconsistency, but rather to the risks incurred with a careless use of the locution "at random". We claim then that these…

History and Overview · Mathematics 2019-08-23 Nicola Cufaro Petroni

The Brier score is frequently used by meteorologists to measure the skill of binary probabilistic forecasts. We show, however, that in simple idealised cases it gives counterintuitive results. We advocate the use of an alternative measure…

Atmospheric and Oceanic Physics · Physics 2007-05-23 Stephen Jewson

Predicting X from Twitter is a popular fad within the Twitter research subculture. It seems both appealing and relatively easy. Among such kind of studies, electoral prediction is maybe the most attractive, and at this moment there is a…

Computers and Society · Computer Science 2015-03-20 Daniel Gayo-Avello

In narrative synthesis of evidence, it can be the case that the only quantitative measures available concerning the efficacy of an intervention is the direction of the effect, i.e. whether it is positive or negative. In such situations, the…

Methodology · Statistics 2021-05-05 Stavros Nikolakopoulos

Multiple-choice tests are a common approach for assessing candidates' comprehension skills. Standard multiple-choice reading comprehension exams require candidates to select the correct answer option from a discrete set based on a question…

Computation and Language · Computer Science 2023-11-09 Vatsal Raina , Adian Liusie , Mark Gales

The prevailing maximum likelihood estimators for inferring power law models from rank-frequency data are biased. The source of this bias is an inappropriate likelihood function. The correct likelihood function is derived and shown to be…

Applications · Statistics 2021-07-27 Charlie Pilgrim , Thomas T Hills

Incidence Calculus and Dempster-Shafer Theory of Evidence are both theories to describe agents' degrees of belief in propositions, thus being appropriate to represent uncertainty in reasoning systems. This paper presents a straightforward…

Artificial Intelligence · Computer Science 2013-04-05 F. Correa da Silva , Alan Bundy

Citation analysis does not generally take the quality of citations into account: all citations are weighted equally irrespective of source. However, a scholar may be highly cited but not highly regarded: popularity and prestige are not…

Digital Libraries · Computer Science 2010-12-23 Ying Ding , Blaise Cronin

Sarcasm can be defined as saying or writing the opposite of what one truly wants to express, usually to insult, irritate, or amuse someone. Because of the obscure nature of sarcasm in textual data, detecting it is difficult and of great…

Computation and Language · Computer Science 2022-09-21 Faria Binte Kader , Nafisa Hossain Nujat , Tasmia Binte Sogir , Mohsinul Kabir , Hasan Mahmud , Kamrul Hasan

Likelihood ratio tests are intuitively appealing. Nevertheless, a number of examples are known in which they perform very poorly. The present paper discusses a large class of situations in which this is the case, and analyzes just how…

Statistics Theory · Mathematics 2007-06-13 Erich L. Lehmann

I think we can agree that dealing with uncertainty is not easy. Probability is the main tool for dealing with uncertainty, and we know there are many probability-related puzzles and paradoxes. Here I describe a rather idiosyncratic…

Other Statistics · Statistics 2022-01-19 Yudi Pawitan

This is a report about the use and misuse of citation data in the assessment of scientific research. The idea that research assessment must be done using ``simple and objective'' methods is increasingly prevalent today. The ``simple and…

Methodology · Statistics 2009-10-20 Robert Adler , John Ewing , Peter Taylor

An individual's variation in writing style is often a function of both social and personal attributes. While structured social variation has been extensively studied, e.g., gender based variation, far less is known about how to characterize…

Computation and Language · Computer Science 2021-09-13 Jian Zhu , David Jurgens

(To appear in The American Statistician.) Distance covariance (Sz\'ekely, Rizzo, and Bakirov, 2007) is a fascinating recent notion, which is popular as a test for dependence of any type between random variables $X$ and $Y$. This approach…

Methodology · Statistics 2024-07-08 Jakob Raymaekers , Peter J. Rousseeuw

The impact of text length on the estimation of lexical diversity has captured the attention of the scientific community for more than a century. Numerous indices have been proposed, and many studies have been conducted to evaluate them, but…

Computation and Language · Computer Science 2023-08-01 Yves Bestgen

Texts exhibit considerable stylistic variation. This paper reports an experiment where a corpus of documents (N= 75 000) is analyzed using various simple stylistic metrics. A subset (n = 1000) of the corpus has been previously assessed to…

cmp-lg · Computer Science 2008-02-03 Jussi Karlgren