English
Related papers

Related papers: The Accuracy of Confidence Intervals for Field Nor…

200 papers

We consider the problem of interval estimation of the odds ratio. An asymptotic confidence interval is widely applied in medical research. Unfortunately that confidence interval has a poor coverage probability: it is significantly smaller…

Methodology · Statistics 2020-11-19 Zofia Zielińska-Kolasińska , Wojciech Zieliński

In this paper, we show that citation counts work better than a random baseline (by a margin of 10%) in distinguishing excellent research, while Mendeley reader counts don't work better than the baseline. Specifically, we study the potential…

Digital Libraries · Computer Science 2018-02-15 Drahomira Herrmannova , Robert M. Patton , Petr Knoth , Christopher G. Stahl

In Natural Language Processing (NLP), binary classification algorithms are often evaluated using the F1 score. Because the sample F1 score is an estimate of the population F1 score, it is not sufficient to report the sample F1 score without…

Methodology · Statistics 2024-06-11 Kevin Fu Yuan Lam , Vikneswaran Gopal , Jiang Qian

As large language models (LLMs) are increasingly used in high-stakes domains, accurately assessing their confidence is crucial. Humans typically express confidence through epistemic markers (e.g., "fairly confident") instead of numerical…

Computation and Language · Computer Science 2026-04-14 Jiayu Liu , Qing Zong , Weiqi Wang , Yangqiu Song

A prediction interval covers a future observation from a random process in repeated sampling, and is typically constructed by identifying a pivotal quantity that is also an ancillary statistic. Analogously, a tolerance interval covers a…

Methodology · Statistics 2022-01-19 Geoffrey S Johnson

We propose a methodology for constructing confidence regions with partially identified models of general form. The region is obtained by inverting a test of internal consistency of the econometric structure. We develop a dilation bootstrap…

Econometrics · Economics 2021-02-10 Alfred Galichon , Marc Henry

The Leiden Rankings can be used for grouping research universities by considering universities which are not statistically significantly different as homogeneous sets. The groups and intergroup relations can be analyzed and visualized using…

Digital Libraries · Computer Science 2018-10-16 Loet Leydesdorff , Lutz Bornmann , John Mingers

Citation averages, and Impact Factors (IFs) in particular, are sensitive to sample size. We apply the Central Limit Theorem (CLT) to IFs to understand their scale-dependent behavior. For a journal of $n$ randomly selected papers from a…

Physics and Society · Physics 2018-09-24 Manolis Antonoyiannakis

Clustered data arise naturally in many scientific and applied research settings where units are grouped within clusters. They are commonly analyzed using linear mixed models to account for within-cluster correlations. This article focuses…

Methodology · Statistics 2025-10-10 Zhi Yang Tho , Raymond Chambers , A. H. Welsh

Benchmarking models is a key factor for the rapid progress in machine learning (ML) research. Thus, further progress depends on improving benchmarking metrics. A standard metric to measure the behavioral alignment between ML models and…

Neurons and Cognition · Quantitative Biology 2025-11-10 Thomas Klein , Sascha Meyen , Wieland Brendel , Felix A. Wichmann , Kristof Meding

Bootstrapping and other resampling methods are increasingly appearing in the textbooks and curricula of courses that introduce undergraduate students to statistical methods. In order to teach the bootstrap well, students and instructors…

Other Statistics · Statistics 2024-05-30 Njesa Totty , James Molyneux , Claudio Fuentes

The ISO 5725 series frames interlaboratory precision through repeatability, between-laboratory, and reproducibility variances, yet practical guidance on deploying bootstrap methods within this one-way random-effects setting remains limited.…

Applications · Statistics 2026-02-10 Jun-ichi Takeshita , Kazuhiro Morita , Tomomichi Suzuki

Fractional counting of citations can improve on ranking of multi-disciplinary research units (such as universities) by normalizing the differences among fields of science in terms of differences in citation behavior. Furthermore,…

Digital Libraries · Computer Science 2010-10-13 Loet Leydesdorff , Jung C. Shin

The purpose of this study is to compare the changing behavior of two counting methods (whole counting and whole-normalized counting) and inflation rate at country level research productivity and impact. For this, publication data on…

Digital Libraries · Computer Science 2014-10-14 B. Elango , P. Rajendran

Confidence calibration assumes a unique ground-truth label per input, yet this assumption fails wherever annotators genuinely disagree. Post-hoc calibrators fitted on majority-voted labels, the standard single-label targets used in…

Machine Learning · Computer Science 2026-03-25 Linwei Tao , Haoyang Luo , Minjing Dong , Chang Xu

In the analysis of survey data it is of interest to estimate and quantify uncertainty about means or totals for each of several non-overlapping subpopulations, or areas. When the sample size for a given area is small, standard confidence…

Methodology · Statistics 2018-09-26 Kyle Burris , Peter Hoff

Citation counts are widely used as indicators of research quality to support or replace human peer review and for lists of top cited papers, researchers, and institutions. Nevertheless, the relationship between citations and research…

Digital Libraries · Computer Science 2023-07-31 Mike Thelwall , Kayvan Kousha , Mahshid Abdoli , Emma Stuart , Meiko Makita , Paul Wilson , Jonathan Levitt

Journal Impact Factors (IFs) can be considered historically as the first attempt to normalize citation distributions by using averages over two years. However, it has been recognized that citation distributions vary among fields of science…

Digital Libraries · Computer Science 2012-02-07 Loet Leydesdorff

We investigate the calibration of large language models' (LLMs') confidence across diverse tasks. The results of our preregistered study show that the current crop of LLMs are, like people, too sure they are right: confidence exceeds…

Artificial Intelligence · Computer Science 2026-05-26 Noam Michael , Daniel BenShushan , Jacob Bien , Don A. Moore

The recent paper "Simple confidence intervals for MCMC without CLTs" by J.S. Rosenthal, showed the derivation of a simple MCMC confidence interval using only Chebyshev's inequality, not CLT. That result required certain assumptions about…

Statistics Theory · Mathematics 2021-07-01 Yu Hang Jiang , Tong Liu , Zhiya Lou , Jeffrey S. Rosenthal , Shanshan Shangguan , Fei Wang , Zixuan Wu
‹ Prev 1 8 9 10 Next ›