English
Related papers

Related papers: On the Confidence Intervals in Bioequivalence Stud…

200 papers

We consider a Bayesian framework for estimating the sample size of a clinical trial. The new approach, called BESS, is built upon three pillars: Sample size of the trial, Evidence from the observed data, and Confidence of the final decision…

Methodology · Statistics 2026-01-21 Dehua Bi , Yuan Ji

We introduce a publication policy that incorporates conditional equivalence testing (CET), a two-stage testing scheme in which standard NHST is followed conditionally by testing for equivalence. The idea of CET is carefully considered as it…

Methodology · Statistics 2018-07-04 Harlan Campbell , Paul Gustafson

Testing the (in)equality of variances is an important problem in many statistical applications. We develop default Bayes factor tests to assess the (in)equality of two or more population variances, as well as a test for whether the…

Methodology · Statistics 2022-08-02 Fabian Dablander , Don van den Bergh , Eric-Jan Wagenmakers , Alexander Ly

Linear combinations of multinomial probabilities, such as those resulting from contingency tables, are of use when evaluating classification system performance. While large sample inference methods for these combinations exist, small sample…

Methodology · Statistics 2021-04-20 Katherine A. Batterton , Christine M. Schubert , Richard L. Warr

Quality assessment algorithms measure the quality of a captured biometric sample. Since the sample quality strongly affects the recognition performance of a biometric system, it is essential to only process samples of sufficient quality and…

Computer Vision and Pattern Recognition · Computer Science 2024-08-22 André Dörsch , Torsten Schlett , Peter Munch , Christian Rathgeb , Christoph Busch

To make informative public policy decisions in battling the ongoing COVID-19 pandemic, it is important to know the disease prevalence in a population. There are two intertwined difficulties in estimating this prevalence based on testing…

Methodology · Statistics 2020-12-01 Bryan Cai , John P. A. Ioannidis , Eran Bendavid , Lu Tian

We investigate the calibration of large language models' (LLMs') confidence across diverse tasks. The results of our preregistered study show that the current crop of LLMs are, like people, too sure they are right: confidence exceeds…

Artificial Intelligence · Computer Science 2026-05-26 Noam Michael , Daniel BenShushan , Jacob Bien , Don A. Moore

Many experiments can be interpreted in terms of random processes operating according to some internal protocols. When experiments are costly or cannot be repeated only one or a few finite samples are available. In this paper we study data…

Data Analysis, Statistics and Probability · Physics 2016-02-02 Marian Kupczynski , Hans De Raedt

Biomedical research encompasses diverse types of activities, from basic science ("bench") to clinical medicine ("bedside") to bench-to-bedside translational research. It, however, remains unclear whether different types of research receive…

Digital Libraries · Computer Science 2019-12-17 Qing Ke

Classical confidence limits are compared to Bayesian error bounds by studying relevant examples. The performance of the two methods is investigated relative to the properties coherence, precision, bias, universality, simplicity. A proposal…

High Energy Physics - Experiment · Physics 2007-05-23 G. Zech

Basket trials are increasingly used for the simultaneous evaluation of a new treatment in various patient subgroups under one overarching protocol. We propose a Bayesian approach to sample size determination in basket trials that permit…

Methodology · Statistics 2022-09-02 Haiyan Zheng , Michael J. Grayling , Pavel Mozgunov , Thomas Jaki , James M. S. Wason

In medical device comparison studies, equivalency test is commonly used to demonstrate two measurement methods agree up to a pre-specified performance goal based on the paired repeated measures. Such equivalency test often involves…

Methodology · Statistics 2019-08-22 Yun Bai , Zengri Wang , Theodore Lystig , Baolin Wu

In order to determine whether or not an effect is absent based on a statistical test, the recommended frequentist tool is the equivalence test. Typically, it is expected that an appropriate equivalence margin has been specified before any…

Methodology · Statistics 2021-02-24 Harlan Campbell , Paul Gustafson

This text is a survey on cross-validation. We define all classical cross-validation procedures, and we study their properties for two different goals: estimating the risk of a given estimator, and selecting the best estimator among a given…

Statistics Theory · Mathematics 2017-03-10 Sylvain Arlot

In clinical trials the comparison of two different populations is a frequently addressed problem. Non-linear (parametric) regression models are commonly used to describe the relationship between covariates as the dose and a response…

Methodology · Statistics 2019-02-12 Kathrin Möllenhoff , Frank Bretz , Holger Dette

Corrected confidence intervals are developed for the mean of the second component of a bivariate normal process when the first component is being monitored sequentially. This is accomplished by constructing a first approximation to a…

Statistics Theory · Mathematics 2007-06-13 R. C. Weng , D. S. Coad

An assurance calculation is a Bayesian alternative to a power calculation. One may be performed to aid the planning of a clinical trial, specifically setting the sample size or to support decisions about whether or not to perform a study.…

Applications · Statistics 2024-03-29 James Salsbury , Jeremy Oakley , Steven Julious , Lisa Hampson

In our chapter we address the statistical analysis of percentiles: How should the citation impact of institutions be compared? In educational and psychological testing, percentiles are already used widely as a standard to evaluate an…

Digital Libraries · Computer Science 2014-04-16 Richard Williams , Lutz Bornmann

We introduce a protocol addressing the conformance test problem, which consists in determining whether a process under test conforms to a reference one. We consider a process to be characterized by the set of end-product it produces, which…

Systems that answer questions by reviewing the scientific literature are becoming increasingly feasible. To draw reliable conclusions, these systems should take into account the quality of available evidence from different studies, placing…

Computation and Language · Computer Science 2025-09-23 Jianyou Wang , Weili Cao , Longtian Bao , Youze Zheng , Gil Pasternak , Kaicheng Wang , Xiaoyue Wang , Ramamohan Paturi , Leon Bergen