English
Related papers

Related papers: Setting the duration of online A/B experiments

200 papers

Online experiments are the gold standard for evaluating impact on user experience and accelerating innovation in software. However, since experiments are typically limited in duration, observed treatment effects are not always permanently…

Human-Computer Interaction · Computer Science 2021-02-26 Soheil Sadeghi , Somit Gupta , Stefan Gramatovici , Jiannan Lu , Hao Ai , Ruhan Zhang

Composite endpoints are widely used in cardiovascular clinical trials to improve statistical efficiency while preserving clinical relevance. The Win Ratio (WR) measure and more general frameworks of Win Statistics have emerged as…

Methodology · Statistics 2025-11-24 Yunhan Mou , Fan Li , Denise Esserman , Yuan Huang

We consider the problem of constructing confidence intervals (CIs) for the population mean of $N$ values $\{x_1, \ldots, x_N\} \subset \Sigma^N$ based on a random sample of size $n$, denoted by $X^n \equiv (X_1, \ldots, X_n)$, drawn…

Statistics Theory · Mathematics 2026-03-17 Shubhanshu Shekhar , Aaditya Ramdas

Turing's estimator allows one to estimate the probabilities of outcomes that either do not appear or only rarely appear in a given random sample. We perform a simulation study to understand the finite sample performance of several related…

Statistics Theory · Mathematics 2025-03-19 Jie Chang , Michael Grabchak , Jialin Zhang

In an A/B test, the typical objective is to measure the total average treatment effect (TATE), which measures the difference between the average outcome if all users were treated and the average outcome if all users were untreated. However,…

Applications · Statistics 2020-04-28 David Holtz , Sinan Aral

There are over 55 different ways to construct a confidence respectively credible interval (CI) for the binomial proportion. Methods to compare them are necessary to decide which should be used in practice. The interval score has been…

Methodology · Statistics 2022-07-08 Lisa J. Hofer , Leonhard Held

Performance uncertainty quantification is essential for reliable validation and eventual clinical translation of medical imaging artificial intelligence (AI). Confidence intervals (CIs) play a central role in this process by indicating how…

Two-sided marketplace platforms often run experiments to test the effect of an intervention before launching it platform-wide. A typical approach is to randomize individuals into the treatment group, which receives the intervention, and the…

Methodology · Statistics 2021-04-27 Hannah Li , Geng Zhao , Ramesh Johari , Gabriel Y. Weintraub

Meta-analysis is a well-established tool used to combine data from several independent studies, each of which usually compares the effect of an experimental treatment with a control group. While meta-analyses are often performed using…

Methodology · Statistics 2025-06-02 Phuc Thien Tran , Long-Hao Xu , Christian Röver , Tim Friede

Participants in online experiments often enroll over time, which can compromise sample representativeness due to temporal shifts in covariates. This issue is particularly critical in A/B tests, online controlled experiments extensively used…

General Economics · Economics 2026-03-30 Chen Wang , Shichao Han , Shan Huang

A comprehensive understanding of data quality is the cornerstone of measurement studies in social media research. This paper presents in-depth measurements on the effects of Twitter data sampling across different timescales and different…

Social and Information Networks · Computer Science 2020-04-07 Siqi Wu , Marian-Andrei Rizoiu , Lexing Xie

Randomized A/B tests within online learning platforms represent an exciting direction in learning sciences. With minimal assumptions, they allow causal effect estimation without confounding bias and exact statistical inference even in small…

Methodology · Statistics 2023-06-13 Adam C. Sales , Ethan B. Prihar , Johann A. Gagnon-Bartsch , Neil T. Heffernan

We consider a discrete time stochastic model with infinite variance and study the mean estimation problem as in Wang and Ramdas (2023). We refine the Catoni-type confidence sequence (abbr. CS) and use an idea of Bhatt et al. (2022) to…

Statistics Theory · Mathematics 2024-09-10 Chengfu Wei , Jordan Stoyanov , Yiming Chen , Zijun Chen

The problem is in the estimation of the fraction of population with a sensitive characteristic. We consider the Item Count Technique an indirect method of questioning designed to protect respondents' privacy. The exact confidence interval…

Methodology · Statistics 2024-10-21 Stanislaw Jaworski , Wojciech Zielinski

Online controlled experiments (A/B testing) are fundamental to data-driven decision-making in many companies. Improving the sensitivity of these experiments under fixed sample size constraints requires reducing the variance of the average…

Methodology · Statistics 2026-03-24 Zhexiao Lin , Pablo Crespo

Regression adjustment, sometimes known as Controlled-experiment Using Pre-Experiment Data (CUPED), is an important technique in internet experimentation. It decreases the variance of effect size estimates, often cutting confidence interval…

Methodology · Statistics 2023-11-30 Daniel Ting , Kenneth Hung

Although temporal persistence, or permanence, is a well understood requirement for optimal biometric features, there is no general agreement on how to assess temporal persistence. We suggest that the best way to assess temporal persistence…

Quantitative Methods · Quantitative Biology 2017-07-05 Lee Friedman , Ioannis Rigas , Mark S. Nixon , Oleg V. Komogortsev

Eliminating the effect of confounding in observational studies typically involves fitting a model for an outcome adjusted for covariates. When, as often, these covariates are high-dimensional, this necessitates the use of sparse estimators…

Methodology · Statistics 2019-03-26 Oliver Dukes , Stijn Vansteelandt

The traditional model specification of stepped-wedge cluster-randomized trials assumes a homogeneous treatment effect across time while adjusting for fixed-time effects. However, when treatment effects vary over time, the constant effect…

Concentration inequalities have become increasingly popular in machine learning, probability, and statistical research. Using concentration inequalities, one can construct confidence intervals (CIs) for many quantities of interest.…

Statistics Theory · Mathematics 2019-03-06 Hien D. Nguyen