中文
相关论文

相关论文: Setting the duration of online A/B experiments

200 篇论文

Nonparametric two-stage procedures to construct fixed-width confidence intervals are studied to quantify uncertainty. It is shown that the validity of the random central limit theorem (RCLT) accompanied by a consistent and asymptotically…

统计理论 · 数学 2019-10-08 Yuan-Tsung Chang , Ansgar Steland

Measurements are generally collected as unilateral or bilateral data in clinical trials or observational studies. For example, in ophthalmology studies, the primary outcome is often obtained from one eye or both eyes of an individual. In…

统计方法学 · 统计学 2021-11-01 Kejia Wang , Chang-Xing Ma

Many online experiments exhibit dependence between users and items. For example, in online advertising, observations that have a user or an ad in common are likely to be associated. Because of this, even in experiments involving millions of…

统计方法学 · 统计学 2017-10-26 Eytan Bakshy , Dean Eckles

Conformal prediction, which makes no distributional assumptions about the data, has emerged as a powerful and reliable approach to uncertainty quantification in practical applications. The nonconformity measure used in conformal prediction…

机器学习 · 计算机科学 2024-10-15 Yuko Kato , David M. J. Tax , Marco Loog

Evaluating treatment effect heterogeneity widely informs treatment decision making. At the moment, much emphasis is placed on the estimation of the conditional average treatment effect via flexible machine learning algorithms. While these…

统计方法学 · 统计学 2021-05-07 Lihua Lei , Emmanuel J. Candès

Experimenters often collect baseline data to study heterogeneity. I propose the first valid confidence intervals for the VCATE, the treatment effect variance explained by observables. Conventional approaches yield incorrect coverage when…

计量经济学 · 经济学 2023-06-07 Alejandro Sanchez-Becerra

We derive the sample size formulae for comparing two negative binomial rates based on both the relative and absolute rate difference metrics in noninferiority and equivalence trials with unequal follow-up times, and establish an approximate…

统计方法学 · 统计学 2017-05-24 Yongqiang Tang

Portraying emotion and trustworthiness is known to increase the appeal of video content. However, the causal relationship between these signals and online user engagement is not well understood. This limited understanding is partly due to a…

多媒体 · 计算机科学 2021-05-05 Lukas Stappen , Alice Baird , Michelle Lienhart , Annalena Bätz , Björn Schuller

In randomized controlled trials (RCTs) of infectious disease interventions, it is well recognized that unmeasured individual heterogeneity at baseline can induce selection bias over time, thereby complicating the interpretation of the…

统计方法学 · 统计学 2026-04-24 Hiroyasu Ando , A. James O'Malley , Akihiro Nishi

Online experiments such as Randomised Controlled Trials (RCTs) or A/B-tests are the bread and butter of modern platforms on the web. They are conducted continuously to allow platforms to estimate the causal effect of replacing system…

机器学习 · 计算机科学 2023-04-24 Olivier Jeunen

Online marketplace designers frequently run A/B tests to measure the impact of proposed product changes. However, given that marketplaces are inherently connected, total average treatment effect estimates obtained through Bernoulli…

统计方法学 · 统计学 2020-04-28 David Holtz , Ruben Lobel , Inessa Liskovich , Sinan Aral

Every design choice will have different effects on different units. However traditional A/B tests are often underpowered to identify these heterogeneous effects. This is especially true when the set of unit-level attributes is…

人工智能 · 计算机科学 2016-11-09 Alexander Peysakhovich , Akos Lada

A/B test, a simple type of controlled experiment, refers to the statistical procedure of experimenting to compare two treatments applied to test subjects. For example, many IT companies frequently conduct A/B tests on their users who are…

统计方法学 · 统计学 2026-05-12 Qiong Zhang , Lulu Kang

Accurate estimation of treatment effects in online A/B testing is challenging with zero-inflated and skewed metrics. Traditional tests, like Welch's t-test, often lack sensitivity with heavy-tailed data due to their reliance on means, as…

统计方法学 · 统计学 2025-10-07 Kevin Charette , Tristan Boudreault

Meta-analysis is an important statistical technique for synthesizing the results of multiple studies regarding the same or closely related research question. So-called meta-regression extends meta-analysis models by accounting for…

统计方法学 · 统计学 2023-02-22 Thilo Welz , Eric S. Knop , Tim Friede , Markus Pauly

Methods for random-effects meta-analysis require an estimate of the between-study variance, $\tau^2$. The performance of estimators of $\tau^2$ (measured by bias and coverage) affects their usefulness in assessing heterogeneity of…

统计方法学 · 统计学 2019-04-04 Ilyas Bakbergenuly , David C. Hoaglin , Elena Kulinskaya

We describe how to calculate standard errors for A/B tests that include clustered data, ratio metrics, and/or covariate adjustment. We may do this for power analysis/sample size calculations prior to running an experiment using historical…

统计方法学 · 统计学 2024-06-12 Tim Hesterberg , Ben Knight

In many industry settings, online controlled experimentation (A/B test) has been broadly adopted as the gold standard to measure product or feature impacts. Most research has primarily focused on user engagement type metrics, specifically…

统计方法学 · 统计学 2020-10-30 Weinan Wang , Xi Zhang

Tech companies (e.g., Google or Facebook) often use randomized online experiments and/or A/B testing primarily based on the average treatment effects to compare their new product with an old one. However, it is also critically important to…

统计方法学 · 统计学 2021-11-09 Chengchun Shi , Shikai Luo , Hongtu Zhu , Rui Song

With the passage of more time from the original date of publication, the measure of the impact of scientific works using subsequent citation counts becomes more accurate. However the measurement of individual and organizational research…

数字图书馆 · 计算机科学 2018-11-06 Giovanni Abramo , Tindaro Cicero , Ciriaco Andrea D'Angelo