English
Related papers

Related papers: A modified $\chi^2$-test for uplift models with ap…

200 papers

The ROC curve is the gold standard for measuring the performance of a test/scoring statistic regarding its capacity to discriminate between two statistical populations in a wide variety of applications, ranging from anomaly detection in…

Statistics Theory · Mathematics 2023-01-25 Stéphan Clémençon , Myrto Limnios , Nicolas Vayatis

We investigate the large-sample behavior of change-point tests based on weighted two-sample U-statistics, in the case of short-range dependent data. Under some mild mixing conditions, we establish convergence of the test statistic to an…

Statistics Theory · Mathematics 2023-04-04 Herold Dehling , Kata Vuk , Martin Wendler

Statistical NLP systems are frequently evaluated and compared on the basis of their performances on a single split of training and test data. Results obtained using a single split are, however, subject to sampling noise. In this paper we…

Computation and Language · Computer Science 2007-05-23 Yuval Krymolowski

As data shift or new data become available, updating clinical machine learning models may be necessary to maintain or improve performance over time. However, updating a model can introduce compatibility issues when the behavior of the…

Machine Learning · Statistics 2023-08-11 Erkin Ötleş , Brian T. Denton , Jenna Wiens

Models that top leaderboards often perform unsatisfactorily when deployed in real world applications; this has necessitated rigorous and expensive pre-deployment model testing. A hitherto unexplored facet of model performance is: Are our…

Computation and Language · Computer Science 2021-06-11 Swaroop Mishra , Anjana Arunkumar

We study a tractable two-player contest built on a truncated cubic contest success function. Its defining feature is a strategic-feedback parameter whose sign determines whether a leading player's effort lowers (suppression) or raises…

Theoretical Economics · Economics 2026-05-27 Alexander Matros , Constantine Sorokin

In this paper, four new Chi-Square type statistics are presented for testing the hypothesis of a uniform null versus specified trend alternatives. The powers of these test statistics are compared with the powers of the statistics considered…

Statistics Theory · Mathematics 2010-07-30 Clement Ampadu

Experimental comparisons of performance represent an important aspect of research on optimization algorithms. In this work we present a methodology for defining the required sample sizes for designing experiments with desired statistical…

Neural and Evolutionary Computing · Computer Science 2018-10-16 Felipe Campelo , Fernanda Takahashi

Surveys are commonly used to facilitate research in epidemiology, health, and the social and behavioral sciences. Often, these surveys are not simple random samples, and respondents are given weights reflecting their probability of…

Methodology · Statistics 2024-08-20 Adway S. Wadekar , Jerome P. Reiter

This study replicates and adapts the experiment of Hoelzl and Rustichini (2005), which examined overplacement, i.e., overconfidence in relative self-assessments, by analyzing individuals' voting preferences between a performance-based and a…

General Economics · Economics 2025-07-22 Marius Protte

The causal effect of showing an ad to a potential customer versus not, commonly referred to as "incrementality", is the fundamental question of advertising effectiveness. In digital advertising three major puzzle pieces are central to…

Machine Learning · Computer Science 2022-08-30 Randall Lewis , Jeffrey Wong

The test statistics of two powerful tests for normality \citep{lm1,mud2} are estimators of the correlation coefficient between certain sample moments. We derive new versions of the test statistics that are functions of the sample skewness…

Statistics Theory · Mathematics 2011-08-03 Måns Thulin

The raking-ratio method is a statistical and computational method which adjusts the empirical measure to match the true probability of sets of a finite partition. We study the asymptotic behavior of the raking-ratio empirical process…

Statistics Theory · Mathematics 2019-05-07 Mickael Albertus

We consider the conditional randomization test as a way to account for covariate imbalance in randomized experiments. The test accounts for covariate imbalance by comparing the observed test statistic to the null distribution of the test…

We introduce a new conservative test for quantifying the consistency of two or more datasets. The test is based on the Bayesian answer to the question, ``How much more probable is it that all my data were generated from the same model…

Astrophysics · Physics 2008-11-26 Phil Marshall , Nutan Rajguru , Anze Slosar

Estimating the strength of dependency between two variables is fundamental for exploratory analysis and many other applications in data mining. For example: non-linear dependencies between two continuous variables can be explored with the…

Machine Learning · Statistics 2016-01-21 Simone Romano , Nguyen Xuan Vinh , James Bailey , Karin Verspoor

This paper introduces a marketing decision framework that optimizes customer targeting by integrating heterogeneous treatment effect estimation with explicit business guardrails. The objective is to maximize revenue and retention while…

Machine Learning · Computer Science 2026-02-05 Deepit Sapru

The median statistic has recently been discussed by Gott \textit{et al.} as a more reliable alternative to the standard $\chi^2$ likelihood analysis, in the sense of requiring fewer assumptions about the data and being almost as…

Astrophysics · Physics 2009-11-07 P. P. Avelino , C. J. A. P. Martins , P. Pinto

We consider whether the asymptotic distributions for the log-likelihood ratio test statistic are expected to be Gaussian or chi-squared. Two straightforward examples provide insight on the difference.

Data Analysis, Statistics and Probability · Physics 2017-12-13 Louis Lyons

The odds ratio measure is used in health and social surveys where the odds of a certain event is to be compared between two populations. It is defined using logistic regression, and requires that data from surveys are accompanied by their…

Methodology · Statistics 2014-07-01 C. Goga , A Ruiz-Gazen