English
Related papers

Related papers: Removal of zero-point drift from AB data and the s…

200 papers

Drift analysis is one of the state-of-the-art techniques for the runtime analysis of randomized search heuristics (RSHs) such as evolutionary algorithms (EAs), simulated annealing etc. The vast majority of existing drift theorems yield…

Neural and Evolutionary Computing · Computer Science 2018-05-30 Per Kristian Lehre , Carsten Witt

Machine learning systems are often trained and evaluated for fairness on historical data, yet deployed in environments where conditions have shifted. A particularly common form of shift occurs when the prevalence of positive outcomes…

Machine Learning · Computer Science 2026-02-06 Amir Asiaee , Kaveh Aryan

While many statistical procedures rely on a fixed sample size, sequential methods allow a decision-maker to adapt the sample size to achieve a given precision. In this way, sequential tests reduce the average number of observations required…

Statistics Theory · Mathematics 2026-03-03 Henri Doerks , Erik Ekström , Yuqiong Wang

We discuss a simple extension of the Ho and Lee model with generic time-dependent drift in which: 1) we compute bond prices analytically; 2) the yield curve is sensible and the asymptotic yield is positive; and 3) our analytical solution…

Mathematical Finance · Quantitative Finance 2016-01-26 Zura Kakushadze

In data streams, the data distribution of arriving observations at different time points may change - a phenomenon called concept drift. While detecting concept drift is a relatively mature area of study, solutions to the uncertainty…

Machine Learning · Computer Science 2020-08-11 Anjin Liu , Jie Lu , Guangquan Zhang

Functional data often arise as sequential temporal observations over a continuous state-space. A set of functional data with a possible change in its structure may lead to a wrong conclusion if it is not taken in to account. So, sometimes,…

Methodology · Statistics 2015-03-18 Buddhananda Banerjee , Satyaki Mazumder

Motivated by problems from statistical analysis for discretely sampled SPDEs, first we derive central limit theorems for higher order finite differences applied to stochastic process with arbitrary finitely regular paths. These results are…

Probability · Mathematics 2021-03-09 Igor Cialenco , Hyun-Jung Kim , Gregor Pasemann

We address the problem of detection and estimation of one or two change-points in the mean of a series of random variables. We use the formalism of set estimation in regression: To each point of a design is attached a binary label that…

Statistics Theory · Mathematics 2018-09-07 Victor-Emmanuel Brunel

Friedman test is a nonparametric method that proposed for analyzing data from a randomized complete block design as a robust alternative to parametric method and widely applied in many fields such as agriculture, biology, business,…

Methodology · Statistics 2022-02-21 Elsayed A. H. Elamir

Testing for association or dependence between pairs of random variables is a fundamental problem in statistics. In some applications, data are subject to selection bias that causes dependence between observations even when it is absent from…

Methodology · Statistics 2020-10-13 Yaniv Tenzer , Micha Mandel , Or Zuk

Under the null hypothesis, the marginal probability of the positive response is symmetric at any specified correlated coefficient, and the discordance probability is also symmetric to the positive response probability. The marginal…

Methodology · Statistics 2022-11-11 Guanghui Huang

In real-world applications, the process generating the data might suffer from nonstationary effects (e.g., due to seasonality, faults affecting sensors or actuators, and changes in the users' behaviour). These changes, often called concept…

Machine Learning · Computer Science 2025-08-25 Kleanthis Malialis , Manuel Roveri , Cesare Alippi , Christos G. Panayiotou , Marios M. Polycarpou

Gorman and Bedrick (2019) argued for using random splits rather than standard splits in NLP experiments. We argue that random splits, like standard splits, lead to overly optimistic performance estimates. We can also split data in biased or…

Computation and Language · Computer Science 2021-04-27 Anders Søgaard , Sebastian Ebert , Jasmijn Bastings , Katja Filippova

Hypothesis testing results often rely on simple, yet important assumptions about the behaviour of the distribution of p-values under the null and the alternative. We examine tests for one dimensional parameters of interest that converge to…

Statistics Theory · Mathematics 2021-08-06 Yanbo Tang , Radu Craiu , Lei Sun

Assessment of risk prediction models has primarily utilized measures of discrimination, the ROC curve AUC and C-statistic. These derive from the risk distributions of patients and nonpatients, which in turn are derived from a population…

Quantitative Methods · Quantitative Biology 2023-12-05 Ralph H. Stern

Empirical work often uses treatment assigned following geographic boundaries. When the effects of treatment cross over borders, classical difference-in-differences estimation produces biased estimates for the average treatment effect. In…

Econometrics · Economics 2023-06-13 Kyle Butts

Distance correlation is a new class of multivariate dependence coefficients applicable to random vectors of arbitrary and not necessarily equal dimension. Distance covariance and distance correlation are analogous to product-moment…

Applications · Statistics 2010-10-07 Gábor J. Székely , Maria L. Rizzo

This paper introduces a statistical test inferring whether a variable allows separating two classes by means of a single critical value. Its test statistic is the prediction error of a nonparametric threshold classifier. While this approach…

Methodology · Statistics 2017-07-17 Fabian Schroeder

When analyzing data from randomized clinical trials, covariate adjustment can be used to account for chance imbalance in baseline covariates and to increase precision of the treatment effect estimate. A practical barrier to covariate…

Methodology · Statistics 2023-07-04 Chia-Rui Chang , Yue Song , Fan Li , Rui Wang

The performance of deep neural networks is strongly influenced by the training dataset setup. In particular, when attributes having a strong correlation with the target attribute are present, the trained model can provide unintended…

Machine Learning · Computer Science 2023-02-14 Sumyeong Ahn , Seongyoon Kim , Se-young Yun