English
Related papers

Related papers: GSSMD: A new standardized effect size measure to i…

200 papers

In meta-analysis, the random-effects models are standard tools to address between-study heterogeneity in evidence synthesis analyses. For the random-effects distribution models, the normal distribution model has been adopted in most…

Applications · Statistics 2021-07-28 Hisashi Noma , Kengo Nagashima , Shogo Kato , Satoshi Teramukai , Toshi A. Furukawa

We introduce a generalized formulation of mutual information (MI) based on the extended Bregman divergence, a framework that subsumes the generalized S-Bregman (GSB) divergence family. The GSB divergence unifies two important classes of…

Methodology · Statistics 2026-02-05 Arijit Pyne

Variable selection remains a difficult problem, especially for generalized linear mixed models (GLMMs). While some frequentist approaches to simultaneously select joint fixed and random effects exist, primarily through the use of…

Methodology · Statistics 2024-12-03 Feng Ding , Ian Laga

We combine two recently proposed nonparametric difference-in-differences methods, extending them to enable the examination of treatment effect heterogeneity in the staggered adoption setting using machine learning. The proposed method,…

Econometrics · Economics 2023-10-19 Julia Hatamyar , Noemi Kreif , Rudi Rocha , Martin Huber

The magnitude-based decisions (MBD) procedure was developed within sports science as an alternative to null hypothesis significance tests. It aimed to emphasise effect sizes and discourage dichotomous decision-making. The use of MBD was…

Applications · Statistics 2020-11-26 Janet Aisbett , Eric J. Drinkwater , Kenneth L. Quarrie , Stephen Woodcock

This paper introduces a novel test for conditional stochastic dominance (CSD) at specific values of the conditioning covariates, referred to as target points. The test is relevant for analyzing income inequality, evaluating treatment…

Econometrics · Economics 2025-11-20 Federico A. Bugni , Ivan A. Canay , Deborah Kim

Beyond conditional average treatment effects, treatments may impact the entire outcome distribution in covariate-dependent ways, for example, by altering the variance or tail risks for specific subpopulations. We propose a novel estimand to…

Machine Learning · Statistics 2026-03-18 Saksham Jain , Alex Luedtke

Generalizing treatment effects from a randomized trial to a target population requires the assumption that potential outcome distributions are invariant across populations after conditioning on observed covariates. This assumption fails…

Methodology · Statistics 2026-04-16 Amir Asiaee , Samhita Pal , Cole Beck , Jared D. Huling

Recent work has focused on nonparametric estimation of conditional treatment effects, but inference has remained relatively unexplored. We propose a class of nonparametric tests for both quantitative and qualitative treatment effect…

Methodology · Statistics 2026-04-07 Oliver Dukes , Mats J. Stensrud , Riccardo Brioschi , Aaron Hudson

Regression discontinuity (RD) designs are popular quasi-experimental studies in which treatment assignment depends on whether the value of a running variable exceeds a cutoff. RD designs are increasingly popular in educational applications…

Methodology · Statistics 2024-07-23 Daryl Swartzentruber , Eloise Kaizar

In this paper, we explore the modified Greenwood statistic, which, in contrast to the classical Greenwood statistic, is properly defined for random samples from any distribution. The classical Greenwood statistic, extensively examined in…

Statistics Theory · Mathematics 2024-05-21 Katarzyna Skowronek , Marek Arendarczyk , Radosław Zimroz , Agnieszka Wyłomańska

Recently, self-supervised metric learning has raised attention for the potential to learn a generic distance function. It overcomes the limitations of conventional supervised one, e.g., scalability and label biases. Despite progress in this…

Computer Vision and Pattern Recognition · Computer Science 2023-12-05 Jiantao Wu , Shentong Mo , Sara Atito , Josef Kittler , Zhenhua Feng , Muhammad Awais

This paper revisits datasets and evaluation criteria for Symbolic Regression (SR), specifically focused on its potential for scientific discovery. Focused on a set of formulas used in the existing datasets based on Feynman Lectures on…

Machine Learning · Computer Science 2025-01-06 Yoshitomo Matsubara , Naoya Chiba , Ryo Igarashi , Yoshitaka Ushiku

Testing for normality is a widely used procedure in statistics and data analysis, often applied prior to employing methods that rely on the assumption of normally distributed data. While several existing tests target distributional…

Methodology · Statistics 2026-04-07 Akin Anarat , Holger Schwender

Linear mixed models (LMMs) are used as an important tool in the data analysis of repeated measures and longitudinal studies. The most common form of LMMs utilize a normal distribution to model the random effects. Such assumptions can often…

Methodology · Statistics 2016-02-16 Hien D. Nguyen , Geoffrey J. McLachlan

Traditional methods for linear regression generally assume that the underlying error distribution, equivalently the distribution of the responses, is normal. Yet, sometimes real life response data may exhibit a skewed pattern, and assuming…

Methodology · Statistics 2025-01-07 Amarnath Nandy , Ayanendranath Basu , Abhik Ghosh

Bias evaluation is fundamental to trustworthy AI, both in terms of checking data quality and in terms of checking the outputs of AI systems. In testing data quality, for example, one may study the distance of a given dataset, viewed as a…

Machine Learning · Computer Science 2025-06-12 Jiří Němeček , Mark Kozdoba , Illia Kryvoviaz , Tomáš Pevný , Jakub Mareček

Propensity score (PS) methods are widely used to estimate treatment effects in non-randomized studies. Variance is typically estimated using sandwich or bootstrap methods, which can either treat the PS as estimated or fixed. The latter is…

Methodology · Statistics 2025-11-17 Baoshan Zhang , Sean M. O'Brien , Yuan Wu , Laine E. Thomas

The notion of signal sparsity has been gaining increasing interest in information theory and signal processing communities. As a consequence, a plethora of sparsity metrics has been presented in the literature. The appropriateness of these…

Information Theory · Computer Science 2016-02-08 Anastasios Maronidis , Elisavet Chatzilari , Spiros Nikolopoulos , Ioannis Kompatsiaris

Causal inference plays an important role in under standing the underlying mechanisation of the data generation process across various domains. It is challenging to estimate the average causal effect and individual causal effects from…

Data Structures and Algorithms · Computer Science 2023-01-05 Haoran Zhao , Yinghao Zhang , Debo Cheng , Chen Li , Zaiwen Feng