English
Related papers

Related papers: Eliminating Systematic Bias from Difference-in-Dif…

200 papers

Difference-in-differences (DID) is commonly used to estimate treatment effects but is infeasible in settings where data are unpoolable due to privacy concerns or legal restrictions on data sharing, particularly across jurisdictions. In this…

Econometrics · Economics 2025-07-28 Sunny Karim , Matthew D. Webb , Nichole Austin , Erin Strumpf

Difference-in-differences (DiD) is one of the most popular approaches for empirical research in economics, political science, and beyond. Identification in these models is based on the conditional parallel trends assumption: In the absence…

Econometrics · Economics 2025-10-13 Philipp Bach , Sven Klaassen , Jannis Kueck , Mara Mattes , Martin Spindler

Causal inference, as a major research area in statistics and data science, plays a central role across diverse fields such as medicine, economics, education, and the social sciences. Design-based causal inference begins with randomized…

Methodology · Statistics 2025-12-01 Xin Lu , Wanjia Fu , Hongzi Li , Haoyang Yu , Honghao Zhang , Ke Zhu , Hanzhong Liu

The difference-in-differences (DID) design is widely used in observational studies to estimate the causal effect of a treatment when repeated observations over time are available. Yet, almost all existing methods assume linearity in the…

Applications · Statistics 2020-09-29 Soichiro Yamauchi

Triple Differences (DDD) designs are widely used in empirical work to relax parallel trends assumptions in Difference-in-Differences (DiD) settings. This paper highlights that common DDD implementations -- such as taking the difference…

Econometrics · Economics 2025-07-21 Marcelo Ortiz-Villavicencio , Pedro H. C. Sant'Anna

Bipartite experiments arise in various fields, in which the treatments are randomized over one set of units, while the outcomes are measured over another separate set of units. However, existing methods often rely on strong model…

Methodology · Statistics 2025-04-16 Sizhu Lu , Lei Shi , Yue Fang , Wenxin Zhang , Peng Ding

Digital twins have been actively explored in many engineering applications, such as manufacturing and autonomous systems. However, model discrepancy is ubiquitous in most digital twin models and has significant impacts on the performance of…

Machine Learning · Computer Science 2025-08-12 Huchen Yang , Chuanqi Chen , Jin-Long Wu

Consider a researcher estimating the parameters of a regression function based on data for all 50 states in the United States or on data for all visits to a website. What is the interpretation of the estimated parameters and the standard…

Statistics Theory · Mathematics 2019-06-25 Alberto Abadie , Susan Athey , Guido W. Imbens , Jeffrey M. Wooldridge

The boundary discontinuity (BD) design is a non-experimental method for identifying causal effects that exploits a thresholding rule based on a bivariate score and a boundary curve. This widely used method generalizes the univariate…

Econometrics · Economics 2026-02-16 Matias D. Cattaneo , Rocio Titiunik , Ruiqi Rae Yu

Statistics is sometimes described as the science of reasoning under uncertainty. Statistical models provide one view of this uncertainty, but what is frequently neglected is the 'invisible' portion of uncertainty: that assumed not to exist…

Methodology · Statistics 2026-03-18 Oliver L. Pescott , Robin J. Boyd , Gary D. Powney , Gavin B. Stewart

In many science and engineering settings, system dynamics are characterized by governing PDEs, and a major challenge is to solve inverse problems (IPs) where unknown PDE parameters are inferred based on observational data gathered under…

Machine Learning · Computer Science 2025-03-11 Apivich Hemachandra , Gregory Kang Ruey Lau , See-Kiong Ng , Bryan Kian Hsiang Low

Bayesian experimental design (BED) is to answer the question that how to choose designs that maximize the information gathering. For implicit models, where the likelihood is intractable but sampling is possible, conventional BED methods…

Machine Learning · Computer Science 2021-03-16 Jiaxin Zhang , Sirui Bi , Guannan Zhang

We propose a new estimation method for heterogeneous causal effects which utilizes a regression discontinuity (RD) design for multiple datasets with different thresholds. The standard RD design is frequently used in applied researches, but…

Econometrics · Economics 2019-05-14 Takayuki Toda , Ayako Wakano , Takahiro Hoshino

Dataset bias is a significant challenge in machine learning, where specific attributes, such as texture or color of the images are unintentionally learned resulting in detrimental performance. To address this, previous efforts have focused…

Computer Vision and Pattern Recognition · Computer Science 2024-06-11 Donggeun Ko , Sangwoo Jo , Dongjun Lee , Namjun Park , Jaekwang Kim

This article develops a covariate balancing approach for the estimation of treatment effects on the treated (ATT) in a difference-in-differences (DID) research design when panel data are available. We show that the proposed covariate…

Econometrics · Economics 2025-08-05 Junjie Li , Yukitoshi Matsushita

We present a novel extension of the influential changes-in-changes (CiC) framework of Athey and Imbens (2006) for estimating the average treatment effect on the treated (ATT) and distributional causal effects in panel data with unmeasured…

Methodology · Statistics 2025-08-20 Jinghao Sun , Eric J. Tchetgen Tchetgen

Managers, employers, policymakers, and others often seek to understand whether decisions are biased against certain groups. One popular analytic strategy is to estimate disparities after adjusting for observed covariates, typically with a…

Applications · Statistics 2024-01-29 Jongbin Jung , Sam Corbett-Davies , Johann D. Gaebler , Ravi Shroff , Sharad Goel

Traditional regression and prediction tasks often only provide deterministic point estimates. To estimate the distribution or uncertainty of the response variable, traditional methods either assume that the posterior distribution of samples…

Machine Learning · Computer Science 2025-01-08 Daojun Liang , Haixia Zhang , Dongfeng Yuan

This paper proposes a novel approach for estimating treatment effects in panel data settings, addressing key limitations of the standard difference-in-differences (DID) approach. The standard approach relies on the parallel trends…

Econometrics · Economics 2026-01-14 Shoya Ishimaru

We describe the DISC (Different Individuals, Same Clusters) design, a sampling scheme that can improve the precision of difference-in-differences (DID) estimators in settings involving repeated sampling of a population at multiple time…

Methodology · Statistics 2025-08-21 Jordan Downey , Avi Kenny