English
Related papers

Related papers: Two-stage Least Squares with Clustered Data under …

200 papers

Plausible identification of conditional average treatment effects (CATEs) may rely on controlling for a large number of variables to account for confounding factors. In these high-dimensional settings, estimation of the CATE requires…

Econometrics · Economics 2023-01-18 Adam Baybutt , Manu Navjeevan

We explore the capability of transformers to address endogeneity in in-context linear regression. Our main finding is that transformers inherently possess a mechanism to handle endogeneity effectively using instrumental variables (IV).…

Machine Learning · Statistics 2025-05-13 Haodong Liang , Krishnakumar Balasubramanian , Lifeng Lai

In medical, social, and behavioral research we often encounter datasets with a multilevel structure and multiple correlated dependent variables. These data are frequently collected from a study population that distinguishes several…

Methodology · Statistics 2023-12-18 Xynthia Kavelaars , Joris Mulder , Maurits Kaptein

In this paper, we provide efficient estimators and honest confidence bands for a variety of treatment effects including local average (LATE) and local quantile treatment effects (LQTE) in data-rich environments. We can handle very many…

Statistics Theory · Mathematics 2018-01-08 Alexandre Belloni , Victor Chernozhukov , Ivan Fernández-Val , Christian Hansen

The paper proposes an estimator to make inference of heterogeneous treatment effects sorted by impact groups (GATES) for non-randomised experiments. The groups can be understood as a broader aggregation of the conditional average treatment…

Econometrics · Economics 2020-03-30 Daniel Jacob

This paper estimates individual treatment effects in a triangular model with binary--valued endogenous treatments. Following the identification strategy established in Vuong and Xu (2014), we propose a two--stage estimation approach. First,…

Methodology · Statistics 2016-10-28 Qian Feng , Quang Vuong , Haiqing Xu

Conventional survival analysis approaches estimate risk scores or individualized time-to-event distributions conditioned on covariates. In practice, there is often great population-level phenotypic heterogeneity, resulting from (unknown)…

Machine Learning · Statistics 2020-03-03 Paidamoyo Chapfuwa , Chunyuan Li , Nikhil Mehta , Lawrence Carin , Ricardo Henao

Two-stage hierarchical models have been widely used in small area estimation to produce indirect estimates of areal means. When the areas are treated exchangeably and the model parameters are assumed to be the same over all areas, we might…

Methodology · Statistics 2020-01-10 Shonosuke Sugasawa , Yuki Kawakubo , Kota Ogasawara

We revisit the classical causal inference problem of estimating the average treatment effect in the presence of fully observed confounding variables using two-stage semiparametric methods. In existing theoretical studies of methods such as…

Methodology · Statistics 2022-05-23 Steve Yadlowsky

Evaluating heterogeneity of treatment effects (HTE) across subgroups is common in both randomized trials and observational studies. Although several statistical challenges of HTE analyses including low statistical power and multiple…

Methodology · Statistics 2024-07-10 Noorie Hyun , Abisola E. Idu , Andrea J. Cook , Jennifer F. Bobb

Stepped wedge designs (SWDs) are increasingly used to evaluate longitudinal cluster-level interventions but pose substantial challenges for valid inference. Because crossover times are randomized, intervention effects are intrinsically…

Methodology · Statistics 2026-05-12 Fan Xia , K. C. Gary Chan , Emily Voldal , Avi Kenny , Patrick J. Heagerty , James P. Hughes

The two-stage least-squares (2SLS) estimator is known to be biased when its first-stage fit is poor. I show that better first-stage prediction can alleviate this bias. In a two-stage linear regression model with Normal noise, I consider…

Statistics Theory · Mathematics 2017-11-01 Jann Spiess

Causal inference from observational data requires untestable identification assumptions. If these assumptions apply, machine learning (ML) methods can be used to study complex forms of causal effect heterogeneity. Recently, several ML…

Methodology · Statistics 2023-12-20 Richard Post , Isabel van den Heuvel , Marko Petkovic , Edwin van den Heuvel

We consider least squares estimation in a general nonparametric regression model. The rate of convergence of the least squares estimator (LSE) for the unknown regression function is well studied when the errors are sub-Gaussian. We find…

Statistics Theory · Mathematics 2021-04-12 Arun K. Kuchibhotla , Rohit K. Patra

In this paper, a statistical model for panel data with unobservable grouped factor structures which are correlated with the regressors and the group membership can be unknown. The factor loadings are assumed to be in different subspaces and…

Econometrics · Economics 2021-02-26 Jiangtao Duan , Wei Gao , Hao Qu , Hon Keung Tony

High-dimensional health and surveillance studies often involve many collinear predictors, multiple correlated outcomes of different types, and latent heterogeneity across observational units. We propose a Bayesian latent-cluster…

Methodology · Statistics 2026-05-13 Hsin-Hsiung Huang , Suyeon Kang

From personalised medicine to targeted advertising, it is an inherent task to provide a sequence of decisions with historical covariates and outcome data. This requires understanding of both the dynamics and heterogeneity of treatment…

Methodology · Statistics 2022-06-22 Oscar Hernan Madrid Padilla , Yi Yu

A new meta-algorithm for estimating the conditional average treatment effects is proposed in the paper. The main idea underlying the algorithm is to consider a new dataset consisting of feature vectors produced by means of concatenation of…

Machine Learning · Statistics 2019-09-10 Lev V. Utkin , Mikhail V. Kots , Viacheslav S. Chukanov

Analyzing data from multiple sources offers valuable opportunities to improve the estimation efficiency of causal estimands. However, this analysis also poses many challenges due to population heterogeneity and data privacy constraints.…

Methodology · Statistics 2025-10-23 Rong Zhao , Jason Falvey , Xu Shi , Vernon M. Chinchilli , Chixiang Chen

Feature selection is a vital technique in machine learning, as it can reduce computational complexity, improve model performance, and mitigate the risk of overfitting. However, the increasing complexity and dimensionality of datasets pose…

Machine Learning · Computer Science 2024-07-24 Yuepeng Chen , Weiping Ding , Hengrong Ju , Jiashuang Huang , Tao Yin
‹ Prev 1 3 4 5 6 7 10 Next ›