English
Related papers

Related papers: Subgroup analysis in multi level hierarchical clus…

200 papers

Clustering analysis is one of the most widely used statistical tools in many emerging areas such as microarray data analysis. For microarray and other high-dimensional data, the presence of many noise variables may mask underlying…

Machine Learning · Statistics 2008-03-26 Benhuai Xie , Wei Pan , Xiaotong Shen

The manuscript discusses how to incorporate random effects for quantile regression models for clustered data with focus on settings with many but small clusters. The paper has three contributions: (i) documenting that existing methods may…

Methodology · Statistics 2022-02-24 Maria Laura Battagliola , Helle Sørensen , Anders Tolver , Ana-Maria Staicu

I introduce a simple permutation procedure to test conventional (non-sharp) hypotheses about the effect of a binary treatment in the presence of a finite number of large, heterogeneous clusters when the treatment effect is identified by…

Econometrics · Economics 2023-02-08 Andreas Hagemann

This paper studies the design of cluster experiments to estimate the global treatment effect in the presence of network spillovers. We provide a framework to choose the clustering that minimizes the worst-case mean-squared error of the…

Econometrics · Economics 2025-01-29 Davide Viviano , Lihua Lei , Guido Imbens , Brian Karrer , Okke Schrijvers , Liang Shi

Technological advancements in mobile devices have made it possible to deliver mobile health interventions to individuals. A novel intervention framework that emerges from such advancements is the just-in-time adaptive intervention (JITAI),…

Methodology · Statistics 2020-07-29 Jing Xu , Xiaoxi Yan , Caroline Figueroa , Joseph Jay Williams , Bibhas Chakraborty

Estimating the effects of interventions in networks is complicated when the units are interacting, such that the outcomes for one unit may depend on the treatment assignment and behavior of many or all other units (i.e., there is…

Methodology · Statistics 2014-08-15 Dean Eckles , Brian Karrer , Johan Ugander

In cluster-randomized trials (CRTs), missing data can occur in various ways, including missing values in outcomes and baseline covariates at the individual or cluster level, or completely missing information for non-participants. Among the…

Methodology · Statistics 2025-11-06 Bingkai Wang , Fan Li , Rui Wang

Spatial clustering has important implications in various fields. In particular, disease clustering is of major public concern in epidemiology. In this article, we propose the use of two distance-based segregation indices to test the…

Methodology · Statistics 2013-10-03 Elvan Ceyhan

Although combination antiretroviral therapy (ART) is highly effective in suppressing viral load for people with HIV (PWH), many ART agents may exacerbate central nervous system (CNS)-related adverse effects including depression. Therefore,…

Methodology · Statistics 2020-04-14 Wei Jin , Yang Ni , Leah H. Rubin , Amanda B. Spence , Yanxun Xu

Randomization tests are a popular method for testing causal effects in clinical trials with finite-sample validity. In the presence of heterogeneous treatment effects, it is often of interest to select a subgroup that benefits from the…

Methodology · Statistics 2025-04-29 Zijun Gao

Mediation analysis has been comprehensively studied for independent data but relatively little work has been done for correlated data, especially for the increasingly adopted stepped wedge cluster randomized trials (SW-CRTs). Motivated by…

Methodology · Statistics 2025-04-03 Zhiqiang Cao , Fan Li

Distribution shifts remain a fundamental problem for the safe application of machine learning systems. If undetected, they may impact the real-world performance of such systems or will at least render original performance claims invalid. In…

Machine Learning · Computer Science 2023-03-10 Lisa M. Koch , Christian M. Schürch , Christian F. Baumgartner , Arthur Gretton , Philipp Berens

In cluster randomized controlled trials (CRCT) with a finite populations, the exact design-based variance of the Horvitz-Thompson (HT) estimator for the average treatment effect (ATE) depends on the joint distribution of unobserved…

Econometrics · Economics 2025-12-17 Yue Fang , Geert Ridder

Behavioral health interventions, such as trainings or incentives, are implemented in settings where individuals are interconnected, and the intervention assigned to some individuals may also affect others within their network. Evaluating…

Methodology · Statistics 2025-02-17 Zhibing He , Junhan Fan , Ashley Buchanan , Donna Spiegelman , Laura Forastiere

Most cluster randomized trials (CRTs) randomize fewer than 30-40 clusters in total. When performing inference for such ``small'' CRTs, it is important to use methods that appropriately account for the small sample size. When the generalized…

Methodology · Statistics 2025-12-01 Shifeng Sun , Xueqi Wang , Zhuoran Hou , Elizabeth L. Turner

In cluster randomized trials, the average treatment effect among individuals (i-ATE) can be different from the cluster average treatment effect (c-ATE) when informative cluster size is present, i.e., when treatment effects or participant…

Methodology · Statistics 2025-10-02 Bryan S. Blette , Zhe Chen , Brennan C. Kahan , Andrew Forbes , Michael O. Harhay , Fan Li

Randomized experiments ensure robust causal inference that are critical to effective learning analytics research and practice. However, traditional randomized experiments, like A/B tests, are limiting in large scale digital learning…

Applications · Statistics 2019-02-04 Timothy NeCamp , Josh Gardner , Christopher Brooks

Investigators often use multi-source data (e.g., multi-center trials, meta-analyses of randomized trials, pooled analyses of observational cohorts) to learn about the effects of interventions in subgroups of some well-defined target…

Methodology · Statistics 2024-02-06 Guanbo Wang , Alexander Levis , Jon Steingrimsson , Issa Dahabreh

Adaptive approaches, allowing for more flexible trial design, have been proposed for individually randomized trials to save time or reduce sample size. However, adaptive designs for cluster-randomized trials in which groups of participants…

Methodology · Statistics 2022-01-10 Junwei Shen , Shirin Golchi , Erica E. M. Moodie , David Benrimoh

Heterogeneous treatment effects (HTEs) are increasingly estimated using machine learning models that produce highly personalized predictions of treatment effects. In practice, however, predicted treatment effects are rarely interpreted,…