English
Related papers

Related papers: A causal inference framework for cancer cluster in…

200 papers

Competing risk analysis considers event times due to multiple causes, or of more than one event types. Commonly used regression models for such data include 1) cause-specific hazards model, which focuses on modeling one type of event while…

Applications · Statistics 2017-04-27 Jiayi Hou , Anthony Paravati , Ronghui Xu , James Murphy

In epidemiology research with cancer registry data, it is often of primary interest to make inference on cancer death, not overall survival. Since cause of death is not easy to collect or is not necessarily reliable in cancer registries,…

Methodology · Statistics 2023-03-17 Sho Komukai , Satoshi Hattori , Bernard Rachet

The standard approach to causal modelling especially in social and health sciences is the potential outcomes framework due to Neyman and Rubin. In this framework, observations are thought to be drawn from a distribution over variables of…

Methodology · Statistics 2025-07-18 Benedikt Höltgen , Robert C. Williamson

Longitudinal cohort studies, which follow a group of individuals over time, provide the opportunity to examine causal effects of complex exposures on long-term health outcomes. Utilizing data from multiple cohorts has the potential to add…

The estimation of causal effects is a fundamental goal in the field of causal inference. However, it is challenging for various reasons. One reason is that the exposure (or treatment) is naturally continuous in many real-world scenarios.…

Methodology · Statistics 2023-12-01 Suhwan Bong , Kwonsang Lee

Subspace clustering refers to the problem of segmenting high dimensional data drawn from a union of subspaces into the respective subspaces. In some applications, partial side-information to indicate "must-link" or "cannot-link" in…

Computer Vision and Pattern Recognition · Computer Science 2018-05-23 Chun-Guang Li , Junjian Zhang , Jun Guo

Cluster sampling is common in survey practice, and the corresponding inference has been predominantly design-based. We develop a Bayesian framework for cluster sampling and account for the design effect in the outcome modeling. We consider…

Methodology · Statistics 2020-06-24 Susanna Makela , Yajuan Si , Andrew Gelman

Most existing causal structure learning methods assume data collected from one environment and independent and identically distributed (i.i.d.). In some cases, data are collected from different subjects from multiple environments, which…

Machine Learning · Computer Science 2023-02-07 Wei Chen , Yunjin Wu , Ruichu Cai , Yueguo Chen , Zhifeng Hao

Identifying covariates that modify treatment effects is a central problem in causal inference. Yet existing data-adaptive procedures do not provide finite-sample control over the expected number of false discoveries, risking spurious…

Methodology · Statistics 2026-05-12 Falco J. Bargagli-Stoffi , Omar Melikechi

Female breast cancer (FBC) incidence rate (IR) varies greatly by counties across the United States (US). Factors responsible for such high spatial disparities are not well understood, making it challenging to design effective intervention…

Applications · Statistics 2024-01-19 Tingting Zhao , Qing Han , Jinfeng Zhang

Causal identification of treatment effects for infectious disease outcomes in interconnected populations is challenging because infection outcomes may be transmissible to others, and treatment given to one individual may affect others'…

Methodology · Statistics 2021-05-11 Xiaoxuan Cai , Eben Kenah , Forrest W. Crawford

In most nonrandomized observational studies, differences between treatment groups may arise not only due to the treatment but also because of the effect of confounders. Therefore, causal inference regarding the treatment effect is not as…

Methodology · Statistics 2018-07-04 Debashis Ghosh

We discuss a shift in perspective from traditional approaches to breast cancer risk prediction: modelling families rather than individuals as unit of analysis. By investigating the latent familial risk underlying breast cancer diagnoses, we…

Applications · Statistics 2025-08-25 Maria Veronica Vinattieri , Marco Bonetti , Kamila Czene

The task of clustering a set of objects based on multiple sources of data arises in several modern applications. We propose an integrative statistical model that permits a separate clustering of the objects for each data source. These…

Machine Learning · Statistics 2015-12-01 Eric F. Lock , David B. Dunson

In cluster-randomized crossover (CRXO) trials, groups of individuals are randomly assigned to two or more sequences of alternating treatments. Since clusters serve as their own control, the CRXO design is typically more statistically…

Methodology · Statistics 2026-01-30 Dane Isenberg , Michael O. Harhay , Andrew B. Forbes , Paul J. Young , Fan Li , Nandita Mitra

Causal inference is a critical research topic across many domains, such as statistics, computer science, education, public policy and economics, for decades. Nowadays, estimating causal effect from observational data has become an appealing…

Methodology · Statistics 2020-02-10 Liuyi Yao , Zhixuan Chu , Sheng Li , Yaliang Li , Jing Gao , Aidong Zhang

Modern bio-technologies have produced a vast amount of high-throughput data with the number of predictors far greater than the sample size. In order to identify more novel biomarkers and understand biological mechanisms, it is vital to…

Machine Learning · Statistics 2018-05-18 Kevin He , Jian Kang , Hyokyoung Grace Hong , Ji Zhu , Yanming Li , Huazhen Lin , Han Xu , Yi Li

Scientists regularly pose questions about treatment effects on outcomes conditional on a post-treatment event. However, causal inference in such settings requires care, even in perfectly executed randomized experiments. Recently, the…

Methodology · Statistics 2026-02-19 Chan Park , Mats Stensrud , Eric Tchetgen Tchetgen

Understanding the factors that trigger or prevent undesirable health outcomes across patient subpopulations is essential for designing targeted interventions. While randomized controlled trials and expert-led patient interviews are standard…

Artificial Intelligence · Computer Science 2026-05-28 Shishir Adhikari , Guido Muscioni , Mark Shapiro , Plamen Petrov , Elena Zheleva

In the absence of a randomized experiment, a key assumption for drawing causal inference about treatment effects is the ignorable treatment assignment. Violations of the ignorability assumption may lead to biased treatment effect estimates.…

Methodology · Statistics 2021-08-17 Liangyuan Hu , Jungang Zou , Chenyang Gu , Jiayi Ji , Michael Lopez , Minal Kale