English
Related papers

Related papers: Propensity score weighting under limited overlap a…

200 papers

One of the primary goals of statistical precision medicine is to learn optimal individualized treatment rules (ITRs). The classification-based, or machine learning-based, approach to estimating optimal ITRs was first introduced in…

Methodology · Statistics 2024-06-18 Sophia Yazzourh , Nikki L. B. Freeman

In observational studies, researchers must select a method to control for confounding. Options include propensity score methods and regression. It remains unclear how dataset characteristics (size, overlap in propensity scores, exposure…

Methodology · Statistics 2022-10-21 J. Wilkinson , M. A. Mamas , E. Kontopantelis

Evaluating the performance of a prediction model is a common task in medical statistics. Standard accuracy metrics require the observation of the true outcomes. This is typically not possible in the setting with time-to-event outcomes due…

Methodology · Statistics 2025-07-18 Zhenwei Yang , Dimitris Rizopoulos , Lisa F. Newcomb , Nicole S. Erler

The general aim of the recommender system is to provide personalized suggestions to users, which is opposed to suggesting popular items. However, the normal training paradigm, i.e., fitting a recommender model to recover the user behavior…

Information Retrieval · Computer Science 2021-06-08 Tianxin Wei , Fuli Feng , Jiawei Chen , Ziwei Wu , Jinfeng Yi , Xiangnan He

Causal inference analyses often use existing observational data, which in many cases has some clustering of individuals. In this paper we discuss propensity score weighting methods in a multilevel setting where within clusters individuals…

Applications · Statistics 2020-12-24 Youjin Lee , Trang Q. Nguyen , Elizabeth A. Stuart

Nested case-control (NCC) is a sampling method widely used for developing and evaluating risk models with expensive biomarkers on large prospective cohort studies. The biomarker values are typically obtained on a sub-cohort, consisting of…

Methodology · Statistics 2021-04-07 Qian M. Zhou , Xuan Wang , Yingye Zheng , Tianxi Cai

The inverse probability (IPW) and doubly robust (DR) estimators are often used to estimate the average causal effect (ATE), but are vulnerable to outliers. The IPW/DR median can be used for outlier-resistant estimation of the ATE, but the…

Methodology · Statistics 2024-09-16 Kazuharu Harada , Hironori Fujisawa

RCTs sometimes test interventions that aim to improve existing services targeted to a subset of individuals identified after randomization. Accordingly, the treatment could affect the composition of service recipients and the offered…

Methodology · Statistics 2022-05-18 Peter Z. Schochet

Offline reinforcement learning refers to the process of learning policies from fixed datasets, without requiring additional environment interaction. However, it often relies on well-defined reward functions, which are difficult and…

Artificial Intelligence · Computer Science 2025-10-13 Xiancheng Gao , Yufeng Shi , Wengang Zhou , Houqiang Li

We study the problem of estimating the average treatment effect (ATE) under sequentially adaptive treatment assignment mechanisms. In contrast to classical completely randomized designs, we consider a setting in which the probability of…

Statistics Theory · Mathematics 2026-05-12 Saikat Sengupta , Koulik Khamaru , Suvrojit Ghosh , Tirthankar Dasgupta

We address measurement error bias in propensity score (PS) analysis due to covariates that are latent variables. In the setting where latent covariate $X$ is measured via multiple error-prone items $\mathbf{W}$, PS analysis using several…

Methodology · Statistics 2020-02-13 Trang Quynh Nguyen , Elizabeth A. Stuart

Learning and evaluating recommender systems from logged implicit feedback is challenging due to exposure bias. While inverse propensity scoring (IPS) corrects this bias, it often suffers from high variance and instability. In this paper, we…

Machine Learning · Computer Science 2025-09-03 Rahul Raja , Arpita Vats

Many observational studies feature irregular longitudinal data, where the observation times are not common across individuals in the study. Further, the observation times may be related to the longitudinal outcome. In this setting, failing…

Methodology · Statistics 2024-05-27 Grace Tompkins , Joel A Dubin , Michael Wallace

The large-scale multiple testing inherent to high throughput biological data necessitates very high statistical stringency and thus true effects in data are difficult to detect unless they have high effect sizes. One solution to this…

Methodology · Statistics 2017-12-21 Mohamad S. Hasan

Estimating the mean counterfactual outcome under a treatment rule is a central problem in causal inference and policy evaluation. Standard estimators, including inverse probability weighting (IPW), augmented IPW (AIPW), and targeted maximum…

Methodology · Statistics 2026-05-06 Yichen Xu , Mark J. van der Laan

There has been growing attention on how to effectively and objectively use covariate information when the primary goal is to estimate the average treatment effect (ATE) in randomized clinical trials (RCTs). In this paper, we propose an…

Methodology · Statistics 2020-09-01 Yuanyao Tan , Xialing Wen , Wei Liang , Ying Yan

Weighting and trimming are popular methods for addressing positivity violations in causal inference. While well-studied with single-timepoint data, standard methods do not easily generalize to address non-baseline positivity violations in…

Methodology · Statistics 2025-06-12 Alec McClean , Alexander W. Levis , Nicholas Williams , Ivan Diaz

Selection bias is a major obstacle toward valid causal inference in epidemiology. Over the past decade, several graphical rules based on causal diagrams have been proposed as the sufficient identification conditions for addressing selection…

Methodology · Statistics 2025-12-18 Yichi Zhang , Haidong Lu

Personalized decision-making, tailored to individual characteristics, is gaining significant attention. The optimal treatment regime aims to provide the best-expected outcome in the entire population, known as the value function. One…

Methodology · Statistics 2024-05-28 Yuwen Cheng , Shu Yang

Semi-parametric methods are often used for the estimation of intervention effects on correlated outcomes in cluster-randomized trials (CRTs). When outcomes are missing at random (MAR), Inverse Probability Weighted (IPW) methods…

Methodology · Statistics 2016-01-27 Melanie Prague , Rui Wang , Alisa Stephens , Eric Tchetgen Tchetgen , Victor DeGruttola