English
Related papers

Related papers: Variable Selection in Covariate Dependent Random P…

200 papers

Clustering is one of the most widely used procedures in the analysis of microarray data, for example with the goal of discovering cancer subtypes based on observed heterogeneity of genetic marks between different tissues. It is well-known…

Methodology · Statistics 2009-04-21 Heng Lian

Tuberculosis (TB), caused by Mycobacterium tuberculosis, remains a critical global health issue, necessitating timely diagnosis and treatment. Current methods for detecting tuberculosis bacilli from bright field microscopic sputum smear…

Image and Video Processing · Electrical Eng. & Systems 2025-01-08 Greeshma K , Vishnukumar S

We propose a new method to estimate causal effects from nonexperimental data. Each pair of sample units is first associated with a stochastic 'treatment' - differences in factors between units - and an effect - a resultant outcome…

Methodology · Statistics 2022-11-08 Andre F. Ribeiro , Frank Neffke , Ricardo Hausmann

When surveillance data of infectious disease incidence (e.g. weekly case counts) are disaggregated by demographic indicators, disparities in long-run health outcomes between these groups become apparent. Accurate identification of high-risk…

Methodology · Statistics 2026-05-29 Miles Moran , Rob Trangucci , Lisa Madsen

Semi-parametric methods are often used for the estimation of intervention effects on correlated outcomes in cluster-randomized trials (CRTs). When outcomes are missing at random (MAR), Inverse Probability Weighted (IPW) methods…

Methodology · Statistics 2016-01-27 Melanie Prague , Rui Wang , Alisa Stephens , Eric Tchetgen Tchetgen , Victor DeGruttola

Parkinson's disease (PD) is a common neurodegenerative disease with a high degree of heterogeneity in its clinical features, rate of progression, and change of variables over time. In this work, we present a novel data-driven, network-based…

Applications · Statistics 2020-07-01 Sanjukta Krishnagopal , Rainer Von Coelln , Lisa M. Shulman , Michelle Girvan

Survey data are often collected under multistage sampling designs where units are binned to clusters that are sampled in a first stage. The unit-indexed population variables of interest are typically dependent within cluster. We propose a…

Methodology · Statistics 2021-08-26 Luis G. Leon-Novelo , Terrance D. Savitsky

The average treatment effect can obscure important heterogeneity when individuals respond differently to a treatment. While the conditional average treatment effect (CATE) function captures such heterogeneity, it is difficult to communicate…

Methodology · Statistics 2026-05-18 Anders Munch , Thomas A. Gerds

Many diseases display heterogeneity in clinical features and their progression, indicative of the existence of disease subtypes. Extracting patterns of disease variable progression for subtypes has tremendous application in medicine, for…

Quantitative Methods · Quantitative Biology 2020-08-04 Sanjukta Krishnagopal

Infectious disease dynamics operate across multiple biological scales, with within-host viral dynamics being a key driver of between-host transmission. However, while models that explicitly link these scales exist, none have been developed…

Applications · Statistics 2026-04-23 Dylan J. Morris , Lauren Kennedy , Andrew J. Black

A stepped wedge design is a unidirectional crossover design where clusters are randomized to distinct treatment sequences. While model-based analysis of stepped wedge designs is standard practice to evaluate treatment effects accounting for…

Methodology · Statistics 2024-09-13 Bingkai Wang , Xueqi Wang , Fan Li

PURPOSE: Clinical examinations are performed on the basis of necessity. However, our decisions to investigate and document are influenced by various other factors, such as workload and preconceptions. Data missingness patterns may contain…

Applications · Statistics 2019-12-19 Robert O'Shea

Cluster-randomized trials (CRTs) on fragile populations frequently encounter complex attrition problems where the reasons for missing outcomes can be heterogeneous, with participants who are known alive, known to have died, or with unknown…

Methodology · Statistics 2025-05-06 Guangyu Tong , Chenxi Li , Eric Velazquez , Michael O. Harhay , Fan Li

The regression discontinuity (RD) design is a popular approach to causal inference in non-randomized studies. This is because it can be used to identify and estimate causal effects under mild conditions. Specifically, for each subject, the…

Methodology · Statistics 2014-02-11 George Karabatsos , Stephen G. Walker

High-throughput microarray and sequencing technology have been used to identify disease subtypes that could not be observed otherwise by using clinical variables alone. The classical unsupervised clustering strategy concerns primarily the…

Methodology · Statistics 2020-07-23 Peng Liu , Yusi Fang , Zhao Ren , Lu Tang , George C. Tseng

Difference-in-differences is undoubtedly one of the most widely used methods for evaluating the causal effect of an intervention in observational (i.e., nonrandomized) settings. The approach is typically used when pre- and post-exposure…

Methodology · Statistics 2023-08-21 Eric Tchetgen Tchetgen , Chan Park , David Richardson

$\textbf{Objective}$ Develop an automatic diagnostic system which only uses textual admission information from Electronic Health Records (EHRs) and assist clinicians with a timely and statistically proved decision tool. The hope is that the…

Computation and Language · Computer Science 2017-12-08 Christy Li , Dimitris Konomis , Graham Neubig , Pengtao Xie , Carol Cheng , Eric Xing

Real world observational data, together with causal inference, allow the estimation of causal effects when randomized controlled trials are not available. To be accepted into practice, such predictive models must be validated for the…

The Regression Discontinuity Design (RDD) is a quasi-experimental design that estimates the causal effect of a treatment when its assignment is defined by a threshold value for a continuous assignment variable. The RDD assumes that subjects…

Applications · Statistics 2020-03-27 Federico Ricciardi , Silvia Liverani , Gianluca Baio

Covariate-specific treatment effects (CSTEs) represent heterogeneous treatment effects across subpopulations defined by certain selected covariates. In this article, we consider marginal structural models where CSTEs are linearly…

Methodology · Statistics 2021-05-25 Peng Wu , Zhiqiang Tan , Wenjie Hu , Xiao-Hua Zhou