中文
相关论文

相关论文: Using Missing Types to Improve Partial Identificat…

200 篇论文

Electronic health records are a valuable data source for investigating health-related questions, and propensity score analysis has become an increasingly popular approach to address confounding bias in such investigations. However, because…

We describe the Bedside Patient Rescue (BPR) project, the goal of which is risk prediction of adverse events for non-ICU patients using ~200 variables (vitals, lab results, assessments, ...). There are several missing predictor values for…

Estimating the causal effects of interventions is crucial to policy and decision-making, yet outcome data are often missing or subject to non-standard measurement error. While ground-truth outcomes can sometimes be obtained through costly…

机器学习 · 统计学 2026-04-22 Ezinne Nwankwo , Lauri Goldkind , Angela Zhou

Routinely collected nation-wide registers contain socio-economic and health-related information from a large number of individuals. However, important information on lifestyle, biological and other risk factors is available at most for…

统计方法学 · 统计学 2024-09-23 Tommi Härkänen , Sangita Kulathinal , Arya Panthalanickal Vijayakumar

Understanding the prevalence of key demographic and health indicators in small geographic areas and domains is of global interest, especially in low- and middle-income countries (LMICs), where vital registration data is sparse and household…

应用统计 · 统计学 2025-04-24 Qianyu Dong , Yunhan Wu , Zehang Richard Li , Jon Wakefield

Background: Although the missing covariate indicator method (MCIM) has been shown to be biased under extreme conditions, the degree and determinants of bias have not been formally assessed. We derived the formula for the relative bias in…

应用统计 · 统计学 2025-08-01 Gang Xu , Mingyang Song , Xin Zhou , Yilun Wu , Mathew Pazaris , Donna Spiegelman

Missing data is a common problem in clinical data collection, which causes difficulty in the statistical analysis of such data. In this article, we consider the problem under a framework of a semiparametric partially linear model when…

统计方法学 · 统计学 2022-06-13 Zishu Zhan , Xiangjie Li , Jingxiao Zhang

In designed experiments and surveys, known laws or design feat ures provide checks on the most relevant aspects of a model and identify the target parameters. In contrast, in most observational studies in the health and social sciences, the…

统计方法学 · 统计学 2010-01-18 Sander Greenland

The cause of failure in cohort studies that involve competing risks is frequently incompletely observed. To address this, several methods have been proposed for the semiparametric proportional cause-specific hazards model under a missing at…

统计方法学 · 统计学 2020-02-24 Giorgos Bakoyannis , Ying Zhang , Constantin T. Yiannoutsos

Reliable estimation of spatio-temporal trends in population-level HIV incidence is becoming an increasingly critical component of HIV prevention policy-making. However, direct measurement is nearly impossible. Current, widely used models…

应用统计 · 统计学 2019-12-04 Timothy M Wolock , Seth R Flaxman , Jeffrey W Eaton

Identifying significant sites in sequence data and analogous data is of fundamental importance in many biological fields. Fisher's exact test is a popular technique, however this approach to sparse count data is not appropriate due to…

This paper addresses one of the most prevalent problems encountered by political scientists working with difference-in-differences (DID) design: missingness in panel data. A common practice for handling missing data, known as complete case…

统计方法学 · 统计学 2024-12-02 Sooahn Shin

Measurement error arises commonly in clinical research settings that rely on data from electronic health records or large observational cohorts. In particular, self-reported outcomes are typical in cohort studies for chronic diseases such…

统计方法学 · 统计学 2021-02-08 Lillian A. Boe , Lesley F. Tinker , Pamela A. Shaw

Interval-censored competing risks data arise when each study subject may experience an event or failure from one of several causes and the failure time is not observed exactly but rather known to lie in an interval between two successive…

统计方法学 · 统计学 2016-03-02 Lu Mao , D. Y. Lin , Donglin Zeng

Many clinical studies require the follow-up of patients over time. This is challenging: apart from frequently observed drop-out, there are often also organizational and financial challenges, which can lead to reduced data collection and, in…

机器学习 · 计算机科学 2022-10-26 Fateme Nateghi Haredasht , Celine Vens

This paper addresses the sample selection model within the context of the gender gap problem, where even random treatment assignment is affected by selection bias. By offering a robust alternative free from distributional or specification…

计量经济学 · 经济学 2024-10-04 Xiaolin Sun , Xueyan Zhao , D. S. Poskitt

Social context plays an important role in perpetuating or reducing HIV risk behaviors. This study analyzed the network and individual attributes that were associated with the likelihood that people who inject drugs (PWID) will engage in HIV…

Econometricians have usefully separated study of estimation into identification and statistical components. Identification analysis, which assumes knowledge of the probability distribution generating observable data, places an upper bound…

计量经济学 · 经济学 2025-09-03 Charles F. Manski

Noncompliance and missing data often occur in randomized trials, which complicate the inference of causal effects. When both noncompliance and missing data are present, previous papers proposed moment and maximum likelihood estimators for…

统计方法学 · 统计学 2014-09-04 Hua Chen , Peng Ding , Zhi Geng , Xiao-Hua Zhou

Survival analysis is a widely-used technique for analyzing time-to-event data in the presence of censoring. In recent years, numerous survival analysis methods have emerged which scale to large datasets and relax traditional assumptions…

机器学习 · 计算机科学 2023-11-06 Mert Ketenci , Shreyas Bhave , Noémie Elhadad , Adler Perotte