中文
相关论文

相关论文: Correcting for non-ignorable missingness in smokin…

200 篇论文

Exposure to air pollution is associated with increased morbidity and mortality. Recent technological advancements permit the collection of time-resolved personal exposure data. Such data are often incomplete with missing observations and…

Electronic health record (EHR) data has emerged as a valuable resource for analyzing patient health status. However, the prevalence of missing data in EHR poses significant challenges to existing methods, leading to spurious correlations…

机器学习 · 计算机科学 2024-05-16 Zhihao Yu , Xu Chu , Yujie Jin , Yasha Wang , Junfeng Zhao

Decision trees are a popular family of models due to their attractive properties such as interpretability and ability to handle heterogeneous data. Concurrently, missing data is a prevalent occurrence that hinders performance of machine…

机器学习 · 计算机科学 2020-07-01 Pasha Khosravi , Antonio Vergari , YooJung Choi , Yitao Liang , Guy Van den Broeck

Flexible laryngoscopy is commonly performed by otolaryngologists to detect laryngeal diseases and to recognize potentially malignant lesions. Recently, researchers have introduced machine learning techniques to facilitate automated…

计算机视觉与模式识别 · 计算机科学 2023-05-29 Tianxiao Zhang , Andrés M. Bur , Shannon Kraft , Hannah Kavookjian , Bryan Renslo , Xiangyu Chen , Bo Luo , Guanghui Wang

Electronic health records (EHR) are rich heterogeneous collection of patient health information, whose broad adoption provides great opportunities for systematic health data mining. However, heterogeneous EHR data types and biased…

机器学习 · 计算机科学 2018-11-02 Yue Li , Manolis Kellis

We illustrate a class of conditional models for the analysis of longitudinal data suffering attrition in random effects models framework, where the subject-specific random effects are assumed to be discrete and to follow a time-dependent…

统计方法学 · 统计学 2014-04-28 Antonello Maruotti

Despite the recent development of methods dealing with partially observed epidemic dynamics (unobserved model coordinates, discrete and noisy outbreak data), limitations remain in practice, mainly related to the quantity of augmented data…

应用统计 · 统计学 2021-07-26 Romain Narci , Maud Delattre , Catherine Larédo , Elisabeta Vergu

Longitudinal data are essential for studying within subject change and between subject differences in change. However, missing data, especially when the observed variables are nonnormal, remain a significant challenge in longitudinal…

统计方法学 · 统计学 2025-04-21 Dandan Tang , Xin Tong , Jianhui Zhou

PURPOSE: Clinical examinations are performed on the basis of necessity. However, our decisions to investigate and document are influenced by various other factors, such as workload and preconceptions. Data missingness patterns may contain…

应用统计 · 统计学 2019-12-19 Robert O'Shea

We offer a natural and extensible measure-theoretic treatment of missingness at random. Within the standard missing data framework, we give a novel characterisation of the observed data as a stopping-set sigma algebra. We demonstrate that…

统计方法学 · 统计学 2018-01-23 Daniel Farewell , Rhian Daniel , Shaun Seaman

Deep Learning (DL) methods have dramatically increased in popularity in recent years, with significant growth in their application to supervised learning problems in the biomedical sciences. However, the greater prevalence and complexity of…

机器学习 · 统计学 2023-10-30 David K Lim , Naim U Rashid , Junier B Oliva , Joseph G Ibrahim

Electronic health records (EHR) are characterized as non-stationary, heterogeneous, noisy, and sparse data; therefore, it is challenging to learn the regularities or patterns inherent within them. In particular, sparseness caused mostly by…

机器学习 · 计算机科学 2020-03-03 Eunji Jun , Ahmad Wisnu Mulyadi , Jaehun Choi , Heung-Il Suk

Identifiability is the property in mathematical modelling that determines if model parameters can be uniquely estimated from data. For infectious disease models, failure to ensure identifiability can lead to misleading parameter estimates…

统计方法学 · 统计学 2025-06-10 Fanny Bergström , Martina Favero , Tom Britton

Causal inference in observational studies can be challenging when confounders are subject to missingness. Generally, the identification of causal effects is not guaranteed even under restrictive parametric model assumptions when confounders…

统计方法学 · 统计学 2023-03-23 Jian Sun , Bo Fu

Data harmonization is the process by which an equivalence is developed between two variables measuring a common trait. Our problem is motivated by dementia research in which multiple tests are used in practice to measure the same underlying…

统计方法学 · 统计学 2021-10-13 Steven Wilkins-Reeves , Yen-Chi Chen , Kwun Chuen Gary Chan

This paper contributes a set of quality metrics for identification and visual analysis of structured missingness in high-dimensional data. Missing values in data are a frequent challenge in most data generating domains and may cause a range…

图形学 · 计算机科学 2025-05-30 Sara Johansson Fernstad , Sarah Alsufyani , Silvia Del Din , Alison Yarnall , Lynn Rochester

Longitudinal electronic health record (EHR) data offer opportunities to study biomarker trajectories; however, association estimates-the primary inferential target-from standard models designed for regular observation times may be biased by…

统计方法学 · 统计学 2026-02-18 Cheng-Han Yang , Xu Shi , Bhramar Mukherjee

We expand Mendelian Randomization (MR) methodology to deal with randomly missing data on either the exposure or the outcome variable, and furthermore with data from nonindependent individuals (eg components of a family). Our method rests on…

Dyadic data are common in the social and behavioral sciences, in which members of dyads are correlated due to the interdependence structure within dyads. The analysis of longitudinal dyadic data becomes complex when nonignorable dropouts…

应用统计 · 统计学 2012-06-29 Guangyu Zhang , Ying Yuan

A large fraction of the electronic health records (EHRs) consists of clinical measurements collected over time, such as lab tests and vital signs, which provide important information about a patient's health status. These sequences of…

机器学习 · 统计学 2020-03-02 Karl Øyvind Mikalsen , Cristina Soguero-Ruiz , Robert Jenssen