English
Related papers

Related papers: Semi-supervised Estimation of Event Rate with Doub…

200 papers

Predicting time-to-event outcomes in large databases can be a challenging but important task. One example of this is in predicting the time to a clinical outcome for patients in intensive care units (ICUs), which helps to support critical…

Computation · Statistics 2019-08-06 Yingying Xu , Joon Lee , Joel A. Dubin

While the volume of electronic health records (EHR) data continues to grow, it remains rare for hospital systems to capture dense physiological data streams, even in the data-rich intensive care unit setting. Instead, typical EHR records…

Machine Learning · Computer Science 2018-12-04 Satya Narayan Shukla , Benjamin M. Marlin

Meta-analysis is a powerful tool for assessing drug safety by combining treatment-related toxicological findings across multiple studies, as clinical trials are typically underpowered for detecting adverse drug effects. However, incomplete…

Employing a machine learning approach we predict, up to 24 hours prior, a diagnosis of severe sepsis. Strongly predictive models are possible that use only text reports from the Electronic Health Record (EHR), and omit structured numerical…

Computers and Society · Computer Science 2017-12-01 Phil Culliton , Michael Levinson , Alice Ehresman , Joshua Wherry , Jay S. Steingrub , Stephen I. Gallant

Electronic Health Records (EHR) have been heavily used in modern healthcare systems for recording patients' admission information to hospitals. Many data-driven approaches employ temporal features in EHR for predicting specific diseases,…

Machine Learning · Computer Science 2021-12-07 Chang Lu , Chandan K. Reddy , Yue Ning

Representative risk estimation is fundamental to clinical decision-making. However, risks are often estimated from non-representative epidemiologic studies, which usually underrepresent minorities. "Model-based" methods use population…

Methodology · Statistics 2023-04-12 Lingxiao Wang , Yan Li , Barry I. Graubard , Hormuzd A. Katki

Electronic Health Records (EHRs) enable deep learning for clinical predictions, but the optimal method for representing patient data remains unclear due to inconsistent evaluation practices. We present the first systematic benchmark to…

Machine Learning · Computer Science 2025-10-13 Tianyi Chen , Mingcheng Zhu , Zhiyao Luo , Tingting Zhu

Electronic health records (EHRs) contain patients' heterogeneous data that are collected from medical providers involved in the patient's care, including medical notes, clinical events, laboratory test results, symptoms, and diagnoses. In…

Artificial Intelligence · Computer Science 2024-11-12 Shuai Niu , Yunya Song , Qing Yin , Yike Guo , Xian Yang

Interval censoring arises frequently in clinical, epidemiological, financial, and sociological studies, where the event or failure of interest is known only to occur within an interval induced by periodic monitoring. We formulate the…

Methodology · Statistics 2016-03-01 Donglin Zeng , Lu Mao , D. Y. Lin

The onset of several silent, chronic diseases such as diabetes can be detected only through diagnostic tests. Due to cost considerations, self-reported outcomes are routinely collected in lieu of expensive diagnostic tests in large-scale…

Applications · Statistics 2015-09-15 Xiangdong Gu , Yunsheng Ma , Raji Balasubramanian

It is often of interest to study the association between covariates and the cumulative incidence of a right-censored time-to-event outcome. When time-varying covariates are measured on a fixed discrete time scale, it is desirable to account…

Methodology · Statistics 2026-04-28 Hongxiang Qiu , Marco Carone , Alex Luedtke , Peter B. Gilbert

Electronic healthcare records (EHR) contain a huge wealth of data that can support the prediction of clinical outcomes. EHR data is often stored and analysed using clinical codes (ICD10, SNOMED), however these can differ across registries…

Machine Learning · Computer Science 2024-12-03 Elizabeth Remfry , Rafael Henkin , Michael R Barnes , Aakanksha Naik

In many modern machine learning applications, the outcome is expensive or time-consuming to collect while the predictor information is easy to obtain. Semi-supervised learning (SSL) aims at utilizing large amounts of `unlabeled' data along…

Methodology · Statistics 2017-11-16 Jessica Gronsbell , Tianxi Cai

We propose a semiparametric framework for causal inference with right-censored survival outcomes and many weak invalid instruments, motivated by Mendelian randomization in biobank studies where classical methods may fail. We adopt an…

Methodology · Statistics 2025-10-06 Qiushi Bu , Wen Su , Xingqiu Zhao , Zhonghua Liu

Electronic health records (EHRs) are increasingly recognized as a cost-effective resource for patient recruitment in clinical research. However, how to optimally select a cohort from millions of individuals to answer a scientific question…

Methodology · Statistics 2023-12-14 Guanghao Zhang , Lauren J. Beesley , Bhramar Mukherjee , Xu Shi

Motivation: Electronic health record (EHR) data provides a new venue to elucidate disease comorbidities and latent phenotypes for precision medicine. To fully exploit its potential, a realistic data generative process of the EHR data needs…

Machine Learning · Computer Science 2021-05-05 Ziyang Song , Xavier Sumba Toral , Yixin Xu , Aihua Liu , Liming Guo , Guido Powell , Aman Verma , David Buckeridge , Ariane Marelli , Yue Li

This paper focuses on quantifying and estimating the predictive accuracy of prognostic models for time-to-event outcomes with competing events. We consider the time-dependent discrimination and calibration metrics, including the receiver…

Methodology · Statistics 2017-07-14 Cai Wu , Liang Li

The electronic health record (EHR) provides an unprecedented opportunity to build actionable tools to support physicians at the point of care. In this paper, we investigate survival analysis in the context of EHR data. We introduce deep…

Machine Learning · Statistics 2016-09-20 Rajesh Ranganath , Adler Perotte , Noémie Elhadad , David Blei

Measurement error arises commonly in clinical research settings that rely on data from electronic health records or large observational cohorts. In particular, self-reported outcomes are typical in cohort studies for chronic diseases such…

Methodology · Statistics 2021-02-08 Lillian A. Boe , Lesley F. Tinker , Pamela A. Shaw

The growing availability of observational databases like electronic health records (EHR) provides unprecedented opportunities for secondary use of such data in biomedical research. However, these data can be error-prone and need to be…

Methodology · Statistics 2024-05-28 Sarah C. Lotspeich , Gustavo G. C. Amorim , Pamela A. Shaw , Ran Tao , Bryan E. Shepherd