中文
相关论文

相关论文: Outcome identification in electronic health record…

200 篇论文

Difference-in-differences (DID) approaches are widely used for estimating causal effects with observational data before and after an intervention. DID traditionally estimates the average treatment effect among the treated after making a…

统计方法学 · 统计学 2025-06-24 Julia C. Thome , Andrew J. Spieker , Peter F. Rebeiro , Chun Li , Tong Li , Bryan E. Shepherd

Dirichlet process mixtures are flexible non-parametric models, particularly suited to density estimation and probabilistic clustering. In this work we study the posterior distribution induced by Dirichlet process mixtures as the sample size…

统计理论 · 数学 2022-11-29 Filippo Ascolani , Antonio Lijoi , Giovanni Rebaudo , Giacomo Zanella

Diabetes Mellitus has no permanent cure to date and is one of the leading causes of death globally. The alarming increase in diabetes calls for the need to take precautionary measures to avoid/predict the occurrence of diabetes. This paper…

机器学习 · 计算机科学 2023-01-26 Alain Hennebelle , Huned Materwala , Leila Ismail

Driven by the multi-level structure of human intracranial electroencephalogram (iEEG) recordings of epileptic seizures, we introduce a new variant of a hierarchical Dirichlet Process---the multi-level clustering hierarchical Dirichlet…

应用统计 · 统计学 2012-06-22 Drausin Wulsin , Shane Jensen , Brian Litt

Motivation: With the development of droplet based systems, massive single cell transcriptome data has become available, which enables analysis of cellular and molecular processes at single cell resolution and is instrumental to…

机器学习 · 计算机科学 2018-12-27 Tiehang Duan , José P. Pinto , Xiaohui Xie

The modelling of action potentials from extracellular recordings, or spike sorting, is a rich area of neuroscience research in which latent variable models are often used. Two such models, Overfitted Finite Mixture models (OFMs) and…

应用统计 · 统计学 2016-02-08 Zoé van Havre , Nicole White , Judith Rousseau , Kerrie Mengersen

In public health management there is a need to produce subnational estimates of health outcomes. Often, however, funds are not available to collect samples large enough to produce traditional survey sample estimates for each subnational…

应用统计 · 统计学 2008-12-18 Donald Malec , Peter Müller

We develop a new Gibbs sampler for a linear mixed model with a Dirichlet process random effect term, which is easily extended to a generalized linear mixed model with a probit link function. Our Gibbs sampler exploits the properties of the…

统计理论 · 数学 2010-02-26 Minjung Kyung , Jeff Gill , George Casella

We propose a general modeling framework for marked Poisson processes observed over time or space. The modeling approach exploits the connection of the nonhomogeneous Poisson process intensity with a density function. Nonparametric Dirichlet…

统计方法学 · 统计学 2011-11-02 Matthew A. Taddy , Athanasios Kottas

The use of hierarchical mixture priors with shared atoms has recently flourished in the Bayesian literature for partially exchangeable data. Leveraging on nested levels of mixtures, these models allow the estimation of a two-layered data…

统计方法学 · 统计学 2024-06-21 Laura D'Angelo , Francesco Denti

Dirichlet Process(DP) is a Bayesian non-parametric prior for infinite mixture modeling, where the number of mixture components grows with the number of data items. The Hierarchical Dirichlet Process (HDP), is an extension of DP for grouped…

机器学习 · 统计学 2015-09-02 Lavanya Sita Tekumalla , Priyanka Agrawal , Indrajit Bhattacharya

Health risk prediction is one of the fundamental tasks under predictive modeling in the medical domain, which aims to forecast the potential health risks that patients may face in the future using their historical Electronic Health Records…

机器学习 · 计算机科学 2023-10-09 Yuan Zhong , Suhan Cui , Jiaqi Wang , Xiaochen Wang , Ziyi Yin , Yaqing Wang , Houping Xiao , Mengdi Huai , Ting Wang , Fenglong Ma

Marked point process data arise when events occur in a space with event-level marks. We study clustering of replicated marked Poisson point processes and introduce Dirichlet process mixtures of marked Poisson point processes, a Bayesian…

统计方法学 · 统计学 2026-05-12 Minsung Choi , Seonghyun Jeong

We present the \textit{hierarchical Dirichlet scaling process} (HDSP), a Bayesian nonparametric mixed membership model. The HDSP generalizes the hierarchical Dirichlet process (HDP) to model the correlation structure between metadata in the…

机器学习 · 计算机科学 2017-07-10 Dongwoo Kim , Alice Oh

Modeling the dynamics of probability distributions from time-dependent data samples is a fundamental problem in many fields, including digital health. The goal is to analyze how the distribution of a biomarker, such as glucose, changes over…

机器学习 · 统计学 2025-09-18 Antonio Álvarez-López , Marcos Matabuena

Neural network classifiers trained with cross-entropy loss achieve strong predictive accuracy but lack the capability to provide inherent predictive uncertainty estimates, thus requiring external techniques to obtain these estimates. In…

机器学习 · 统计学 2026-04-08 Courtney Franzen , Farhad Pourkamali-Anaraki

Electronic health records (EHRs) are invaluable for clinical research, yet privacy concerns severely restrict data sharing. Synthetic data generation offers a promising solution, but EHRs present unique challenges: they contain both…

机器学习 · 计算机科学 2026-03-26 Shaonan Liu , Yuichiro Iwashita , Soichiro Nakako , Masakazu Iwamura , Koichi Kise

We present BEEP (Biomedical Evidence-Enhanced Predictions), a novel approach for clinical outcome prediction that retrieves patient-specific medical literature and incorporates it into predictive models. Based on each individual patient's…

计算与语言 · 计算机科学 2022-11-17 Aakanksha Naik , Sravanthi Parasa , Sergey Feldman , Lucy Lu Wang , Tom Hope

This study develops a cloud-based deep learning system for early prediction of diabetes, leveraging the distributed computing capabilities of the AWS cloud platform and deep learning technologies to achieve efficient and accurate risk…

分布式、并行与集群计算 · 计算机科学 2025-01-07 Yang Zhang , Fa Wang , Xin Huang , Xintao Li , Sibei Liu , Hansong Zhang

Time series data may exhibit clustering over time and, in a multiple time series context, the clustering behavior may differ across the series. This paper is motivated by the Bayesian non--parametric modeling of the dependence between the…

统计理论 · 数学 2011-09-23 Federico Bassetti , Roberto Casarin , Fabrizio Leisen