中文
相关论文

相关论文: Semiparametric count data regression for self-repo…

200 篇论文

Missing values in datasets are common in applied statistics. For regression problems, theoretical work thus far has largely considered the issue of missing covariates as distinct from missing responses. However, in practice, many datasets…

统计理论 · 数学 2026-02-17 Benedict M. Risebrow , Thomas B. Berrett

Recent advances in remote health monitoring systems have significantly benefited patients and played a crucial role in improving their quality of life. However, while physiological health-focused solutions have demonstrated increasing…

Count data frequently arises in biomedical applications, such as the length of hospital stay. However, their discrete nature poses significant challenges for appropriately modeling conditional quantiles, which are crucial for understanding…

统计方法学 · 统计学 2025-07-28 Yuta Yamauchi , Genya Kobayashi , Shonosuke Sugasawa

Prognostics is concerned with predicting the future health of the equipment and any potential failures. With the advances in the Internet of Things (IoT), data-driven approaches for prognostics that leverage the power of machine learning…

机器学习 · 计算机科学 2020-06-09 Qiyao Wang , Ahmed Farahat , Chetan Gupta , Haiyan Wang

Cohort studies of the onset of a disease often encounter left-truncation on the event time of interest in addition to right-censoring due to variable enrollment times of study participants. Analysis of such event time data can be biased if…

统计方法学 · 统计学 2025-04-11 Spencer Matthews , Bin Nan

Across a variety of scientific disciplines, sparse inverse covariance estimation is a popular tool for capturing the underlying dependency relationships in multivariate data. Unfortunately, most estimators are not scalable enough to handle…

Longitudinal electronic health record (EHR) data offer opportunities to study biomarker trajectories; however, association estimates-the primary inferential target-from standard models designed for regular observation times may be biased by…

统计方法学 · 统计学 2026-02-18 Cheng-Han Yang , Xu Shi , Bhramar Mukherjee

Deep learning has revolutionized medical imaging, but its effectiveness is severely limited by insufficient labeled training data. This paper introduces a novel GAN-based semi-supervised learning framework specifically designed for low…

计算机视觉与模式识别 · 计算机科学 2025-08-11 Guido Manni , Clemente Lauretti , Loredana Zollo , Paolo Soda

A plethora of machine learning methods have been applied to imaging data, enabling the construction of clinically relevant imaging signatures of neurological and neuropsychiatric diseases. Oftentimes, such methods don't explicitly model the…

机器学习 · 计算机科学 2022-05-11 Zhijian Yang , Junhao Wen , Christos Davatzikos

This paper focuses on the estimation of distributional treatment effects in randomized experiments that use covariate-adaptive randomization (CAR). These include designs such as Efron's biased-coin design and stratified block randomization,…

计量经济学 · 经济学 2025-06-09 Undral Byambadalai , Tomu Hirata , Tatsushi Oka , Shota Yasui

Missing data occur frequently in empirical studies in health and social sciences, often compromising our ability to make accurate inferences. An outcome is said to be missing not at random (MNAR) if, conditional on the observed variables,…

统计方法学 · 统计学 2019-01-23 BaoLuo Sun , Lan Liu , Wang Miao , Kathleen Wirth , James Robins , Eric Tchetgen Tchetgen

Sequencing-based technologies provide an abundance of high-dimensional biological datasets with skewed and zero-inflated measurements. Classification of such data with linear discriminant analysis leads to poor performance due to the…

统计方法学 · 统计学 2022-08-09 Hee Cheol Chung , Yang Ni , Irina Gaynanova

Panel count data describes aggregated counts of recurrent events observed at discrete time points. To understand dynamics of health behaviors, the field of quantitative behavioral research has evolved to increasingly rely upon panel count…

In many biomedical problems, data are often heterogeneous, with samples spanning multiple patient subgroups, where different subgroups may have different disease subtypes, stages, or other medical contexts. These subgroups may be related,…

统计方法学 · 统计学 2022-11-30 Zihan Li , Ziye Luo , Yifan Sun

A semiparametric copula-based two-part quantile regression framework is developed for the analysis of semicontinuous outcomes characterized by a point mass at zero and a continuous positive component. The proposed approach models the…

统计方法学 · 统计学 2026-03-17 Guanjie Lyu , Mohamed Belalia , Abdulkadir Hussein

Data-driven methods for mental health treatment and surveillance have become a major focus in computational science research in the last decade. However, progress in the domain, in terms of both medical understanding and system performance,…

计算与语言 · 计算机科学 2021-04-27 Keith Harrigian , Carlos Aguirre , Mark Dredze

Panel count data arise in clinical trials when patients are asked to report their occurrences of events of interest periodically but the exact event times are unknown, only the count of events between two successive examinations are…

统计方法学 · 统计学 2025-05-29 Jiangjie Zhou , Baosheng Liang

Surveys often ask respondents to report nonnegative counts, but respondents may misremember or round to a nearby multiple of 5 or 10. This phenomenon is called heaping, and the error inherent in heaped self-reported numbers can bias…

应用统计 · 统计学 2015-09-15 Forrest W. Crawford , Robert E. Weiss , Marc A. Suchard

Multistate process data are common in studies of chronic diseases such as cancer. These data are ideal for precision medicine purposes as they can be leveraged to improve more refined health outcomes, compared to standard survival outcomes,…

统计方法学 · 统计学 2022-11-28 Giorgos Bakoyannis

Generating step-by-step "chain-of-thought" rationales improves language model performance on complex reasoning tasks like mathematics or commonsense question-answering. However, inducing language model rationale generation currently…

机器学习 · 计算机科学 2022-05-23 Eric Zelikman , Yuhuai Wu , Jesse Mu , Noah D. Goodman