English
Related papers

Related papers: Bayesian Copula Density Deconvolution for Zero-Inf…

200 papers

This paper presents an improved implicit sampling method for hierarchical Bayesian inverse problems. A widely used approach for sampling posterior distribution is based on Markov chain Monte Carlo (MCMC). However, the samples generated by…

Numerical Analysis · Mathematics 2018-11-27 Xiaoyan Song , Lijian Jiang , Guanghui Zheng

We propose reinterpreting copula density estimation as a discriminative task. Under this novel estimation scheme, we train a classifier to distinguish samples from the joint density from those of the product of independent marginals,…

Methodology · Statistics 2025-03-20 David Huk , Mark Steel , Ritabrata Dutta

Wearable devices collect time-varying biobehavioral data, offering opportunities to investigate how behaviors influence health outcomes. However, these data often contain measurement error and excess zeros (due to nonwear, sedentary…

Methodology · Statistics 2026-02-06 Caihong Qin , Lan Xue , Ufuk Beyaztas , Roger S. Zoh , Mark Benden , Jeff Goldsmith , Carmen D. Tekwe

Covariate measurement error in nonparametric regression is a common problem in nutritional epidemiology and geostatistics, and other fields. Over the last two decades, this problem has received substantial attention in the frequentist…

Statistics Theory · Mathematics 2023-01-27 Shuang Zhou , Debdeep Pati , Tianying Wang , Yun Yang , Raymond J. Carroll

When surveillance data of infectious disease incidence (e.g. weekly case counts) are disaggregated by demographic indicators, disparities in long-run health outcomes between these groups become apparent. Accurate identification of high-risk…

Methodology · Statistics 2026-05-29 Miles Moran , Rob Trangucci , Lisa Madsen

Dietary patterns synthesize multiple related diet components, which can be used by nutrition researchers to examine diet-disease relationships. Latent class models (LCMs) have been used to derive dietary patterns from dietary intake…

Methodology · Statistics 2025-04-11 Mengbing Li , Briana Stephenson , Zhenke Wu

Understanding the links between diet, metabolic changes, and health outcomes is a key focus in nutritional science and broader biological research. Analyzing relationships, such as those between ultra-processed food (UPF) intake and…

Methodology · Statistics 2026-05-19 Sang Kyu Lee , Erikka Loftfield , Hyokyoung G. Hong , Haolei Weng

Air pollution is a serious issue that currently affects many industrial cities in the world and can cause severe illness to the population. In particular, it has been proven that extreme high levels of airborne contaminants have dangerous…

Applications · Statistics 2019-11-12 Alexander Kreuzer , Luciana Dalla Valle , Claudia Czado

Density level sets can be estimated using plug-in methods, excess mass algorithms or a hybrid of the two previous methodologies. The plug-in algorithms are based on replacing the unknown density by some nonparametric estimator, usually the…

Statistics Theory · Mathematics 2016-11-26 A. Rodríguez-Casal , P. Saavedra-Nieves

The infant microbiome undergoes rapid changes in composition over time and is associated with long-term risks of conditions such as immune strength, allergy, asthma, and other health outcomes. Modeling the associations between exposures or…

Methodology · Statistics 2026-03-31 Brody Erlandson , Ander Wilson , Matthew D. Koslovsky

Products manufactured from the same batch or utilized in the same region often exhibit correlated lifetime observations due to the latent heterogeneity caused by the influence of shared but unobserved covariates. The unavailable…

Methodology · Statistics 2021-07-15 Xuxue Sun , Mingyang Li

Copula models are flexible tools to represent complex structures of dependence for multivariate random variables. According to Sklar's theorem (Sklar, 1959), any d-dimensional absolutely continuous density can be uniquely represented as the…

Methodology · Statistics 2021-03-05 Clara Grazian , Luciana Dalla Valle , Brunero Liseo

Replicated weighted networks often exhibit many structural zeros alongside heterogeneous non-zero edge strengths. In structural connectomics, this zero-inflation coincides with subjects expressing overlapping, rather than discrete,…

Methodology · Statistics 2026-05-14 Hsin-Hsiung Huang , Yuh-Haur Chen , Teng Zhang

Missing data imputation forms the first critical step of many data analysis pipelines. The challenge is greatest for mixed data sets, including real, Boolean, and ordinal data, where standard techniques for imputation fail basic sanity…

Methodology · Statistics 2020-06-17 Yuxuan Zhao , Madeleine Udell

In many applications, survey data are collected from different survey centers in different regions. It happens that in some circumstances, response variables are completely observed while the covariates have missing values. In this paper,…

Methodology · Statistics 2020-07-07 Zhihua Ma , Guanyu Hu , Ming-Hui Chen

Bayesian nonparametric (BNP) models provide elegant methods for discovering underlying latent features within a data set, but inference in such models can be slow. We exploit the fact that completely random measures, which commonly used…

Machine Learning · Statistics 2020-07-17 Avinava Dubey , Michael Minyi Zhang , Eric P. Xing , Sinead A. Williamson

Unsupervised estimation of the dimensionality of hyperspectral microspectroscopy datasets containing pure and mixed spectral features, and extraction of their representative endmember spectra, remains a challenge in biochemical data mining.…

Copulas, generalized estimating equations, and generalized linear mixed models promote the analysis of grouped data where non-normal responses are correlated. Unfortunately, parameter estimation remains challenging in these three…

Methodology · Statistics 2024-10-16 Sarah S. Ji , Benjamin B. Chu , Hua Zhou , Kenneth Lange

Joint modelling of longitudinal and time-to-event data is usually described by a joint model which uses shared or correlated latent effects to capture associations between the two processes. Under this framework, the joint distribution of…

Methodology · Statistics 2022-03-07 Zili Zhang , Christiana Charalambous , Peter Foster

Multilevel compositional data are data that are repeatedly measured or clustered within groups and are non-negative and sum to a constant value. These data arise in various settings, such as intensive, longitudinal studies using ecological…

Methodology · Statistics 2025-02-21 Flora Le , Tyman E. Stanford , Dorothea Dumuid , Joshua F. Wiley