English
Related papers

Related papers: A Bayesian Nonparametric Approach for Identifying …

200 papers

Missing data represents a fundamental challenge in machine learning applications, often reducing model performance and reliability. This problem is particularly acute in fields like bioinformatics and clinical machine learning, where…

Machine Learning · Computer Science 2025-09-04 Fatemeh Azad , Zoran Bosnić , Matjaž Kukar

Background: Single-cell RNA sequencing (scRNA-seq) is a powerful profiling technique at the single-cell resolution. Appropriate analysis of scRNA-seq data can characterize molecular heterogeneity and shed light into the underlying cellular…

Applications · Statistics 2019-08-05 Siamak Zamani Dadaneh , Paul de Figueiredo , Sing-Hoi Sze , Mingyuan Zhou , Xiaoning Qian

The human microbiome is a complex ecological system, and describing its structure and function under different environmental conditions is important from both basic scientific and medical perspectives. Viewed through a biostatistical lens,…

Applications · Statistics 2017-11-17 Kris Sankaran , Susan P. Holmes

In recent years microbiome studies have become increasingly prevalent and large-scale. Through high-throughput sequencing technologies and well-established analytical pipelines, relative abundance data of operational taxonomic units and…

Methodology · Statistics 2022-05-12 Yan Li , Gen Li , Kun Chen

We wish to formally test for changes in the taxonomic diversity of a community, especially in the presence of high latent diversity. Drawing on the meta-analysis literature, we construct a model for diversity that accounts for covariate…

Methodology · Statistics 2015-06-19 Amy Willis , John Bunge , Thea Whitman

Statistical inference on biodiversity has a rich history going back to RA Fisher. An influential ecological theory suggests the existence of a fundamental biodiversity number, denoted $\alpha$, which coincides with the precision parameter…

Methodology · Statistics 2025-02-04 Tommaso Rigon , Ching-Lung Hsu , David B. Dunson

Differential abundance (DA) analysis in microbiome studies has recently been used to uncover a plethora of associations between microbial composition and various health conditions. While current approaches to DA typically apply only to…

Methodology · Statistics 2026-01-19 Simon Fontaine , Nisha J. D'Silva , Marcell Costa de Medeiros , Grace Y. Chen , Ji Zhu , Gen Li

Missing covariates are not uncommon in capture-recapture studies. When covariate information is missing at random in capture-recapture data, an empirical full likelihood method has been demonstrated to outperform…

Methodology · Statistics 2025-07-15 Yang Liu , Yukun Liu , Pengfei Li , Riquan Zhang

We consider the problem of inferring the values of an arbitrary set of variables (e.g., risk of diseases) given other observed variables (e.g., symptoms and diagnosed diseases) and high-dimensional signals (e.g., MRI images or EEG). This is…

Machine Learning · Statistics 2019-02-07 Hao Wang , Chengzhi Mao , Hao He , Mingmin Zhao , Tommi S. Jaakkola , Dina Katabi

In many life science experiments or medical studies, subjects are repeatedly observed and measurements are collected in factorial designs with multivariate data. The analysis of such multivariate data is typically based on multivariate…

Methodology · Statistics 2023-05-24 Lubna Amro , Frank Konietschke , Markus Pauly

Clustering mixed-type data remains a major challenge in biomedical research to uncover clinically meaningful subgroups within heterogeneous patient populations. Most existing clustering methods impose restrictive assumptions like local…

Applications · Statistics 2026-04-23 Yueting Wang , Shu Wang , Jonathan G. Yabes , Chung-Chou H. Chang

Gene-gene and gene-environment interactions are widely believed to play significant roles in explaining the variability of complex traits. While substantial research exists in this area, a comprehensive statistical framework that addresses…

Methodology · Statistics 2026-02-18 Durba Bhattacharya , Sourabh Bhattacharya

Covariate-dependent graph learning has gained increasing interest in the graphical modeling literature for the analysis of heterogeneous data. This task, however, poses challenges to modeling, computational efficiency, and interpretability.…

Methodology · Statistics 2025-04-04 Zijian Zeng , Meng Li , Marina Vannucci

Background: Parkinson's disease remains a major neurodegenerative disorder with high misdiagnosis rates, primarily due to reliance on clinical rating scales. Recent studies have demonstrated a strong association between gut microbiota and…

Artificial Intelligence · Computer Science 2025-09-10 Bo Yu , Zhixiu Hua , Bo Zhao

A key challenge in building effective regression models for large and diverse populations is accounting for patient heterogeneity. An example of such heterogeneity is in health system risk modeling efforts where different combinations of…

Methodology · Statistics 2022-12-26 Jared D. Huling , Menggang Yu

Comparative meta-analyses of groups of subjects by integrating multiple observational studies rely on estimated propensity scores (PSs) to mitigate covariate imbalances. However, PS estimation grapples with the theoretical and practical…

Methodology · Statistics 2024-05-09 Subharup Guha , Yi Li

Advances in cellular imaging technologies, especially those based on fluorescence in situ hybridization (FISH) now allow detailed visualization of the spatial organization of human or bacterial cells. Quantifying this spatial organization…

Differential analysis is a routine procedure in the statistical analysis toolbox across many applied fields, including quantitative proteomics, the main illustration of the present paper. The state-of-the-art limma approach uses a…

Methodology · Statistics 2025-12-12 Marie Chion , Arthur Leroy

This paper develops a Bayesian graphical model for fusing disparate types of count data. The motivating application is the study of bacterial communities from diverse high dimensional features, in this case transcripts, collected from…

Graph-based machine learning methods are useful tools in the identification and prediction of variation in genetic data. In particular, the comprehension of phenotypic effects at the cellular level is an accelerating research area in…

Quantitative Methods · Quantitative Biology 2024-12-06 Nandini Gadhia , Michalis Smyrnakis , Po-Yu Liu , Damer Blake , Melanie Hay , Anh Nguyen , Dominic Richards , Dong Xia , Ritesh Krishna