中文
相关论文

相关论文: A Bayesian Feature Allocation Model for Identifica…

200 篇论文

Partially recorded data are frequently encountered in many applications and usually clustered by first removing incomplete cases or features with missing values, or by imputing missing values, followed by application of a clustering…

统计方法学 · 统计学 2021-10-20 Emily M. Goren , Ranjan Maitra

The significant morphological and distributional variability among subcellular components poses a long-standing challenge for learning-based organelle segmentation models, significantly increasing the risk of biased feature learning.…

计算机视觉与模式识别 · 计算机科学 2025-07-24 Bo Fang , Jianan Fan , Dongnan Liu , Hang Chang , Gerald J. Shami , Filip Braet , Weidong Cai

Label-free metabolic dynamics contrast is highly appealing but difficult to achieve in biomedical imaging. Interference offers a highly sensitive mechanism for capturing the metabolic dynamics of the subcellular scatterers. However,…

One of the focal points of the modern literature on Bayesian nonparametrics has been the problem of clustering, or partitioning, where each data point is modeled as being associated with one and only one of some collection of groups called…

统计理论 · 数学 2013-10-02 Tamara Broderick , Michael I. Jordan , Jim Pitman

Surface parameterization plays an essential role in numerous computer graphics and geometry processing applications. Traditional parameterization approaches are designed for high-quality meshes laboriously created by specialized 3D…

计算机视觉与模式识别 · 计算机科学 2024-10-29 Qijian Zhang , Junhui Hou , Wenping Wang , Ying He

In systems biomedicine, an experimenter encounters different potential sources of variation in data such as individual samples, multiple experimental conditions, and multi-variable network-level responses. In multiparametric cytometry,…

The development of high throughput single-cell sequencing technologies now allows the investigation of the population level diversity of cellular transcriptomes. This diversity has shown two faces. First, the expression dynamics (gene to…

统计方法学 · 统计学 2021-04-10 Ghislain Durif , Laurent Modolo , Jeff E. Mold , Sophie Lambert-Lacroix , Franck Picard

We believe that a wide range of physical processes conspire to shape the observed galaxy population but we remain unsure of their detailed interactions. The semi-analytic model (SAM) of galaxy formation uses multi-dimensional…

宇宙学与河外天体物理 · 物理学 2011-11-07 Yu Lu , H. J. Mo , Martin D. Weinberg , Neal Katz

Finite mixture model is an important branch of clustering methods and can be applied on data sets with mixed types of variables. However, challenges exist in its applications. First, it typically relies on the EM algorithm which could be…

机器学习 · 统计学 2019-05-10 Shu Wang , Jonathan G. Yabes , Chung-Chou H. Chang

Factorial hidden Markov models (FHMMs) are powerful tools of modeling sequential data. Learning FHMMs yields a challenging simultaneous model selection issue, i.e., selecting the number of multiple Markov chains and the dimensionality of…

机器学习 · 统计学 2015-06-29 Shaohua Li , Ryohei Fujimaki , Chunyan Miao

We propose a homotopy sampling procedure, loosely based on importance sampling. Starting from a known probability distribution, the homotopy procedure generates the unknown normalization of a target distribution. In the context of…

统计计算 · 统计学 2021-05-05 Juan M. Restrepo , Jorge M. Ramirez

In many fields, researchers are interested in large and complex biological processes. Two important examples are gene expression and DNA methylation in genetics. One key problem is to identify aberrant patterns of these processes and…

应用统计 · 统计学 2012-10-03 Matthias Kormaksson , James G. Booth , Maria E. Figueroa , Ari Melnick

Researchers are often interested in predicting outcomes, conducting clustering analysis to detect distinct subgroups of their data, or computing causal treatment effects. Pathological data distributions that exhibit skewness and…

统计方法学 · 统计学 2020-08-24 Arman Oganisian , Nandita Mitra , Jason Roy

For the past few years, deep generative models have increasingly been used in biological research for a variety of tasks. Recently, they have proven to be valuable for uncovering subtle cell phenotypic differences that are not directly…

图像与视频处理 · 电气工程与系统科学 2026-01-28 Anis Bourou , Thomas Boyer , Kévin Daupin , Véronique Dubreuil , Aurélie De Thonel , Valérie Mezger , Auguste Genovesio

Pathomics is a recent approach that offers rich quantitative features beyond what black-box deep learning can provide, supporting more reproducible and explainable biomarkers in digital pathology. However, many derived features (e.g.,…

计算机视觉与模式识别 · 计算机科学 2026-02-06 Yuechen Yang , Junlin Guo , Ruining Deng , Junchao Zhu , Zhengyi Lu , Chongyu Qu , Yanfan Zhu , Xingyi Guo , Yu Wang , Shilin Zhao , Haichun Yang , Yuankai Huo

This paper presents a new statistical method for clustering step data, a popular form of health record data easily obtained from wearable devices. Since step data are high-dimensional and zero-inflated, classical methods such as K-means and…

统计方法学 · 统计学 2020-10-16 Wookyeong Song , Hee-Seok Oh , Yaeji Lim , Ying Kuen Cheung

Typical state of the art flow cytometry data samples consists of measures of more than 100.000 cells in 10 or more features. AI systems are able to diagnose such data with almost the same accuracy as human experts. However, there is one…

To quantify how well theoretical predictions of structural ensembles agree with experimental measurements, we depend on the accuracy of forward models. These models are computational frameworks that generate observable quantities from…

生物物理 · 物理学 2025-11-04 Robert M. Raddi , Tim Marshall , Vincent A. Voelz

Cluster analysis of biological samples using gene expression measurements is a common task which aids the discovery of heterogeneous biological sub-populations having distinct mRNA profiles. Several model-based clustering algorithms have…

统计方法学 · 统计学 2012-01-30 Alberto Cozzini , Ajay Jasra , Giovanni Montana

Real-world applications may be affected by outlying values. In the model-based clustering literature, several methodologies have been proposed to detect units that deviate from the majority of the data (rowwise outliers) and trim them from…