中文
相关论文

相关论文: Bayesian Hidden Markov Tree Models for Clustering …

200 篇论文

Clustering is a popular data mining technique that aims to partition an input space into multiple homogeneous regions. There exist several clustering algorithms in the literature. The performance of a clustering algorithm depends on its…

人机交互 · 计算机科学 2020-08-20 Sudip Poddar , Anirban Mukhopadhyay

When statistical analyses consider multiple data sources, Markov melding provides a method for combining the source-specific Bayesian models. Markov melding joins together submodels that have a common quantity. One challenge is that the…

统计方法学 · 统计学 2022-03-17 Andrew A. Manderson , Robert J. B. Goudie

In recent work, robust mixture modelling approaches using skewed distributions have been explored to accommodate asymmetric data. We introduce parsimony by developing skew-t and skew-normal analogues of the popular GPCM family that employ…

统计方法学 · 统计学 2013-11-12 Irene Vrbik , Paul D. McNicholas

Binned data often appears in different fields of research, and it is generated after summarizing the original data in a sequence of pairs of bins (or their midpoints) and frequencies. There may exist different reasons to only provide this…

统计方法学 · 统计学 2024-09-13 Asael Fabian Martínez , Carlos Díaz-Avalos

Genetic sequence data are well described by hidden Markov models (HMMs) in which latent states correspond to clusters of similar mutation patterns. Theory from statistical genetics suggests that these HMMs are nonhomogeneous (their…

应用统计 · 统计学 2016-11-03 Lloyd T. Elliott , Yee Whye Teh

Various and ubiquitous information systems are being used in monitoring, exchanging, and collecting information. These systems are generating massive amount of event sequence logs that may help us understand underlying phenomenon. By…

机器学习 · 统计学 2018-07-13 Yihuang Kang , Vladimir Zadorozhny

The recent advances in single-cell technologies have enabled us to profile genomic features at unprecedented resolution and datasets from multiple domains are available, including datasets that profile different types of genomic features…

机器学习 · 统计学 2020-06-09 Pengcheng Zeng , Zhixiang Lin

We develop a latent variable model and an efficient spectral algorithm motivated by the recent emergence of very large data sets of chromatin marks from multiple human cell types. A natural model for chromatin data in one cell type is a…

机器学习 · 统计学 2015-06-09 Chicheng Zhang , Jimin Song , Kevin C Chen , Kamalika Chaudhuri

Evolutionary accumulation models (EvAMs) are an emerging class of machine learning methods designed to infer the evolutionary pathways by which features are acquired. Applications include cancer evolution (accumulation of mutations),…

种群与进化 · 定量生物学 2026-03-13 Iain G. Johnston

Classically, Bayesian clustering interprets each component of a mixture model as a cluster. The inferred clustering posterior is highly sensitive to any inaccuracies in the kernel within each component. As this kernel is made more flexible,…

统计方法学 · 统计学 2025-12-12 David Buch , Miheer Dewaskar , David B. Dunson

A model based clustering procedure for data of mixed type, clustMD, is developed using a latent variable model. It is proposed that a latent variable, following a mixture of Gaussian distributions, generates the observed data of mixed type.…

统计方法学 · 统计学 2015-11-06 Damien McParland , Isobel Claire Gormley

The modeling of time series is becoming increasingly critical in a wide variety of applications. Overall, data evolves by following different patterns, which are generally caused by different user behaviors. Given a time series, we define…

机器学习 · 计算机科学 2022-07-13 Wenjie Hu , Jianping Huang , Liang Wu , Yang Yang , Zongtao Liu , Zhanlin Sun , Bingshen Yao , Ke Chen

An evolutionary algorithm (EA) is developed as an alternative to the EM algorithm for parameter estimation in model-based clustering. This EA facilitates a different search of the fitness landscape, i.e., the likelihood surface, utilizing…

统计计算 · 统计学 2020-06-09 Sharon M. McNicholas , Paul D. McNicholas , Daniel A. Ashlock

The ongoing explosion of genome sequence data is transforming how we reconstruct and understand the histories of biological systems. Across biological scales, from individual cells to populations and species, trees-based models provide a…

种群与进化 · 定量生物学 2025-12-08 Yun Deng , Shing H. Zhan , Yulin Zhang , Chao Zhang , Bingjie Chen

A phylogenetic tree is an important way in Bioinformatics to find the evolutionary relationship among biological species. In this research, a proposed model is described for the estimation of a phylogenetic tree for a given set of data. To…

种群与进化 · 定量生物学 2025-09-03 S M Rafiuddin

Bayesian inference for factorial hidden Markov models is challenging due to the exponentially sized latent variable space. Standard Monte Carlo samplers can have difficulties effectively exploring the posterior landscape and are often…

统计计算 · 统计学 2019-02-28 Kaspar Märtens , Michalis K Titsias , Christopher Yau

Clustering under pairwise constraints is an important knowledge discovery tool that enables the learning of appropriate kernels or distance metrics to improve clustering performance. These pairwise constraints, which come in the form of…

机器学习 · 计算机科学 2022-03-24 Benedikt Boecking , Vincent Jeanselme , Artur Dubrawski

Researchers are often interested in predicting outcomes, conducting clustering analysis to detect distinct subgroups of their data, or computing causal treatment effects. Pathological data distributions that exhibit skewness and…

统计方法学 · 统计学 2020-08-24 Arman Oganisian , Nandita Mitra , Jason Roy

We consider an extension of model-based clustering to the semi-supervised case, where some of the data are pre-labeled. We provide a derivation of the Bayesian Information Criterion (BIC) approximation to the Bayes factor in this setting.…

统计方法学 · 统计学 2016-04-28 Jordan Yoder , Carey E. Priebe

Bayesian models have become very popular over the last years in several fields such as signal processing, statistics, and machine learning. Bayesian inference requires the approximation of complicated integrals involving posterior…

统计计算 · 统计学 2021-07-20 Luca Martino , Víctor Elvira
‹ 上一页 1 8 9 10 下一页 ›