中文
相关论文

相关论文: Bayesian mixture of Plackett-Luce models for parti…

200 篇论文

Given the joint chances of a pair of random variables one can compute quantities of interest, like the mutual information. The Bayesian treatment of unknown chances involves computing, from a second order prior distribution and the data…

机器学习 · 计算机科学 2007-05-23 Marcus Hutter , Marco Zaffalon

Mixtures of product distributions are a powerful device for learning about heterogeneity within data populations. In this class of latent structure models, de Finetti's mixing measure plays the central role for describing the uncertainty…

统计理论 · 数学 2021-09-27 Yun Wei , XuanLong Nguyen

In computational biology, gene expression datasets are characterized by very few individual samples compared to a large number of measurements per sample. Thus, it is appealing to merge these datasets in order to increase the number of…

统计方法学 · 统计学 2011-08-18 Meili Baragatti

Finite mixture models are a useful statistical model class for clustering and density approximation. In the Bayesian framework finite mixture models require the specification of suitable priors in addition to the data model. These priors…

统计方法学 · 统计学 2024-07-09 Bettina Grün , Gertraud Malsiner-Walli

Finite mixture distributions arise in sampling a heterogeneous population. Data drawn from such a population will exhibit extra variability relative to any single subpopulation. Statistical models based on finite mixtures can assist in the…

统计方法学 · 统计学 2024-01-19 Andrew M. Raim , Nagaraj K. Neerchal , Jorge G. Morel

Machine learning models offer the potential to understand diverse datasets in a data-driven way, powering insights into individual disease experiences and ensuring equitable healthcare. In this study, we explore Bayesian inference for…

机器学习 · 计算机科学 2023-11-23 Beatrice Taylor , Cameron Shand , Chris J. D. Hardy , Neil Oxtoby

Discrete data such as counts of microbiome taxa resulting from next-generation sequencing are routinely encountered in bioinformatics. Taxa count data in microbiome studies are typically high-dimensional, over-dispersed, and can only reveal…

统计方法学 · 统计学 2022-06-23 Yuan Fang , Sanjeena Subedi

Rank data arises frequently in marketing, finance, organizational behavior, and psychology. Most analysis of rank data reported in the literature assumes the presence of one or more variables (sometimes latent) based on whose values the…

统计方法学 · 统计学 2017-09-08 Arnab Kumar Laha , Somak Dutta , Vivekananda Roy

Bayesian methods are often optimal, yet increasing pressure for fast computations, especially with streaming data, brings renewed interest in faster, possibly sub-optimal, solutions. The extent to which these algorithms approximate Bayesian…

统计理论 · 数学 2026-02-18 Sandra Fortini , Sonia Petrone

We present a Dirichlet process mixture model over discrete incomplete rankings and study two Gibbs sampling inference techniques for estimating posterior clusterings. The first approach uses a slice sampling subcomponent for estimating…

机器学习 · 计算机科学 2012-03-19 Marina Meila , Harr Chen

Maximum entropy (MAXENT) method has a large number of applications in theoretical and applied machine learning, since it provides a convenient non-parametric tool for estimating unknown probabilities. The method is a major contribution of…

数据分析、统计与概率 · 物理学 2020-12-18 A. E. Allahverdyan , N. H. Martirosyan

Robust statistical data modelling under potential model mis-specification often requires leaving the parametric world for the nonparametric. In the latter, parameters are infinite dimensional objects such as functions, probability…

We propose a Bayesian test of normality for univariate or multivariate data against alternative nonparametric models characterized by Dirichlet process mixture distributions. The alternative models are based on the principles of embedding…

统计理论 · 数学 2023-04-12 Surya T. Tokdar , Ryan Martin

Many latent (factorized) models have been proposed for recommendation tasks like collaborative filtering and for ranking tasks like document or image retrieval and annotation. Common to all those methods is that during inference the items…

机器学习 · 计算机科学 2012-10-19 Jason Weston , John Blitzer

We propose a tractable semiparametric estimation method for structural dynamic discrete choice models. The distribution of additive utility shocks in the proposed framework is modeled by location-scale mixtures of extreme value…

计量经济学 · 经济学 2023-08-15 Andriy Norets , Kenichi Shimizu

Learning the structure of Bayesian networks from data provides insights into underlying processes and the causal relationships that generate the data, but its usefulness depends on the homogeneity of the data population, a condition often…

Mixture model-based clustering has become an increasingly popular data analysis technique since its introduction over fifty years ago, and is now commonly utilized within a family setting. Families of mixture models arise when the component…

统计方法学 · 统计学 2019-11-11 Sanjeena Subedi , Paul D. McNicholas

Microbiome research has immense potential for unlocking insights into human health and disease. A common goal in human microbiome research is identifying subgroups of individuals with similar microbial composition that may be linked to…

统计方法学 · 统计学 2025-08-21 Suppapat Korsurat , Matthew D. Koslovsky

Estimating the model evidence - or mariginal likelihood of the data - is a notoriously difficult task for finite and infinite mixture models and we reexamine here different Monte Carlo techniques advocated in the recent literature, as well…

统计计算 · 统计学 2022-05-12 Adrien Hairault , Christian P. Robert , Judith Rousseau

In this paper we address the identifiability and efficient learning problems of finite mixtures of Plackett-Luce models for rank data. We prove that for any $k\geq 2$, the mixture of $k$ Plackett-Luce models for no more than $2k-1$…

机器学习 · 计算机科学 2020-03-10 Zhibing Zhao , Peter Piech , Lirong Xia