English
Related papers

Related papers: Gibbs-type Indian buffet processes

200 papers

Double-descent refers to the unexpected drop in test loss of a learning algorithm beyond an interpolating threshold with over-parameterization, which is not predicted by information criteria in their classical forms due to the limitations…

Machine Learning · Computer Science 2023-11-15 Haobo Chen , Yuheng Bu , Gregory W. Wornell

Real-world problems, often couched as machine learning applications, involve quantities of interest that have real-world meaning, independent of any statistical model. To avoid potential model misspecification bias or over-complicating the…

Methodology · Statistics 2022-05-10 Ryan Martin , Nicholas Syring

We develop methods for efficient amortized approximate Bayesian inference over posterior distributions of probabilistic clustering models, such as Dirichlet process mixture models. The approach is based on mapping distributed,…

Machine Learning · Statistics 2018-11-27 Ari Pakman , Liam Paninski

The beta-negative binomial process (BNBP), an integer-valued stochastic process, is employed to partition a count vector into a latent random count matrix. As the marginal probability distribution of the BNBP that governs the exchangeable…

Methodology · Statistics 2015-01-05 Mingyuan Zhou

In social science research, understanding latent structures in populations through survey data with categorical responses is a common and important task. Traditional methods like Factor Analysis and Latent Class Analysis have limitations,…

Methodology · Statistics 2024-12-30 Chayut Wongkamthong

We study computational aspects of repulsive Gibbs point processes, which are probabilistic models of interacting particles in a finite-volume region of space. We introduce an approach for reducing a Gibbs point process to the hard-core…

Data Structures and Algorithms · Computer Science 2023-12-15 Tobias Friedrich , Andreas Göbel , Maximilian Katzmann , Martin Krejca , Marcus Pappik

In this paper, we present a novel approach to fitting mixture models based on estimating first the posterior distribution of the auxiliary variables that assign each observation to a group in the mixture. The posterior distributions of the…

Computation · Statistics 2017-12-29 Virgilio Gomez-Rubio

Accurately credit default prediction faces challenges due to imbalanced data and low correlation between features and labels. Existing default prediction studies on the basis of gradient boosting decision trees (GBDT), deep learning…

Computational Engineering, Finance, and Science · Computer Science 2023-12-06 Yandan Tan , Hongbin Zhu , JieWu , Hongfeng Chai

Nonparametric Bayesian models are often based on the assumption that the objects being modeled are exchangeable. While appropriate in some applications (e.g., bag-of-words models for documents), exchangeability is sometimes assumed simply…

Machine Learning · Computer Science 2012-06-18 Kurt T. Miller , Thomas Griffiths , Michael I. Jordan

We introduce Interleaved Gibbs Diffusion (IGD), a novel generative modeling framework for discrete-continuous data, focusing on problems with important, implicit and unspecified constraints in the data. Most prior works on discrete and…

Machine Learning · Computer Science 2025-07-04 Gautham Govind Anil , Sachin Yadav , Dheeraj Nagaraj , Karthikeyan Shanmugam , Prateek Jain

The problem of inferring a clustering of a data set has been the subject of much research in Bayesian analysis, and there currently exists a solid mathematical foundation for Bayesian approaches to clustering. In particular, the class of…

Probability · Mathematics 2013-01-30 Tamara Broderick , Jim Pitman , Michael I. Jordan

A novel unsupervised learning method is proposed in this paper for biclustering large-dimensional matrix-valued time series based on an entirely new latent two-way factor structure. Each block cluster is characterized by its own row and…

Methodology · Statistics 2025-02-11 Yong He , Xiaoyang Ma , Xingheng Wang , Yalin Wang

Models based on preferential attachment have had much success in reproducing the power law degree distributions which seem ubiquitous in both natural and engineered systems. Here, rather than assuming preferential attachment, we give an…

Statistical Mechanics · Physics 2007-05-23 N. Berger , C. Borgs , J. T. Chayes , R. M. D'Souza , R. D. Kleinberg

Current deep learning classifiers, carry out supervised learning and store class discriminatory information in a set of shared network weights. These weights cannot be easily altered to incrementally learn additional classes, since the…

Computer Vision and Pattern Recognition · Computer Science 2022-12-02 Penny Johnston , Keiller Nogueira , Kevin Swingler

It is reasonable to consider, in many cases, that individuals' latent traits have a hierarchical structure such that more general traits are a suitable composition of more specific ones. Existing item response models that account for such…

Methodology · Statistics 2020-07-28 Juliane Venturelli S. L. , Flavio B. Gonçalves , Dalton F. Andrade

The Gibbs sampler (a.k.a. Glauber dynamics and heat-bath algorithm) is a popular Markov Chain Monte Carlo algorithm which iteratively samples from the conditional distributions of a probability measure $\pi$ of interest. Under the…

Probability · Mathematics 2026-01-21 Filippo Ascolani , Hugo Lavenant , Giacomo Zanella

We extend the notion of Gibbsianness for mean-field systems to the set-up of general (possibly continuous) local state spaces. We investigate the Gibbs properties of systems arising from an initial mean-field Gibbs measure by application of…

Probability · Mathematics 2009-11-13 C. Kuelske , A. A. Opoku

It is shown that power law phase space distributions describe marginally stable Gibbsian equilibria far from thermal equilibrium which are expected to occur in collisionless plasmas containing fully developed quasi-stationary turbulence.…

Plasma Physics · Physics 2008-04-22 R. A. Treumann , C. H. Jaroschek

We study a model of spatial random permutations over a discrete set of points. Formally, a permutation $\sigma$ is sampled proportionally to the weight $\exp\{-\alpha \sum_x V(\sigma(x)-x)\},$ where $\alpha>0$ is the temperature and $V$ is…

Probability · Mathematics 2019-04-09 Inés Armendáriz , Pablo A. Ferrari , Nicolás Frevenza

Bayesian nonparametric (BNP) models provide elegant methods for discovering underlying latent features within a data set, but inference in such models can be slow. We exploit the fact that completely random measures, which commonly used…

Machine Learning · Statistics 2020-07-17 Avinava Dubey , Michael Minyi Zhang , Eric P. Xing , Sinead A. Williamson