English
Related papers

Related papers: Poisson Hierarchical Indian Buffet Processes-With …

200 papers

Feature allocation models are an extension of Bayesian nonparametric clustering models, where individuals can share multiple features. We study a broad class of models whose probability distribution has a product form, which includes the…

Methodology · Statistics 2025-11-12 Lorenzo Ghilotti , Federico Camerlenghi , Tommaso Rigon

We propose the attraction Indian buffet distribution (AIBD), a distribution for binary feature matrices influenced by pairwise similarity information. Binary feature matrices are used in Bayesian models to uncover latent variables (i.e.,…

Methodology · Statistics 2021-07-19 Richard L. Warr , David B. Dahl , Jeremy M. Meyer , Arthur Lui

Species sampling processes have long served as the fundamental framework for modeling random discrete distributions and exchangeable sequences. However, data arising from distinct but related sources require a broader notion of…

Statistics Theory · Mathematics 2026-02-03 Beatrice Franzolini , Antonio Lijoi , Igor Prünster , Giovanni Rebaudo

The Indian buffet process (IBP) and phylogenetic Indian buffet process (pIBP) can be used as prior models to infer latent features in a data set. The theoretical properties of these models are under-explored, however, especially in high…

Applications · Statistics 2019-09-23 Tong Li , Tianjian Zhou , Kam-Wah Tsui , Lin Wei , Yuan Ji

Deep generative models (DGMs) have brought about a major breakthrough, as well as renewed interest, in generative latent variable models. However, DGMs do not allow for performing data-driven inference of the number of latent features…

Machine Learning · Computer Science 2018-04-03 Sotirios P. Chatzis

Analyzing multivariate time series data is important to predict future events and changes of complex systems in finance, manufacturing, and administrative decisions. The expressiveness power of Gaussian Process (GP) regression methods has…

Machine Learning · Statistics 2019-05-23 Anh Tong , Jaesik Choi

We build upon probabilistic models for Boolean Matrix and Boolean Tensor factorisation that have recently been shown to solve these problems with unprecedented accuracy and to enable posterior inference to scale to Billions of observation.…

Machine Learning · Statistics 2019-07-02 Tammo Rukat , Christopher Yau

In microbiome studies, it is of interest to use a sample from a population of microbes, such as the gut microbiota community, to estimate the population proportion of these taxa. However, due to biases introduced in sampling and…

Methodology · Statistics 2022-10-11 Roulan Jiang , Xiang Zhan , Tianying Wang

Multivariate Hawkes Processes (MHPs) are an important class of temporal point processes that have enabled key advances in understanding and predicting social information systems. However, due to their complex modeling of temporal…

Machine Learning · Computer Science 2020-03-02 Maximilian Nickel , Matthew Le

We develop a new class of dynamic multivariate Poisson count models that allow for fast online updating and we refer to these models as multivariate Poisson-scaled beta (MPSB). The MPSB model allows for serial dependence in the counts as…

Methodology · Statistics 2016-09-16 Tevfik Aktekin , Nicholas G. Polson , Refik Soyer

In genomics, differential abundance and expression analyses are complicated by the compositional nature of sequence count data, which reflect only relative-not absolute-abundances or expression levels. Many existing methods attempt to…

Methodology · Statistics 2025-12-16 Won Gu , Francesca Chiaromonte , Justin D. Silverman

Inferring concerted changes among biological traits along an evolutionary history remains an important yet challenging problem. Besides adjusting for spurious correlation induced from the shared history, the task also requires sufficient…

Scientific studies in the last two decades have established the central role of the microbiome in disease and health. Differential abundance analysis seeks to identify microbial taxa associated with sample groups defined by a factor such as…

Methodology · Statistics 2023-12-29 Archie Sachdeva , Somnath Datta , Subharup Guha

We study random families of subsets of $\mathbb{N}$ that are similar to exchangeable random partitions, but do not require constituent sets to be disjoint: Each element of ${\mathbb{N}}$ may be contained in multiple subsets. One class of…

Probability · Mathematics 2015-10-27 Lancelot F. James , Peter Orbanz , Yee Whye Teh

This paper introduces a novel family of geostatistical models designed to capture complex features beyond the reach of traditional Gaussian processes. The proposed family, termed the Poisson-Gaussian Mixture Process (POGAMP), is…

Methodology · Statistics 2024-12-09 F. B. Gonçalves , M. O. Prates , G. A. S. Aguilar

Vine copulas allow to build flexible dependence models for an arbitrary number of variables using only bivariate building blocks. The number of parameters in a vine copula model increases quadratically with the dimension, which poses new…

Methodology · Statistics 2018-11-20 Thomas Nagler , Christian Bumann , Claudia Czado

In this paper, we present a new variable selection method for regression and classification purposes. Our method, called Subsampling Ranking Forward selection (SuRF), is based on LASSO penalised regression, subsampling and forward-selection…

Methodology · Statistics 2021-05-25 Lihui Liu , Hong Gu , Johan Van Limbergen , Toby Kenney

We address the problem of the joint statistical inference of phylogenetic trees and multiple sequence alignments from unaligned molecular sequences. This problem is generally formulated in terms of string-valued evolutionary processes along…

Populations and Evolution · Quantitative Biology 2015-06-05 Alexandre Bouchard-Côté , Michael I. Jordan

Distributions over exchangeable matrices with infinitely many columns, such as the Indian buffet process, are useful in constructing nonparametric latent variable models. However, the distribution implied by such models over the number of…

Methodology · Statistics 2012-09-07 Sinead Williamson , Zoubin Ghahramani , Steven N. MacEachern , Eric P. Xing

Latent feature models are widely used to decompose data into a small number of components. Bayesian nonparametric variants of these models, which use the Indian buffet process (IBP) as a prior over latent features, allow the number of…

Machine Learning · Statistics 2012-09-11 Samuel J. Gershman , Peter I. Frazier , David M. Blei