English
Related papers

Related papers: Dependent Multinomial Models Made Easy: Stick Brea…

200 papers

We introduce a generalized Bayesian method for multiple changepoint analysis with a loss function inspired by multinomial logistic regression. The method does not require a specification of the data-generating process and avoids restrictive…

Methodology · Statistics 2026-03-27 Yuhui Wang , Andrew M. Thomas , Michael Jauch

In multi-label classification, where a single example may be associated with several class labels at the same time, the ability to model dependencies between labels is considered crucial to effectively optimize non-decomposable evaluation…

Machine Learning · Computer Science 2021-06-23 Michael Rapp , Eneldo Loza Mencía , Johannes Fürnkranz , Eyke Hüllermeier

We propose a new data-augmentation strategy for fully Bayesian inference in models with binomial likelihoods. The approach appeals to a new class of Polya-Gamma distributions, which are constructed in detail. A variety of examples are…

Methodology · Statistics 2015-03-20 Nicholas G. Polson , James G. Scott , Jesse Windle

We describe a procedure to introduce general dependence structures on a set of Dirichlet processes. Dependence can be in one direction to define a time series or in two directions to define spatial dependencies. More directions can also be…

Methodology · Statistics 2021-10-18 Luis E. Nieto-Barajas

The dynamics of linear stochastic growth equations on growing substrates is studied. The substrate is assumed to grow in time following the power law $t^\gamma$, where the growth index $\gamma$ is an arbitrary positive number. Two different…

Statistical Mechanics · Physics 2015-05-13 Carlos Escudero

A ubiquitous challenge in machine learning is the problem of domain generalisation. This can exacerbate bias against groups or labels that are underrepresented in the datasets used for model development. Model bias can lead to unintended…

The functions of proteins and RNAs are determined by a myriad of interactions between their constituent residues, but most quantitative models of how molecular phenotype depends on genotype must approximate this by simple additive effects.…

Quantitative Methods · Quantitative Biology 2017-12-19 Adam J. Riesselman , John B. Ingraham , Debora S. Marks

Bayesian statistical graphical models are typically classified as either continuous and parametric (Gaussian, parameterized by the graph-dependent precision matrix with Wishart-type priors) or discrete and non-parametric (with…

Probability · Mathematics 2025-04-01 Iza Danielewska , Bartosz Kołodziejek , Jacek Wesołowski , Xiaolin Zeng

The P\'olya tree (PT) process is a general-purpose Bayesian nonparametric model that has found wide application in a range of inference problems. It has a simple analytic form and the posterior computation boils down to beta-binomial…

Methodology · Statistics 2021-12-09 Naoki Awaya , Li Ma

Epigenetic observations are represented by the total number of reads from a given pool of cells and the number of methylated reads, making it reasonable to model this data by a binomial distribution. There are numerous factors that can…

Applications · Statistics 2020-04-29 Aliaksandr Hubin , Geir O Storvik , Paul E Grini , Melinka A Butenko

Topic models, such as latent Dirichlet allocation (LDA), can be useful tools for the statistical analysis of document collections and other discrete data. The LDA model assumes that the words of each document arise from a mixture of topics,…

Applications · Statistics 2009-09-29 David M. Blei , John D. Lafferty

In this paper, we discuss the construction of a multivariate generalisation of the Dirichlet-multinomial distribution. An example from forensic genetics in the statistical analysis of DNA mixtures motivates the study of this multivariate…

Applications · Statistics 2014-11-05 Torben Tvedebrink , Poul Svante Eriksen , Niels Morling

We consider the problem of estimating high-dimensional Gaussian graphical models corresponding to a single set of variables under several distinct conditions. This problem is motivated by the task of recovering transcriptional regulatory…

Machine Learning · Statistics 2014-01-24 Karthik Mohan , Palma London , Maryam Fazel , Daniela Witten , Su-In Lee

Modern cancer genomics datasets involve widely varying sizes and scales, measurement variables, and correlation structures. A fundamental analytical goal in these high-throughput studies is the development of general statistical techniques…

Methodology · Statistics 2022-04-12 Chiyu Gu , Veerabhadran Baladandayuthapani , Subharup Guha

In this paper we introduce a novel Bayesian data augmentation approach for estimating the parameters of the generalised logistic regression model. We propose a P\'olya-Gamma sampler algorithm that allows us to sample from the exact…

Methodology · Statistics 2020-12-22 Luciana Dalla Valle , Fabrizio Leisen , Luca Rossini , Weixuan Zhu

In recent years, conditional copulas, that allow dependence between variables to vary according to the values of one or more covariates, have attracted increasing attention. In high dimension, vine copulas offer greater flexibility compared…

Methodology · Statistics 2021-09-24 Rosario Barone , Luciana Dalla Valle

Bayesian models based on the Dirichlet process and other stick-breaking priors have been proposed as core ingredients for clustering, topic modeling, and other unsupervised learning tasks. Prior specification is, however, relatively…

Methodology · Statistics 2021-10-27 Ryan Giordano , Runjing Liu , Michael I. Jordan , Tamara Broderick

We consider the task of detecting regulatory elements in the human genome directly from raw DNA. Past work has focused on small snippets of DNA, making it difficult to model long-distance dependencies that arise from DNA's 3-dimensional…

Genomics · Quantitative Biology 2017-10-04 Ankit Gupta , Alexander M. Rush

Bayesian models based on the Dirichlet process and other stick-breaking priors have been proposed as core ingredients for clustering, topic modeling, and other unsupervised learning tasks. However, due to the flexibility of these models,…

Methodology · Statistics 2022-01-27 Ryan Giordano , Runjing Liu , Michael I. Jordan , Tamara Broderick

Stick-breaking has a long history and is one of the most popular procedures for constructing random discrete distributions in Statistics and Machine Learning. In particular, due to their intuitive construction and computational tractability…

Statistics Theory · Mathematics 2026-01-26 María F. Gil-Leyva , Antonio Lijoi , Ramsés H. Mena , Igor Prünster