English
Related papers

Related papers: Exponential families from a single KL identity

200 papers

In this letter, we present a novel exponentially embedded families (EEF) based classification method, in which the probability density function (PDF) on raw data is estimated from the PDF on features. With the PDF construction, we show that…

Machine Learning · Statistics 2016-08-24 Bo Tang , Steven Kay , Haibo He , Paul M. Baggenstoss

We describe \textit{deep exponential families} (DEFs), a class of latent variable models that are inspired by the hidden structures used in deep neural networks. DEFs capture a hierarchy of dependencies between latent variables, and are…

Machine Learning · Statistics 2014-11-11 Rajesh Ranganath , Linpeng Tang , Laurent Charlin , David M. Blei

Calibration of mean estimates for predictions is a crucial property in many applications, particularly in the fields of financial and actuarial decision-making. In this paper, we first review classical approaches for validating…

Applications · Statistics 2025-10-29 Łukasz Delong , Mario Wüthrich

Here we propose a novel model family with the objective of learning to disentangle the factors of variation in data. Our approach is based on the spike-and-slab restricted Boltzmann machine which we generalize to include higher-order…

Machine Learning · Statistics 2012-10-22 Guillaume Desjardins , Aaron Courville , Yoshua Bengio

Many real life domains contain a mixture of discrete and continuous variables and can be modeled as hybrid Bayesian Networks. Animportant subclass of hybrid BNs are conditional linear Gaussian (CLG) networks, where the conditional…

Artificial Intelligence · Computer Science 2013-01-14 Uri Lerner , Eran Segal , Daphne Koller

In this paper, we introduce a new four-parameter generalization of the exponentiated Weibull (EW) distribution, called the exponentiated Weibull-logarithmic (EWL) distribution, which obtained by compounding EW and logarithmic distributions.…

Methodology · Statistics 2014-02-24 Eisa Mahmoudi , Afsaneh Sepahdar , Artur Lemonte

Combining discrete probability distributions and combinatorial optimization problems with neural network components has numerous applications but poses several challenges. We propose Implicit Maximum Likelihood Estimation (I-MLE), a…

Machine Learning · Computer Science 2021-10-28 Mathias Niepert , Pasquale Minervini , Luca Franceschi

Reward augmented maximum likelihood (RAML), a simple and effective learning framework to directly optimize towards the reward function in structured prediction tasks, has led to a number of impressive empirical successes. RAML incorporates…

Machine Learning · Computer Science 2017-10-31 Xuezhe Ma , Pengcheng Yin , Jingzhou Liu , Graham Neubig , Eduard Hovy

Evidential Deep Learning (EDL) has emerged as an efficient, sampling-free strategy for uncertainty estimation. A series of EDL variants have been proposed to address specific limitations of the original framework, achieving notable success.…

Machine Learning · Computer Science 2026-05-26 Yuanye Liu , Yibo Gao , Yuanyang Chen , Xiahai Zhuang

The Kullback-Leibler (KL) divergence is a foundational measure for comparing probability distributions. Yet in multivariate settings, its single value often obscures the underlying reasons for divergence, conflating mismatches in individual…

Other Computer Science · Computer Science 2025-05-06 William Cook

Simplex-valued data appear throughout statistics and machine learning, for example in the context of transfer learning and compression of deep networks. Existing models for this class of data rely on the Dirichlet distribution or other…

Machine Learning · Statistics 2020-06-09 Elliott Gordon-Rodriguez , Gabriel Loaiza-Ganem , John P. Cunningham

A new method called "variational sampling" is proposed to estimate integrals under probability distributions that can be evaluated up to a normalizing constant. The key idea is to fit the target distribution with an exponential family model…

Computation · Statistics 2013-10-15 Alexis Roche

The main aim of this article is to characterize and investigate the three parameter exponentiated exponential Poisson probability distribution ${\rm EEP}(\alpha, \beta, \lambda)$ by giving explicit closed form expressions for its…

Statistics Theory · Mathematics 2014-02-04 Tibor K Pogány

We consider the problem of computing the maximum likelihood multivariate log-concave distribution for a set of points. Specifically, we present an algorithm which, given $n$ points in $\mathbb{R}^d$ and an accuracy parameter $\epsilon>0$,…

Data Structures and Algorithms · Computer Science 2019-07-22 Brian Axelrod , Ilias Diakonikolas , Anastasios Sidiropoulos , Alistair Stewart , Gregory Valiant

There is an abundance of useful fluctuation identities for one-sided L\'evy processes observed up to an independent exponentially distributed time horizon. We show that all the fundamental formulas generalize to time horizons having matrix…

Probability · Mathematics 2021-01-21 Mogens Bladt , Jevgenijs Ivanovs

The methods of statistical physics are widely used for modelling complex networks. Building on the recently proposed Equilibrium Expectation approach, we derive a simple and efficient algorithm for maximum likelihood estimation (MLE) of…

Computation · Statistics 2020-02-12 Alexander Borisenko , Maksym Byshkin , Alessandro Lomi

We introduce generalized notions of a divergence function and a Fisher information matrix. We propose to generalize the notion of an exponential family of models by reformulating it in terms of the Fisher information matrix. Our methods are…

Information Theory · Computer Science 2013-02-22 Jan Naudts , Ben Anthonis

In the context of modulated-symmetry distributions, there exist various forms of skew-elliptical families. We present yet another one, but with an unusual feature: the modulation factor of the baseline elliptical density is represented by a…

Probability · Mathematics 2017-10-11 Adelchi Azzalini , Giuliana Regoli

While standard estimation assumes that all datapoints are from probability distribution of the same fixed parameters $\theta$, we will focus on maximum likelihood (ML) adaptive estimation for nonstationary time series: separately estimating…

Machine Learning · Statistics 2020-03-24 Jarek Duda

It is commonly believed that optimizing the reverse KL divergence results in "mode seeking", while optimizing forward KL results in "mass covering", with the latter being preferred if the goal is to sample from multiple diverse modes. We…

Machine Learning · Computer Science 2025-10-24 Anthony GX-Chen , Jatin Prakash , Jeff Guo , Rob Fergus , Rajesh Ranganath