English
Related papers

Related papers: Relaxed Triangle Inequality for Kullback-Leibler D…

200 papers

We study the Unadjusted Langevin Algorithm (ULA) for sampling from a probability distribution $\nu = e^{-f}$ on $\mathbb{R}^n$. We prove a convergence guarantee in Kullback-Leibler (KL) divergence assuming $\nu$ satisfies a log-Sobolev…

Data Structures and Algorithms · Computer Science 2022-03-04 Santosh S. Vempala , Andre Wibisono

This paper investigates the best known bounds on the quadratic Gaussian distortion-rate-perception function with limited common randomness for the Kullback-Leibler divergence-based perception measure, as well as their counterparts for the…

Information Theory · Computer Science 2024-09-05 Li Xie , Liangyan Li , Jun Chen , Lei Yu , Zhongshan Zhang

Statistical distances (SDs), which quantify the dissimilarity between probability distributions, are central to machine learning and statistics. A modern method for estimating such distances from data relies on parametrizing a variational…

Statistics Theory · Mathematics 2021-03-18 Sreejith Sreekumar , Zhengxin Zhang , Ziv Goldfeld

We show that absolute correlation distance satisfies a K-relaxed triangle inequality, with the best K = 2.

Metric Geometry · Mathematics 2022-02-11 Stanislav Dubrovskiy

Two geometrical structures have been extensively studied for a manifold of probability distributions. One is based on the Fisher information metric, which is invariant under reversible transformations of random variables, while the other is…

Optimization and Control · Mathematics 2017-10-02 Shun-ichi Amari , Ryo Karakida , Masafumi Oizumi

The book is structured into four main chapters. Chapter 1 introduces the foundational concepts of divergence measures, including the well-known Kullback-Leibler divergence and its limitations. It then presents a detailed exploration of…

Methodology · Statistics 2024-09-04 Shinto Eguchi

In this work, we present formulations for regularized Kullback-Leibler and R\'enyi divergences via the Alpha Log-Determinant (Log-Det) divergences between positive Hilbert-Schmidt operators on Hilbert spaces in two different settings,…

Machine Learning · Statistics 2022-07-19 Minh Ha Quang

We study the problem of learning mixtures of linear classifiers under Gaussian covariates. Given sample access to a mixture of $r$ distributions on $\mathbb{R}^n$ of the form $(\mathbf{x},y_{\ell})$, $\ell\in [r]$, where…

Machine Learning · Computer Science 2023-10-19 Ilias Diakonikolas , Daniel M. Kane , Yuxin Sun

In this note, we characterize the Gompertz distribution in terms of extreme value distributions and point out that it implicitly models the interplay of two antagonistic growth processes. In addition, we derive a closed form expressions for…

Information Theory · Computer Science 2014-02-14 Christian Bauckhage

We study the relation between the total variation (TV) and Hellinger distances between two Gaussian location mixtures. Our first result establishes a general upper bound: for any two mixing distributions supported on a compact set, the…

Statistics Theory · Mathematics 2026-05-27 Joonhyuk Jung , Chao Gao

We consider the question of estimating multi-dimensional Gaussian mixtures (GM) with compactly supported or subgaussian mixing distributions. Minimax estimation rate for this class (under Hellinger, TV and KL divergences) is a long-standing…

Statistics Theory · Mathematics 2023-06-28 Zeyu Jia , Yury Polyanskiy , Yihong Wu

Recent research has revealed that deep generative models including flow-based models and Variational Autoencoders may assign higher likelihoods to out-of-distribution (OOD) data than in-distribution (ID) data. However, we cannot sample OOD…

Machine Learning · Computer Science 2023-03-03 Yufeng Zhang , Jialu Pan , Wanwei Liu , Zhenbang Chen , Ji Wang , Zhiming Liu , Kenli Li , Hongmei Wei

We study density estimation in Kullback-Leibler divergence: given an i.i.d. sample from an unknown density $p^\star$, the goal is to construct an estimator $\widehat{p}$ such that $\mathrm{KL}(p^\star,\widehat{p})$ is small with high…

Statistics Theory · Mathematics 2026-04-03 Spencer Compton , Gábor Lugosi , Jaouad Mourtada , Jian Qian , Nikita Zhivotovskiy

Statistical modelling of covariate distributions allows to generate virtual populations or to impute missing values in a covariate dataset. Covariate distributions typically have non-Gaussian margins and show nonlinear correlation…

Applications · Statistics 2025-03-20 Niklas Hartung , Aleksandra Khatova

In Bayesian statistics probability distributions express beliefs. However, for many problems the beliefs cannot be computed analytically and approximations of beliefs are needed. We seek a loss function that quantifies how "embarrassing" it…

Statistics Theory · Mathematics 2017-08-07 Reimar H. Leike , Torsten A. Enßlin

We review recent results about the maximal values of the Kullback-Leibler information divergence from statistical models defined by neural networks, including naive Bayes models, restricted Boltzmann machines, deep belief networks, and…

Statistics Theory · Mathematics 2014-06-18 Guido Montufar , Johannes Rauh , Nihat Ay

Mixture distributions arise in many parametric and non-parametric settings -- for example, in Gaussian mixture models and in non-parametric estimation. It is often necessary to compute the entropy of a mixture, but, in most cases, this…

Information Theory · Computer Science 2022-11-22 Artemy Kolchinsky , Brendan D. Tracey

Many interesting machine learning problems are best posed by considering instances that are distributions, or sample sets drawn from distributions. Previous work devoted to machine learning tasks with distributional inputs has done so…

Machine Learning · Statistics 2021-01-15 Danica J. Sutherland , Junier B. Oliva , Barnabás Póczos , Jeff Schneider

Consider the Langevin diffusion process $\mathrm{d} X_t = \nabla \log p_t(X_t) + \sqrt{2}\mathrm{d} W_t$ guided by the time-dependent probability density $p_t(x)$. Let $q_t$ be the density of $X_t$. Recently, in order to analyze convergence…

Optimization and Control · Mathematics 2025-11-19 Andreas Habring

We show that the moment generating function of the Kullback-Leibler divergence (relative entropy) between the empirical distribution of $n$ independent samples from a distribution $P$ over a finite alphabet of size $k$ (i.e. a multinomial…

Information Theory · Computer Science 2020-10-06 Rohit Agrawal
‹ Prev 1 8 9 10 Next ›