中文
相关论文

相关论文: KL Divergence Between Gaussians: A Step-by-Step De…

200 篇论文

Discriminator Guidance has become a popular method for efficiently refining pre-trained Score-Matching Diffusion models. However, in this paper, we demonstrate that the standard implementation of this technique does not necessarily lead to…

It has been discovered that latent-Euclidean variational autoencoders (VAEs) admit, in various capacities, Riemannian structure. We adapt these arguments but for complex VAEs with a complex latent stage. We show that complex VAEs reveal to…

机器学习 · 计算机科学 2026-01-01 Andrew Gracyk

We study empirical Bayes (EB) predictive density estimation in linear mixed models (LMMs) with large number of units, which induce a high dimensional random effects space. Focusing on Kullback Leibler (KL) risk minimization, we develop a…

统计方法学 · 统计学 2026-03-31 Abir Sarkar , Gourab Mukherjee , Keisuke Yano

We propose a general algorithm for approximating nonstandard Bayesian posterior distributions. The algorithm minimizes the Kullback-Leibler divergence of an approximating distribution to the intractable posterior distribution. Our method…

统计计算 · 统计学 2014-07-29 Tim Salimans , David A. Knowles

We show that the Kullback-Leibler distance is a good measure of the statistical uncertainty of correlation matrices estimated by using a finite set of data. For correlation matrices of multivariate Gaussian variables we analytically…

数据分析、统计与概率 · 物理学 2008-12-02 Michele Tumminello , Fabrizio Lillo , Rosario Nunzio Mantegna

Divergences are quantities that measure discrepancy between two probability distributions and play an important role in various fields such as statistics and machine learning. Divergences are non-negative and are equal to zero if and only…

统计理论 · 数学 2019-10-22 Tomohiro Nishiyama

Expectation maximization (EM) is the default algorithm for fitting probabilistic models with missing or latent variables, yet we lack a full understanding of its non-asymptotic convergence properties. Previous works show results along the…

机器学习 · 计算机科学 2022-03-01 Frederik Kunstner , Raunak Kumar , Mark Schmidt

In knowledge distillation, a primary focus has been on transforming and balancing multiple distillation components. In this work, we emphasize the importance of thoroughly examining each distillation component, as we observe that not all…

机器学习 · 计算机科学 2024-10-22 Zao Zhang , Huaming Chen , Pei Ning , Nan Yang , Dong Yuan

We derive a new variational formula for the R\'enyi family of divergences, $R_\alpha(Q\|P)$, between probability measures $Q$ and $P$. Our result generalizes the classical Donsker-Varadhan variational formula for the Kullback-Leibler…

机器学习 · 统计学 2021-07-21 Jeremiah Birrell , Paul Dupuis , Markos A. Katsoulakis , Luc Rey-Bellet , Jie Wang

In a first part, we present a mathematical analysis of a general methodology of a probabilistic learning inference that allows for estimating a posterior probability model for a stochastic boundary value problem from a prior probability…

机器学习 · 统计学 2022-06-08 Christian Soize

Small-scale intermittency is studied as the deviation of the probability distributions of pseudodissipation, dissipation and enstrophy in turbulence from those of a Gaussian random velocity field. This deviation is quantified using…

流体动力学 · 物理学 2026-05-26 Shreyashri Sarkar , Rishita Das

Deep kernel learning (DKL) leverages the connection between Gaussian process (GP) and neural networks (NN) to build an end-to-end, hybrid model. It combines the capability of NN to learn rich representations under massive data and the…

机器学习 · 统计学 2020-08-20 Haitao Liu , Yew-Soon Ong , Xiaomo Jiang , Xiaofang Wang

To achieve scalable and accurate inference for latent Gaussian processes, we propose a variational approximation based on a family of Gaussian distributions whose covariance matrices have sparse inverse Cholesky (SIC) factors. We combine…

机器学习 · 统计学 2023-05-30 Jian Cao , Myeongjong Kang , Felix Jimenez , Huiyan Sang , Florian Schafer , Matthias Katzfuss

To ensure stability of learning, state-of-the-art generalized policy iteration algorithms augment the policy improvement step with a trust region constraint bounding the information loss. The size of the trust region is commonly determined…

机器学习 · 计算机科学 2018-04-05 Boris Belousov , Jan Peters

The Kullback-Leibler divergence, the Kullback-Leibler variation, and the Bernstein "norm" are used to quantify discrepancies among probability distributions in likelihood models such as nonparametric maximum likelihood and nonparametric…

统计理论 · 数学 2026-01-27 Tetsuya Kaji

We present a statistical mechanical framework based on the Kullback-Leibler divergence (KLD) to analyze the relativistic limits of decoding time-encoded information from a moving source. By modeling the symbol durations as…

数学物理 · 物理学 2025-07-31 Tatsuaki Tsuruyama

We introduce an improved variational autoencoder (VAE) for text modeling with topic information explicitly modeled as a Dirichlet latent variable. By providing the proposed model topic awareness, it is more superior at reconstructing input…

计算与语言 · 计算机科学 2018-11-02 Yijun Xiao , Tiancheng Zhao , William Yang Wang

In this paper, we propose some estimators for the parameters of a statistical model based on Kullback-Leibler divergence of the survival function in continuous setting. We prove that the proposed estimators are subclass of "generalized…

统计理论 · 数学 2016-07-01 Yaser Mehrali , Majid Asadi

Obtaining an accurate estimate of the underlying covariance matrix from finite sample size data is challenging due to sample size noise. In recent years, sophisticated covariance-cleaning techniques based on random matrix theory have been…

统计计算 · 统计学 2024-11-11 Christian Bongiorno , Lamia Lamrani

Diffusion models have achieved great success in generating high-dimensional samples across various applications. While the theoretical guarantees for continuous-state diffusion models have been extensively studied, the convergence analysis…

机器学习 · 计算机科学 2025-04-15 Zikun Zhang , Zixiang Chen , Quanquan Gu
‹ 上一页 1 8 9 10 下一页 ›