English

Explicitly Minimizing the Blur Error of Variational Autoencoders

Computer Vision and Pattern Recognition 2023-04-13 v1 Machine Learning Image and Video Processing

Abstract

Variational autoencoders (VAEs) are powerful generative modelling methods, however they suffer from blurry generated samples and reconstructions compared to the images they have been trained on. Significant research effort has been spent to increase the generative capabilities by creating more flexible models but often flexibility comes at the cost of higher complexity and computational cost. Several works have focused on altering the reconstruction term of the evidence lower bound (ELBO), however, often at the expense of losing the mathematical link to maximizing the likelihood of the samples under the modeled distribution. Here we propose a new formulation of the reconstruction term for the VAE that specifically penalizes the generation of blurry images while at the same time still maximizing the ELBO under the modeled distribution. We show the potential of the proposed loss on three different data sets, where it outperforms several recently proposed reconstruction losses for VAEs.

Keywords

Cite

@article{arxiv.2304.05939,
  title  = {Explicitly Minimizing the Blur Error of Variational Autoencoders},
  author = {Gustav Bredell and Kyriakos Flouris and Krishna Chaitanya and Ertunc Erdil and Ender Konukoglu},
  journal= {arXiv preprint arXiv:2304.05939},
  year   = {2023}
}

Comments

Accepted to ICLR 2023

R2 v1 2026-06-28T10:02:26.995Z