English
Related papers

Related papers: Unifying and extending Precision Recall metrics fo…

200 papers

Recent advances in generative modeling have led to an increased interest in the study of statistical divergences as means of model comparison. Commonly used evaluation methods, such as the Frechet Inception Distance (FID), correlate well…

Machine Learning · Statistics 2018-10-30 Mehdi S. M. Sajjadi , Olivier Bachem , Mario Lucic , Olivier Bousquet , Sylvain Gelly

In this article we revisit the definition of Precision-Recall (PR) curves for generative models proposed by Sajjadi et al. (arXiv:1806.00035). Rather than providing a scalar for generative quality, PR curves distinguish mode-collapse (poor…

Machine Learning · Computer Science 2023-07-03 Loïc Simon , Ryan Webster , Julien Rabin

Devising indicative evaluation metrics for the image generation task remains an open problem. The most widely used metric for measuring the similarity between real and generated images has been the Fr\'echet Inception Distance (FID) score.…

Computer Vision and Pattern Recognition · Computer Science 2020-06-30 Muhammad Ferjad Naeem , Seong Joon Oh , Youngjung Uh , Yunjey Choi , Jaejun Yoo

While generative models have become increasingly prevalent across various domains, fundamental concerns regarding their reliability persist. A crucial yet understudied aspect of these models is the uncertainty quantification surrounding…

Machine Learning · Computer Science 2025-11-17 Giorgio Morales , Frederic Jurie , Jalal Fadili

With the recent success of generative models in image and text, the question of their evaluation has recently gained a lot of attention. While most methods from the state of the art rely on scalar metrics, the introduction of Precision and…

Artificial Intelligence · Computer Science 2026-05-19 Benjamin Sykes , Loïc Simon , Julien Rabin , Jalal Fadili

Precision and Recall are two prominent metrics of generative performance, which were proposed to separately measure the fidelity and diversity of generative models. Given their central role in comparing and improving generative models,…

Machine Learning · Computer Science 2023-07-20 Mahyar Khayatkhoei , Wael AbdAlmageed

Despite the tremendous progress in the estimation of generative models, the development of tools for diagnosing their failures and assessing their performance has advanced at a much slower pace. Recent developments have investigated metrics…

Machine Learning · Computer Science 2020-06-09 Josip Djolonga , Mario Lucic , Marco Cuturi , Olivier Bachem , Olivier Bousquet , Sylvain Gelly

The recent advent of powerful generative models has triggered the renewed development of quantitative measures to assess the proximity of two probability distributions. As the scalar Frechet inception distance remains popular, several…

Machine Learning · Computer Science 2022-10-14 Rodrigue Siry , Ryan Webster , Loic Simon , Julien Rabin

Achieving a balance between image quality (precision) and diversity (recall) is a significant challenge in the domain of generative models. Current state-of-the-art models primarily rely on optimizing heuristics, such as the Fr\'echet…

Machine Learning · Computer Science 2023-11-02 Alexandre Verine , Benjamin Negrevergne , Muni Sreenivas Pydi , Yann Chevaleyre

For some classification scenarios, it is desirable to use only those classification instances that a trained model associates with a high certainty. To obtain such high-certainty instances, previous work has proposed accuracy-reject curves.…

Machine Learning · Computer Science 2024-03-15 Lydia Fischer , Patricia Wollstadt

Generative models can have distinct mode of failures like mode dropping and low quality samples, which cannot be captured by a single scalar metric. To address this, recent works propose evaluating generative models using precision and…

Machine Learning · Computer Science 2023-02-03 Alexandre Verine , Benjamin Negrevergne , Muni Sreenivas Pydi , Yann Chevaleyre

Assessing the fidelity and diversity of the generative model is a difficult but important issue for technological advancement. So, recent papers have introduced k-Nearest Neighbor ($k$NN) based precision-recall metrics to break down the…

Machine Learning · Computer Science 2024-01-25 Dogyun Park , Suhyun Kim

Implicit generative models, which do not return likelihood values, such as generative adversarial networks and diffusion models, have become prevalent in recent years. While it is true that these models have shown remarkable results,…

Machine Learning · Computer Science 2022-06-23 Eyal Betzalel , Coby Penso , Aviv Navon , Ethan Fetaya

Considering the difficulty of interpreting generative model output, there is significant current research focused on determining meaningful evaluation metrics. Several recent approaches utilize "precision" and "recall," borrowed from the…

Machine Learning · Computer Science 2025-02-28 Alexis Fox , Samarth Swarup , Abhijin Adiga

The evaluation of deep generative models has been extensively studied in the centralized setting, where the reference data are drawn from a single probability distribution. On the other hand, several applications of generative models…

Machine Learning · Computer Science 2024-06-12 Zixiao Wang , Farzan Farnia , Zhenghao Lin , Yunheng Shen , Bei Yu

Protein structure generative models have seen a recent surge of interest, but meaningfully evaluating them computationally is an active area of research. While current metrics have driven useful progress, they do not capture how well models…

Biomolecules · Quantitative Biology 2025-07-25 Felix Faltings , Hannes Stark , Tommi Jaakkola , Regina Barzilay

We propose a robust and reliable evaluation metric for generative models by introducing topological and statistical treatments for rigorous support estimation. Existing metrics, such as Inception Score (IS), Frechet Inception Distance…

Machine Learning · Computer Science 2024-01-25 Pum Jun Kim , Yoojin Jang , Jisu Kim , Jaejun Yoo

Deep generative models are powerful tools that have produced impressive results in recent years. These advances have been for the most part empirically driven, making it essential that we use high quality evaluation metrics. In this paper,…

Machine Learning · Statistics 2018-06-22 Shane Barratt , Rishi Sharma

Although generative models have made remarkable progress in recent years, their use in critical applications has been hindered by an inability to reliably evaluate the quality of their generated samples. Quality refers to at least two…

Machine Learning · Computer Science 2026-02-18 Nicolas Salvy , Hugues Talbot , Bertrand Thirion

This work is an update of a previous paper on the same topic published a few years ago. With the dramatic progress in generative modeling, a suite of new quantitative and qualitative techniques to evaluate models has emerged. Although some…

Machine Learning · Computer Science 2021-10-05 Ali Borji
‹ Prev 1 2 3 10 Next ›