中文
相关论文

相关论文: A Pragmatic Note on Evaluating Generative Models w…

200 篇论文

Text-to-image generation models have progressed considerably in recent years, which can now generate impressive realistic images from arbitrary text. Most of such models are trained on web-scale image-text paired datasets, which may not be…

计算机视觉与模式识别 · 计算机科学 2022-10-26 Yufan Zhou , Chunyuan Li , Changyou Chen , Jianfeng Gao , Jinhui Xu

Discovering crystal structures with specific chemical properties has become an increasingly important focus in material science. However, current models are limited in their ability to generate new crystal lattices, as they only consider…

材料科学 · 物理学 2024-01-12 Astrid Klipfel , Yaël Fregier , Adlane Sayede , Zied Bouraoui

Experts use retinal images and vessel trees to detect and diagnose various eye, blood circulation, and brain-related diseases. However, manual segmentation of retinal images is a time-consuming process that requires high expertise and is…

图像与视频处理 · 电气工程与系统科学 2023-08-17 Alnur Alimanov , Md Baharul Islam

Early detection of cancer is key to a good prognosis and requires frequent testing, especially in pediatrics. Whole-body magnetic resonance imaging (wbMRI) is an essential part of several well-established screening protocols, with screening…

图像与视频处理 · 电气工程与系统科学 2020-06-02 Alex Chang , Vinith M. Suriyakumar , Abhishek Moturu , Nipaporn Tewattanarat , Andrea Doria , Anna Goldenberg

Modern medical image translation methods use generative models for tasks such as the conversion of CT images to MRI. Evaluating these methods typically relies on some chosen downstream task in the target domain, such as segmentation. On the…

图像与视频处理 · 电气工程与系统科学 2024-04-12 Nicholas Konz , Yuwen Chen , Hanxue Gu , Haoyu Dong , Maciej A. Mazurowski

It is well known that the reconstruction FID (rFID) of a VAE is poorly correlated with the generation FID (gFID) of a latent diffusion model. We propose interpolated FID (iFID), a simple variant of rFID that exhibits a strong correlation…

计算机视觉与模式识别 · 计算机科学 2026-05-08 Tongda Xu , Mingwei He , Shady Abu-Hussein , Jose Miguel Hernandez-Lobato , Chunhang Zheng , Kai Zhao , Chao Zhou , Ya-Qin Zhang , Yan Wang

We introduce a new generative approach for synthesizing 3D geometry and images from single-view collections. Most existing approaches predict volumetric density to render multi-view consistent images. By employing volumetric rendering using…

计算机视觉与模式识别 · 计算机科学 2024-06-17 Salvatore Esposito , Qingshan Xu , Kacper Kania , Charlie Hewitt , Octave Mariotti , Lohit Petikam , Julien Valentin , Arno Onken , Oisin Mac Aodha

Generative Adversarial Networks (GANs) have high computational costs to train their complex architectures. Throughout the training process, GANs' output is analyzed qualitatively based on the loss and synthetic images' diversity and…

计算机视觉与模式识别 · 计算机科学 2024-06-03 Muhammad Muneeb Saad , Mubashir Husain Rehmani , Ruairi O'Reilly

Deep generative models are powerful tools that have produced impressive results in recent years. These advances have been for the most part empirically driven, making it essential that we use high quality evaluation metrics. In this paper,…

机器学习 · 统计学 2018-06-22 Shane Barratt , Rishi Sharma

The 3D-aware image synthesis focuses on conserving spatial consistency besides generating high-resolution images with fine details. Recently, Neural Radiance Field (NeRF) has been introduced for synthesizing novel views with low…

计算机视觉与模式识别 · 计算机科学 2024-01-10 Jiwook Kim , Minhyeok Lee

Devising domain- and model-agnostic evaluation metrics for generative models is an important and as yet unresolved problem. Most existing metrics, which were tailored solely to the image synthesis setup, exhibit a limited capacity for…

机器学习 · 计算机科学 2022-07-14 Ahmed M. Alaa , Boris van Breugel , Evgeny Saveliev , Mihaela van der Schaar

Recently, learning-based approaches for 3D model reconstruction have attracted attention owing to its modern applications such as Extended Reality(XR), robotics and self-driving cars. Several approaches presented good performance on…

计算机视觉与模式识别 · 计算机科学 2021-04-30 Luoyang Lin , Dihong Tian

We present novel approaches involving generative adversarial networks and diffusion models in order to synthesize high quality, live and spoof fingerprint images while preserving features such as uniqueness and diversity. We generate live…

计算机视觉与模式识别 · 计算机科学 2024-03-22 W. Tang , D. Figueroa , D. Liu , K. Johnsson , A. Sopasakis

In this paper we discuss a class of AutoEncoder based generative models based on one dimensional sliced approach. The idea is based on the reduction of the discrimination between samples to one-dimensional case. Our experiments show that…

机器学习 · 计算机科学 2019-01-30 Szymon Knop , Marcin Mazur , Jacek Tabor , Igor Podolak , Przemysław Spurek

Although masked image generation models and masked diffusion models are designed with different motivations and objectives, we observe that they can be unified within a single framework. Building upon this insight, we carefully explore the…

计算机视觉与模式识别 · 计算机科学 2026-03-03 Zebin You , Jingyang Ou , Xiaolu Zhang , Jun Hu , Jun Zhou , Chongxuan Li

Automatic evaluation of the goodness of Generative Adversarial Networks (GANs) has been a challenge for the field of machine learning. In this work, we propose a distance complementary to existing measures: Topology Distance (TD), the main…

机器学习 · 计算机科学 2020-02-28 Danijela Horak , Simiao Yu , Gholamreza Salimi-Khorshidi

The performance of diagnostic Computer-Aided Design (CAD) systems for retinal diseases depends on the quality of the retinal images being screened. Thus, many studies have been developed to evaluate and assess the quality of such retinal…

图像与视频处理 · 电气工程与系统科学 2024-09-17 Saif Khalid , Hatem A. Rashwan , Saddam Abdulwahab , Mohamed Abdel-Nasser , Facundo Manuel Quiroga , Domenec Puig

Existing state-of-the-art techniques in exemplar-based image-to-image translation hold several critical concerns. Existing methods related to exemplar-based image-to-image translation are impossible to translate on an image tuple input…

计算机视觉与模式识别 · 计算机科学 2021-08-20 Taewon Kang

Many recent developments on generative models for natural images have relied on heuristically-motivated metrics that can be easily gamed by memorizing a small sample from the true distribution or training a model directly to improve the…

机器学习 · 计算机科学 2021-06-08 Ching-Yuan Bai , Hsuan-Tien Lin , Colin Raffel , Wendy Chih-wen Kan

A good metric, which promises a reliable comparison between solutions, is essential for any well-defined task. Unlike most vision tasks that have per-sample ground-truth, image synthesis tasks target generating unseen data and hence are…

计算机视觉与模式识别 · 计算机科学 2023-10-24 Mengping Yang , Ceyuan Yang , Yichi Zhang , Qingyan Bai , Yujun Shen , Bo Dai