基于模型的 QUILT-1M 病理数据集清洗用于文本条件图像合成
计算机视觉与模式识别
2024-04-12 v1 人工智能
摘要
QUILT-1M 数据集是首个公开的数据集,包含来自 various 在线来源的图像。虽然提供了海量数据多样性,但图像质量和构成高度异构,影响其在文本条件图像合成中的实用性。我们提出了一个自动化管道,用于预测图像中最常见的杂质,如讲者可见性、桌面环境和病理软件,或图像中的文本。此外,我们提出使用语义对齐过滤图像文本对。我们的发现表明,通过严格过滤数据集,可在文本到图像任务中显著提升图像保真度。
关键词
引用
@article{arxiv.2404.07676,
title = {Model-based Cleaning of the QUILT-1M Pathology Dataset for Text-Conditional Image Synthesis},
author = {Marc Aubreville and Jonathan Ganz and Jonas Ammeling and Christopher C. Kaltenecker and Christof A. Bertram},
journal= {arXiv preprint arXiv:2404.07676},
year = {2024}
}
备注
4 pages (short paper)