中文

基于模型的 QUILT-1M 病理数据集清洗用于文本条件图像合成

计算机视觉与模式识别 2024-04-12 v1 人工智能

摘要

QUILT-1M 数据集是首个公开的数据集,包含来自 various 在线来源的图像。虽然提供了海量数据多样性,但图像质量和构成高度异构,影响其在文本条件图像合成中的实用性。我们提出了一个自动化管道,用于预测图像中最常见的杂质,如讲者可见性、桌面环境和病理软件,或图像中的文本。此外,我们提出使用语义对齐过滤图像文本对。我们的发现表明,通过严格过滤数据集,可在文本到图像任务中显著提升图像保真度。

关键词

引用

@article{arxiv.2404.07676,
  title  = {Model-based Cleaning of the QUILT-1M Pathology Dataset for Text-Conditional Image Synthesis},
  author = {Marc Aubreville and Jonathan Ganz and Jonas Ammeling and Christopher C. Kaltenecker and Christof A. Bertram},
  journal= {arXiv preprint arXiv:2404.07676},
  year   = {2024}
}

备注

4 pages (short paper)