无需超越物体即可生成图像描述
计算机视觉与模式识别
2016-10-19 v2 计算与语言
摘要
本文探索了图像描述的新评估视角,并引入了一项名词翻译任务,该任务通过将一组名词翻译为描述,实现了可比较的图像描述生成性能。这意味着在图像描述中,除名词外的所有词类都可以由一个强大的语言模型唤起,而不会牺牲n-gram精度。本文还研究了描述中各个词类对最终BLEU分数贡献的下界和上界。名词、动词和介词存在较大的可能改进空间。
引用
@article{arxiv.1610.03708,
title = {Generating captions without looking beyond objects},
author = {Hendrik Heuer and Christof Monz and Arnold W. M. Smeulders},
journal= {arXiv preprint arXiv:1610.03708},
year = {2016}
}
备注
This paper was presented at the ECCV2016 2nd Workshop on Storytelling with Images and Videos (VisStory)