中文

量化神经描述生成器所使用的视觉信息量

神经与进化计算 2019-02-05 v1 计算与语言

摘要

本文探讨神经图像描述生成器对其视觉输入的敏感性。报告了基于图像干扰项的敏感性分析与遗漏分析,表明图像描述架构保留并对视觉信息敏感的程度,随所生成词的类型及在整句描述中的位置而变化。我们在该领域实现更高 AI 可解释性的更广义目标背景下阐述此项工作。

关键词

引用

@article{arxiv.1810.05475,
  title  = {Quantifying the amount of visual information used by neural caption generators},
  author = {Marc Tanti and Albert Gatt and Kenneth P. Camilleri},
  journal= {arXiv preprint arXiv:1810.05475},
  year   = {2019}
}

备注

10 pages, 4 figures This publication will appear in the Proceedings of the First Workshop on Shortcomings in Vision and Language (2018). DOI to be inserted later