中文

VILLAIN 在 AVerImaTeC 中的表现:基于多智能体协作的图像-文本声明验证

计算与语言 2026-02-23 v2 人工智能 计算机与社会

摘要

本文描述了 VILLAIN,一个通过基于提示的多智能体协作来验证图像-文本声明的多模态事实核查系统。为完成 AVerImaTeC 共享任务,VILLAIN 采用跨多个阶段的事实核查中使用的视觉-语言模型智能体。文本与视觉证据从经由额外网络收集丰富的知识库中检索。为识别关键信息并解决证据条目之间的不一致性,模态特定智能体与跨模态智能体生成分析报告。在后续阶段,基于这些报告生成问答对。最终, verdict 预测智能体根据图像-文本声明及生成的问答对输出核查结果。我们的系统在所有评估指标上均位居榜单第一。源代码已公开于 https://github.com/ssu-humane/VILLAIN。

关键词

引用

@article{arxiv.2602.04587,
  title  = {VILLAIN at AVerImaTeC: Verifying Image-Text Claims via Multi-Agent Collaboration},
  author = {Jaeyoon Jung and Yejun Yoon and Kunwoo Park},
  journal= {arXiv preprint arXiv:2602.04587},
  year   = {2026}
}

备注

A system description paper for the AVerImaTeC shared task at the Ninth FEVER Workshop (co-located with EACL 2026)