English

Image Matters: A New Dataset and Empirical Study for Multimodal Hyperbole Detection

Computer Vision and Pattern Recognition 2024-03-12 v3 Artificial Intelligence Computation and Language

Abstract

Hyperbole, or exaggeration, is a common linguistic phenomenon. The detection of hyperbole is an important part of understanding human expression. There have been several studies on hyperbole detection, but most of which focus on text modality only. However, with the development of social media, people can create hyperbolic expressions with various modalities, including text, images, videos, etc. In this paper, we focus on multimodal hyperbole detection. We create a multimodal detection dataset from Weibo (a Chinese social media) and carry out some studies on it. We treat the text and image from a piece of weibo as two modalities and explore the role of text and image for hyperbole detection. Different pre-trained multimodal encoders are also evaluated on this downstream task to show their performance. Besides, since this dataset is constructed from five different topics, we also evaluate the cross-domain performance of different models. These studies can serve as a benchmark and point out the direction of further study on multimodal hyperbole detection.

Keywords

Cite

@article{arxiv.2307.00209,
  title  = {Image Matters: A New Dataset and Empirical Study for Multimodal Hyperbole Detection},
  author = {Huixuan Zhang and Xiaojun Wan},
  journal= {arXiv preprint arXiv:2307.00209},
  year   = {2024}
}

Comments

Accepted by LREC-COLING 2024

R2 v1 2026-06-28T11:19:32.281Z