English

Multimodal Storytelling via Generative Adversarial Imitation Learning

Artificial Intelligence 2017-12-06 v1 Computation and Language Computer Vision and Pattern Recognition

Abstract

Deriving event storylines is an effective summarization method to succinctly organize extensive information, which can significantly alleviate the pain of information overload. The critical challenge is the lack of widely recognized definition of storyline metric. Prior studies have developed various approaches based on different assumptions about users' interests. These works can extract interesting patterns, but their assumptions do not guarantee that the derived patterns will match users' preference. On the other hand, their exclusiveness of single modality source misses cross-modality information. This paper proposes a method, multimodal imitation learning via generative adversarial networks(MIL-GAN), to directly model users' interests as reflected by various data. In particular, the proposed model addresses the critical challenge by imitating users' demonstrated storylines. Our proposed model is designed to learn the reward patterns given user-provided storylines and then applies the learned policy to unseen data. The proposed approach is demonstrated to be capable of acquiring the user's implicit intent and outperforming competing methods by a substantial margin with a user study.

Keywords

Cite

@article{arxiv.1712.01455,
  title  = {Multimodal Storytelling via Generative Adversarial Imitation Learning},
  author = {Zhiqian Chen and Xuchao Zhang and Arnold P. Boedihardjo and Jing Dai and Chang-Tien Lu},
  journal= {arXiv preprint arXiv:1712.01455},
  year   = {2017}
}

Comments

IJCAI 2017

R2 v1 2026-06-22T23:06:51.724Z