English

Face Translation between Images and Videos using Identity-aware CycleGAN

Computer Vision and Pattern Recognition 2017-12-05 v1

Abstract

This paper presents a new problem of unpaired face translation between images and videos, which can be applied to facial video prediction and enhancement. In this problem there exist two major technical challenges: 1) designing a robust translation model between static images and dynamic videos, and 2) preserving facial identity during image-video translation. To address such two problems, we generalize the state-of-the-art image-to-image translation network (Cycle-Consistent Adversarial Networks) to the image-to-video/video-to-image translation context by exploiting a image-video translation model and an identity preservation model. In particular, we apply the state-of-the-art Wasserstein GAN technique to the setting of image-video translation for better convergence, and we meanwhile introduce a face verificator to ensure the identity. Experiments on standard image/video face datasets demonstrate the effectiveness of the proposed model in both terms of qualitative and quantitative evaluations.

Keywords

Cite

@article{arxiv.1712.00971,
  title  = {Face Translation between Images and Videos using Identity-aware CycleGAN},
  author = {Zhiwu Huang and Bernhard Kratzwald and Danda Pani Paudel and Jiqing Wu and Luc Van Gool},
  journal= {arXiv preprint arXiv:1712.00971},
  year   = {2017}
}