English

Context-Preserving Two-Stage Video Domain Translation for Portrait Stylization

Computer Vision and Pattern Recognition 2023-05-31 v1

Abstract

Portrait stylization, which translates a real human face image into an artistically stylized image, has attracted considerable interest and many prior works have shown impressive quality in recent years. However, despite their remarkable performances in the image-level translation tasks, prior methods show unsatisfactory results when they are applied to the video domain. To address the issue, we propose a novel two-stage video translation framework with an objective function which enforces a model to generate a temporally coherent stylized video while preserving context in the source video. Furthermore, our model runs in real-time with the latency of 0.011 seconds per frame and requires only 5.6M parameters, and thus is widely applicable to practical real-world applications.

Keywords

Cite

@article{arxiv.2305.19135,
  title  = {Context-Preserving Two-Stage Video Domain Translation for Portrait Stylization},
  author = {Doyeon Kim and Eunji Ko and Hyunsu Kim and Yunji Kim and Junho Kim and Dongchan Min and Junmo Kim and Sung Ju Hwang},
  journal= {arXiv preprint arXiv:2305.19135},
  year   = {2023}
}

Comments

5 pages, 3 figures, CVPR 2023 Workshop on AI for Content Creation