English
Related papers

Related papers: Few-Shot Head Swapping in the Wild

200 papers

Recently, few-shot object detection~(FSOD) has received much attention from the community, and many methods are proposed to address this problem from a knowledge transfer perspective. Though promising results have been achieved, these…

Computer Vision and Pattern Recognition · Computer Science 2022-10-06 Zhiyuan Zhao , Qingjie Liu , Yunhong Wang

The task of face reenactment is to transfer the head motion and facial expressions from a driving video to the appearance of a source image, which may be of a different person (cross-reenactment). Most existing methods are CNN-based and…

Computer Vision and Pattern Recognition · Computer Science 2024-06-11 Andre Rochow , Max Schwarz , Sven Behnke

Speech-driven 3D facial animation is important for many multimedia applications. Recent work has shown promise in using either Diffusion models or Transformer architectures for this task. However, their mere aggregation does not lead to…

Computer Vision and Pattern Recognition · Computer Science 2024-02-09 Zhiyuan Ma , Xiangyu Zhu , Guojun Qi , Chen Qian , Zhaoxiang Zhang , Zhen Lei

Makeup transfer is the task of applying on a source face the makeup style from a reference image. Real-life makeups are diverse and wild, which cover not only color-changing but also patterns, such as stickers, blushes, and jewelries.…

Computer Vision and Pattern Recognition · Computer Science 2021-04-06 Thao Nguyen , Anh Tran , Minh Hoai

We aim to bridge the gap between typical human and machine-learning environments by extending the standard framework of few-shot learning to an online, continual setting. In this setting, episodes do not have separate training and testing…

Machine Learning · Computer Science 2021-04-26 Mengye Ren , Michael L. Iuzzolino , Michael C. Mozer , Richard S. Zemel

In this work, we propose a semantic flow-guided two-stage framework for shape-aware face swapping, namely FlowFace. Unlike most previous methods that focus on transferring the source inner facial features but neglect facial contours, our…

Computer Vision and Pattern Recognition · Computer Science 2022-12-07 Hao Zeng , Wei Zhang , Changjie Fan , Tangjie Lv , Suzhen Wang , Zhimeng Zhang , Bowen Ma , Lincheng Li , Yu Ding , Xin Yu

Very recently, Window-based Transformers, which computed self-attention within non-overlapping local windows, demonstrated promising results on image classification, semantic segmentation, and object detection. However, less study has been…

Computer Vision and Pattern Recognition · Computer Science 2021-06-08 Zilong Huang , Youcheng Ben , Guozhong Luo , Pei Cheng , Gang Yu , Bin Fu

We introduce an industrial Head Blending pipeline for the task of seamlessly integrating an actor's head onto a target body in digital content creation. The key challenge stems from discrepancies in head shape and hair structure, which lead…

Computer Vision and Pattern Recognition · Computer Science 2024-11-04 Hah Min Lew , Sahng-Min Yoo , Hyunwoo Kang , Gyeong-Moon Park

Cross-subject EEG emotion recognition is challenged by significant inter-subject variability and intricately entangled intra-subject variability. Existing works have primarily addressed these challenges through domain adaptation or…

Image and Video Processing · Electrical Eng. & Systems 2025-03-26 Haiqi Liu , C. L. Philip Chen , Tong Zhang

The human face is central to communication. For immersive applications, the digital presence of a person should mirror the physical reality, capturing the users idiosyncrasies and detailed facial expressions. However, current 3D head avatar…

Computer Vision and Pattern Recognition · Computer Science 2026-04-16 Jalees Nehvi , Timo Bolkart , Thabo Beeler , Justus Thies

Few-shot image generation and few-shot image translation are two related tasks, both of which aim to generate new images for an unseen category with only a few images. In this work, we make the first attempt to adapt few-shot image…

Computer Vision and Pattern Recognition · Computer Science 2022-07-25 Yan Hong , Li Niu , Jianfu Zhang , Liqing Zhang

Existing face swap methods rely heavily on large-scale networks for adequate capacity to generate visually plausible results, which inhibits its applications on resource-constraint platforms. In this work, we propose MobileFSGAN, a novel…

Computer Vision and Pattern Recognition · Computer Science 2022-04-19 Haiming Yu , Hao Zhu , Xiangju Lu , Junhui Liu

In this paper, we propose a novel encoder, called ShapeEditor, for high-resolution, realistic and high-fidelity face exchange. First of all, in order to ensure sufficient clarity and authenticity, our key idea is to use an advanced…

Computer Vision and Pattern Recognition · Computer Science 2021-06-29 Shuai Yang , Kai Qiao

Large-scale in-the-wild speech datasets have become more prevalent in recent years due to increased interest in models that can learn useful features from unlabelled data for tasks such as speech recognition or synthesis. These datasets…

This paper presents FSNet, a deep generative model for image-based face swapping. Traditionally, face-swapping methods are based on three-dimensional morphable models (3DMMs), and facial textures are replaced between the estimated…

Computer Vision and Pattern Recognition · Computer Science 2022-07-04 Ryota Natsume , Tatsuya Yatagawa , Shigeo Morishima

Deep fake technology became a hot field of research in the last few years. Researchers investigate sophisticated Generative Adversarial Networks (GAN), autoencoders, and other approaches to establish precise and robust algorithms for face…

Computer Vision and Pattern Recognition · Computer Science 2022-02-08 Daniil Chesakov , Anastasia Maltseva , Alexander Groshev , Andrey Kuznetsov , Denis Dimitrov

Few-shot classification which aims to recognize unseen classes using very limited samples has attracted more and more attention. Usually, it is formulated as a metric learning problem. The core issue of few-shot classification is how to…

Computer Vision and Pattern Recognition · Computer Science 2022-08-29 Xixi Wang , Xiao Wang , Bo Jiang , Bin Luo

Multi-modal reasoning plays a vital role in bridging the gap between textual and visual information, enabling a deeper understanding of the context. This paper presents the Feature Swapping Multi-modal Reasoning (FSMR) model, designed to…

Computer Vision and Pattern Recognition · Computer Science 2024-04-01 Shuang Li , Jiahua Wang , Lijie Wen

Conventional voice conversion modifies voice characteristics from a source speaker to a target speaker, relying on audio input from both sides. However, this process becomes infeasible when clean audio is unavailable, such as in silent…

Sound · Computer Science 2025-08-05 Yifan Liu , Yu Fang , Zhouhan Lin

Face swapping has gained significant attention for its varied applications. Most previous face swapping approaches have relied on the seesaw game training scheme, also known as the target-oriented approach. However, this often leads to…

Computer Vision and Pattern Recognition · Computer Science 2024-07-23 Jaeseong Lee , Junha Hyung , Sohyun Jeong , Jaegul Choo