English
Related papers

Related papers: FaceOff: A Video-to-Video Face Swapping System

200 papers

We present a novel approach for the detection of deepfake videos using a pair of vision transformers pre-trained by a self-supervised masked autoencoding setup. Our method consists of two distinct components, one of which focuses on…

Computer Vision and Pattern Recognition · Computer Science 2024-02-12 Sayantan Das , Mojtaba Kolahdouzi , Levent Özparlak , Will Hickie , Ali Etemad

The increasing demand for large-scale visual data, coupled with strict privacy regulations, has driven research into anonymization methods that hide personal identities without seriously degrading data quality. In this paper, we explore the…

Computer Vision and Pattern Recognition · Computer Science 2025-05-28 Mustafa İzzet Muştu , Hazım Kemal Ekenel

Despite recent advances in text-to-speech (TTS) models, audio-visual-to-audio-visual (AV2AV) translation still faces a critical challenge: maintaining speaker consistency between the original and translated vocal and facial features. To…

Audio and Speech Processing · Electrical Eng. & Systems 2025-07-31 Sungwoo Cho , Jeongsoo Choi , Sungnyun Kim , Se-Young Yun

Deepfakes are a recent off-the-shelf manipulation technique that allows anyone to swap two identities in a single video. In addition to Deepfakes, a variety of GAN-based face swapping methods have also been published with accompanying code.…

Computer Vision and Pattern Recognition · Computer Science 2020-10-29 Brian Dolhansky , Joanna Bitton , Ben Pflaum , Jikuo Lu , Russ Howes , Menglin Wang , Cristian Canton Ferrer

Over the past years, a substantial amount of work has been done on the problem of facial reenactment, with the solutions coming mainly from the graphics community. Head reenactment is an even more challenging task, which aims at…

Computer Vision and Pattern Recognition · Computer Science 2021-03-31 Michail Christos Doukas , Mohammad Rami Koujan , Viktoriia Sharmanska , Stefanos Zafeiriou

Many modern online 3D applications and videogames rely on parametric models of human faces for creating believable avatars. However, manual reproduction of someone's facial likeness with a parametric model is difficult and time-consuming.…

Computer Vision and Pattern Recognition · Computer Science 2022-08-08 Igor Borovikov , Karine Levonyan , Jon Rein , Pawel Wrotek , Nitish Victor

In recent years, the explosive advancement of deepfake technology has posed a critical and escalating threat to public security: diffusion-based digital human generation. Unlike traditional face manipulation methods, such models can…

Computer Vision and Pattern Recognition · Computer Science 2026-01-29 Jiaxin Liu , Jia Wang , Saihui Hou , Min Ren , Huijia Wu , Long Ma , Renwang Pei , Zhaofeng He

This paper addresses the problem of face video inpainting. Existing video inpainting methods target primarily at natural scenes with repetitive patterns. They do not make use of any prior knowledge of the face to help retrieve…

Computer Vision and Pattern Recognition · Computer Science 2023-02-14 Wenqi Yang , Zhenfang Chen , Chaofeng Chen , Guanying Chen , Kwan-Yee K. Wong

This paper introduces StreamV2V, a diffusion model that achieves real-time streaming video-to-video (V2V) translation with user prompts. Unlike prior V2V methods using batches to process limited frames, we opt to process frames in a…

Computer Vision and Pattern Recognition · Computer Science 2025-02-18 Feng Liang , Akio Kodaira , Chenfeng Xu , Masayoshi Tomizuka , Kurt Keutzer , Diana Marculescu

With the rapid progress of video generation, demand for customized video editing is surging, where subject swapping constitutes a key component yet remains under-explored. Prevailing swapping approaches either specialize in narrow…

Computer Vision and Pattern Recognition · Computer Science 2025-08-21 Weitao Wang , Zichen Wang , Hongdeng Shen , Yulei Lu , Xirui Fan , Suhui Wu , Jun Zhang , Haoqian Wang , Hao Zhang

Text-guided image-to-video (I2V) generation aims to generate a coherent video that preserves the identity of the input image and semantically aligns with the input prompt. Existing methods typically augment pretrained text-to-video (T2V)…

Computer Vision and Pattern Recognition · Computer Science 2024-06-28 Xun Guo , Mingwu Zheng , Liang Hou , Yuan Gao , Yufan Deng , Pengfei Wan , Di Zhang , Yufan Liu , Weiming Hu , Zhengjun Zha , Haibin Huang , Chongyang Ma

While the abuse of deepfake technology has caused serious concerns recently, how to detect deepfake videos is still a challenge due to the high photo-realistic synthesis of each frame. Existing image-level approaches often focus on single…

Computer Vision and Pattern Recognition · Computer Science 2022-07-15 Daichi Zhang , Fanzhao Lin , Yingying Hua , Pengju Wang , Dan Zeng , Shiming Ge

Deep learning has been successfully appertained to solve various complex problems in the area of big data analytics to computer vision. A deep learning-powered application recently emerged is Deep Fake. It helps to create fake images and…

Computer Vision and Pattern Recognition · Computer Science 2021-09-28 Jatin Sharma , Sahil Sharma

With rapid advancements in image generation technology, face swapping for privacy protection has emerged as an active area of research. The ultimate benefit is improved access to video datasets, e.g. in healthcare settings. Recent…

Computer Vision and Pattern Recognition · Computer Science 2022-04-14 Ethan Wilson , Frederick Shic , Jenny Skytta , Eakta Jain

It is becoming increasingly easy to automatically replace a face of one person in a video with the face of another person by using a pre-trained generative adversarial network (GAN). Recent public scandals, e.g., the faces of celebrities…

Computer Vision and Pattern Recognition · Computer Science 2018-12-21 Pavel Korshunov , Sebastien Marcel

Face forgery by deepfake is widely spread over the internet and this raises severe societal concerns. In this paper, we propose a novel video transformer with incremental learning for detecting deepfake videos. To better align the input…

Computer Vision and Pattern Recognition · Computer Science 2021-08-12 Sohail A. Khan , Hang Dai

Modeling the reactive tempo of human conversation remains difficult because most audio-visual datasets portray isolated speakers delivering short monologues. We introduce \textbf{Face-to-Face with Jimmy Fallon (F2F-JF)}, a 70-hour, 14k-clip…

Computer Vision and Pattern Recognition · Computer Science 2026-04-01 Ernie Chu , Vishal M. Patel

Face aging has become a crucial task in computer vision, with applications ranging from entertainment to healthcare. However, existing methods struggle with achieving a realistic and seamless transformation across the entire lifespan,…

Computer Vision and Pattern Recognition · Computer Science 2025-10-21 Tao Liu , Dafeng Zhang , Gengchen Li , Shizhuo Liu , Yongqi Song , Senmao Li , Shiqi Yang , Boqian Li , Kai Wang , Yaxing Wang

Do our facial expressions change when we speak over video calls? Given two unpaired sets of videos of people, we seek to automatically find spatio-temporal patterns that are distinctive of each set. Existing methods use discriminative…

Computer Vision and Pattern Recognition · Computer Science 2024-06-04 Sumit Sarin , Utkarsh Mall , Purva Tendulkar , Carl Vondrick

We consider the task of Image-to-Video (I2V) generation, which involves transforming static images into realistic video sequences based on a textual description. While recent advancements produce photorealistic outputs, they frequently…

Computer Vision and Pattern Recognition · Computer Science 2025-01-07 Guy Yariv , Yuval Kirstain , Amit Zohar , Shelly Sheynin , Yaniv Taigman , Yossi Adi , Sagie Benaim , Adam Polyak