English
Related papers

Related papers: Lips Don't Lie: A Generalisable and Robust Approac…

200 papers

This paper presents a generalized and robust face manipulation detection method based on the edge region features appearing in images. Most contemporary face synthesis processes include color awkwardness reduction but damage the natural…

Computer Vision and Pattern Recognition · Computer Science 2021-12-02 Dong-Keon Kim , Kwangsu Kim

The rise of manipulated media has made deepfakes a particularly insidious threat, involving various generative manipulations such as lip-sync modifications, face-swaps, and avatar-driven facial synthesis. Conventional detection methods,…

Computer Vision and Pattern Recognition · Computer Science 2025-10-17 Soumyya Kanti Datta , Tanvi Ranga , Chengzhe Sun , Siwei Lyu

Lip-reading models have been significantly improved recently thanks to powerful deep learning architectures. However, most works focused on frontal or near frontal views of the mouth. As a consequence, lip-reading performance seriously…

Computer Vision and Pattern Recognition · Computer Science 2019-11-15 Shiyang Cheng , Pingchuan Ma , Georgios Tzimiropoulos , Stavros Petridis , Adrian Bulat , Jie Shen , Maja Pantic

As deepfake technologies continue to advance, passive detection methods struggle to generalize with various forgery manipulations and datasets. Proactive defense techniques have been actively studied with the primary aim of preventing…

Computer Vision and Pattern Recognition · Computer Science 2025-04-16 Hongbo Li , Shangchao Yang , Ruiyang Xia , Lin Yuan , Xinbo Gao

Recent advances in artificial intelligence make it progressively hard to distinguish between genuine and counterfeit media, especially images and videos. One recent development is the rise of deepfake videos, based on manipulating videos…

Computer Vision and Pattern Recognition · Computer Science 2021-08-19 Rashmiranjan Das , Gaurav Negi , Alan F. Smeaton

Existing deepfake detection research has primarily focused on scenarios where the manipulated subject is actively speaking, i.e., generating fabricated content by altering the speaker's appearance or voice. However, in realistic interaction…

Computer Vision and Pattern Recognition · Computer Science 2026-04-15 Miao Liu , Fangda Wei , Jing Wang , Xinyuan Qian

Recent DeepFake detection methods have shown excellent performance on public datasets but are significantly degraded on new forgeries. Solving this problem is important, as new forgeries emerge daily with the continuously evolving…

Computer Vision and Pattern Recognition · Computer Science 2024-08-20 Qingxuan Lv , Yuezun Li , Junyu Dong , Sheng Chen , Hui Yu , Huiyu Zhou , Shu Zhang

Deepfake has emerged for several years, yet efficient detection techniques could generalize over different manipulation methods require further research. While current image-level detection method fails to generalize to unseen domains,…

Computer Vision and Pattern Recognition · Computer Science 2023-11-21 Beilin Chu , Xuan Xu , Weike You , Linna Zhou

Video editing-based talking face generation aims to preserve video details such as pose, lighting, and gestures while modifying only lip motion, often using an identity reference image to maintain speaker consistency. However, this…

Computer Vision and Pattern Recognition · Computer Science 2026-02-11 Dogucan Yaman , Fevziye Irem Eyiokur , Hazım Kemal Ekenel , Alexander Waibel

The rapid emergence of multimodal deepfakes (visual and auditory content are manipulated in concert) undermines the reliability of existing detectors that rely solely on modality-specific artifacts or cross-modal inconsistencies. In this…

Computer Vision and Pattern Recognition · Computer Science 2025-05-22 Yuxuan Du , Zhendong Wang , Yuhao Luo , Caiyong Piao , Zhiyuan Yan , Hao Li , Li Yuan

Lip-reading aims to recognize speech content from videos via visual analysis of speakers' lip movements. This is a challenging task due to the existence of homophemes-words which involve identical or highly similar lip movements, as well as…

Computer Vision and Pattern Recognition · Computer Science 2019-09-04 Chenhao Wang

Visual speech recognition (VSR), commonly known as lip reading, has garnered significant attention due to its wide-ranging practical applications. The advent of deep learning techniques and advancements in hardware capabilities have…

Computer Vision and Pattern Recognition · Computer Science 2025-01-09 Bowen Hao , Dongliang Zhou , Xiaojie Li , Xingyu Zhang , Liang Xie , Jianlong Wu , Erwei Yin

Multimodal deepfakes involving audiovisual manipulations are a growing threat because they are difficult to detect with the naked eye or using unimodal deep learningbased forgery detection methods. Audiovisual forensic models, while more…

Computer Vision and Pattern Recognition · Computer Science 2024-11-15 Sahibzada Adil Shahzad , Ammarah Hashmi , Yan-Tsung Peng , Yu Tsao , Hsin-Min Wang

Deep Learning as a field has been successfully used to solve a plethora of complex problems, the likes of which we could not have imagined a few decades back. But as many benefits as it brings, there are still ways in which it can be used…

Computer Vision and Pattern Recognition · Computer Science 2021-06-25 Samay Pashine , Sagar Mandiya , Praveen Gupta , Rashid Sheikh

Face manipulation techniques have achieved significant advances, presenting serious challenges to security and social trust. Recent works demonstrate that leveraging multimodal models can enhance the generalization and interpretability of…

Computer Vision and Pattern Recognition · Computer Science 2025-03-03 Ke Sun , Shen Chen , Taiping Yao , Ziyin Zhou , Jiayi Ji , Xiaoshuai Sun , Chia-Wen Lin , Rongrong Ji

Face enhancement techniques are widely used to enhance facial appearance. However, they can inadvertently distort biometric features, leading to significant decrease in the accuracy of deepfake detectors. This study hypothesizes that these…

Computer Vision and Pattern Recognition · Computer Science 2025-09-10 Muhammad Saad Saeed , Ijaz Ul Haq , Khalid Malik

The misuse of deepfake technology by malicious actors poses a potential threat to nations, societies, and individuals. However, existing methods for detecting deepfakes primarily focus on uncompressed videos, such as noise characteristics,…

Computer Vision and Pattern Recognition · Computer Science 2024-04-30 Zongmei Chen , Xin Liao , Xiaoshuai Wu , Yanxiang Chen

Although modern face verification systems are accessible and accurate, they are not always robust to pose variance and occlusions. Moreover, accurate models require a large amount of data to train. We structure our experiments to operate on…

Computer Vision and Pattern Recognition · Computer Science 2018-11-16 Kaushal Bhogale , Nishant Shankar , Adheesh Juvekar , Asutosh Padhi

We introduce Forensim, an attention-based state-space framework for image forgery detection that jointly localizes both manipulated (target) and source regions. Unlike traditional approaches that rely solely on artifact cues to detect…

Computer Vision and Pattern Recognition · Computer Science 2026-02-11 Soumyaroop Nandi , Prem Natarajan

We aim to edit the lip movements in talking video according to the given speech while preserving the personal identity and visual details. The task can be decomposed into two sub-problems: (1) speech-driven lip motion generation and (2)…

Computer Vision and Pattern Recognition · Computer Science 2024-06-18 Runyi Yu , Tianyu He , Ailing Zhang , Yuchi Wang , Junliang Guo , Xu Tan , Chang Liu , Jie Chen , Jiang Bian