中文
相关论文

相关论文: Pindrop it! Audio and Visual Deepfake Countermeasu…

200 篇论文

Deep generative models can create remarkably photorealistic fake images while raising concerns about misinformation and copyright infringement, known as deepfake threats. Deepfake detection technique is developed to distinguish between real…

计算机视觉与模式识别 · 计算机科学 2024-08-22 You-Ming Chang , Chen Yeh , Wei-Chen Chiu , Ning Yu

Since the invention of cinema, the manipulated videos have existed. But generating manipulated videos that can fool the viewer has been a time-consuming endeavor. With the dramatic improvements in the deep generative modeling, generating…

计算机视觉与模式识别 · 计算机科学 2020-12-09 Ivan Kukanov , Janne Karttunen , Hannu Sillanpää , Ville Hautamäki

Visual Place recognition is commonly addressed as an image retrieval problem. However, retrieval methods are impractical to scale to large datasets, densely sampled from city-wide maps, since their dimension impact negatively on the…

计算机视觉与模式识别 · 计算机科学 2023-12-08 Gabriele Trivigno , Gabriele Berton , Juan Aragon , Barbara Caputo , Carlo Masone

We introduce FakeParts, a new class of deepfakes characterized by subtle, localized manipulations to specific spatial regions or temporal segments of otherwise authentic videos. Unlike fully synthetic content, these partial manipulations -…

计算机视觉与模式识别 · 计算机科学 2025-12-22 Ziyi Liu , Firas Gabetni , Awais Hussain Sani , Xi Wang , Soobash Daiboo , Gaetan Brison , Gianni Franchi , Vicky Kalogeiton

Deepfakes are computer manipulated videos where the face of an individual has been replaced with that of another. Software for creating such forgeries is easy to use and ever more popular, causing serious threats to personal reputation and…

计算机视觉与模式识别 · 计算机科学 2021-05-14 Samuele Pino , Mark James Carman , Paolo Bestagini

Audio-visual deepfakes have reached a level of realism that makes perceptual detection unreliable, threatening media integrity and biometric security. While multimodal detection has shown promise, most approaches are binary classification…

计算机视觉与模式识别 · 计算机科学 2026-04-30 Wasim Ahmad , Wei Zhang , Xuerui Mao

Due to the development of facial manipulation techniques in recent years deepfake detection in video stream became an important problem for face biometrics, brand monitoring or online video conferencing solutions. In case of a biometric…

计算机视觉与模式识别 · 计算机科学 2024-10-11 Kirill Vyshegorodtsev , Dmitry Kudiyarov , Alexander Balashov , Alexander Kuzmin

With the rapid advancement of generative audio models, distinguishing between human-composed and generated music is becoming increasingly challenging. As a response, models for detecting fake music have been proposed. In this work, we…

声音 · 计算机科学 2025-07-15 Tomasz Sroka , Tomasz Wężowicz , Dominik Sidorczuk , Mateusz Modrzejewski

This study introduces LENS-DF, a novel and comprehensive recipe for training and evaluating audio deepfake detection and temporal localization under complicated and realistic audio conditions. The generation part of the recipe outputs…

声音 · 计算机科学 2025-07-25 Xuechen Liu , Wanying Ge , Xin Wang , Junichi Yamagishi

Partially spoofed audio detection is a challenging task, lying in the need to accurately locate the authenticity of audio at the frame level. To address this issue, we propose a fine-grained partially spoofed audio detection method, namely…

声音 · 计算机科学 2023-11-22 Yuankun Xie , Haonan Cheng , Yutian Wang , Long Ye

Multimodal deepfakes involving audiovisual manipulations are a growing threat because they are difficult to detect with the naked eye or using unimodal deep learningbased forgery detection methods. Audiovisual forensic models, while more…

计算机视觉与模式识别 · 计算机科学 2024-11-15 Sahibzada Adil Shahzad , Ammarah Hashmi , Yan-Tsung Peng , Yu Tsao , Hsin-Min Wang

Generalizing deepfake detection to unseen manipulations remains a key challenge. A recent approach to tackle this issue is to train a network with pristine face images that have been manipulated with hand-crafted artifacts to extract more…

计算机视觉与模式识别 · 计算机科学 2026-04-13 Alejandro Cobo , Roberto Valle , José Miguel Buenaposada , Luis Baumela

Recent rapid advancements in deepfake technology have allowed the creation of highly realistic fake media, such as video, image, and audio. These materials pose significant challenges to human authentication, such as impersonation,…

密码学与安全 · 计算机科学 2023-09-12 Binh Le , Shahroz Tariq , Alsharif Abuadbba , Kristen Moore , Simon Woo

In the face of a new era of generative models, the detection of artificially generated content has become a matter of utmost importance. The ability to create credible minute-long music deepfakes in a few seconds on user-friendly platforms…

声音 · 计算机科学 2024-05-24 Darius Afchar , Gabriel Meseguer-Brocal , Romain Hennequin

Audio deepfake detection is an emerging topic in the artificial intelligence community. The second Audio Deepfake Detection Challenge (ADD 2023) aims to spur researchers around the world to build new innovative technologies that can further…

Online media data, in the forms of images and videos, are becoming mainstream communication channels. However, recent advances in deep learning, particularly deep generative models, open the doors for producing perceptually convincing…

计算机视觉与模式识别 · 计算机科学 2022-12-13 Junke Wang , Zhenxin Li , Chao Zhang , Jingjing Chen , Zuxuan Wu , Larry S. Davis , Yu-Gang Jiang

This research explores the positive application of deepfake technology for upper body generation, specifically sign language for the Deaf and Hard of Hearing (DHoH) community. Given the complexity of sign language and the scarcity of…

计算机视觉与模式识别 · 计算机科学 2025-02-18 Shahzeb Naeem , Muhammad Riyyan Khan , Usman Tariq , Abhinav Dhall , Carlos Ivan Colon , Hasan Al-Nashash

How to visually localize multiple sound sources in unconstrained videos is a formidable problem, especially when lack of the pairwise sound-object annotations. To solve this problem, we develop a two-stage audiovisual learning framework…

计算机视觉与模式识别 · 计算机科学 2020-07-15 Rui Qian , Di Hu , Heinrich Dinkel , Mengyue Wu , Ning Xu , Weiyao Lin

Synthetically-generated audios and videos -- so-called deep fakes -- continue to capture the imagination of the computer-graphics and computer-vision communities. At the same time, the democratization of access to technology that can create…

计算机视觉与模式识别 · 计算机科学 2021-01-29 Shruti Agarwal , Tarek El-Gaaly , Hany Farid , Ser-Nam Lim

It is increasingly easy to automatically swap faces in images and video or morph two faces into one using generative adversarial networks (GANs). The high quality of the resulted deep-morph raises the question of how vulnerable the current…

计算机视觉与模式识别 · 计算机科学 2019-10-07 Pavel Korshunov , Sébastien Marcel