中文
相关论文

相关论文: Continual Audio Deepfake Detection via Universal A…

200 篇论文

With just a few speech samples, it is possible to perfectly replicate a speaker's voice in recent years, while malicious voice exploitation (e.g., telecom fraud for illegal financial gain) has brought huge hazards in our daily lives.…

声音 · 计算机科学 2024-10-29 Zhisheng Zhang , Qianyi Yang , Derui Wang , Pengyang Huang , Yuxin Cao , Kai Ye , Jie Hao

Deep learning-based time series models are being extensively utilized in engineering and manufacturing industries for process control and optimization, asset monitoring, diagnostic and predictive maintenance. These models have shown great…

机器学习 · 计算机科学 2021-09-16 Arghya Basak , Pradeep Rathore , Sri Harsha Nistala , Sagar Srinivas , Venkataramana Runkana

Ensuring the reliability of face recognition systems against presentation attacks necessitates the deployment of face anti-spoofing techniques. Despite considerable advancements in this domain, the ability of even the most state-of-the-art…

计算机视觉与模式识别 · 计算机科学 2024-04-26 Jiawei Chen , Xiao Yang , Heng Yin , Mingzhi Ma , Bihui Chen , Jianteng Peng , Yandong Guo , Zhaoxia Yin , Hang Su

Deepfake detectors face growing challenges in generalization as new image synthesis techniques emerge. In particular, deepfakes generated by diffusion models are highly photorealistic and often evade detectors trained on GAN-based…

计算机视觉与模式识别 · 计算机科学 2026-04-17 Hongyuan Qi , Wenjin Hou , Hehe Fan , Jun Xiao

Voice spoofing attacks pose a significant threat to automated speaker verification systems. Existing anti-spoofing methods often simulate specific attack types, such as synthetic or replay attacks. However, in real-world scenarios, the…

声音 · 计算机科学 2023-09-19 Awais Khan , Khalid Mahmood Malik , Shah Nawaz

Various forefront countermeasure methods for automatic speaker verification (ASV) with considerable performance in anti-spoofing are proposed in the ASVspoof 2019 challenge. However, previous work has shown that countermeasure models are…

音频与语音处理 · 电气工程与系统科学 2020-03-09 Haibin Wu , Songxiang Liu , Helen Meng , Hung-yi Lee

It is now well established from a variety of studies that there is a significant benefit from combining video and audio data in detecting active speakers. However, either of the modalities can potentially mislead audiovisual fusion by…

Recently, studies show that deep learning-based automatic speech recognition (ASR) systems are vulnerable to adversarial examples (AEs), which add a small amount of noise to the original audio examples. These AE attacks pose new challenges…

音频与语音处理 · 电气工程与系统科学 2023-04-21 Feng Guo , Zheng Sun , Yuxuan Chen , Lei Ju

This work uses adversarial perturbations to enhance deepfake images and fool common deepfake detectors. We created adversarial perturbations using the Fast Gradient Sign Method and the Carlini and Wagner L2 norm attack in both blackbox and…

计算机视觉与模式识别 · 计算机科学 2020-05-18 Apurva Gandhi , Shomik Jain

The SAFE Challenge evaluates synthetic speech detection across three tasks: unmodified audio, processed audio with compression artifacts, and laundered audio designed to evade detection. We systematically explore self-supervised learning…

音频与语音处理 · 电气工程与系统科学 2025-10-08 Hashim Ali , Surya Subramani , Lekha Bollinani , Nithin Sai Adupa , Sali El-Loh , Hafiz Malik

Deep neural networks (DNNs) have risen to prominence as key solutions in numerous AI applications for earth observation (AI4EO). However, their susceptibility to adversarial examples poses a critical challenge, compromising the reliability…

计算机视觉与模式识别 · 计算机科学 2024-05-28 Weikang Yu , Yonghao Xu , Pedram Ghamisi

We propose a method to generate audio adversarial examples that can attack a state-of-the-art speech recognition model in the physical world. Previous work assumes that generated adversarial examples are directly fed to the recognition…

机器学习 · 计算机科学 2019-08-20 Hiromu Yakura , Jun Sakuma

With the rise in manipulated media, deepfake detection has become an imperative task for preserving the authenticity of digital content. In this paper, we present a novel multi-modal audio-video framework designed to concurrently process…

计算机视觉与模式识别 · 计算机科学 2023-09-14 Aaditya Kharel , Manas Paranjape , Aniket Bera

We study universal deepfake detection. Our goal is to detect synthetic images from a range of generative AI approaches, particularly from emerging ones which are unseen during training of the deepfake detector. Universal deepfake detection…

计算机视觉与模式识别 · 计算机科学 2024-01-18 Chandler Timm Doloriel , Ngai-Man Cheung

The widespread use of generative AI has shown remarkable success in producing highly realistic deepfakes, posing a serious threat to various voice biometric applications, including speaker verification, voice biometrics, audio conferencing,…

声音 · 计算机科学 2025-09-10 Kutub Uddin , Muhammad Umar Farooq , Awais Khan , Khalid Mahmood Malik

Audio deepfakes are acquiring an unprecedented level of realism with advanced AI. While current research focuses on discerning real speech from spoofed speech, tracing the source system is equally crucial. This work proposes a novel audio…

声音 · 计算机科学 2025-06-04 Ajinkya Kulkarni , Sandipana Dowerah , Tanel Alumae , Mathew Magimai. -Doss

Perception module of Autonomous vehicles (AVs) are increasingly susceptible to be attacked, which exploit vulnerabilities in neural networks through adversarial inputs, thereby compromising the AI safety. Some researches focus on creating…

计算机视觉与模式识别 · 计算机科学 2025-02-13 Yuanhao Huang , Qinfan Zhang , Jiandong Xing , Mengyue Cheng , Haiyang Yu , Yilong Ren , Xiao Xiong

To develop a machine sound monitoring system, a method for detecting anomalous sound is proposed. In this paper, we explore a method for multiple clients to collaboratively learn an anomalous sound detection model while keeping their raw…

音频与语音处理 · 电气工程与系统科学 2024-03-26 Kota Dohi , Yohei Kawaguchi

In real-world applications, it is challenging to build a speaker verification system that is simultaneously robust against common threats, including spoofing attacks, channel mismatch, and domain mismatch. Traditional automatic speaker…

音频与语音处理 · 电气工程与系统科学 2024-09-11 Chang Zeng , Xiaoxiao Miao , Xin Wang , Erica Cooper , Junichi Yamagishi

Deepfakes are synthetically generated media often devised with malicious intent. They have become increasingly more convincing with large training datasets advanced neural networks. These fakes are readily being misused for slander,…

密码学与安全 · 计算机科学 2022-03-30 Nicolas M. Müller , Franziska Dieckmann , Jennifer Williams
‹ 上一页 1 8 9 10 下一页 ›