中文
相关论文

相关论文: Voice Spoofing Detection Corpus for Single and Mul…

200 篇论文

Recent speech and audio coding standards such as 3GPP Enhanced Voice Services match the foreseeable needs and requirements in transmission of speech and audio, when using current transmission infrastructure and applications. Trends in…

音频与语音处理 · 电气工程与系统科学 2018-11-15 Tom Bäckström

Deep speech classification has achieved tremendous success and greatly promoted the emergence of many real-world applications. However, backdoor attacks present a new security threat to it, particularly with untrustworthy third-party…

声音 · 计算机科学 2023-08-15 Zhe Ye , Terui Mao , Li Dong , Diqun Yan

For practical automatic speaker verification (ASV) systems, replay attack poses a true risk. By replaying a pre-recorded speech signal of the genuine speaker, ASV systems tend to be easily fooled. An effective replay detection method is…

声音 · 计算机科学 2017-06-08 Lantian Li , Yixiang Chen , Dong Wang , Thomas Fang Zheng

IoT networks are increasingly becoming target of sophisticated new cyber-attacks. Anomaly-based detection methods are promising in finding new attacks, but there are certain practical challenges like false-positive alarms, hard to explain,…

密码学与安全 · 计算机科学 2023-04-12 Ayyoob Hamza , Hassan Habibi Gharakheili , Theophilus A. Benson , Gustavo Batista , Vijay Sivaraman

Internet-of-Things (IoT) devices are increasingly deployed at home, at work, and in other shared and public spaces. IoT devices collect and share data with service providers and third parties, which poses privacy concerns. Although privacy…

人机交互 · 计算机科学 2024-09-12 Jad Al Aaraj , Olivia Figueira , Tu Le , Isabela Figueira , Rahmadi Trimananda , Athina Markopoulou

Internet of Things (IoT) devices have grown in popularity since they can directly interact with the real world. Home automation systems automate these interactions. IoT events are crucial to these systems' decision-making but are often…

密码学与安全 · 计算机科学 2024-07-30 Uzma Maroof , Gustavo Batista , Arash Shaghaghi , Sanjay Jha

ASVspoof5, the fifth edition of the ASVspoof series, is one of the largest global audio security challenges. It aims to advance the development of countermeasure (CM) to discriminate bonafide and spoofed speech utterances. In this paper, we…

声音 · 计算机科学 2024-08-14 Yuankun Xie , Xiaopeng Wang , Zhiyong Wang , Ruibo Fu , Zhengqi Wen , Haonan Cheng , Long Ye

The state-of-art models for speech synthesis and voice conversion are capable of generating synthetic speech that is perceptually indistinguishable from bonafide human speech. These methods represent a threat to the automatic speaker…

机器学习 · 计算机科学 2019-07-11 Moustafa Alzantot , Ziqi Wang , Mani B. Srivastava

Voice-Controllable Devices (VCDs) have seen an increasing trend towards their adoption due to the small form factor of the MEMS microphones and their easy integration into modern gadgets. Recent studies have revealed that MEMS microphones…

音频与语音处理 · 电气工程与系统科学 2023-10-17 Hashim Ali , Dhimant Khuttan , Rafi Ud Daula Refat , Hafiz Malik

Modern voice cloning, also known as zero-shot text-to-speech (TTS), can synthesize speech that closely matches a target speaker from only seconds of reference audio, enabling applications such as personalized speech interfaces and dubbing.…

声音 · 计算机科学 2026-05-26 Ruinan Jin , Xinting Liao , Hanlin Yu , Deval Pandya , Xiaoxiao Li

IoT devices fundamentally lack built-in security mechanisms to protect themselves from security attacks. Existing works on improving IoT security mostly focus on detecting anomalous behaviors of IoT devices. However, these existing anomaly…

密码学与安全 · 计算机科学 2024-04-23 Md Mainuddin , Zhenhai Duan , Yingfei Dong

Audio deepfake detection is an emerging active topic. A growing number of literatures have aimed to study deepfake detection algorithms and achieved effective performance, the problem of which is far from being solved. Although there are…

声音 · 计算机科学 2023-08-30 Jiangyan Yi , Chenglong Wang , Jianhua Tao , Xiaohui Zhang , Chu Yuan Zhang , Yan Zhao

Internet-of-Things (IoT) devices that are limited in power and processing are susceptible to physical layer (PHY) spoofing (signal exploitation) attacks owing to their inability to implement a full-blown protocol stack for security. The…

机器学习 · 计算机科学 2021-03-18 Alireza Nooraiepour , Waheed U. Bajwa , Narayan B. Mandayam

Voice Authentication (VA), also known as Automatic Speaker Verification (ASV), is a widely adopted authentication method, particularly in automated systems like banking services, where it serves as a secondary layer of user authentication.…

密码学与安全 · 计算机科学 2025-02-14 Eshaq Jamdar , Amith Kamath Belman

Advancements in AI-synthesized human voices have created a growing threat of impersonation and disinformation, making it crucial to develop methods to detect synthetic human voices. This study proposes a new approach to identifying…

声音 · 计算机科学 2023-04-28 Chengzhe Sun , Shan Jia , Shuwei Hou , Siwei Lyu

Neural speech editing advancements have raised concerns about their misuse in spoofing attacks. Traditional partially edited speech corpora primarily focus on cut-and-paste edits, which, while maintaining speaker consistency, often…

The current speech anti-spoofing countermeasures (CMs) show excellent performance on specific datasets. However, removing the silence of test speech through Voice Activity Detection (VAD) can severely degrade performance. In this paper, the…

音频与语音处理 · 电气工程与系统科学 2023-09-22 Yuxiang Zhang , Zhuo Li , Jingze Lu , Hua Hua , Wenchao Wang , Pengyuan Zhang

In this paper, we initiate the concern of enhancing the spoofing robustness of the automatic speaker verification (ASV) system, without the primary presence of a separate countermeasure module. We start from the standard ASV framework of…

声音 · 计算机科学 2022-04-27 Xuechen Liu , Md Sahidullah , Tomi Kinnunen

This paper describes our submitted systems to the 2022 ADD challenge withing the tracks 1 and 2. Our approach is based on the combination of a pre-trained wav2vec2 feature extractor and a downstream classifier to detect spoofed audio. This…

音频与语音处理 · 电气工程与系统科学 2022-03-04 Juan M. Martín-Doñas , Aitor Álvarez

Mainstream zero-shot TTS production systems like Voicebox and Seed-TTS achieve human parity speech by leveraging Flow-matching and Diffusion models, respectively. Unfortunately, human-level audio synthesis leads to identity misuse and…