中文
相关论文

相关论文: Proactive Detection of Voice Cloning with Localize…

200 篇论文

We introduce TextSeal, a state-of-the-art watermark for large language models. Building on Gumbel-max sampling, TextSeal introduces dual-key generation to restore output diversity, along with entropy-weighted scoring and multi-region…

Voice cloning (VC)-resistant watermarking is an emerging technique for tracing and preventing unauthorized cloning. Existing methods effectively trace traditional VC models by training them on watermarked audio but fail in zero-shot VC…

声音 · 计算机科学 2025-06-02 Haiyun Li , Zhiyong Wu , Xiaofeng Xie , Jingran Xie , Yaoxun Xu , Hanyang Peng

Generative models have rapidly evolved to generate realistic outputs. However, their synthetic outputs increasingly challenge the clear distinction between natural and AI-generated content, necessitating robust watermarking techniques.…

机器学习 · 计算机科学 2026-05-20 Kasra Arabi , R. Teal Witter , Chinmay Hegde , Niv Cohen

Nowadays, it is common to release audio content to the public. However, with the rise of voice cloning technology, attackers have the potential to easily impersonate a specific person by utilizing his publicly released audio without any…

声音 · 计算机科学 2023-12-07 Chang Liu , Jie Zhang , Tianwei Zhang , Xi Yang , Weiming Zhang , Nenghai Yu

The rapid advancement of generative AI has made it increasingly challenging to distinguish between deepfake audio and authentic human speech. To overcome the limitations of passive detection methods, we propose StreamMark, a novel deep…

音频与语音处理 · 电气工程与系统科学 2026-04-15 Zhentao Liu , Milos Cernak

The availability of high-quality, AI-generated audio raises security challenges such as misinformation campaigns and voice-cloning fraud. A key defense against the misuse of AI-generated audio is by watermarking it, so that it can be easily…

声音 · 计算机科学 2026-05-20 Kexin Li , Xiao Hu , Ilya Grishchenko , David Lie

The advancements in audio generative models have opened up new challenges in their responsible disclosure and the detection of their misuse. In response, we introduce a method to watermark latent generative models by a specific watermarking…

声音 · 计算机科学 2024-09-05 Robin San Roman , Pierre Fernandez , Antoine Deleforge , Yossi Adi , Romain Serizel

Audio watermarking is increasingly used to verify the provenance of AI-generated content, enabling applications such as detecting AI-generated speech, protecting music IP, and defending against voice cloning. To be effective, audio…

密码学与安全 · 计算机科学 2025-03-28 Yizhu Wen , Ashwin Innuganti , Aaron Bien Ramos , Hanqing Guo , Qiben Yan

Recent breakthroughs in zero-shot voice synthesis have enabled imitating a speaker's voice using just a few seconds of recording while maintaining a high level of realism. Alongside its potential benefits, this powerful technology…

声音 · 计算机科学 2024-01-09 Guangyu Chen , Yu Wu , Shujie Liu , Tao Liu , Xiaoyong Du , Furu Wei

The rapid proliferation of generative audio synthesis and editing technologies has raised serious concerns about copyright infringement, data provenance, and the spread of misinformation via deepfake audio. Watermarking offers a proactive…

声音 · 计算机科学 2026-05-25 Yixin Liu , Lie Lu , Jihui Jin , Lichao Sun , Andrea Fanelli

The increasing realism of synthetic speech, driven by advancements in text-to-speech models, raises ethical concerns regarding impersonation and disinformation. Audio watermarking offers a promising solution via embedding…

机器学习 · 计算机科学 2024-11-14 Hongbin Liu , Moyang Guo , Zhengyuan Jiang , Lun Wang , Neil Zhenqiang Gong

Advances in neural speech synthesis have brought us technology that is not only close to human naturalness, but is also capable of instant voice cloning with little data, and is highly accessible with pre-trained models available.…

音频与语音处理 · 电气工程与系统科学 2024-01-03 Lauri Juvela , Xin Wang

The rapid advancement of generative artificial intelligence (GenAI) has revolutionized content creation across text, visual, and audio domains, simultaneously introducing significant risks such as misinformation, identity fraud, and content…

密码学与安全 · 计算机科学 2025-04-08 Lele Cao

As deep learning advances in audio generation, challenges in audio security and copyright protection highlight the need for robust audio watermarking. Recent neural network-based methods have made progress but still face three main issues:…

声音 · 计算机科学 2025-06-09 Yaoxun Xu , Jianwei Yu , Hangting Chen , Zhiyong Wu , Xixin Wu , Dong Yu , Rongzhi Gu , Yi Luo

Speech watermarking techniques can proactively mitigate the potential harmful consequences of instant voice cloning techniques. These techniques involve the insertion of signals into speech that are imperceptible to humans but can be…

音频与语音处理 · 电气工程与系统科学 2024-12-19 Shengpeng Ji , Ziyue Jiang , Jialong Zuo , Minghui Fang , Yifu Chen , Tao Jin , Zhou Zhao

The rapid advancement of voice generation technologies has enabled the synthesis of speech that is perceptually indistinguishable from genuine human voices. While these innovations facilitate beneficial applications such as personalized…

密码学与安全 · 计算机科学 2025-07-30 Aditya Pujari , Ajita Rattani

There are growing implications surrounding generative AI in the speech domain that enable voice cloning and real-time voice conversion from one individual to another. This technology poses a significant ethical threat and could lead to…

声音 · 计算机科学 2023-08-25 Jordan J. Bird , Ahmad Lotfi

Automatic detection of synthetic speech is becoming increasingly important as current synthesis methods are both near indistinguishable from human speech and widely accessible to the public. Audio watermarking and other active disclosure…

声音 · 计算机科学 2024-09-23 Lauri Juvela , Xin Wang

This paper presents AquaSignal, a modular and scalable pipeline for preprocessing, denoising, classification, and novelty detection of underwater acoustic signals. Designed to operate effectively in noisy and dynamic marine environments,…

声音 · 计算机科学 2025-05-21 Eirini Panteli , Paulo E. Santos , Nabil Humphrey

Diffusion models have gained prominence in the image domain for their capabilities in data generation and transformation, achieving state-of-the-art performance in various tasks in both image and audio domains. In the rapidly evolving field…

声音 · 计算机科学 2023-11-02 Xirong Cao , Xiang Li , Divyesh Jadav , Yanzhao Wu , Zhehui Chen , Chen Zeng , Wenqi Wei
‹ 上一页 1 2 3 10 下一页 ›