中文
相关论文

相关论文: AudioMarkBench: Benchmarking Robustness of Audio W…

200 篇论文

With the development of large models, watermarks are increasingly employed to assert copyright, verify authenticity, or monitor content distribution. As applications become more multimodal, the utility of watermarking techniques becomes…

计算机视觉与模式识别 · 计算机科学 2024-06-07 Jielin Qiu , William Han , Xuandong Zhao , Shangbang Long , Christos Faloutsos , Lei Li

While text-to-image models offer numerous benefits, they also pose significant societal risks. Detecting AI-generated images is crucial for mitigating these risks. Detection methods can be broadly categorized into passive and…

密码学与安全 · 计算机科学 2025-01-13 Moyang Guo , Yuepeng Hu , Zhengyuan Jiang , Zeyu Li , Amir Sadovnik , Arka Daw , Neil Gong

To mitigate the potential misuse of large language models (LLMs), recent research has developed watermarking algorithms, which restrict the generation process to leave an invisible trace for watermark detection. Due to the two-stage nature…

计算与语言 · 计算机科学 2024-07-02 Shangqing Tu , Yuliang Sun , Yushi Bai , Jifan Yu , Lei Hou , Juanzi Li

The surging demand for large-scale datasets in deep learning has heightened the need for effective copyright protection, given the risks of unauthorized use to data owners. Although the dataset watermark technique holds promise for auditing…

密码学与安全 · 计算机科学 2026-02-17 Xiao Ren , Xinyi Yu , Linkang Du , Min Chen , Yuanchao Shu , Zhou Su , Yunjun Gao , Zhikun Zhang

Advances in neural speech synthesis have brought us technology that is not only close to human naturalness, but is also capable of instant voice cloning with little data, and is highly accessible with pre-trained models available.…

音频与语音处理 · 电气工程与系统科学 2024-01-03 Lauri Juvela , Xin Wang

This paper presents the first study on the impact of audio watermarking on spoofing countermeasures. While anti-spoofing systems are essential for securing speech-based applications, the influence of widely used audio watermarking,…

机器学习 · 计算机科学 2025-09-26 Zhenshan Zhang , Xueping Zhang , Yechen Wang , Liwei Jin , Ming Li

Nowadays, it is common to release audio content to the public. However, with the rise of voice cloning technology, attackers have the potential to easily impersonate a specific person by utilizing his publicly released audio without any…

声音 · 计算机科学 2023-12-07 Chang Liu , Jie Zhang , Tianwei Zhang , Xi Yang , Weiming Zhang , Nenghai Yu

The rapid advancement of next-token-prediction models has led to widespread adoption across modalities, enabling the creation of realistic synthetic media. In the audio domain, while autoregressive speech models have propelled…

声音 · 计算机科学 2025-10-27 Yihan Wu , Georgios Milis , Ruibo Chen , Heng Huang

Audio watermarking has been widely applied in copyright protection and source tracing. However, due to the inherent characteristics of audio signals, watermark localization and resistance to desynchronization attacks remain significant…

密码学与安全 · 计算机科学 2025-09-03 Zhenliang Gan , Xiaoxiao Hu , Sheng Li , Zhenxing Qian , Xinpeng Zhang

Recent progress in large language models enables the creation of realistic machine-generated content. Watermarking is a promising approach to distinguish machine-generated text from human text, embedding statistical signals in the output…

密码学与安全 · 计算机科学 2026-02-25 Patrick Chao , Yan Sun , Edgar Dobriban , Hamed Hassani

The rapid advancement of voice generation technologies has enabled the synthesis of speech that is perceptually indistinguishable from genuine human voices. While these innovations facilitate beneficial applications such as personalized…

密码学与安全 · 计算机科学 2025-07-30 Aditya Pujari , Ajita Rattani

The advancement of artificial intelligence generated content (AIGC) has created a pressing need for robust image watermarking that can withstand both conventional signal processing and novel semantic editing attacks. Current deep…

计算机视觉与模式识别 · 计算机科学 2025-11-17 Yichao Tang , Mingyang Li , Di Miao , Sheng Li , Zhenxing Qian , Xinpeng Zhang

Speech watermarking techniques can proactively mitigate the potential harmful consequences of instant voice cloning techniques. These techniques involve the insertion of signals into speech that are imperceptible to humans but can be…

音频与语音处理 · 电气工程与系统科学 2024-12-19 Shengpeng Ji , Ziyue Jiang , Jialong Zuo , Minghui Fang , Yifu Chen , Tao Jin , Zhou Zhao

Recent advances in audio-aware large language models have shown strong performance on audio question answering. However, existing benchmarks mainly cover answerable questions and overlook the challenge of unanswerable ones, where no…

音频与语音处理 · 电气工程与系统科学 2026-05-12 Chun-Yi Kuan , Hung-yi Lee

With the surge of social media, maliciously tampered public speeches, especially those from influential figures, have seriously affected social stability and public trust. Existing speech tampering detection methods remain insufficient:…

密码学与安全 · 计算机科学 2025-06-03 Lingfeng Yao , Chenpei Huang , Shengyao Wang , Junpei Xue , Hanqing Guo , Jiang Liu , Xun Chen , Miao Pan

The rapid advancement of generative artificial intelligence (GenAI) has revolutionized content creation across text, visual, and audio domains, simultaneously introducing significant risks such as misinformation, identity fraud, and content…

密码学与安全 · 计算机科学 2025-04-08 Lele Cao

As generative audio models are rapidly evolving, AI-generated audios increasingly raise concerns about copyright infringement and misinformation spread. Audio watermarking, as a proactive defense, can embed secret messages into audio for…

密码学与安全 · 计算机科学 2025-12-08 Lingfeng Yao , Chenpei Huang , Shengyao Wang , Junpei Xue , Hanqing Guo , Jiang Liu , Phone Lin , Tomoaki Ohtsuki , Miao Pan

We propose a novel audio watermarking system that is robust to the distortion due to the indoor acoustic propagation channel between the loudspeaker and the receiving microphone. The system utilizes a set of new algorithms that effectively…

多媒体 · 计算机科学 2019-03-21 Yuan-Yen Tai , Mohamed F. Mansour

While recent audio-visual models have demonstrated impressive performance, their robustness to distributional shifts at test-time remains not fully understood. Existing robustness benchmarks mainly focus on single modalities, making them…

Generated speech achieves human-level naturalness but escalates security risks of misuse. However, existing watermarking methods fail to reconcile fidelity with robustness, as they rely either on simple superposition in the noise space or…

密码学与安全 · 计算机科学 2026-02-02 Weizhi Liu , Yue Li , Zhaoxia Yin