English
Related papers

Related papers: WavMark: Watermarking for Audio Generation

200 papers

In recent years, there has been significant advancement in the field of model watermarking techniques. However, the protection of image-processing neural networks remains a challenge, with only a limited number of methods being developed.…

Cryptography and Security · Computer Science 2023-02-20 Huajie Chen , Tianqing Zhu , Chi Liu , Shui Yu , Wanlei Zhou

Backdoor data poisoning is a crucial technique for ownership protection and defending against malicious attacks. Embedding hidden triggers in training data can manipulate model outputs, enabling provenance verification, and deterring…

Audio and Speech Processing · Electrical Eng. & Systems 2026-03-24 Kuan-Yu Chen , Yi-Cheng Lin , Jeng-Lin Li , Jian-Jiun Ding

Modern generative diffusion models rely on vast training datasets, often including images with uncertain ownership or usage rights. Radioactive watermarks -- marks that transfer to a model's outputs -- can help detect when such unauthorized…

Cryptography and Security · Computer Science 2025-12-02 Kexin Li , Guozhen Ding , Ilya Grishchenko , David Lie

As generative AI models produce increasingly realistic output, both academia and industry are focusing on the ability to detect whether an output was generated by an AI model or not. Many of the research efforts and policy discourse are…

Cryptography and Security · Computer Science 2025-04-21 Houssam Kherraz

In this paper a new approach to image watermarking in wavelet domain is presented. The idea is to hide the watermark data in blocks of the block segmented image. Two schemes are presented based on this idea by embedding the watermark data…

Information Theory · Computer Science 2010-01-05 Hamed Dehghan , S. Ebrahim Safavi

We introduce the Robust Audio Watermarking Benchmark (RAW-Bench), a benchmark for evaluating deep learning-based audio watermarking methods with standardized and systematic comparisons. To simulate real-world usage, we introduce a…

We propose a watermarking method for protecting the Intellectual Property (IP) of Generative Adversarial Networks (GANs). The aim is to watermark the GAN model so that any image generated by the GAN contains an invisible watermark…

Computer Vision and Pattern Recognition · Computer Science 2022-09-09 Jianwei Fei , Zhihua Xia , Benedetta Tondi , Mauro Barni

AI-generated video has revolutionized short video production, filmmaking, and personalized media, making video local editing an essential tool. However, this progress also blurs the line between reality and fiction, posing challenges in…

Computer Vision and Pattern Recognition · Computer Science 2024-11-15 Xuanyu Zhang , Youmin Xu , Runyi Li , Jiwen Yu , Weiqi Li , Zhipei Xu , Jian Zhang

In the era of big data, remarkable advancements have been achieved in personalized speech generation techniques that utilize speaker attributes, including voice and speaking style, to generate deepfake speech. This has also amplified global…

Audio and Speech Processing · Electrical Eng. & Systems 2025-09-10 Liping Chen , Kong Aik Lee , Zhen-Hua Ling , Xin Wang , Rohan Kumar Das , Tomoki Toda , Haizhou Li

To mitigate potential risks associated with language models, recent AI detection research proposes incorporating watermarks into machine-generated text through random vocabulary restrictions and utilizing this information for detection.…

Computation and Language · Computer Science 2024-02-14 Yu Fu , Deyi Xiong , Yue Dong

Watermarking is an important copyright protection technology which generally embeds the identity information into the carrier imperceptibly. Then the identity can be extracted to prove the copyright from the watermarked carrier even after…

Computer Vision and Pattern Recognition · Computer Science 2022-03-01 Sulong Ge , Zhihua Xia , Jianwei Fei , Xingming Sun , Jian Weng

Modern audio is created by mixing stems from different sources, raising the question: can we independently watermark each stem and recover all watermarks after separation? We study a separation-first, multi-stream watermarking…

Sound · Computer Science 2026-03-18 Houmin Sun , Zi Hu , Linxi Li , Yechen Wang , Liwei Jin , Ming Li

The explosive growth of generative video models has amplified the demand for reliable copyright preservation of AI-generated content. Despite its popularity in image synthesis, invisible generative watermarking remains largely underexplored…

Computer Vision and Pattern Recognition · Computer Science 2025-09-23 Zihan Su , Xuerui Qiu , Hongbin Xu , Tangyu Jiang , Junhao Zhuang , Chun Yuan , Ming Li , Shengfeng He , Fei Richard Yu

In this paper, we introduce SSR-Speech, a neural codec autoregressive model designed for stable, safe, and robust zero-shot textbased speech editing and text-to-speech synthesis. SSR-Speech is built on a Transformer decoder and incorporates…

Audio and Speech Processing · Electrical Eng. & Systems 2025-01-03 Helin Wang , Meng Yu , Jiarui Hai , Chen Chen , Yuchen Hu , Rilin Chen , Najim Dehak , Dong Yu

Generative Artificial Intelligence (Gen-AI) models are increasingly used to produce content across domains, including text, images, and audio. While these models represent a major technical breakthrough, they gain their generative…

Machine Learning · Computer Science 2024-12-13 Pascal Epple , Igor Shilov , Bozhidar Stevanoski , Yves-Alexandre de Montjoye

The rapid advancement of large language models (LLMs) has made it increasingly difficult to distinguish between text written by humans and machines. Addressing this, we propose a novel method for generating watermarks that strategically…

Computation and Language · Computer Science 2024-05-15 Georg Niess , Roman Kern

Existing watermarking methods for audio generative models only enable model-level attribution, allowing the identification of the originating generation model, but are unable to trace the underlying training dataset. This significant…

Sound · Computer Science 2025-08-22 Xuefeng Yang , Jian Guan , Feiyang Xiao , Congyi Fan , Haohe Liu , Qiaoxi Zhu , Dongli Xu , Youtian Lin

We propose a learning-based filter that allows us to directly modify a synthetic speech waveform into a natural speech waveform. Speech-processing systems using a vocoder framework such as statistical parametric speech synthesis and voice…

Audio and Speech Processing · Electrical Eng. & Systems 2018-10-02 Kou Tanaka , Takuhiro Kaneko , Nobukatsu Hojo , Hirokazu Kameoka

In this work, we propose a speaker anonymization pipeline that leverages high quality automatic speech recognition and synthesis systems to generate speech conditioned on phonetic transcriptions and anonymized speaker embeddings. Using…

Sound · Computer Science 2022-07-12 Sarina Meyer , Florian Lux , Pavel Denisov , Julia Koch , Pascal Tilli , Ngoc Thang Vu

Imperceptible digital watermarking is important in copyright protection, misinformation prevention, and responsible generative AI. We propose TrustMark - a GAN-based watermarking method with novel design in architecture and spatio-spectra…

Computer Vision and Pattern Recognition · Computer Science 2023-12-01 Tu Bui , Shruti Agarwal , John Collomosse
‹ Prev 1 4 5 6 7 8 10 Next ›