English
Related papers

Related papers: GROOT: Generating Robust Watermark for Diffusion-M…

200 papers

Deep generative models can generate high-fidelity audio conditioned on various types of representations (e.g., mel-spectrograms, Mel-frequency Cepstral Coefficients (MFCC)). Recently, such models have been used to synthesize audio waveforms…

The proliferation of autoregressive (AR) image generators demands reliable detection and attribution of their outputs to mitigate misinformation, and to filter synthetic images from training data to prevent model collapse. To address this…

Computer Vision and Pattern Recognition · Computer Science 2026-04-14 Andreas Müller , Denis Lukovnikov , Shingo Kodama , Minh Pham , Anubhav Jain , Jonathan Petit , Niv Cohen , Asja Fischer

In the realm of audio watermarking, it is challenging to simultaneously encode imperceptible messages while enhancing the message capacity and robustness. Although recent advancements in deep learning-based methods bolster the message…

Sound · Computer Science 2024-11-05 Mayank Kumar Singh , Naoya Takahashi , Weihsiang Liao , Yuki Mitsufuji

With recent advances in speech synthesis including text-to-speech (TTS) and voice conversion (VC) systems enabling the generation of ultra-realistic audio deepfakes, there is growing concern about their potential misuse. However, most…

Sound · Computer Science 2024-04-24 Zuheng Kang , Yayun He , Botao Zhao , Xiaoyang Qu , Junqing Peng , Jing Xiao , Jianzong Wang

As deep learning advances in audio generation, challenges in audio security and copyright protection highlight the need for robust audio watermarking. Recent neural network-based methods have made progress but still face three main issues:…

Sound · Computer Science 2025-06-09 Yaoxun Xu , Jianwei Yu , Hangting Chen , Zhiyong Wu , Xixin Wu , Dong Yu , Rongzhi Gu , Yi Luo

As 3D Gaussian Splatting (3D-GS) gains significant attention and its commercial usage increases, the need for watermarking technologies to prevent unauthorized use of the 3D-GS models and rendered images has become increasingly important.…

Computer Vision and Pattern Recognition · Computer Science 2025-04-01 Youngdong Jang , Hyunje Park , Feng Yang , Heeju Ko , Euijin Choo , Sangpil Kim

Recent progress in large language models enables the creation of realistic machine-generated content. Watermarking is a promising approach to distinguish machine-generated text from human text, embedding statistical signals in the output…

Cryptography and Security · Computer Science 2026-02-25 Patrick Chao , Yan Sun , Edgar Dobriban , Hamed Hassani

Artificial Intelligence Generated Content (AIGC) has advanced significantly, particularly with the development of video generation models such as text-to-video (T2V) models and image-to-video (I2V) models. However, like other AIGC types,…

Computer Vision and Pattern Recognition · Computer Science 2025-02-26 Runyi Hu , Jie Zhang , Yiming Li , Jiwei Li , Qing Guo , Han Qiu , Tianwei Zhang

As the outputs of generative AI (GenAI) techniques improve in quality, it becomes increasingly challenging to distinguish them from human-created content. Watermarking schemes are a promising approach to address the problem of…

Watermarking is an important copyright protection technology which generally embeds the identity information into the carrier imperceptibly. Then the identity can be extracted to prove the copyright from the watermarked carrier even after…

Computer Vision and Pattern Recognition · Computer Science 2022-03-01 Sulong Ge , Zhihua Xia , Jianwei Fei , Xingming Sun , Jian Weng

Diffusion-based watermarking methods embed verifiable marks by manipulating the initial noise or the reverse diffusion trajectory. However, these methods share a critical assumption: verification can succeed only if the diffusion trajectory…

Computer Vision and Pattern Recognition · Computer Science 2026-04-02 Rui Bao , Zheng Gao , Xiaoyu Li , Xiaoyan Feng , Yang Song , Jiaojiao Jiang

The rapid advancement of generative artificial intelligence (GenAI) has revolutionized content creation across text, visual, and audio domains, simultaneously introducing significant risks such as misinformation, identity fraud, and content…

Cryptography and Security · Computer Science 2025-04-08 Lele Cao

Nowadays, it is common to release audio content to the public. However, with the rise of voice cloning technology, attackers have the potential to easily impersonate a specific person by utilizing his publicly released audio without any…

Sound · Computer Science 2023-12-07 Chang Liu , Jie Zhang , Tianwei Zhang , Xi Yang , Weiming Zhang , Nenghai Yu

The rapid advancement of text-to-image generation systems, exemplified by models like Stable Diffusion, Midjourney, Imagen, and DALL-E, has heightened concerns about their potential misuse. In response, companies like Meta and Google have…

Computer Vision and Pattern Recognition · Computer Science 2024-08-21 Niyar R Barman , Krish Sharma , Ashhar Aziz , Shashwat Bajpai , Shwetangshu Biswas , Vasu Sharma , Vinija Jain , Aman Chadha , Amit Sheth , Amitava Das

Mainstream zero-shot TTS production systems like Voicebox and Seed-TTS achieve human parity speech by leveraging Flow-matching and Diffusion models, respectively. Unfortunately, human-level audio synthesis leads to identity misuse and…

The explosive growth of generative video models has amplified the demand for reliable copyright preservation of AI-generated content. Despite its popularity in image synthesis, invisible generative watermarking remains largely underexplored…

Computer Vision and Pattern Recognition · Computer Science 2025-09-23 Zihan Su , Xuerui Qiu , Hongbin Xu , Tangyu Jiang , Junhao Zhuang , Chun Yuan , Ming Li , Shengfeng He , Fei Richard Yu

The rapid development of video generative models has led to a surge in highly realistic synthetic videos, raising ethical concerns related to disinformation and copyright infringement. Recently, video watermarking has been proposed as a…

Cryptography and Security · Computer Science 2025-05-29 Zhengyuan Jiang , Moyang Guo , Kecen Li , Yuepeng Hu , Yupu Wang , Zhicong Huang , Cheng Hong , Neil Zhenqiang Gong

Ethical concerns surrounding copyright protection and inappropriate content generation pose challenges for the practical implementation of diffusion models. One effective solution involves watermarking the generated images. Existing methods…

Computer Vision and Pattern Recognition · Computer Science 2025-05-14 Zijin Yang , Xin Zhang , Kejiang Chen , Kai Zeng , Qiyi Yao , Han Fang , Weiming Zhang , Nenghai Yu

Methods for watermarking large language models have been proposed that distinguish AI-generated text from human-generated text by slightly altering the model output distribution, but they also distort the quality of the text, exposing the…

Computation and Language · Computer Science 2024-02-27 Massieh Kordi Boroujeny , Ya Jiang , Kai Zeng , Brian Mark

The audio watermarking technique embeds messages into audio and accurately extracts messages from the watermarked audio. Traditional methods develop algorithms based on expert experience to embed watermarks into the time-domain or…

Multimedia · Computer Science 2024-10-01 Pengcheng Li , Xulong Zhang , Jing Xiao , Jianzong Wang
‹ Prev 1 8 9 10 Next ›