中文
相关论文

相关论文: GROOT: Generating Robust Watermark for Diffusion-M…

200 篇论文

Deep generative models can generate high-fidelity audio conditioned on various types of representations (e.g., mel-spectrograms, Mel-frequency Cepstral Coefficients (MFCC)). Recently, such models have been used to synthesize audio waveforms…

The proliferation of autoregressive (AR) image generators demands reliable detection and attribution of their outputs to mitigate misinformation, and to filter synthetic images from training data to prevent model collapse. To address this…

计算机视觉与模式识别 · 计算机科学 2026-04-14 Andreas Müller , Denis Lukovnikov , Shingo Kodama , Minh Pham , Anubhav Jain , Jonathan Petit , Niv Cohen , Asja Fischer

In the realm of audio watermarking, it is challenging to simultaneously encode imperceptible messages while enhancing the message capacity and robustness. Although recent advancements in deep learning-based methods bolster the message…

声音 · 计算机科学 2024-11-05 Mayank Kumar Singh , Naoya Takahashi , Weihsiang Liao , Yuki Mitsufuji

With recent advances in speech synthesis including text-to-speech (TTS) and voice conversion (VC) systems enabling the generation of ultra-realistic audio deepfakes, there is growing concern about their potential misuse. However, most…

声音 · 计算机科学 2024-04-24 Zuheng Kang , Yayun He , Botao Zhao , Xiaoyang Qu , Junqing Peng , Jing Xiao , Jianzong Wang

As deep learning advances in audio generation, challenges in audio security and copyright protection highlight the need for robust audio watermarking. Recent neural network-based methods have made progress but still face three main issues:…

声音 · 计算机科学 2025-06-09 Yaoxun Xu , Jianwei Yu , Hangting Chen , Zhiyong Wu , Xixin Wu , Dong Yu , Rongzhi Gu , Yi Luo

As 3D Gaussian Splatting (3D-GS) gains significant attention and its commercial usage increases, the need for watermarking technologies to prevent unauthorized use of the 3D-GS models and rendered images has become increasingly important.…

计算机视觉与模式识别 · 计算机科学 2025-04-01 Youngdong Jang , Hyunje Park , Feng Yang , Heeju Ko , Euijin Choo , Sangpil Kim

Recent progress in large language models enables the creation of realistic machine-generated content. Watermarking is a promising approach to distinguish machine-generated text from human text, embedding statistical signals in the output…

密码学与安全 · 计算机科学 2026-02-25 Patrick Chao , Yan Sun , Edgar Dobriban , Hamed Hassani

Artificial Intelligence Generated Content (AIGC) has advanced significantly, particularly with the development of video generation models such as text-to-video (T2V) models and image-to-video (I2V) models. However, like other AIGC types,…

计算机视觉与模式识别 · 计算机科学 2025-02-26 Runyi Hu , Jie Zhang , Yiming Li , Jiwei Li , Qing Guo , Han Qiu , Tianwei Zhang

As the outputs of generative AI (GenAI) techniques improve in quality, it becomes increasingly challenging to distinguish them from human-created content. Watermarking schemes are a promising approach to address the problem of…

Watermarking is an important copyright protection technology which generally embeds the identity information into the carrier imperceptibly. Then the identity can be extracted to prove the copyright from the watermarked carrier even after…

计算机视觉与模式识别 · 计算机科学 2022-03-01 Sulong Ge , Zhihua Xia , Jianwei Fei , Xingming Sun , Jian Weng

Diffusion-based watermarking methods embed verifiable marks by manipulating the initial noise or the reverse diffusion trajectory. However, these methods share a critical assumption: verification can succeed only if the diffusion trajectory…

计算机视觉与模式识别 · 计算机科学 2026-04-02 Rui Bao , Zheng Gao , Xiaoyu Li , Xiaoyan Feng , Yang Song , Jiaojiao Jiang

The rapid advancement of generative artificial intelligence (GenAI) has revolutionized content creation across text, visual, and audio domains, simultaneously introducing significant risks such as misinformation, identity fraud, and content…

密码学与安全 · 计算机科学 2025-04-08 Lele Cao

Nowadays, it is common to release audio content to the public. However, with the rise of voice cloning technology, attackers have the potential to easily impersonate a specific person by utilizing his publicly released audio without any…

声音 · 计算机科学 2023-12-07 Chang Liu , Jie Zhang , Tianwei Zhang , Xi Yang , Weiming Zhang , Nenghai Yu

The rapid advancement of text-to-image generation systems, exemplified by models like Stable Diffusion, Midjourney, Imagen, and DALL-E, has heightened concerns about their potential misuse. In response, companies like Meta and Google have…

计算机视觉与模式识别 · 计算机科学 2024-08-21 Niyar R Barman , Krish Sharma , Ashhar Aziz , Shashwat Bajpai , Shwetangshu Biswas , Vasu Sharma , Vinija Jain , Aman Chadha , Amit Sheth , Amitava Das

Mainstream zero-shot TTS production systems like Voicebox and Seed-TTS achieve human parity speech by leveraging Flow-matching and Diffusion models, respectively. Unfortunately, human-level audio synthesis leads to identity misuse and…

The explosive growth of generative video models has amplified the demand for reliable copyright preservation of AI-generated content. Despite its popularity in image synthesis, invisible generative watermarking remains largely underexplored…

计算机视觉与模式识别 · 计算机科学 2025-09-23 Zihan Su , Xuerui Qiu , Hongbin Xu , Tangyu Jiang , Junhao Zhuang , Chun Yuan , Ming Li , Shengfeng He , Fei Richard Yu

The rapid development of video generative models has led to a surge in highly realistic synthetic videos, raising ethical concerns related to disinformation and copyright infringement. Recently, video watermarking has been proposed as a…

密码学与安全 · 计算机科学 2025-05-29 Zhengyuan Jiang , Moyang Guo , Kecen Li , Yuepeng Hu , Yupu Wang , Zhicong Huang , Cheng Hong , Neil Zhenqiang Gong

Ethical concerns surrounding copyright protection and inappropriate content generation pose challenges for the practical implementation of diffusion models. One effective solution involves watermarking the generated images. Existing methods…

计算机视觉与模式识别 · 计算机科学 2025-05-14 Zijin Yang , Xin Zhang , Kejiang Chen , Kai Zeng , Qiyi Yao , Han Fang , Weiming Zhang , Nenghai Yu

Methods for watermarking large language models have been proposed that distinguish AI-generated text from human-generated text by slightly altering the model output distribution, but they also distort the quality of the text, exposing the…

计算与语言 · 计算机科学 2024-02-27 Massieh Kordi Boroujeny , Ya Jiang , Kai Zeng , Brian Mark

The audio watermarking technique embeds messages into audio and accurately extracts messages from the watermarked audio. Traditional methods develop algorithms based on expert experience to embed watermarks into the time-domain or…

多媒体 · 计算机科学 2024-10-01 Pengcheng Li , Xulong Zhang , Jing Xiao , Jianzong Wang
‹ 上一页 1 8 9 10 下一页 ›