English
Related papers

Related papers: Saliency-Aware Diffusion Reconstruction for Effect…

200 papers

Blind face restoration (BFR) is a highly challenging problem due to the uncertainty of degradation patterns. Current methods have low generalization across photorealistic and heterogeneous domains. In this paper, we propose a…

Computer Vision and Pattern Recognition · Computer Science 2024-03-18 Nan Gao , Jia Li , Huaibo Huang , Zhi Zeng , Ke Shang , Shuwu Zhang , Ran He

Deep neural networks have recently achieved significant advancements in remote sensing superresolu-tion (SR). However, most existing methods are limited to low magnification rates (e.g., 2 or 4) due to the escalating ill-posedness at higher…

Image and Video Processing · Electrical Eng. & Systems 2024-12-30 Yue Shi , Liangxiu Han , Darren Dancy , Lianghao Han

Seismic impedance inversion is one of the most important part of geophysical exploration. However, due to random noise, the traditional semi-supervised learning (SSL) methods lack generalization and stability. To solve this problem, some…

Geophysics · Physics 2024-06-26 Yingtian Liu , Yong Li , Xingan Hao , Huating Li , Zhangquan Liao , Junheng Peng

Text-to-Speech (TTS) diffusion models generate high-quality speech, which raises challenges for the model intellectual property protection and speech tracing for legal use. Audio watermarking is a promising solution. However, due to the…

Sound · Computer Science 2025-12-23 Yichuan Zhang , Chengxin Li , Yujie Gu

Diffusion models have demonstrated impressive performance in face restoration. Yet, their multi-step inference process remains computationally intensive, limiting their applicability in real-world scenarios. Moreover, existing methods often…

Computer Vision and Pattern Recognition · Computer Science 2026-04-14 Jingkai Wang , Jue Gong , Lin Zhang , Zheng Chen , Xing Liu , Hong Gu , Yutong Liu , Yulun Zhang , Xiaokang Yang

Magnetic Resonance Imaging (MRI) produces excellent soft tissue contrast, albeit it is an inherently slow imaging modality. Promising deep learning methods have recently been proposed to reconstruct accelerated MRI scans. However, existing…

Image and Video Processing · Electrical Eng. & Systems 2024-04-17 Yilmaz Korkmaz , Tolga Cukur , Vishal M. Patel

The present paper introduces a method for substantial reduction of the number of diffusion encoding gradients required for reliable reconstruction of HARDI signals. The method exploits the theory of compressed sensing (CS), which…

Information Theory · Computer Science 2010-09-21 Oleg Michailovich , Yogesh Rathi , Sudipto Dolui

Due to storage and bandwidth limitations, videos transmitted over the Internet often exhibit low quality, characterized by low-resolution and compression artifacts. Although video super-resolution (VSR) is an efficient video enhancing…

Computer Vision and Pattern Recognition · Computer Science 2025-06-30 Hongyu An , Xinfeng Zhang , Shijie Zhao , Li Zhang , Ruiqin Xiong

Latent-based diffusion model watermarking embeds watermarks into generated images' latent space to enable content attribution, offering a training-free solution for intellectual property protection and digital forensics. However, these…

Computer Vision and Pattern Recognition · Computer Science 2026-05-05 Jiewei Lai , Lan Zhang , Chen Tang , Pengcheng Sun , Zhaopeng Zhang , Yunhao Wang , Hui Jin

We introduce a novel, training-free system for reconstructing, understanding, and rendering 3D indoor scenes from a sparse set of unposed RGB images. Unlike traditional radiance field approaches that require dense views and per-scene…

Computer Vision and Pattern Recognition · Computer Science 2026-03-24 Jiatong Xia , Lingqiao Liu

The rapid advancement of diffusion models has enhanced their image inpainting and editing capabilities but also introduced significant societal risks. Adversaries can exploit user images from social media to generate misleading or harmful…

Computer Vision and Pattern Recognition · Computer Science 2025-05-27 Yuhao He , Jinyu Tian , Haiwei Wu , Jianqing Li

Digital watermarking is the process of embedding secret information by altering images in an undetectable way to the human eye. To increase the robustness of the model, many deep learning-based watermarking methods use the…

Multimedia · Computer Science 2024-05-20 Sijing Xie , Chengxin Zhao , Nan Sun , Wei Li , Hefei Ling

This study presents a new image super-resolution (SR) technique based on diffusion inversion, aiming at harnessing the rich image priors encapsulated in large pre-trained diffusion models to improve SR performance. We design a Partial noise…

Computer Vision and Pattern Recognition · Computer Science 2025-03-14 Zongsheng Yue , Kang Liao , Chen Change Loy

DNN-based watermarking methods are rapidly developing and delivering impressive performances. Recent advances achieve resolution-agnostic image watermarking by reducing the variant resolution watermarking problem to a fixed resolution…

Cryptography and Security · Computer Science 2024-09-04 Yuchen Wang , Xingyu Zhu , Guanhui Ye , Shiyao Zhang , Xuetao Wei

Due to the high potential for abuse of GenAI systems, the task of detecting synthetic images has recently become of great interest to the research community. Unfortunately, existing image-space detectors quickly become obsolete as new…

Computer Vision and Pattern Recognition · Computer Science 2024-06-14 George Cazenavette , Avneesh Sud , Thomas Leung , Ben Usman

As AI-generated images become widespread, reliable watermarking is essential for content verification, copyright enforcement, and combating disinformation. Existing techniques rely on heuristic approaches and lack formal guarantees of…

Cryptography and Security · Computer Science 2025-09-01 Kareem Shehata , Aashish Kolluri , Prateek Saxena

The Semantic Layered Embedding Diffusion (SLED) mechanism redefines the representation of hierarchical semantics within transformer-based architectures, enabling enhanced contextual consistency across a wide array of linguistic tasks. By…

Computation and Language · Computer Science 2025-03-26 Irin Kabakum , Thomas Montgomery , Daniel Ravenwood , Genevieve Harrington

We present a novel approach to leverage prior knowledge encapsulated in pre-trained text-to-image diffusion models for blind super-resolution (SR). Specifically, by employing our time-aware encoder, we can achieve promising restoration…

Computer Vision and Pattern Recognition · Computer Science 2024-07-01 Jianyi Wang , Zongsheng Yue , Shangchen Zhou , Kelvin C. K. Chan , Chen Change Loy

While deep learning-based methods for blind face restoration have achieved unprecedented success, they still suffer from two major limitations. First, most of them deteriorate when facing complex degradations out of their training data.…

Computer Vision and Pattern Recognition · Computer Science 2024-07-23 Zongsheng Yue , Chen Change Loy

With the rapid development of image generation technologies, especially the advancement of Diffusion Models, the quality of synthesized images has significantly improved, raising concerns among researchers about information security. To…

Computer Vision and Pattern Recognition · Computer Science 2025-06-23 Weinan Guan , Wei Wang , Bo Peng , Ziwen He , Jing Dong , Haonan Cheng