中文
相关论文

相关论文: SpiralDiff: Spiral Diffusion with LoRA for RGB-to-…

200 篇论文

Foundation models for vision are predominantly trained on RGB data, while many safety-critical applications rely on non-visible modalities such as infrared (IR) and synthetic aperture radar (SAR). We study whether a single flow-matching…

计算机视觉与模式识别 · 计算机科学 2026-01-09 Maxim Clouser , Kia Khezeli , John Kalantari

Synthetic Aperture Radar (SAR) imagery provides robust environmental and temporal coverage (e.g., during clouds, seasons, day-night cycles), yet its noise and unique structural patterns pose interpretation challenges, especially for…

计算机视觉与模式识别 · 计算机科学 2025-03-11 Jeonghyeok Do , Jaehyup Lee , Munchurl Kim

The increasing demand for AR/VR applications has highlighted the need for high-quality content, such as 360{\deg} live wallpapers. However, generating high-quality 360{\deg} panoramic contents remains a challenging task due to the severe…

计算机视觉与模式识别 · 计算机科学 2025-11-14 Minho Park , Taewoong Kang , Jooyeol Yun , Sungwon Hwang , Jaegul Choo

Recent diffusion models have exhibited great potential in generative modeling tasks. Part of their success can be attributed to the ability of training stable on huge sets of paired synthetic data. However, adapting these models to…

计算机视觉与模式识别 · 计算机科学 2024-05-02 Yiyang Shen , Mingqiang Wei , Yongzhen Wang , Xueyang Fu , Jing Qin

Recent advancements in image motion deblurring, driven by CNNs and transformers, have made significant progress. Large-scale pre-trained diffusion models, which are rich in real-world modeling, have shown great promise for high-quality…

计算机视觉与模式识别 · 计算机科学 2026-05-07 Xiaoyang Liu , Zhengyan Zhou , Zihang Xu , Jiezhang Cao , Zheng Chen , Yulun Zhang

We present a novel diffusion-based approach for coherent 3D scene reconstruction from a single RGB image. Our method utilizes an image-conditioned 3D scene diffusion model to simultaneously denoise the 3D poses and geometries of all objects…

计算机视觉与模式识别 · 计算机科学 2024-12-16 Manuel Dahnert , Angela Dai , Norman Müller , Matthias Nießner

Image recognition models that work in challenging environments (e.g., extremely dark, blurry, or high dynamic range conditions) must be useful. However, creating training datasets for such environments is expensive and hard due to the…

计算机视觉与模式识别 · 计算机科学 2023-03-28 Masakazu Yoshimura , Junji Otsuka , Atsushi Irie , Takeshi Ohashi

Diffusion models have achieved remarkable success in image generation, with applications broadening across various domains. Inpainting is one such application that can benefit significantly from diffusion models. Existing methods either…

计算机视觉与模式识别 · 计算机科学 2026-02-06 Sora Kim , Sungho Suh , Minsik Lee

Traditional range-instantaneous Doppler (RID) methods for rigid-body target imaging often suffer from low resolution due to the limitations of time-frequency analysis (TFA). To address this challenge, our primary focus is on obtaining high…

计算机视觉与模式识别 · 计算机科学 2025-03-31 Boan Zhang , Hang Dong , Jiongge Zhang , Long Tian , Rongrong Wang , Zhenhua Wu , Xiyang Liu , Hongwei Liu

Removing blur caused by moving objects is challenging, as the moving objects are usually significantly blurry while the static background remains clear. Existing methods that rely on local blur detection often suffer from inaccuracies and…

计算机视觉与模式识别 · 计算机科学 2024-12-13 Zhongbao Yang , Jiangxin Dong , Jinhui Tang , Jinshan Pan

Underwater images are severely degraded by wavelength-dependent light absorption and scattering, resulting in color distortion, low contrast, and loss of fine details that hinder vision-based underwater applications. To address these…

计算机视觉与模式识别 · 计算机科学 2025-12-18 Afrah Shaahid , Muzammil Behzad

Face video restoration (FVR) is a challenging but important problem where one seeks to recover a perceptually realistic face videos from a low-quality input. While diffusion probabilistic models (DPMs) have been shown to achieve remarkable…

计算机视觉与模式识别 · 计算机科学 2023-11-28 Zihao Zou , Jiaming Liu , Shirin Shoushtari , Yubo Wang , Weijie Gan , Ulugbek S. Kamilov

Capturing High Dynamic Range (HDR) scenery using 8-bit cameras often suffers from over-/underexposure, loss of fine details due to low bit-depth compression, skewed color distributions, and strong noise in dark areas. Traditional LDR image…

图像与视频处理 · 电气工程与系统科学 2024-06-14 Baiang Li , Sizhuo Ma , Yanhong Zeng , Xiaogang Xu , Youqing Fang , Zhao Zhang , Jian Wang , Kai Chen

Benefiting from their powerful generative capabilities, pretrained diffusion models have garnered significant attention for real-world image super-resolution (Real-SR). Existing diffusion-based SR approaches typically utilize semantic…

计算机视觉与模式识别 · 计算机科学 2024-12-11 Jiangang Wang , Qingnan Fan , Jinwei Chen , Hong Gu , Feng Huang , Wenqi Ren

Image restoration (IR) aims to recover images degraded by unknown mixtures while preserving semanticsconditions under which discriminative restorers and UNet-based diffusion priors often oversmooth, hallucinate, or drift. We present…

计算机视觉与模式识别 · 计算机科学 2026-05-13 Song Fei , Tian Ye , Lujia Wang , Lei Zhu

Low Dynamic Range (LDR) to High Dynamic Range (HDR) image translation is a fundamental task in many computational vision problems. Numerous data-driven methods have been proposed to address this problem; however, they lack explicit modeling…

图形学 · 计算机科学 2025-09-23 Hrishav Bakul Barua , Kalin Stefanov , Ganesh Krishnasamy , KokSheik Wong , Abhinav Dhall

Recent advancements in diffusion models have significantly improved performance in super-resolution (SR) tasks. However, previous research often overlooks the fundamental differences between SR and general image generation. General image…

图像与视频处理 · 电气工程与系统科学 2024-10-31 Hanlin Wu , Jiangwei Mo , Xiaohui Sun , Jie Ma

X-ray imaging is a rapid and cost-effective tool for visualizing internal human anatomy. While multi-view X-ray imaging provides complementary information that enhances diagnosis, intervention, and education, acquiring images from multiple…

图像与视频处理 · 电气工程与系统科学 2025-10-21 Chun Xie , Yuichi Yoshii , Itaru Kitahara

Burst image super resolution (BISR) aims to construct a single high-resolution (HR) image by aggregating information from multiple low-resolution (LR) frames, relying on temporal redundancy and spatial coherence across the burst. While…

Video deblurring presents a considerable challenge owing to the complexity of blur, which frequently results from a combination of camera shakes, and object motions. In the field of video deblurring, many previous works have primarily…

计算机视觉与模式识别 · 计算机科学 2024-12-03 Haoyang Long , Yan Wang , Wendong Wang