中文
相关论文

相关论文: SpiralDiff: Spiral Diffusion with LoRA for RGB-to-…

200 篇论文

The integration of RGB and depth modalities significantly enhances the accuracy of segmenting complex indoor scenes, with depth data from RGB-D cameras playing a crucial role in this improvement. However, collecting an RGB-D dataset is more…

计算机视觉与模式识别 · 计算机科学 2025-03-25 Xinhua Xu , Hong Liu , Jianbing Wu , Jinfu Liu

Most existing super-resolution methods do not perform well in real scenarios due to lack of realistic training data and information loss of the model input. To solve the first problem, we propose a new pipeline to generate realistic…

图像与视频处理 · 电气工程与系统科学 2019-05-30 Xiangyu Xu , Yongrui Ma , Wenxiu Sun

Multi-modal image fusion aims to consolidate complementary information from diverse source images into a unified representation. The fused image is expected to preserve fine details and maintain high visual fidelity. While diffusion models…

计算机视觉与模式识别 · 计算机科学 2026-01-29 Xingxin Xu , Bing Cao , DongDong Li , Qinghua Hu , Pengfei Zhu

Autonomous driving algorithms usually employ sRGB images as model input due to their compatibility with the human visual system. However, visually pleasing sRGB images are possibly sub-optimal for downstream tasks when compared to RAW…

图像与视频处理 · 电气工程与系统科学 2024-09-05 Anqi Liu , Shiyi Mu , Shugong Xu

Training learning-based deblurring methods demands a tremendous amount of blurred and sharp image pairs. Unfortunately, existing synthetic datasets are not realistic enough, and deblurring models trained on them cannot handle real blurred…

计算机视觉与模式识别 · 计算机科学 2022-07-22 Jaesung Rim , Geonung Kim , Jungeon Kim , Junyong Lee , Seungyong Lee , Sunghyun Cho

The burgeoning field of camouflaged object detection (COD) seeks to identify objects that blend into their surroundings. Despite the impressive performance of recent models, we have identified a limitation in their robustness, where…

计算机视觉与模式识别 · 计算机科学 2023-04-13 Xue-Jing Luo , Shuo Wang , Zongwei Wu , Christos Sakaridis , Yun Cheng , Deng-Ping Fan , Luc Van Gool

Image restoration (IR) has been an indispensable and challenging task in the low-level vision field, which strives to improve the subjective quality of images distorted by various forms of degradation. Recently, the diffusion model has…

计算机视觉与模式识别 · 计算机科学 2025-12-09 Xin Li , Yulin Ren , Xin Jin , Cuiling Lan , Xingrui Wang , Wenjun Zeng , Xinchao Wang , Zhibo Chen

Autonomous systems rely on sensors to estimate the environment around them. However, cameras, LiDARs, and RADARs have their own limitations. In nighttime or degraded environments such as fog, mist, or dust, thermal cameras can provide…

机器人学 · 计算机科学 2025-06-27 Shruti Bansal , Wenshan Wang , Yifei Liu , Parv Maheshwari

Following the remarkable success of diffusion models on image generation, recent works have also demonstrated their impressive ability to address a number of inverse problems in an unsupervised way, by properly constraining the sampling…

计算机视觉与模式识别 · 计算机科学 2023-08-23 Foivos Paraperas Papantoniou , Alexandros Lattas , Stylianos Moschoglou , Stefanos Zafeiriou

Generating high-resolution images with generative models has recently been made widely accessible by leveraging diffusion models pre-trained on large-scale datasets. Various techniques, such as MultiDiffusion and SyncDiffusion, have further…

计算机视觉与模式识别 · 计算机科学 2025-01-08 Stanislav Frolov , Brian B. Moser , Andreas Dengel

Rectified-flow-based diffusion transformers like FLUX and OpenSora have demonstrated outstanding performance in the field of image and video generation. Despite their robust generative capabilities, these models often struggle with…

计算机视觉与模式识别 · 计算机科学 2025-06-16 Jiangshan Wang , Junfu Pu , Zhongang Qi , Jiayi Guo , Yue Ma , Nisha Huang , Yuxin Chen , Xiu Li , Ying Shan

While single-image super-resolution (SISR) has attracted substantial interest in recent years, the proposed approaches are limited to learning image priors in order to add high frequency details. In contrast, multi-frame super-resolution…

计算机视觉与模式识别 · 计算机科学 2021-04-07 Goutam Bhat , Martin Danelljan , Luc Van Gool , Radu Timofte

Person re-identification (re-id) aims to retrieve images of same identities across different camera views. Resolution mismatch occurs due to varying distances between person of interest and cameras, this significantly degrades the…

计算机视觉与模式识别 · 计算机科学 2021-09-17 Asad Munir , Chengjin Lyu , Bart Goossens , Wilfried Philips , Christian Micheloni

Recently, diffusion models have shown remarkable results in image synthesis by gradually removing noise and amplifying signals. Although the simple generative process surprisingly works well, is this the best way to generate image data? For…

计算机视觉与模式识别 · 计算机科学 2022-11-22 Sangyun Lee , Hyungjin Chung , Jaehyeon Kim , Jong Chul Ye

Lens flare significantly degrades image quality, impacting critical computer vision tasks like object detection and autonomous driving. Recent Single Image Flare Removal (SIFR) methods perform poorly when off-frame light sources are…

计算机视觉与模式识别 · 计算机科学 2025-10-20 Shr-Ruei Tsai , Wei-Cheng Chang , Jie-Ying Lee , Chih-Hai Su , Yu-Lun Liu

High dynamic range (HDR) imaging is a crucial task in computational photography, which captures details across diverse lighting conditions. Traditional HDR fusion methods face limitations in dynamic scenes with extreme exposure differences,…

计算机视觉与模式识别 · 计算机科学 2024-12-20 Shi Guo , Zixuan Chen , Ziran Zhang , Yutian Chen , Gangwei Xu , Tianfan Xue

We propose a novel 3d colored shape reconstruction method from a single RGB image through diffusion model. Diffusion models have shown great development potentials for high-quality 3D shape generation. However, most existing work based on…

计算机视觉与模式识别 · 计算机科学 2023-02-14 Bo Li , Xiaolin Wei , Fengwei Chen , Bin Liu

Existing single image reflection removal (SIRR) methods using deep learning tend to miss key low-frequency (LF) and high-frequency (HF) differences in images, affecting their effectiveness in removing reflections. To address this problem,…

计算机视觉与模式识别 · 计算机科学 2025-08-26 Tao Wang , Wanglong Lu , Kaihao Zhang , Tong Lu , Ming-Hsuan Yang

While Low-Rank Adaptation (LoRA) has proven beneficial for efficiently fine-tuning large models, LoRA fine-tuned text-to-image diffusion models lack diversity in the generated images, as the model tends to copy data from the observed…

Enhancing RAW images captured under low light conditions is a challenging task. Recent deep learning based RAW enhancement methods have shifted from using real paired data to relying on synthetic datasets. These synthetic datasets are…

图像与视频处理 · 电气工程与系统科学 2025-09-11 Juntai Zeng
‹ 上一页 1 8 9 10 下一页 ›