English
Related papers

Related papers: LapLoss: Laplacian Pyramid-based Multiscale loss f…

200 papers

An image pyramid can extend many object detection algorithms to solve detection on multiple scales. However, interpolation during the resampling process of an image pyramid causes gradient variation, which is the difference of the gradients…

Computer Vision and Pattern Recognition · Computer Science 2019-09-06 Yonghyun Kim , Bong-Nam Kang , Daijin Kim

Low-light image enhancement (LLIE) techniques attempt to increase the visibility of images captured in low-light scenarios. However, as a result of enhancement, a variety of image degradations such as noise and color bias are revealed.…

Image and Video Processing · Electrical Eng. & Systems 2024-09-10 Savvas Panagiotou , Anna S. Bosman

Despite significant advancements in Vision-Language Models (VLMs), the performance of existing VLMs remains hindered by object hallucination, a critical challenge to achieving accurate visual understanding. To address this issue, we propose…

Computer Vision and Pattern Recognition · Computer Science 2026-03-31 Woohyeon Park , Woojin Kim , Jaeik Kim , Jaeyoung Do

Image restoration, which aims to recover high-quality images from their corrupted counterparts, often faces the challenge of being an ill-posed problem that allows multiple solutions for a single input. However, most deep learning based…

Computer Vision and Pattern Recognition · Computer Science 2024-04-16 Wenyi Lian , Wenjing Lian , Ziwei Luo

Large-scale diffusion models have made significant advances in image generation, particularly through cross-attention mechanisms. While cross-attention has been well-studied in text-to-image tasks, their interpretability in image-to-image…

Computer Vision and Pattern Recognition · Computer Science 2025-03-21 Junseo Park , Hyeryung Jang

Despite recent significant strides achieved by diffusion-based Text-to-Image (T2I) models, current systems are still less capable of ensuring decent compositional generation aligned with text prompts, particularly for the multi-object…

Computer Vision and Pattern Recognition · Computer Science 2024-02-01 Zhipeng Bao , Yijun Li , Krishna Kumar Singh , Yu-Xiong Wang , Martial Hebert

Image Copy Detection (ICD) aims to identify manipulated content between image pairs through robust feature representation learning. While self-supervised learning (SSL) has advanced ICD systems, existing view-level contrastive methods…

Computer Vision and Pattern Recognition · Computer Science 2026-02-26 Yichen Lu , Siwei Nie , Minlong Lu , Xudong Yang , Xiaobo Zhang , Peng Zhang

Due to their highly structured characteristics, faces are easier to recover than natural scenes for blind image super-resolution. Therefore, we can extract the degradation representation of an image from the low-quality and recovered face…

Computer Vision and Pattern Recognition · Computer Science 2023-09-18 Zhicun Yin , Ming Liu , Xiaoming Li , Hui Yang , Longan Xiao , Wangmeng Zuo

Image restoration is a low-level vision task, most CNN methods are designed as a black box, lacking transparency and internal aesthetics. Although some methods combining traditional optimization algorithms with DNNs have been proposed, they…

Computer Vision and Pattern Recognition · Computer Science 2025-08-27 Xiao Feng Zhang , Chao Chen Gu , Shan Ying Zhu

Multi-image super-resolution (MISR) allows to increase the spatial resolution of a low-resolution (LR) acquisition by combining multiple images carrying complementary information in the form of sub-pixel offsets in the scene sampling, and…

Computer Vision and Pattern Recognition · Computer Science 2024-01-31 Luca Savant Aira , Diego Valsesia , Andrea Bordone Molini , Giulia Fracastoro , Enrico Magli , Andrea Mirabile

Capturing the global topology of an image is essential for proposing an accurate segmentation of its domain. However, most of existing segmentation methods do not preserve the initial topology of the given input, which is detrimental for…

Computer Vision and Pattern Recognition · Computer Science 2022-08-19 Minh On Vu Ngoc , Yizi Chen , Nicolas Boutry , Jonathan Fabrizio , Clement Mallet

In Masked Image Modeling (MIM), two primary methods exist: Pixel MIM and Latent MIM, each utilizing different reconstruction targets, raw pixels and latent representations, respectively. Pixel MIM tends to capture low-level visual details…

Computer Vision and Pattern Recognition · Computer Science 2025-01-07 Junmyeong Lee , Eui Jun Hwang , Sukmin Cho , Jong C. Park

In this work we review the coarse-to-fine spatial feature pyramid concept, which is used in state-of-the-art optical flow estimation networks to make exploration of the pixel flow search space computationally tractable and efficient. Within…

Computer Vision and Pattern Recognition · Computer Science 2020-07-21 Markus Hofinger , Samuel Rota Bulò , Lorenzo Porzi , Arno Knapitsch , Thomas Pock , Peter Kontschieder

Light field (LF) depth estimation plays a crucial role in many LF-based applications. Existing LF depth estimation methods consider depth estimation as a regression problem, where a pixel-wise L1 loss is employed to supervise the training…

Computer Vision and Pattern Recognition · Computer Science 2023-11-22 Wentao Chao , Xuechun Wang , Yingqian Wang , Guanghui Wang , Fuqing Duan

While large vision-language models (LVLMs) have shown impressive capabilities in generating plausible responses correlated with input visual contents, they still suffer from hallucinations, where the generated text inaccurately reflects…

Computer Vision and Pattern Recognition · Computer Science 2024-12-10 Yi-Lun Lee , Yi-Hsuan Tsai , Wei-Chen Chiu

2D Gaussian Splatting (2DGS) is an emerging explicit scene representation method with significant potential for image compression due to high fidelity and high compression ratios. However, existing low-light enhancement algorithms operate…

Computer Vision and Pattern Recognition · Computer Science 2026-01-23 Yuhan Chen , Wenxuan Yu , Guofa Li , Yijun Xu , Ying Fang , Yicui Shi , Long Cao , Wenbo Chu , Keqiang Li

Deep learning-based methods have recently demonstrated promising results in deformable image registration for a wide range of medical image analysis tasks. However, existing deep learning-based methods are usually limited to small…

Image and Video Processing · Electrical Eng. & Systems 2020-07-01 Tony C. W. Mok , Albert C. S. Chung

Low-light image enhancement aims to improve the visibility of degraded images to better align with human visual perception. While diffusion-based methods have shown promising performance due to their strong generative capabilities. However,…

Computer Vision and Pattern Recognition · Computer Science 2025-07-25 Jinhong He , Minglong Xue , Zhipu Liu , Mingliang Zhou , Aoxiang Ning , Palaiahnakote Shivakumara

Image-to-image (I2I) translation is a pixel-level mapping that requires a large number of paired training data and often suffers from the problems of high diversity and strong category bias in image scenes. In order to tackle these…

Computer Vision and Pattern Recognition · Computer Science 2019-04-22 Liqian Ma , Qianru Sun , Bernt Schiele , Luc Van Gool

Existing image enhancement methods fall short of expectations because with them it is difficult to improve global and local image contrast simultaneously. To address this problem, we propose a histogram equalization-based method that adapts…

Computer Vision and Pattern Recognition · Computer Science 2022-09-15 Xiaomeng Wu , Takahito Kawanishi , Kunio Kashino
‹ Prev 1 4 5 6 7 8 10 Next ›