English
Related papers

Related papers: DEFormer: DCT-driven Enhancement Transformer for L…

200 papers

We introduce TransformerFusion, a transformer-based 3D scene reconstruction approach. From an input monocular RGB video, the video frames are processed by a transformer network that fuses the observations into a volumetric feature grid…

Computer Vision and Pattern Recognition · Computer Science 2021-07-07 Aljaž Božič , Pablo Palafox , Justus Thies , Angela Dai , Matthias Nießner

Object detection in images has reached unprecedented performances. The state-of-the-art methods rely on deep architectures that extract salient features and predict bounding boxes enclosing the objects of interest. These methods essentially…

Computer Vision and Pattern Recognition · Computer Science 2021-07-15 Benjamin Deguerre , Clement Chatelain , Gilles Gasso

Accurate depth estimation from images is a fundamental task in many applications including scene understanding and reconstruction. Existing solutions for depth estimation often produce blurry approximations of low resolution. This paper…

Computer Vision and Pattern Recognition · Computer Science 2019-03-12 Ibraheem Alhashim , Peter Wonka

Real-world low-light images often suffer from complex degradations such as local overexposure, low brightness, noise, and uneven illumination. Supervised methods tend to overfit to specific scenarios, while unsupervised methods, though…

Computer Vision and Pattern Recognition · Computer Science 2025-03-20 Huaqiu Li , Xiaowan Hu , Haoqian Wang

Digital imaging aims to replicate realistic scenes, but Low Dynamic Range (LDR) cameras cannot represent the wide dynamic range of real scenes, resulting in under-/overexposed images. This paper presents a deep learning-based approach for…

Computer Vision and Pattern Recognition · Computer Science 2023-07-07 Dwip Dalal , Gautam Vashishtha , Prajwal Singh , Shanmuganathan Raman

Objective: Transformers, born to remedy the inadequate receptive fields of CNNs, have drawn explosive attention recently. However, the daunting computational complexity of global representation learning, together with rigid window…

Computer Vision and Pattern Recognition · Computer Science 2023-04-20 Xian Lin , Li Yu , Kwang-Ting Cheng , Zengqiang Yan

Shadows in scanned documents pose significant challenges for document analysis and recognition tasks due to their negative impact on visual quality and readability. Current shadow removal techniques, including traditional methods and deep…

Computer Vision and Pattern Recognition · Computer Science 2024-07-31 Ziyang Zhou , Yingtie Lei , Xuhang Chen , Shenghong Luo , Wenjun Zhang , Chi-Man Pun , Zhen Wang

Low-light image enhancement (LLIE) techniques attempt to increase the visibility of images captured in low-light scenarios. However, as a result of enhancement, a variety of image degradations such as noise and color bias are revealed.…

Image and Video Processing · Electrical Eng. & Systems 2024-09-10 Savvas Panagiotou , Anna S. Bosman

Varicolored haze caused by chromatic casts poses haze removal and depth estimation challenges. Recent learning-based depth estimation methods are mainly targeted at dehazing first and estimating depth subsequently from haze-free scenes.…

Computer Vision and Pattern Recognition · Computer Science 2023-03-14 Sixiang Chen , Tian Ye , Jun Shi , Yun Liu , JingXia Jiang , Erkang Chen , Peng Chen

Image dehazing is fundamental yet not well-solved in computer vision. Most cutting-edge models are trained in synthetic data, leading to the poor performance on real-world hazy scenarios. Besides, they commonly give deterministic dehazed…

Computer Vision and Pattern Recognition · Computer Science 2022-10-31 Ming Tong , Yongzhen Wang , Peng Cui , Xuefeng Yan , Mingqiang Wei

Image light source transfer (LLST), as the most challenging task in the domain of image relighting, has attracted extensive attention in recent years. In the latest research, LLST is decomposed three sub-tasks: scene reconversion, shadow…

Computer Vision and Pattern Recognition · Computer Science 2021-04-22 Yuanzhi Wang , Tao Lu , Yanduo Zhang , Yuntao Wu

Existing RGB-Event visual object tracking approaches primarily rely on conventional feature-level fusion, failing to fully exploit the unique advantages of event cameras. In particular, the high dynamic range and motion-sensitive nature of…

Computer Vision and Pattern Recognition · Computer Science 2026-01-06 Shiao Wang , Xiao Wang , Haonan Zhao , Jiarui Xu , Bo Jiang , Lin Zhu , Xin Zhao , Yonghong Tian , Jin Tang

Vision-centric perception systems for autonomous driving have gained considerable attention recently due to their cost-effectiveness and scalability, especially compared to LiDAR-based systems. However, these systems often struggle in…

Computer Vision and Pattern Recognition · Computer Science 2024-04-09 Jinlong Li , Baolu Li , Zhengzhong Tu , Xinyu Liu , Qing Guo , Felix Juefei-Xu , Runsheng Xu , Hongkai Yu

By hiding the front-facing camera below the display panel, Under-Display Camera (UDC) provides users with a full-screen experience. However, due to the characteristics of the display, images taken by UDC suffer from significant quality…

Computer Vision and Pattern Recognition · Computer Science 2023-12-04 Jingfan Tan , Xiaoxu Chen , Tao Wang , Kaihao Zhang , Wenhan Luo , Xiaocun Cao

The depth completion task is a critical problem in autonomous driving, involving the generation of dense depth maps from sparse depth maps and RGB images. Most existing methods employ a spatial propagation network to iteratively refine the…

Computer Vision and Pattern Recognition · Computer Science 2026-02-03 Ming Yuan , Chuang Zhang , Lei He , Qing Xu , Jianqiang Wang

Many learning-based low-light image enhancement (LLIE) algorithms are based on the Retinex theory. However, the Retinex-based decomposition techniques in such models introduce corruptions which limit their enhancement performance. In this…

Computer Vision and Pattern Recognition · Computer Science 2024-08-13 Zhihao Zheng , Mooi Choo Chuah

Depth enhancement, which uses RGB images as guidance to convert raw signals from dToF into high-precision, dense depth maps, is a critical task in computer vision. Although existing super-resolution-based methods show promising results on…

Computer Vision and Pattern Recognition · Computer Science 2025-07-30 Jijun Xiang , Xuan Zhu , Xianqi Wang , Yu Wang , Hong Zhang , Fei Guo , Xin Yang

This paper investigates the problem of dim frequency line detection and recovery in the so-called lofargram. Theoretically, time integration long enough can always enhance the detection characteristic. But this does not hold for irregularly…

Signal Processing · Electrical Eng. & Systems 2020-12-02 Yina Han , Yuyan Li , Qingyu Liu , Yuanliang Ma

Referring image segmentation is a fundamental vision-language task that aims to segment out an object referred to by a natural language expression from an image. One of the key challenges behind this task is leveraging the referring…

Computer Vision and Pattern Recognition · Computer Science 2022-04-07 Zhao Yang , Jiaqi Wang , Yansong Tang , Kai Chen , Hengshuang Zhao , Philip H. S. Torr

Although recent works based on deep learning have made progress in improving recognition accuracy on scene text recognition, how to handle low-quality text images in end-to-end deep networks remains a research challenge. In this paper, we…

Computer Vision and Pattern Recognition · Computer Science 2021-08-16 Zhiwei Jia , Shugong Xu , Shiyi Mu , Yue Tao , Shan Cao , Zhiyong Chen