English
Related papers

Related papers: Pixel-aligned RGB-NIR Stereo Imaging and Dataset f…

200 papers

The goal of this work is to replace objects in an RGB-D scene with corresponding 3D models from a library. We approach this problem by first detecting and segmenting object instances in the scene using the approach from Gupta et al. [13].…

Computer Vision and Pattern Recognition · Computer Science 2015-02-17 Saurabh Gupta , Pablo Arbeláez , Ross Girshick , Jitendra Malik

Autonomous driving systems rely heavily on robust sensor fusion to perceive complex envi- ronments. Traditional setups using RGB cameras and LiDAR often struggle in high-dynamic- range scenes or high-speed scenarios due to motion blur and…

Computer Vision and Pattern Recognition · Computer Science 2026-05-07 Mustafa Sakhaia , Kaung Sithua , Min Khant Soe Okea , Maciej Wielgosza

The complementary fusion of light detection and ranging (LiDAR) data and image data is a promising but challenging task for generating high-precision and high-density point clouds. This study proposes an innovative LiDAR-guided stereo…

Computer Vision and Pattern Recognition · Computer Science 2022-02-25 Yongjun Zhang , Siyuan Zou , Xinyi Liu , Xu Huang , Yi Wan , Yongxiang Yao

Single-photon light detection and ranging (LiDAR) has been widely applied to 3D imaging in challenging scenarios. However, limited signal photon counts and high noises in the collected data have posed great challenges for predicting the…

Image and Video Processing · Electrical Eng. & Systems 2022-06-30 Gongxin Yao , Yiwei Chen , Yong Liu , Xiaomin Hu , Yu Pan

Current simultaneous localization and mapping (SLAM) algorithms perform well in static environments but easily fail in dynamic environments. Recent works introduce deep learning-based semantic information to SLAM systems to reduce the…

Robotics · Computer Science 2023-04-24 Jianheng Liu , Xuanfu Li , Yueqian Liu , Haoyao Chen

There are two critical sensors for 3D perception in autonomous driving, the camera and the LiDAR. The camera provides rich semantic information such as color, texture, and the LiDAR reflects the 3D shape and locations of surrounding…

Computer Vision and Pattern Recognition · Computer Science 2022-05-31 Kaicheng Yu , Tang Tao , Hongwei Xie , Zhiwei Lin , Zhongwei Wu , Zhongyu Xia , Tingting Liang , Haiyang Sun , Jiong Deng , Dayang Hao , Yongtao Wang , Xiaodan Liang , Bing Wang

There is an emerging trend of using neural implicit functions for map representation in Simultaneous Localization and Mapping (SLAM). Some pioneer works have achieved encouraging results on RGB-D SLAM. In this paper, we present a dense RGB…

Computer Vision and Pattern Recognition · Computer Science 2023-02-21 Heng Li , Xiaodong Gu , Weihao Yuan , Luwei Yang , Zilong Dong , Ping Tan

Pixel binning is a technique, widely used in optical image acquisition and spectroscopy, in which adjacent detector elements of an image sensor are combined into larger pixels. This reduces the amount of data to be processed as well as the…

Image recognition models that work in challenging environments (e.g., extremely dark, blurry, or high dynamic range conditions) must be useful. However, creating training datasets for such environments is expensive and hard due to the…

Computer Vision and Pattern Recognition · Computer Science 2023-03-28 Masakazu Yoshimura , Junji Otsuka , Atsushi Irie , Takeshi Ohashi

Convolutional neural networks have been the focus of research aiming to solve image denoising problems, but their performance remains unsatisfactory for most applications. These networks are trained with synthetic noise distributions that…

Image and Video Processing · Electrical Eng. & Systems 2020-05-06 Benoit Brummer , Christophe De Vleeschouwer

Event based cameras are a new passive sensing modality with a number of benefits over traditional cameras, including extremely low latency, asynchronous data acquisition, high dynamic range and very low power consumption. There has been a…

Robotics · Computer Science 2018-02-21 Alex Zihao Zhu , Dinesh Thakur , Tolga Ozaslan , Bernd Pfrommer , Vijay Kumar , Kostas Daniilidis

With the popularity of dual cameras in recently released smart phones, a growing number of super-resolution (SR) methods have been proposed to enhance the resolution of stereo image pairs. However, the lack of high-quality stereo datasets…

Computer Vision and Pattern Recognition · Computer Science 2019-08-23 Yingqian Wang , Longguang Wang , Jungang Yang , Wei An , Yulan Guo

Text-guided multispectral object detection uses text semantics to guide semantic-aware cross-modal interaction between RGB and IR for more robust perception. However, notable limitations remain: (1) existing methods often use text only as…

Computer Vision and Pattern Recognition · Computer Science 2026-04-14 Jiaqi Wu , Zhen Wang , Enhao Huang , Kangqing Shen , Yulin Wang , Yang Yue , Yifan Pu , Gao Huang

Multi-view imaging systems enable uniform coverage of 3D space and reduce the impact of occlusion, which is beneficial for 3D object detection and tracking accuracy. However, existing imaging systems built with multi-view cameras or depth…

Computer Vision and Pattern Recognition · Computer Science 2023-02-22 Meng Zhang , Wenxuan Guo , Bohao Fan , Yifan Chen , Jianjiang Feng , Jie Zhou

Visual scene understanding is an important capability that enables robots to purposefully act in their environment. In this paper, we propose a novel approach to object-class segmentation from multiple RGB-D views using deep learning. We…

Computer Vision and Pattern Recognition · Computer Science 2017-12-06 Lingni Ma , Jörg Stückler , Christian Kerl , Daniel Cremers

Developing and integrating advanced image sensors with novel algorithms in camera systems is prevalent with the increasing demand for computational photography and imaging on mobile platforms. However, the lack of high-quality data for…

Computer Vision and Pattern Recognition · Computer Science 2022-09-16 Wenxiu Sun , Qingpeng Zhu , Chongyi Li , Ruicheng Feng , Shangchen Zhou , Jun Jiang , Qingyu Yang , Chen Change Loy , Jinwei Gu

In this paper we study the problem of object detection for RGB-D images using semantically rich image and depth features. We propose a new geocentric embedding for depth images that encodes height above ground and angle with gravity for…

Computer Vision and Pattern Recognition · Computer Science 2014-07-23 Saurabh Gupta , Ross Girshick , Pablo Arbeláez , Jitendra Malik

Recent denoising algorithms based on the "blind-spot" strategy show impressive blind image denoising performances, without utilizing any external dataset. While the methods excel in recovering highly contaminated images, we observe that…

Image and Video Processing · Electrical Eng. & Systems 2022-04-07 Chaewon Kim , Jaeho Lee , Jinwoo Shin

It is well known that the passive stereo system cannot adapt well to weak texture objects, e.g., white walls. However, these weak texture targets are very common in indoor environments. In this paper, we present a novel stereo system, which…

Computer Vision and Pattern Recognition · Computer Science 2022-03-22 Yuhua Xu , Xiaoli Yang , Yushan Yu , Wei Jia , Zhaobi Chu , Yulan Guo

While dense visual SLAM methods are capable of estimating dense reconstructions of the environment, they suffer from a lack of robustness in their tracking step, especially when the optimisation is poorly initialised. Sparse visual SLAM…

Robotics · Computer Science 2022-07-25 Tristan Laidlow , Michael Bloesch , Wenbin Li , Stefan Leutenegger