English
Related papers

Related papers: RGB-Multispectral Matching: Dataset, Learning Meth…

200 papers

In this paper, we propose a neural network architecture for scale-invariant semantic segmentation using RGB-D images. We utilize depth information as an additional modality apart from color images only. Especially in an outdoor scene which…

Computer Vision and Pattern Recognition · Computer Science 2022-04-12 Mohammad Dawud Ansari , Alwi Husada , Didier Stricker

Semantic labeling of RGB-D scenes is crucial to many intelligent applications including perceptual robotics. It generates pixelwise and fine-grained label maps from simultaneously sensed photometric (RGB) and depth channels. This paper…

Computer Vision and Pattern Recognition · Computer Science 2016-07-27 Zhen Li , Yukang Gan , Xiaodan Liang , Yizhou Yu , Hui Cheng , Liang Lin

Object segmentation for robotic grasping under dynamic conditions often faces challenges such as occlusion, low light conditions, motion blur and object size variance. To address these challenges, we propose a Deep Learning network that…

Computer Vision and Pattern Recognition · Computer Science 2025-12-09 Sanket Kachole , Xiaoqian Huang , Fariborz Baghaei Naeini , Rajkumar Muthusamy , Dimitrios Makris , Yahya Zweiri

Existing RGB-Thermal Video Object Detection (RGBT VOD) methods predominantly rely on the manual alignment of image pairs, that is both labor-intensive and time-consuming. This dependency significantly restricts the scalability and practical…

Computer Vision and Pattern Recognition · Computer Science 2025-04-21 Qishun Wang , Zhengzheng Tu , Kunpeng Wang , Le Gu , Chuanwang Guo

Successful navigation in outdoor environments requires accurate prediction of the physical interactions between the robot and the terrain. Many prior methods rely on geometric or semantic labels to classify traversable surfaces. However,…

Robotics · Computer Science 2025-12-01 Sarvesh Prajapati , Ananya Trivedi , Nathaniel Hanson , Bruce Maxwell , Taskin Padir

Digital camera pipelines employ color constancy methods to estimate an unknown scene illuminant, in order to re-illuminate images as if they were acquired under an achromatic light source. Fully-supervised learning approaches exhibit…

Computer Vision and Pattern Recognition · Computer Science 2019-04-05 Steven McDonagh , Sarah Parisot , Fengwei Zhou , Xing Zhang , Ales Leonardis , Zhenguo Li , Gregory Slabaugh

RGBT tracking receives a surge of interest in the computer vision community, but this research field lacks a large-scale and high-diversity benchmark dataset, which is essential for both the training of deep RGBT trackers and the…

Computer Vision and Pattern Recognition · Computer Science 2021-12-22 Chenglong Li , Wanlin Xue , Yaqing Jia , Zhichen Qu , Bin Luo , Jin Tang , Dengdi Sun

Self-diagnosis and self-repair are some of the key challenges in deploying robotic platforms for long-term real-world applications. One of the issues that can occur to a robot is miscalibration of its sensors due to aging, environmental…

Robotics · Computer Science 2020-05-26 Andrei Cramariuc , Aleksandar Petrov , Rohit Suri , Mayank Mittal , Roland Siegwart , Cesar Cadena

This paper proposes a new method called Multimodal RNNs for RGB-D scene semantic segmentation. It is optimized to classify image pixels given two input sources: RGB color channels and Depth maps. It simultaneously performs training of two…

Computer Vision and Pattern Recognition · Computer Science 2018-03-14 Abrar H. Abdulnabi , Bing Shuai , Zhen Zuo , Lap-Pui Chau , Gang Wang

We present three multi-scale similarity learning architectures, or DeepSim networks. These models learn pixel-level matching with a contrastive loss and are agnostic to the geometry of the considered scene. We establish a middle ground…

Computer Vision and Pattern Recognition · Computer Science 2023-04-18 Mohamed Ali Chebbi , Ewelina Rupnik , Marc Pierrot-Deseilligny , Paul Lopes

Developing and integrating advanced image sensors with novel algorithms in camera systems is prevalent with the increasing demand for computational photography and imaging on mobile platforms. However, the lack of high-quality data for…

Computer Vision and Pattern Recognition · Computer Science 2022-09-16 Wenxiu Sun , Qingpeng Zhu , Chongyi Li , Ruicheng Feng , Shangchen Zhou , Jun Jiang , Qingyu Yang , Chen Change Loy , Jinwei Gu

Super Resolution is the problem of recovering a high-resolution image from a single or multiple low-resolution images of the same scene. It is an ill-posed problem since high frequency visual details of the scene are completely lost in…

Image and Video Processing · Electrical Eng. & Systems 2020-04-22 Hamid Reza Vaezi Joze , Ilya Zharkov , Karlton Powell , Carl Ringler , Luming Liang , Andy Roulston , Moshe Lutz , Vivek Pradeep

We propose a new deep learning architecture for the tasks of semantic segmentation and depth prediction from RGB-D images. We revise the state of art based on the RGB and depth feature fusion, where both modalities are assumed to be…

Artificial Intelligence · Computer Science 2018-12-18 Giorgio Giannone , Boris Chidlovskii

This paper considers matching images of low-light scenes, aiming to widen the frontier of SfM and visual SLAM applications. Recent image sensors can record the brightness of scenes with more than eight-bit precision, available in their…

Computer Vision and Pattern Recognition · Computer Science 2021-09-15 Wenzheng Song , Masanori Suganuma , Xing Liu , Noriyuki Shimobayashi , Daisuke Maruta , Takayuki Okatani

We present MS-Splatting -- a multi-spectral 3D Gaussian Splatting (3DGS) framework that is able to generate multi-view consistent novel views from images of multiple, independent cameras with different spectral domains. In contrast to…

Graphics · Computer Science 2026-02-17 Lukas Meyer , Josef Grün , Maximilian Weiherer , Bernhard Egger , Marc Stamminger , Linus Franke

Referring Multi-Object Tracking (RMOT) aims to track specific targets based on language descriptions and is vital for interactive AI systems such as robotics and autonomous driving. However, existing RMOT models rely solely on 2D RGB data,…

Computer Vision and Pattern Recognition · Computer Science 2026-02-09 Sijia Chen , Lijuan Ma , Yanqiu Yu , En Yu , Liman Liu , Wenbing Tao

Multiple human tracking (MHT) is a fundamental task in many computer vision applications. Appearance-based approaches, primarily formulated on RGB data, are constrained and affected by problems arising from occlusions and/or illumination…

Computer Vision and Pattern Recognition · Computer Science 2016-06-15 Massimo Camplani , Adeline Paiement , Majid Mirmehdi , Dima Damen , Sion Hannuna , Tilo Burghardt , Lili Tao

Correspondence estimation is one of the most widely researched and yet only partially solved area of computer vision with many applications in tracking, mapping, recognition of objects and environment. In this paper, we propose a novel way…

Computer Vision and Pattern Recognition · Computer Science 2020-04-16 Umashankar Deekshith , Nishit Gajjar , Max Schwarz , Sven Behnke

While traditional methods relies on depth sensors, the current trend leans towards utilizing cost-effective RGB images, despite their absence of depth cues. This paper introduces an interesting approach to detect grasping pose from a single…

Computer Vision and Pattern Recognition · Computer Science 2023-10-31 Zhaocong Li

Recent guided depth super-resolution methods are premised on the assumption of strict spatial alignment between depth and RGB, achieving high-quality depth reconstruction. However, in real-world scenarios, the acquisition of strictly…

Computer Vision and Pattern Recognition · Computer Science 2026-05-19 Zhengxue Wang , Zhiqiang Yan , Yuan Wu , Guangwei Gao , Xiang Li , Jian Yang