English
Related papers

Related papers: FuseSeg: LiDAR Point Cloud Segmentation Fusing Mul…

200 papers

Autonomous driving vehicles and robotic systems rely on accurate perception of their surroundings. Scene understanding is one of the crucial components of perception modules. Among all available sensors, LiDARs are one of the essential…

Computer Vision and Pattern Recognition · Computer Science 2021-03-17 Ryan Razani , Ran Cheng , Ehsan Taghavi , Liu Bingbing

Recently, the RGB images and point clouds fusion methods have been proposed to jointly estimate 2D optical flow and 3D scene flow. However, as both conventional RGB cameras and LiDAR sensors adopt a frame-based data acquisition mechanism,…

Computer Vision and Pattern Recognition · Computer Science 2023-09-27 Zhexiong Wan , Yuxin Mao , Jing Zhang , Yuchao Dai

Accurate moving object segmentation is an essential task for autonomous driving. It can provide effective information for many downstream tasks, such as collision avoidance, path planning, and static map construction. How to effectively…

Computer Vision and Pattern Recognition · Computer Science 2022-07-06 Jiadai Sun , Yuchao Dai , Xianjing Zhang , Jintao Xu , Rui Ai , Weihao Gu , Xieyuanli Chen

To realize low-latency spatial transmission system for immersive telepresence, there are two major problems: capturing dynamic 3D scene densely and processing them in real time. LiDAR sensors capture 3D in real time, but produce sparce…

Computer Vision and Pattern Recognition · Computer Science 2026-01-14 Kazuhiko Murasaki , Shunsuke Konagai , Masakatsu Aoki , Taiga Yoshida , Ryuichi Tanida

Semantic segmentation of indoor point clouds has found various applications in the creation of digital twins for robotics, navigation and building information modeling (BIM). However, most existing datasets of labeled indoor point clouds…

Computer Vision and Pattern Recognition · Computer Science 2025-01-30 Maxime Mérizette , Nicolas Audebert , Pierre Kervella , Jérôme Verdun

By identifying four important components of existing LiDAR-camera 3D object detection methods (LiDAR and camera candidates, transformation, and fusion outputs), we observe that all existing methods either find dense candidates or yield…

Computer Vision and Pattern Recognition · Computer Science 2023-04-28 Yichen Xie , Chenfeng Xu , Marie-Julie Rakotosaona , Patrick Rim , Federico Tombari , Kurt Keutzer , Masayoshi Tomizuka , Wei Zhan

Fusing data from cameras and LiDAR sensors is an essential technique to achieve robust 3D object detection. One key challenge in camera-LiDAR fusion involves mitigating the large domain gap between the two sensors in terms of coordinates…

Computer Vision and Pattern Recognition · Computer Science 2023-02-17 Yecheol Kim , Konyul Park , Minwook Kim , Dongsuk Kum , Jun Won Choi

Cross-modal data registration has long been a critical task in computer vision, with extensive applications in autonomous driving and robotics. Accurate and robust registration methods are essential for aligning data from different…

Computer Vision and Pattern Recognition · Computer Science 2025-03-20 Yuanchao Yue , Hui Yuan , Qinglong Miao , Xiaolong Mao , Raouf Hamzaoui , Peter Eisert

Constructing a point cloud for a large geographic region, such as a state or country, can require multiple years of effort. Often several vendors will be used to acquire LiDAR data, and a single region may be captured by multiple LiDAR…

Computer Vision and Pattern Recognition · Computer Science 2021-05-06 David Jones , Nathan Jacobs

Semantic segmentation of LiDAR point clouds has been widely studied in recent years, with most existing methods focusing on tackling this task using a single scan of the environment. However, leveraging the temporal stream of observations…

Computer Vision and Pattern Recognition · Computer Science 2023-11-06 Enxu Li , Sergio Casas , Raquel Urtasun

We propose a new method for fusing a LIDAR point cloud and camera-captured images in the deep convolutional neural network (CNN). The proposed method constructs a new layer called non-homogeneous pooling layer to transform features between…

Computer Vision and Pattern Recognition · Computer Science 2018-02-15 Zining Wang , Wei Zhan , Masayoshi Tomizuka

The primary requirement for cross-modal data fusion is the precise alignment of data from different sensors. However, the calibration between LiDAR point clouds and camera images is typically time-consuming and needs external calibration…

Computer Vision and Pattern Recognition · Computer Science 2025-07-11 Yuanchao Yue , Hui Yuan , Zhengxin Li , Shuai Li , Wei Zhang

In self-driving applications, LiDAR data provides accurate information about distances in 3D but lacks the semantic richness of camera data. Therefore, state-of-the-art methods for perception in urban scenes fuse data from both sensor…

Computer Vision and Pattern Recognition · Computer Science 2023-06-13 Royden Wagner , Marvin Klemp , Carlos Fernandez Lopez

As camera and LiDAR sensors capture complementary information used in autonomous driving, great efforts have been made to develop semantic segmentation algorithms through multi-modality data fusion. However, fusion-based approaches require…

Computer Vision and Pattern Recognition · Computer Science 2022-10-17 Xu Yan , Jiantao Gao , Chaoda Zheng , Chao Zheng , Ruimao Zhang , Shenghui Cui , Zhen Li

Earlier work demonstrates the promise of deep-learning-based approaches for point cloud segmentation; however, these approaches need to be improved to be practically useful. To this end, we introduce a new model SqueezeSegV2 that is more…

Computer Vision and Pattern Recognition · Computer Science 2018-09-25 Bichen Wu , Xuanyu Zhou , Sicheng Zhao , Xiangyu Yue , Kurt Keutzer

The use of infrastructure sensor technology for traffic detection has already been proven several times. However, extrinsic sensor calibration is still a challenge for the operator. While previous approaches are unable to calibrate the…

Computer Vision and Pattern Recognition · Computer Science 2020-08-04 Laurent Kloeker , Christian Kotulla , Lutz Eckstein

Multiple object tracking (MOT) is a significant task in achieving autonomous driving. Traditional works attempt to complete this task, either based on point clouds (PC) collected by LiDAR, or based on images captured from cameras. However,…

Computer Vision and Pattern Recognition · Computer Science 2022-03-31 Guangming Wang , Chensheng Peng , Jinpeng Zhang , Hesheng Wang

The worldwide commercialization of fifth generation (5G) wireless networks and the exciting possibilities offered by connected and autonomous vehicles (CAVs) are pushing toward the deployment of heterogeneous sensors for tracking dynamic…

Image and Video Processing · Electrical Eng. & Systems 2022-02-03 Francesco Nardo , Davide Peressoni , Paolo Testolina , Marco Giordani , Andrea Zanella

3D LiDAR scene completion from point clouds is a fundamental component of perception systems in autonomous vehicles. Previous methods have predominantly employed diffusion models for high-fidelity reconstruction. However, their multi-step…

Computer Vision and Pattern Recognition · Computer Science 2025-12-02 Wenzhe He , Xiaojun Chen , Ruiqi Wang , Ruihui Li , Huilong Pi , Jiapeng Zhang , Zhuo Tang , Kenli Li

Point cloud datasets for perception tasks in the context of autonomous driving often rely on high resolution 64-layer Light Detection and Ranging (LIDAR) scanners. They are expensive to deploy on real-world autonomous driving sensor…

Computer Vision and Pattern Recognition · Computer Science 2020-05-28 Leonardo Gigli , B Ravi Kiran , Thomas Paul , Andres Serna , Nagarjuna Vemuri , Beatriz Marcotegui , Santiago Velasco-Forero