English
Related papers

Related papers: RGB2LIDAR: Towards Solving Large-Scale Cross-Modal…

200 papers

Global visual localization in LiDAR-maps, crucial for autonomous driving applications, remains largely unexplored due to the challenging issue of bridging the cross-modal heterogeneity gap. Popular multi-modal learning approach Contrastive…

Robotics · Computer Science 2023-12-29 Sai Shubodh Puligilla , Mohammad Omama , Husain Zaidi , Udit Singh Parihar , Madhava Krishna

This article presents an innovative study in exploring, evaluating, and implementing deep learning architectures for the calibration of multi-modal sensor systems. The focus behind this is to leverage the use of sensor fusion to achieve…

Computer Vision and Pattern Recognition · Computer Science 2024-09-24 Venkat Karramreddy , Liam Mitchell

Sensing the medical scenario can ensure the safety during the surgical operations. So, in this regard, a monitor platform which can obtain the accurate location information of the surgery room is desperately needed. Compared to 2D camera…

Computer Vision and Pattern Recognition · Computer Science 2018-09-06 Ke Wang , Han Song , Jiahui Zhang , Xinran Zhang , Hongen Liao

We present a method for finding cross-modal space-time correspondences. Given two images from different visual modalities, such as an RGB image and a depth map, our model identifies which pairs of pixels correspond to the same physical…

Computer Vision and Pattern Recognition · Computer Science 2025-06-04 Ayush Shrivastava , Andrew Owens

This paper addresses the problem of 3D referring expression comprehension (REC) in autonomous driving scenario, which aims to ground a natural language to the targeted region in LiDAR point clouds. Previous approaches for REC usually focus…

Computer Vision and Pattern Recognition · Computer Science 2023-05-26 Wenhao Cheng , Junbo Yin , Wei Li , Ruigang Yang , Jianbing Shen

This work investigates the use of machine learning applied to the beam tracking problem in 5G networks and beyond. The goal is to decrease the overhead associated to MIMO millimeter wave beamforming. In comparison to beam selection (also…

Signal Processing · Electrical Eng. & Systems 2024-12-10 Ailton Oliveira , Daniel Suzuki , Sávio Bastos , Ilan Correa , Aldebaro Klautau

Global localisation from visual data is a challenging problem applicable to many robotics domains. Prior works have shown that neural networks can be trained to map images of an environment to absolute camera pose within that environment,…

Robotics · Computer Science 2024-01-03 Christopher J. Holder , Muhammad Shafique

We introduce a simple yet effective fusion method of LiDAR and RGB data to segment LiDAR point clouds. Utilizing the dense native range representation of a LiDAR sensor and the setup calibration, we establish point correspondences between…

Computer Vision and Pattern Recognition · Computer Science 2019-12-20 Georg Krispel , Michael Opitz , Georg Waltner , Horst Possegger , Horst Bischof

Visual object tracking, as a fundamental task in computer vision, has drawn much attention in recent years. To extend trackers to a wider range of applications, researchers have introduced information from multiple modalities to handle…

Computer Vision and Pattern Recognition · Computer Science 2020-12-09 Pengyu Zhang , Dong Wang , Huchuan Lu

LiDAR mapping is important yet challenging in self-driving and mobile robotics. To tackle such a global point cloud registration problem, DeepMapping converts the complex map estimation into a self-supervised training of simple deep…

Computer Vision and Pattern Recognition · Computer Science 2023-06-16 Chao Chen , Xinhao Liu , Yiming Li , Li Ding , Chen Feng

This work addresses visual cross-view metric localization for outdoor robotics. Given a ground-level color image and a satellite patch that contains the local surroundings, the task is to identify the location of the ground camera within…

Computer Vision and Pattern Recognition · Computer Science 2022-08-19 Zimin Xia , Olaf Booij , Marco Manfredi , Julian F. P. Kooij

This paper addresses the problem of vehicle-mounted camera localization by matching a ground-level image with an overhead-view satellite map. Existing methods often treat this problem as cross-view image retrieval, and use learned deep…

Computer Vision and Pattern Recognition · Computer Science 2022-09-07 Yujiao Shi , Hongdong Li

End-to-end autonomous driving solutions, which directly process multimodal sensory data and output fine-grained control commands, have gradually become a mainstream direction with the development of autonomous driving technology. However,…

Computer Vision and Pattern Recognition · Computer Science 2026-05-25 Runyi Huang , Ni Ding , Ruidan Xing , Yuheng Shi , Lei He , Keqiang Li

A novel relative localization approach for guidance of a micro-scale Unmanned Aerial Vehicle (UAV) by a well-equipped aerial robot fusing Visual-Inertial Odometry (VIO) with Light Detection and Ranging (LiDAR) is proposed in this paper.…

Robotics · Computer Science 2026-03-05 Václav Pritzl , Matouš Vrba , Petr Štěpán , Martin Saska

Stack interchanges are essential components of transportation systems. Mobile laser scanning (MLS) systems have been widely used in road infrastructure mapping, but accurate mapping of complicated multi-layer stack interchanges are still…

Computer Vision and Pattern Recognition · Computer Science 2021-10-22 Weikai Tan , Dedong Zhang , Lingfei Ma , Ying Li , Lanying Wang , Jonathan Li

We present an end-to-end method for object detection and trajectory prediction utilizing multi-view representations of LiDAR returns and camera images. In this work, we recognize the strengths and weaknesses of different view…

Computer Vision and Pattern Recognition · Computer Science 2021-10-20 Sudeep Fadadu , Shreyash Pandey , Darshan Hegde , Yi Shi , Fang-Chieh Chou , Nemanja Djuric , Carlos Vallespi-Gonzalez

We consider the problem of 3D object pose estimation. While much recent work has focused on the RGB domain, the reliance on accurately annotated images limits their generalizability and scalability. On the other hand, the easily available…

Computer Vision and Pattern Recognition · Computer Science 2019-08-01 Georgios Georgakis , Srikrishna Karanam , Ziyan Wu , Jana Kosecka

We present CrossLoc3D, a novel 3D place recognition method that solves a large-scale point matching problem in a cross-source setting. Cross-source point cloud data corresponds to point sets captured by depth sensors with different…

Computer Vision and Pattern Recognition · Computer Science 2023-10-03 Tianrui Guan , Aswath Muthuselvam , Montana Hoover , Xijun Wang , Jing Liang , Adarsh Jagan Sathyamoorthy , Damon Conover , Dinesh Manocha

High-definition (HD) semantic mapping of complex intersections poses significant challenges for traditional vehicle-based approaches due to occlusions and limited perspectives. This paper introduces a novel camera-LiDAR fusion framework…

Robotics · Computer Science 2025-07-15 Zhongzhang Chen , Miao Fan , Shengtong Xu , Mengmeng Yang , Kun Jiang , Xiangzeng Liu , Haoyi Xiong

Despite the significant progress in 6-DoF visual localization, researchers are mostly driven by ground-level benchmarks. Compared with aerial oblique photography, ground-level map collection lacks scalability and complete coverage. In this…

Computer Vision and Pattern Recognition · Computer Science 2024-07-09 Shen Yan , Xiaoya Cheng , Yuxiang Liu , Juelin Zhu , Rouwan Wu , Yu Liu , Maojun Zhang
‹ Prev 1 4 5 6 7 8 10 Next ›