English
Related papers

Related papers: VioLA: Aligning Videos to 2D LiDAR Scans

200 papers

Accurate camera pose estimation result is essential for visual SLAM (VSLAM). This paper presents a novel pose correction method to improve the accuracy of the VSLAM system. Firstly, the relationship between the camera pose estimation error…

Computer Vision and Pattern Recognition · Computer Science 2019-08-27 Zhaobing Kang , Wei Zou , Zheng Zhu

LiDAR-based 3D panoptic segmentation often struggles with the inherent sparsity of data from LiDAR sensors, which makes it challenging to accurately recognize distant or small objects. Recently, a few studies have sought to overcome this…

Computer Vision and Pattern Recognition · Computer Science 2025-06-11 Yining Pan , Qiongjie Cui , Xulei Yang , Na Zhao

It has recently been discovered that using a pre-trained vision-language model (VLM), e.g., CLIP, to align a whole query image with several finer text descriptions generated by a large language model can significantly enhance zero-shot…

Computer Vision and Pattern Recognition · Computer Science 2024-06-06 Jinhao Li , Haopeng Li , Sarah Erfani , Lei Feng , James Bailey , Feng Liu

Reliable image correspondences form the foundation of vision-based spatial perception, enabling recovery of 3D structure and camera poses. However, unconstrained feature matching across domains such as aerial, indoor, and outdoor scenes…

Computer Vision and Pattern Recognition · Computer Science 2025-12-30 Zhimin Shao , Abhay Yadav , Rama Chellappa , Cheng Peng

Reconstructing an accurate and consistent large-scale LiDAR point cloud map is crucial for robotics applications. The existing solution, pose graph optimization, though it is time-efficient, does not directly optimize the mapping…

Robotics · Computer Science 2022-09-27 Xiyuan Liu , Zheng Liu , Fanze Kong , Fu Zhang

Vision-Language-Action (VLA) models offer a compelling framework for tackling complex robotic manipulation tasks, but they are often expensive to train. In this paper, we propose a novel VLA approach that leverages the competitive…

Robotics · Computer Science 2025-12-23 Max Argus , Jelena Bratulic , Houman Masnavi , Maxim Velikanov , Nick Heppert , Abhinav Valada , Thomas Brox

Visual relocalization has been a widely discussed problem in 3D vision: given a pre-constructed 3D visual map, the 6 DoF (Degrees-of-Freedom) pose of a query image is estimated. Relocalization in large-scale indoor environments enables…

Computer Vision and Pattern Recognition · Computer Science 2022-07-27 Jiahui Zhang , Shitao Tang , Kejie Qiu , Rui Huang , Chuan Fang , Le Cui , Zilong Dong , Siyu Zhu , Ping Tan

Place recognition is a core component of Simultaneous Localization and Mapping (SLAM) algorithms. Particularly in visual SLAM systems, previously-visited places are recognized by measuring the appearance similarity between images…

Computer Vision and Pattern Recognition · Computer Science 2020-07-28 Jiawei Mo , Junaed Sattar

Information inside visual and LiDAR data is well complementary derived from the fine-grained texture of images and massive geometric information in point clouds. However, it remains challenging to explore effective visual-LiDAR fusion,…

Computer Vision and Pattern Recognition · Computer Science 2024-07-18 Jiuming Liu , Dong Zhuo , Zhiheng Feng , Siting Zhu , Chensheng Peng , Zhe Liu , Hesheng Wang

Vision-language models (VLMs), such as CLIP and ALIGN, are generally trained on datasets consisting of image-caption pairs obtained from the web. However, real-world multimodal datasets, such as healthcare data, are significantly more…

Computer Vision and Pattern Recognition · Computer Science 2023-08-23 Maya Varma , Jean-Benoit Delbrouck , Sarah Hooper , Akshay Chaudhari , Curtis Langlotz

Vision-Language-Action (VLA) models provide a promising paradigm for robot learning by integrating visual perception with language-guided policy learning. However, most existing approaches rely on 2D visual inputs to perform actions in 3D…

Robotics · Computer Science 2025-12-16 Yicheng Feng , Wanpeng Zhang , Ye Wang , Hao Luo , Haoqi Yuan , Sipeng Zheng , Zongqing Lu

In this letter, we present a novel method for automatic extrinsic calibration of high-resolution LiDARs and RGB cameras in targetless environments. Our approach does not require checkerboards but can achieve pixel-level accuracy by aligning…

Robotics · Computer Science 2021-06-28 Chongjian Yuan , Xiyuan Liu , Xiaoping Hong , Fu Zhang

We propose CAL (Complete Anything in Lidar) for Lidar-based shape-completion in-the-wild. This is closely related to Lidar-based semantic/panoptic scene completion. However, contemporary methods can only complete and recognize objects from…

Computer Vision and Pattern Recognition · Computer Science 2025-04-17 Ayca Takmaz , Cristiano Saltori , Neehar Peri , Tim Meinhardt , Riccardo de Lutio , Laura Leal-Taixé , Aljoša Ošep

In this paper, an automatic method is proposed to perform image registration in visible and infrared pair of video sequences for multiple targets. In multimodal image analysis like image fusion systems, color and IR sensors are placed close…

Computer Vision and Pattern Recognition · Computer Science 2014-03-18 Tanushri Chakravorty , Guillaume-Alexandre Bilodeau , Eric Granger

Lidar point cloud distortion from moving object is an important problem in autonomous driving, and recently becomes even more demanding with the emerging of newer lidars, which feature back-and-forth scanning patterns. Accurately estimating…

Robotics · Computer Science 2022-07-05 Wen Yang , Zheng Gong , Baifu Huang , Xiaoping Hong

Autonomous navigation is one of the key requirements for every potential application of mobile robots in the real-world. Besides high-accuracy state estimation, a suitable and globally consistent representation of the 3D environment is…

Robotics · Computer Science 2024-03-05 Simon Boche , Sebastián Barbas Laina , Stefan Leutenegger

Calibrating a robot simulator's physics parameters (friction, damping, material stiffness) to match real hardware is often done by hand or with black-box optimizers that reduce error but cannot explain which physical discrepancies drive the…

Robotics · Computer Science 2026-02-24 Kevin Qiu , Yu Zhang , Marek Cygan , Josie Hughes

The precise estimation of camera poses within large camera networks is a foundational problem in computer vision and robotics, with broad applications spanning autonomous navigation, surveillance, and augmented reality. In this paper, we…

Computer Vision and Pattern Recognition · Computer Science 2025-04-15 Gabriel Moreira , Manuel Marques , João Paulo Costeira , Alexander Hauptmann

We propose a novel real-time LiDAR intensity image-based simultaneous localization and mapping method , which addresses the geometry degeneracy problem in unstructured environments. Traditional LiDAR-based front-end odometry mostly relies…

Computer Vision and Pattern Recognition · Computer Science 2023-06-21 Wenqiang Du , Giovanni Beltrame

Calibration of multi-camera systems is a key task for accurate object tracking. However, it remains a challenging problem in real-world conditions, where traditional methods are not applicable due to the lack of accurate floor plans,…

Image and Video Processing · Electrical Eng. & Systems 2025-12-08 Aleksandr Abramov