English
Related papers

Related papers: Monocular Visual Teach and Repeat Aided by Local G…

200 papers

We propose a novel monocular visual odometry (VO) system called UnDeepVO in this paper. UnDeepVO is able to estimate the 6-DoF pose of a monocular camera and the depth of its view by using deep neural networks. There are two salient…

Computer Vision and Pattern Recognition · Computer Science 2018-02-22 Ruihao Li , Sen Wang , Zhiqiang Long , Dongbing Gu

Relative localization is an important ability for multiple robots to perform cooperative tasks in GPS-denied environment. This paper presents a novel autonomous positioning framework for monocular relative localization of multiple tiny…

Robotics · Computer Science 2021-09-23 Shushuai Li , Christophe De Wagter , Guido C. H. E. de Croon

Convolutions on monocular dash cam videos capture spatial invariances in the image plane but do not explicitly reason about distances and depth. We propose a simple transformation of observations into a bird's eye view, also known as plan…

Computer Vision and Pattern Recognition · Computer Science 2019-05-17 Dequan Wang , Coline Devin , Qi-Zhi Cai , Philipp Krähenbühl , Trevor Darrell

While expensive LiDAR and stereo camera rigs have enabled the development of successful 3D object detection methods, monocular RGB-only approaches lag much behind. This work advances the state of the art by introducing MoVi-3D, a novel,…

Computer Vision and Pattern Recognition · Computer Science 2020-04-03 Andrea Simonelli , Samuel Rota Bulò , Lorenzo Porzi , Elisa Ricci , Peter Kontschieder

This paper introduces a cost effective localization system combining monocular visual odometry , augmented reality (AR) poses, and integrated INS-GPS data. We address monocular VO scale factor issues using AR poses and enhance accuracy with…

Robotics · Computer Science 2024-11-26 Ankit Shaw

The application of monocular dense Simultaneous Localization and Mapping (SLAM) is often hindered by high latency, large GPU memory consumption, and reliance on camera calibration. To relax this constraint, we propose EC3R-SLAM, a novel…

Robotics · Computer Science 2025-10-03 Lingxiang Hu , Naima Ait Oufroukh , Fabien Bonardi , Raymond Ghandour

Teach and repeat is a rapid way to achieve autonomy in challenging terrain and off-road environments. A human operator pilots the vehicles to create a network of paths that are mapped and associated with odometry. Immediately after…

Robotics · Computer Science 2025-06-25 Matěj Boxan , Alexander Krawciw , Timothy D. Barfoot , François Pomerleau

Self-supervised monocular depth estimation enables robots to learn 3D perception from raw video streams. This scalable approach leverages projective geometry and ego-motion to learn via view synthesis, assuming the world is mostly static.…

Computer Vision and Pattern Recognition · Computer Science 2022-06-14 Vitor Guizilini , Kuan-Hui Lee , Rares Ambrus , Adrien Gaidon

Off-road autonomous navigation demands reliable 3D perception for robust obstacle detection in challenging unstructured terrain. While LiDAR is accurate, it is costly and power-intensive. Monocular depth estimation using foundation models…

Monocular visual navigation methods have seen significant advances in the last decade, recently producing several real-time solutions for autonomously navigating small unmanned aircraft systems without relying on GPS. This is critical for…

Computer Vision and Pattern Recognition · Computer Science 2020-09-24 Kyung Kim , Robert C. Leishman , Scott L. Nykl

We propose a complete pipeline that allows object detection and simultaneously estimate the pose of these multiple object instances using just a single image. A novel "keypoint regression" scheme with a cross-ratio term is introduced that…

Computer Vision and Pattern Recognition · Computer Science 2018-09-28 Ankit Dhall

Cooperative Simultaneous Localization and Mapping (C-SLAM) enables multiple agents to work together in mapping unknown environments while simultaneously estimating their own positions. This approach enhances robustness, scalability, and…

Robotics · Computer Science 2025-08-28 Joshua Bird , Jan Blumenkamp , Amanda Prorok

Visual SLAM (Simultaneous Localization and Mapping) methods typically rely on handcrafted visual features or raw RGB values for establishing correspondences between images. These features, while suitable for sparse mapping, often lead to…

Computer Vision and Pattern Recognition · Computer Science 2018-11-21 Chamara Saroj Weerasekera , Ravi Garg , Yasir Latif , Ian Reid

This letter introduces a novel framework for dense Visual Simultaneous Localization and Mapping (VSLAM) based on Gaussian Splatting. Recently, SLAM based on Gaussian Splatting has shown promising results. However, in monocular scenarios,…

Computer Vision and Pattern Recognition · Computer Science 2024-09-11 Pengcheng Zhu , Yaoming Zhuang , Baoquan Chen , Li Li , Chengdong Wu , Zhanlin Liu

Monocular cameras are extensively employed in indoor robotics, but their performance is limited in visual odometry, depth estimation, and related applications due to the absence of scale information.Depth estimation refers to the process of…

Robotics · Computer Science 2023-09-15 Yehao Liu , Ruoyan Xia , Xiaosu Xu , Zijian Wang , Yiqing Ya , Mingze Fan

This paper proposes a multi-sensor based approach to detect, track, and localize a quadcopter unmanned aerial vehicle (UAV). Specifically, a pipeline is developed to process monocular RGB and thermal video (captured from a fixed platform)…

Computer Vision and Pattern Recognition · Computer Science 2020-12-10 Zhibo Zhang , Chen Zeng , Maulikkumar Dhameliya , Souma Chowdhury , Rahul Rai

Semantic aware reconstruction is more advantageous than geometric-only reconstruction for future robotic and AR/VR applications because it represents not only where things are, but also what things are. Object-centric mapping is a task to…

Computer Vision and Pattern Recognition · Computer Science 2021-02-16 Kejie Li , Hamid Rezatofighi , Ian Reid

A monocular 3D object tracking system generally has only up-to-scale pose estimation results without any prior knowledge of the tracked object. In this paper, we propose a novel idea to recover the metric scale of an arbitrary dynamic…

Robotics · Computer Science 2018-08-22 Kejie Qiu , Tong Qin , Hongwen Xie , Shaojie Shen

The monocular visual-inertial system (VINS), which consists one camera and one low-cost inertial measurement unit (IMU), is a popular approach to achieve accurate 6-DOF state estimation. However, such locally accurate visual-inertial…

Computer Vision and Pattern Recognition · Computer Science 2018-03-06 Tong Qin , Perliang Li , Shaojie Shen

Autonomous exploration of unknown environments is a key capability for mobile robots, but it is largely unsolved for robots equipped with only a single monocular camera and no dense range sensors. In this paper, we present a novel approach…

Robotics · Computer Science 2026-05-15 Tomáš Musil , Matěj Petrlík , Martin Saska