English
Related papers

Related papers: DINO-VO: A Feature-based Visual Odometry Leveragin…

200 papers

Monocular visual-inertial odometry (VIO) is a low-cost solution to provide high-accuracy, low-drifting pose estimation. However, it has been meeting challenges in vehicular scenarios due to limited dynamics and lack of stable features. In…

Robotics · Computer Science 2023-06-21 Yuxuan Zhou , Xingxing Li , Shengyu Li , Xuanbin Wang , Zhiheng Shen

This paper studies monocular visual odometry (VO) problem. Most of existing VO algorithms are developed under a standard pipeline including feature extraction, feature matching, motion estimation, local optimisation, etc. Although some of…

Computer Vision and Pattern Recognition · Computer Science 2017-09-26 Sen Wang , Ronald Clark , Hongkai Wen , Niki Trigoni

A key task in embedded vision is visual odometry (VO), which estimates camera motion from visual sensors, and it is a core component in many embedded power-constrained systems, from autonomous robots to augmented and virtual reality…

Image and Video Processing · Electrical Eng. & Systems 2026-04-10 Alessandro Marchei , Lorenzo Lamberti , Daniele Palossi , Luca Benini

We introduce ZeroVO, a novel visual odometry (VO) algorithm that achieves zero-shot generalization across diverse cameras and environments, overcoming limitations in existing methods that depend on predefined or static camera calibration…

Computer Vision and Pattern Recognition · Computer Science 2025-06-10 Lei Lai , Zekai Yin , Eshed Ohn-Bar

State-of-the-art forward facing monocular visual-inertial odometry algorithms are often brittle in practice, especially whilst dealing with initialisation and motion in directions that render the state unobservable. In such cases having a…

Robotics · Computer Science 2019-05-15 Bo Fu , Kumar Shaurya Shankar , Nathan Michael

Most feature-based stereo visual odometry (SVO) approaches estimate the motion of mobile robots by matching and tracking point features along a sequence of stereo images. However, in dynamic scenes mainly comprising moving pedestrians,…

Computer Vision and Pattern Recognition · Computer Science 2023-06-27 Baosheng Zhang , Xiaoguang Ma , Hongjun Ma , Chunbo Luo

Robust feature matching forms the backbone for most Visual Simultaneous Localization and Mapping (vSLAM), visual odometry, 3D reconstruction, and Structure from Motion (SfM) algorithms. However, recovering feature matches from texture-poor…

Computer Vision and Pattern Recognition · Computer Science 2023-08-03 Shenbagaraj Kannapiran , Nalin Bendapudi , Ming-Yuan Yu , Devarth Parikh , Spring Berman , Ankit Vora , Gaurav Pandey

In the field of multi-sensor fusion for simultaneous localization and mapping (SLAM), monocular cameras and IMUs are widely used to build simple and effective visual-inertial systems. However, limited research has explored the integration…

We present the first learning-based visual odometry (VO) model, which generalizes to multiple datasets and real-world scenarios and outperforms geometry-based methods in challenging scenes. We achieve this by leveraging the SLAM dataset…

Computer Vision and Pattern Recognition · Computer Science 2020-11-03 Wenshan Wang , Yaoyu Hu , Sebastian Scherer

Visual-inertial odometry (VIO) is widely used in various fields, such as robots, drones, and autonomous vehicles. However, real-world scenes often feature dynamic objects, compromising the accuracy of VIO. The diversity and partial…

Computer Vision and Pattern Recognition · Computer Science 2025-03-04 Rui Zhou , Jingbin Liu , Junbin Xie , Jianyu Zhang , Yingze Hu , Jiele Zhao

Self-supervised visual foundation models produce powerful embeddings that achieve remarkable performance on a wide range of downstream tasks. However, unlike vision-language models such as CLIP, self-supervised visual features are not…

Generally, high-level features provide more geometrical information compared to point features, which can be exploited to further constrain motions. Planes are commonplace in man-made environments, offering an active means to reduce drift,…

Robotics · Computer Science 2025-05-20 Yidi Zhang , Fulin Tang , Zewen Xu , Yihong Wu , Pengju Ma

Visual-inertial odometry (VIO) has demonstrated remarkable success due to its low-cost and complementary sensors. However, existing VIO methods lack the generalization ability to adjust to different environments and sensor attributes. In…

Robotics · Computer Science 2024-05-28 Youqi Pan , Wugen Zhou , Yingdian Cao , Hongbin Zha

We introduce OpenVO, a novel framework for Open-world Visual Odometry (VO) with temporal awareness under limited input conditions. OpenVO effectively estimates real-world-scale ego-motion from monocular dashcam footage with varying…

Computer Vision and Pattern Recognition · Computer Science 2026-04-28 Phuc D. A. Nguyen , Anh N. Nhu , Ming C. Lin

Effectively localizing an agent in a realistic, noisy setting is crucial for many embodied vision tasks. Visual Odometry (VO) is a practical substitute for unreliable GPS and compass sensors, especially in indoor environments. While…

Computer Vision and Pattern Recognition · Computer Science 2023-05-02 Marius Memmel , Roman Bachmann , Amir Zamir

Visual Odometry (VO) is a method to estimate self-motion of a mobile robot using visual sensors. Unlike odometry based on integrating differential measurements that can accumulate errors, such as inertial sensors or wheel encoders, visual…

Data-driven visual odometry (VO) is a critical subroutine for autonomous edge robotics, and recent progress in the field has produced highly accurate point predictions in complex environments. However, emerging autonomous edge robotics…

Computer Vision and Pattern Recognition · Computer Science 2023-03-07 Alex C. Stutts , Danilo Erricolo , Theja Tulabandhula , Amit Ranjan Trivedi

Although cluttered indoor scenes have a lot of useful high-level semantic information which can be used for mapping and localization, most Visual Odometry (VO) algorithms rely on the usage of geometric features such as points, lines and…

Computer Vision and Pattern Recognition · Computer Science 2018-03-02 Huai-Jen Liang , Nitin J. Sanket , Cornelia Fermüller , Yiannis Aloimonos

Point cloud registration is a fundamental task in 3D computer vision. Most existing methods rely solely on geometric information for feature extraction and matching. Recently, several studies have incorporated color information from RGB-D…

Computer Vision and Pattern Recognition · Computer Science 2025-09-30 Congjia Chen , Yufu Qu

We present a novel method for scene change detection that leverages the robust feature extraction capabilities of a visual foundational model, DINOv2, and integrates full-image cross-attention to address key challenges such as varying…

Computer Vision and Pattern Recognition · Computer Science 2025-03-05 Chun-Jung Lin , Sourav Garg , Tat-Jun Chin , Feras Dayoub