中文
相关论文

相关论文: Keyframe-Based Feed-Forward Visual Odometry

200 篇论文

As an essential part of structure from motion (SfM) and Simultaneous Localization and Mapping (SLAM) systems, motion averaging has been extensively studied in the past years and continues to attract surging research attention. While…

计算机视觉与模式识别 · 计算机科学 2020-11-03 Xinyi Li , Lin Yuan , Longin Jan Latecki , Haibin Ling

In modern visual SLAM systems, it is a standard practice to retrieve potential candidate map points from overlapping keyframes for further feature matching or direct tracking. In this work, we argue that keyframes are not the optimal choice…

机器人学 · 计算机科学 2020-03-05 Manasi Muglikar , Zichao Zhang , Davide Scaramuzza

Traditional Visual Odometry (VO) and Visual Inertial Odometry (VIO) methods rely on a 'pose-centric' paradigm, which computes absolute camera poses from the local map thus requires large-scale landmark maintenance and continuous map…

机器人学 · 计算机科学 2025-11-13 Sangheon Yang , Yeongin Yoon , Hong Mo Jung , Jongwoo Lim

Reliable displacement measurement is fundamental for structural health monitoring and digital engineering workflows, as it provides direct structural response information. Vision-based measurement has emerged as a promising approach for…

计算机视觉与模式识别 · 计算机科学 2026-05-12 Qingyu Xian , Hao Cheng , Berend Jan van der Zwaag , Rolands Kromanis , Ozlem Durmaz Incel

This paper presents a novel visual-LiDAR odometry and mapping method with low-drift characteristics. The proposed method is based on two popular approaches, ORB-SLAM and A-LOAM, with monocular scale correction and visual-bootstrapped LiDAR…

计算机视觉与模式识别 · 计算机科学 2023-07-11 Hanyu Cai , Ni Ou , Junzheng Wang

We introduce OpenVO, a novel framework for Open-world Visual Odometry (VO) with temporal awareness under limited input conditions. OpenVO effectively estimates real-world-scale ego-motion from monocular dashcam footage with varying…

计算机视觉与模式识别 · 计算机科学 2026-04-28 Phuc D. A. Nguyen , Anh N. Nhu , Ming C. Lin

LiDAR SLAM has become one of the major localization systems for ground vehicles since LiDAR Odometry And Mapping (LOAM). Many extension works on LOAM mainly leverage one specific constraint to improve the performance, e.g., information from…

机器人学 · 计算机科学 2024-04-03 Jiaying Chen , Han Wang , Minghui Hu , Ponnuthurai Nagaratnam Suganthan

Visual Odometry (VO) is a method to estimate self-motion of a mobile robot using visual sensors. Unlike odometry based on integrating differential measurements that can accumulate errors, such as inertial sensors or wheel encoders, visual…

Visual Odometry (VO) accumulates a positional drift in long-term robot navigation tasks. Although Convolutional Neural Networks (CNNs) improve VO in various aspects, VO still suffers from moving obstacles, discontinuous observation of…

计算机视觉与模式识别 · 计算机科学 2020-06-25 Felix Ott , Tobias Feigl , Christoffer Löffler , Christopher Mutschler

LiDAR-based SLAM is a core technology for autonomous vehicles and robots. One key contribution of this work to 3D LiDAR SLAM and localization is a fierce defense of view-based maps (pose graphs with time-stamped sensor readings) as the…

机器人学 · 计算机科学 2025-08-19 José Luis Blanco-Claraco

Visual Odometry (VO) estimation is an important source of information for vehicle state estimation and autonomous driving. Recently, deep learning based approaches have begun to appear in the literature. However, in the context of driving,…

计算机视觉与模式识别 · 计算机科学 2021-12-28 Nimet Kaygusuz , Oscar Mendez , Richard Bowden

In recent years, deep learning-based approaches for visual-inertial odometry (VIO) have shown remarkable performance outperforming traditional geometric methods. Yet, all existing methods use both the visual and inertial measurements for…

计算机视觉与模式识别 · 计算机科学 2022-10-21 Mingyu Yang , Yu Chen , Hun-Seok Kim

This paper introduces a fully deep learning approach to monocular SLAM, which can perform simultaneous localization using a neural network for learning visual odometry (L-VO) and dense 3D mapping. Dense 2D flow and a depth image are…

机器人学 · 计算机科学 2018-07-26 Cheng Zhao , Li Sun , Pulak Purkait , Tom Duckett , Rustam Stolkin

In this paper, an efficient closed-form solution for the state initialization in visual-inertial odometry (VIO) and simultaneous localization and mapping (SLAM) is presented. Unlike the state-of-the-art, we do not derive linear equations…

计算机视觉与模式识别 · 计算机科学 2021-01-29 Georgios Evangelidis , Branislav Micusik

Unlike loose coupling approaches and the EKF-based approaches in the literature, we propose an optimization-based visual-inertial SLAM tightly coupled with raw Global Navigation Satellite System (GNSS) measurements, a first attempt of this…

机器人学 · 计算机科学 2021-10-26 Jinxu Liu , Wei Gao , Zhanyi Hu

This paper proposes an illumination-robust visual odometry (VO) system that incorporates both accelerated learning-based corner point algorithms and an extended line feature algorithm. To be robust to dynamic illumination, the proposed…

机器人学 · 计算机科学 2023-12-20 Kuan Xu , Yuefan Hao , Shenghai Yuan , Chen Wang , Lihua Xie

While 3D Vision Foundation Models (3DVFMs) have demonstrated remarkable zero-shot capabilities in visual geometry estimation, their direct application to generalizable novel view synthesis (NVS) remains challenging. In this paper, we…

计算机视觉与模式识别 · 计算机科学 2026-03-27 Minh-Quan Viet Bui , Jaeho Moon , Munchurl Kim

Efficient fine-tuning of vision-language models (VLMs) like CLIP for specific downstream tasks is gaining significant attention. Previous works primarily focus on prompt learning to adapt the CLIP into a variety of downstream tasks,…

计算机视觉与模式识别 · 计算机科学 2024-10-17 Jinlong Li , Dong Zhao , Zequn Jie , Elisa Ricci , Lin Ma , Nicu Sebe

Although quadcopters boast impressive traversal capabilities enabled by their omnidirectional maneuverability, the need for continuous pilot control in complex environments impedes their application in GNSS and telemetry-denied scenarios.…

机器人学 · 计算机科学 2026-05-26 Shiladitya Dutta , Aayush Gupta , Varun Saran , Avideh Zakhor

Vision language models (VLMs) demonstrate strong capabilities in jointly processing visual and textual data. However, they often incur substantial computational overhead due to redundant visual information, particularly in long-form video…

机器学习 · 计算机科学 2025-04-25 Yudong Liu , Jingwei Sun , Yueqian Lin , Jingyang Zhang , Ming Yin , Qinsi Wang , Jianyi Zhang , Hai Li , Yiran Chen