中文
相关论文

相关论文: Modality-invariant Visual Odometry for Embodied Vi…

200 篇论文

We propose a self-supervised learning framework that uses unlabeled monocular video sequences to generate large-scale supervision for training a Visual Odometry (VO) frontend, a network which computes pointwise data associations across…

计算机视觉与模式识别 · 计算机科学 2018-12-11 Daniel DeTone , Tomasz Malisiewicz , Andrew Rabinovich

Event-based cameras are bio-inspired vision sensors whose pixels work independently from each other and respond asynchronously to brightness changes, with microsecond resolution. Their advantages make it possible to tackle challenging…

计算机视觉与模式识别 · 计算机科学 2021-10-19 Yi Zhou , Guillermo Gallego , Shaojie Shen

Monocular visual odometry (VO) suffers severely from error accumulation during frame-to-frame pose estimation. In this paper, we present a self-supervised learning method for VO with special consideration for consistency over longer…

计算机视觉与模式识别 · 计算机科学 2020-07-22 Yuliang Zou , Pan Ji , Quoc-Huy Tran , Jia-Bin Huang , Manmohan Chandraker

We address automotive odometry for low-speed driving and parking, where centimeter-level accuracy is required due to tight spaces and nearby obstacles. Traditional methods using inertial-measurement units and wheel encoders require…

机器人学 · 计算机科学 2025-11-05 Luis Diener , Jens Kalkkuhl , Markus Enzweiler

To deal with the degeneration caused by the incomplete constraints of single sensor, multi-sensor fusion strategies especially in LiDAR-vision-inertial fusion area have attracted much interest from both the industry and the research…

机器人学 · 计算机科学 2023-08-08 Bingqi Shen , Yuyin Chen , Fuzhang Han , Shuwei Dai , Rong Xiong , Yue Wang

Simultaneous Localization and Mapping (SLAM) is considered to be an essential capability for intelligent vehicles and mobile robots. However, most of the current lidar SLAM approaches are based on the assumption of a static environment.…

机器人学 · 计算机科学 2022-06-22 Chenglong Qian , Zhaohong Xiang , Zhuoran Wu , Hongbin Sun

Recently, vision transformer based multimodal learning methods have been proposed to improve the robustness of face anti-spoofing (FAS) systems. However, multimodal face data collected from the real world is often imperfect due to missing…

计算机视觉与模式识别 · 计算机科学 2023-07-27 Zitong Yu , Rizhao Cai , Yawen Cui , Ajian Liu , Changsheng Chen

With an ever-widening domain of aerial robotic applications, including many mission critical tasks such as disaster response operations, search and rescue missions and infrastructure inspections taking place in GPS-denied environments, the…

机器人学 · 计算机科学 2019-07-02 Shehryar Khattak , Christos Papachristos , Kostas Alexis

Traveling at constant velocity is the most efficient trajectory for most robotics applications. Unfortunately without accelerometer excitation, monocular Visual-Inertial Odometry (VIO) cannot observe scale and suffers severe error drift.…

机器人学 · 计算机科学 2021-03-30 Jeff Delaune , David S. Bayard , Roland Brockers

Trajectory prediction is a fundamental problem in computer vision, vision-language-action models, world models, and autonomous systems, with broad impact on autonomous driving, robotics, and surveillance. However, most existing methods…

计算机视觉与模式识别 · 计算机科学 2026-03-30 Haichao Zhang , Yi Xu , Yun Fu

Recent learning-based approaches have achieved impressive results in the field of single-shot camera localization. However, how best to fuse multiple modalities (e.g., image and depth) and to deal with degraded or missing input are less…

计算机视觉与模式识别 · 计算机科学 2025-01-22 Kaichen Zhou , Changhao Chen , Bing Wang , Muhamad Risqi U. Saputra , Niki Trigoni , Andrew Markham

Visual Simultaneous Localisation and Mapping (VSLAM) is a key enabling technology for small embedded robotic systems such as aerial vehicles. Recent advances in equivariant filter and observer design offer the potential of a new generation…

机器人学 · 计算机科学 2020-06-01 Pieter van Goor , Robert Mahony , Tarek Hamel , Jochen Trumpf

Learning to navigate in unstructured environments is a challenging task for robots. While reinforcement learning can be effective, it often requires extensive data collection and can pose risk. Learning from expert demonstrations, on the…

机器人学 · 计算机科学 2024-12-31 Nimrod Curtis , Osher Azulay , Avishai Sintov

Pavement condition is crucial for civil infrastructure maintenance. This task usually requires efficient road damage localization, which can be accomplished by the visual odometry system embedded in unmanned aerial vehicles (UAVs). However,…

机器人学 · 计算机科学 2019-10-30 Huaiyang Huang , Rui Fan , Yilong Zhu , Ming Liu , Ioannis Pitas

Multi-modal fusion of sensors is a commonly used approach to enhance the performance of odometry estimation, which is also a fundamental module for mobile robots. However, the question of \textit{how to perform fusion among different…

机器人学 · 计算机科学 2025-03-20 Leyuan Sun , Guanqun Ding , Yue Qiu , Yusuke Yoshiyasu , Fumio Kanehiro

Visual Inertial Odometry (VIO) is the task of estimating the movement trajectory of an agent from an onboard camera stream fused with additional Inertial Measurement Unit (IMU) measurements. A crucial subtask within VIO is the tracking of…

计算机视觉与模式识别 · 计算机科学 2024-06-21 Jonas Kühne , Michele Magno , Luca Benini

In this paper, an efficient closed-form solution for the state initialization in visual-inertial odometry (VIO) and simultaneous localization and mapping (SLAM) is presented. Unlike the state-of-the-art, we do not derive linear equations…

计算机视觉与模式识别 · 计算机科学 2021-01-29 Georgios Evangelidis , Branislav Micusik

Integrating multiple LiDAR sensors can significantly enhance a robot's perception of the environment, enabling it to capture adequate measurements for simultaneous localization and mapping (SLAM). Indeed, solid-state LiDARs can bring in…

机器人学 · 计算机科学 2023-03-07 Li Qingqing , Yu Xianjia , Jorge Peña Queralta , Tomi Westerlund

Accurate, infrastructure-less sensor systems for motion tracking are essential for mobile robotics and augmented reality (AR) applications. The most popular state-of-the-art visual-inertial odometry (VIO) systems, however, are too…

计算机视觉与模式识别 · 计算机科学 2026-02-04 Jonas Kühne , Christian Vogt , Michele Magno , Luca Benini

This paper introduces a fully deep learning approach to monocular SLAM, which can perform simultaneous localization using a neural network for learning visual odometry (L-VO) and dense 3D mapping. Dense 2D flow and a depth image are…

机器人学 · 计算机科学 2018-07-26 Cheng Zhao , Li Sun , Pulak Purkait , Tom Duckett , Rustam Stolkin