中文
相关论文

相关论文: Robust Monocular SLAM for Egocentric Videos

200 篇论文

Simultaneous localization and mapping (SLAM) in slowly varying scenes is important for long-term robot task completion. Failing to detect scene changes may lead to inaccurate maps and, ultimately, lost robots. Classical SLAM algorithms…

This paper presents a visual SLAM system that uses both points and lines for robust camera localization, and simultaneously performs a piece-wise planar reconstruction (PPR) of the environment to provide a structural map in real-time. One…

计算机视觉与模式识别 · 计算机科学 2023-01-31 Fangwen Shu , Jiaxuan Wang , Alain Pagani , Didier Stricker

Crowd-sourced cooperative mapping from monocular cameras promises scalable 3D reconstruction without specialized sensors, yet remains hindered by two scale-specific failure modes: abrupt scale collapse from false-positive loop closures in…

机器人学 · 计算机科学 2026-04-16 Hyoseok Ju , Giseop Kim

Monocular Simultaneous Localization and Mapping (SLAM) aims to estimate a robot's pose while simultaneously reconstructing an unknown 3D scene using a single camera. While existing monocular SLAM systems generate detailed 3D geometry…

机器人学 · 计算机科学 2025-11-27 Yuchen Zhou , Haihang Wu

While egocentric cameras like GoPro are gaining popularity, the videos they capture are long, boring, and difficult to watch from start to end. Fast forwarding (i.e. frame sampling) is a natural choice for faster video browsing. However,…

计算机视觉与模式识别 · 计算机科学 2017-01-04 Yair Poleg , Tavi Halperin , Chetan Arora , Shmuel Peleg

Simultaneous Localization and Mapping (SLAM) systems are fundamental building blocks for any autonomous robot navigating in unknown environments. The SLAM implementation heavily depends on the sensor modality employed on the mobile…

机器人学 · 计算机科学 2022-03-25 Luca Di Giammarino , Leonardo Brizi , Tiziano Guadagnino , Cyrill Stachniss , Giorgio Grisetti

Map-centric SLAM is emerging as an alternative of conventional graph-based SLAM for its accuracy and efficiency in long-term mapping problems. However, in map-centric SLAM, the process of loop closure differs from that of conventional SLAM…

机器人学 · 计算机科学 2019-01-31 Chanoh Park , Soohwan Kim , Peyman Moghadam , Jiadong Guo , Sridha Sridharan , Clinton Fookes

Recently there has been a growing interest in category-level object pose and size estimation, and prevailing methods commonly rely on single view RGB-D images. However, one disadvantage of such methods is that they require accurate depth…

计算机视觉与模式识别 · 计算机科学 2024-03-25 Jiaqi Yang , Yucong Chen , Xiangting Meng , Chenxin Yan , Min Li , Ran Cheng , Lige Liu , Tao Sun , Laurent Kneip

We present an inverse image-formation module that can enhance the robustness of existing visual SLAM pipelines for casually captured scenarios. Casual video captures often suffer from motion blur and varying appearances, which degrade the…

计算机视觉与模式识别 · 计算机科学 2024-07-17 Gwangtak Bae , Changwoon Choi , Hyeongjun Heo , Sang Min Kim , Young Min Kim

We present a dataset for evaluating the tracking accuracy of monocular visual odometry and SLAM methods. It contains 50 real-world sequences comprising more than 100 minutes of video, recorded across dozens of different environments --…

计算机视觉与模式识别 · 计算机科学 2016-10-11 Jakob Engel , Vladyslav Usenko , Daniel Cremers

Gaussian splatting has recently gained traction as a compelling map representation for SLAM systems, enabling dense and photo-realistic scene modeling. However, its application to monocular SLAM remains challenging due to the lack of…

机器人学 · 计算机科学 2026-04-20 Dong-Uk Seo , Jinwoo Jeon , Eungchang Mason Lee , Hyun Myung

Simultaneous localization and mapping (SLAM) is an essential component of robotic systems. In this work we perform a feasibility study of RGB-D SLAM for the task of indoor robot navigation. Recent visual SLAM methods, e.g. ORBSLAM2…

计算机视觉与模式识别 · 计算机科学 2019-10-14 David Prokhorov , Dmitry Zhukov , Olga Barinova , Anna Vorontsova , Anton Konushin

This paper presents a hybrid real-time camera pose estimation framework with a novel partitioning scheme and introduces motion averaging to monocular Simultaneous Localization and Mapping (SLAM) systems. Breaking through the limitations of…

计算机视觉与模式识别 · 计算机科学 2020-11-04 Xinyi Li , Haibin Ling

3D Gaussian Splatting has emerged as a promising technique for high-quality 3D rendering, leading to increasing interest in integrating 3DGS into realism SLAM systems. However, existing methods face challenges such as Gaussian primitives…

机器人学 · 计算机科学 2024-12-16 Lizhi Bai , Chunqi Tian , Jun Yang , Siyu Zhang , Masanori Suganuma , Takayuki Okatani

The recently developed Neural Radiance Fields (NeRF) and 3D Gaussian Splatting (3DGS) have shown encouraging and impressive results for visual SLAM. However, most representative methods require RGBD sensors and are only available for indoor…

计算机视觉与模式识别 · 计算机科学 2025-05-16 Zhe Xin , Chenyang Wu , Penghui Huang , Yanyong Zhang , Yinian Mao , Guoquan Huang

Localization and mapping with heterogeneous multi-sensor fusion have been prevalent in recent years. To adequately fuse multi-modal sensor measurements received at different time instants and different frequencies, we estimate the…

机器人学 · 计算机科学 2023-02-16 Jiajun Lv , Xiaolei Lang , Jinhong Xu , Mengmeng Wang , Yong Liu , Xingxing Zuo

In this paper, we present BirdSLAM, a novel simultaneous localization and mapping (SLAM) system for the challenging scenario of autonomous driving platforms equipped with only a monocular camera. BirdSLAM tackles challenges faced by other…

机器人学 · 计算机科学 2020-11-17 Swapnil Daga , Gokul B. Nair , Anirudha Ramesh , Rahul Sajnani , Junaid Ahmed Ansari , K. Madhava Krishna

Conventional visual simultaneous localization and mapping (SLAM) algorithms often fail under rapid motion, low illumination, or abrupt lighting transitions due to motion blur and limited dynamic range. Event cameras mitigate these issues…

计算机视觉与模式识别 · 计算机科学 2026-03-10 Şebnem Sarıözkan , Hürkan Şahin , Olaya Álvarez-Tuñón , Erdal Kayacan

When adapting Simultaneous Mapping and Localization (SLAM) to real-world applications, such as autonomous vehicles, drones, and augmented reality devices, its memory footprint and computing cost are the two main factors limiting the…

机器人学 · 计算机科学 2022-11-04 Yeonsoo Park , Soohyun Bae

Egocentric video-language understanding demands both high efficiency and accurate spatial-temporal modeling. Existing approaches face three key challenges: 1) Excessive pre-training cost arising from multi-stage pre-training pipelines, 2)…

计算机视觉与模式识别 · 计算机科学 2025-06-18 Xiaoqi Wang , Yi Wang , Lap-Pui Chau