English
Related papers

Related papers: VITAMIN-E: VIsual Tracking And MappINg with Extrem…

200 papers

Gaussian Splatting SLAM (GS-SLAM) offers a notable improvement over traditional SLAM methods, enabling photorealistic 3D reconstruction that conventional approaches often struggle to achieve. However, existing GS-SLAM systems perform poorly…

Robotics · Computer Science 2025-08-12 Siyu Chen , Shenghai Yuan , Thien-Minh Nguyen , Zhuyu Huang , Chenyang Shi , Jin Jing , Lihua Xie

In the field of multi-sensor fusion for simultaneous localization and mapping (SLAM), monocular cameras and IMUs are widely used to build simple and effective visual-inertial systems. However, limited research has explored the integration…

This paper presents a novel dense optical-flow algorithm to solve the monocular simultaneous localization and mapping (SLAM) problem for ground or aerial robots. Dense optical flow can effectively provide the ego-motion of the vehicle while…

Robotics · Computer Science 2021-10-01 Yonhon Ng , Hongdong Li , Jonghyuk Kim

In this paper, we present a monocular Simultaneous Localization and Mapping (SLAM) algorithm using high-level object and plane landmarks. The built map is denser, more compact and semantic meaningful compared to feature point based SLAM. We…

Robotics · Computer Science 2019-07-01 Shichao Yang , Sebastian Scherer

Initial applications of 3D Gaussian Splatting (3DGS) in Visual Simultaneous Localization and Mapping (VSLAM) demonstrate the generation of high-quality volumetric reconstructions from monocular video streams. However, despite these…

Robotics · Computer Science 2024-10-23 Yan Song Hu , Dayou Mao , Yuhao Chen , John Zelek

Robust and fast motion estimation and mapping is a key prerequisite for autonomous operation of mobile robots. The goal of performing this task solely on a stereo pair of video cameras is highly demanding and bears conflicting objectives:…

Robotics · Computer Science 2018-10-19 Nicola Krombach , David Droeschel , Sebastian Houben , Sven Behnke

Semantic-aware 3D scene reconstruction is essential for autonomous robots to perform complex interactions. Semantic SLAM, an online approach, integrates pose tracking, geometric reconstruction, and semantic mapping into a unified framework,…

Robotics · Computer Science 2025-05-20 Zuxing Lu , Xin Yuan , Shaowen Yang , Jingyu Liu , Changyin Sun

The visual SLAM method is widely used for self-localization and mapping in complex environments. Visual-inertia SLAM, which combines a camera with IMU, can significantly improve the robustness and enable scale weak-visibility, whereas…

Robotics · Computer Science 2020-03-06 Peng Gang , Lu Zezao , Chen Bocheng , Chen Shanliang , He Dingxin

Numerical resolution of high-dimensional nonlinear PDEs remains a huge challenge due to the curse of dimensionality. Starting from the weak formulation of the Lawson-Euler scheme, this paper proposes a stochastic particle method (SPM) by…

Numerical Analysis · Mathematics 2025-02-11 Zhengyang Lei , Sihong Shao , Yunfeng Xiong

Monocular visual SLAM has become an attractive practical approach for robot localization and 3D environment mapping, since cameras are small, lightweight, inexpensive, and produce high-rate, high-resolution data streams. Although numerous…

Computer Vision and Pattern Recognition · Computer Science 2018-06-08 Hasnain Vohra , Maxim Bazik , Matthew Antone , Joseph Mundy , William Stephenson

We propose the first 4D tracking and mapping method that jointly performs camera localization and non-rigid surface reconstruction via differentiable rendering. Our approach captures 4D scenes from an online stream of color images with…

Computer Vision and Pattern Recognition · Computer Science 2025-05-30 Hidenobu Matsuki , Gwangbin Bae , Andrew J. Davison

Visual SLAM has regained attention due to its ability to provide perceptual capabilities and simulation test data for Embodied AI. However, traditional SLAM methods struggle to meet the demands of high-quality scene reconstruction, and…

Robotics · Computer Science 2025-09-03 Fan Zhu , Yifan Zhao , Ziyu Chen , Biao Yu , Hui Zhu

Pre-trained Vision-Language Models (VLMs), such as CLIP, have shown enhanced performance across a range of tasks that involve the integration of visual and linguistic modalities. When CLIP is used for depth estimation tasks, the patches,…

Computer Vision and Pattern Recognition · Computer Science 2023-11-03 Xueting Hu , Ce Zhang , Yi Zhang , Bowen Hai , Ke Yu , Zhihai He

SLAM technology plays a crucial role in indoor mapping and localization. A common challenge in indoor environments is the "double-sided mapping issue", where closely positioned walls, doors, and other surfaces are mistakenly identified as a…

Robotics · Computer Science 2025-04-14 Chengwei Zhao , Yixuan Li , Yina Jian , Jie Xu , Linji Wang , Yongxin Ma , Xinglai Jin

Recent work has shown impressive localization performance using only images of ground textures taken with a downward facing monocular camera. This provides a reliable navigation method that is robust to feature sparse environments and…

Robotics · Computer Science 2023-03-13 Kyle M. Hart , Brendan Englot , Ryan P. O'Shea , John D. Kelly , David Martinez

Direct SLAM methods have shown exceptional performance on odometry tasks. However, they are susceptible to dynamic lighting and weather changes while also suffering from a bad initialization on large baselines. To overcome this, we propose…

Computer Vision and Pattern Recognition · Computer Science 2019-11-28 Lukas von Stumberg , Patrick Wenzel , Qadeer Khan , Daniel Cremers

This paper proposes a novel simultaneous localization and mapping (SLAM) approach, namely Attention-SLAM, which simulates human navigation mode by combining a visual saliency model (SalNavNet) with traditional monocular visual SLAM. Most…

Computer Vision and Pattern Recognition · Computer Science 2020-09-16 Jinquan Li , Ling Pei , Danping Zou , Songpengcheng Xia , Qi Wu , Tao Li , Zhen Sun , Wenxian Yu

This paper addresses the problem of learning to complete a scene's depth from sparse depth points and images of indoor scenes. Specifically, we study the case in which the sparse depth is computed from a visual-inertial simultaneous…

Computer Vision and Pattern Recognition · Computer Science 2020-08-18 Kourosh Sartipi , Tien Do , Tong Ke , Khiem Vuong , Stergios I. Roumeliotis

Novel view synthesis from a sparse set of input images is a challenging problem of great practical interest, especially when camera poses are absent or inaccurate. Direct optimization of camera poses and usage of estimated depths in neural…

Computer Vision and Pattern Recognition · Computer Science 2024-06-12 Kaiwen Jiang , Yang Fu , Mukund Varma T , Yash Belhe , Xiaolong Wang , Hao Su , Ravi Ramamoorthi

This paper introduces a fully deep learning approach to monocular SLAM, which can perform simultaneous localization using a neural network for learning visual odometry (L-VO) and dense 3D mapping. Dense 2D flow and a depth image are…

Robotics · Computer Science 2018-07-26 Cheng Zhao , Li Sun , Pulak Purkait , Tom Duckett , Rustam Stolkin