English
Related papers

Related papers: Training Deep SLAM on Single Frames

200 papers

We present ViSTA-SLAM as a real-time monocular visual SLAM system that operates without requiring camera intrinsics, making it broadly applicable across diverse camera setups. At its core, the system employs a lightweight symmetric two-view…

Computer Vision and Pattern Recognition · Computer Science 2026-01-07 Ganlin Zhang , Shenhan Qian , Xi Wang , Daniel Cremers

In endoscopy, many applications (e.g., surgical navigation) would benefit from a real-time method that can simultaneously track the endoscope and reconstruct the dense 3D geometry of the observed anatomy from a monocular endoscopic video.…

Computer Vision and Pattern Recognition · Computer Science 2022-02-23 Xingtong Liu , Zhaoshuo Li , Masaru Ishii , Gregory D. Hager , Russell H. Taylor , Mathias Unberath

Simultaneous localization and mapping (SLAM) is a critical capability for autonomous systems. Traditional SLAM approaches, which often rely on visual or LiDAR sensors, face significant challenges in adverse conditions such as low light or…

Robotics · Computer Science 2026-02-06 Dong Wang , Hannes Haag , Daniel Casado Herraez , Stefan May , Cyrill Stachniss , Andreas Nüchter

Handling the dynamic environments is a significant research challenge in Visual Simultaneous Localization and Mapping (SLAM). Recent research combines 3D Gaussian Splatting (3DGS) with SLAM to achieve both robust camera pose estimation and…

Computer Vision and Pattern Recognition · Computer Science 2026-04-29 Yunsong Wang , Gim Hee Lee

In recent years, visual SLAM has achieved great progress and development, but in complex scenes, especially rotating scenes, the error of mapping will increase significantly, and the slam system is easy to lose track. In this article, we…

Robotics · Computer Science 2021-10-07 Zhenkun Zhu , Jikai Wang

Moving objects can greatly jeopardize the performance of a visual simultaneous localization and mapping (vSLAM) system which relies on the static-world assumption. Motion removal have seen successful on solving this problem. Two main…

Robotics · Computer Science 2019-08-01 Ting Sun , Yuxiang Sun , Ming Liu , Dit-Yan Yeung

We present a self-supervised learning framework to estimate the individual object motion and monocular depth from video. We model the object motion as a 6 degree-of-freedom rigid-body transformation. The instance segmentation mask is…

Computer Vision and Pattern Recognition · Computer Science 2020-05-14 Qi Dai , Vaishakh Patil , Simon Hecker , Dengxin Dai , Luc Van Gool , Konrad Schindler

This paper presents GeoFlow-SLAM, a robust and effective Tightly-Coupled RGBD-inertial SLAM for legged robotics undergoing aggressive and high-frequency motions.By integrating geometric consistency, legged odometry constraints, and…

Robotics · Computer Science 2025-07-23 Tingyang Xiao , Xiaolin Zhou , Liu Liu , Wei Sui , Wei Feng , Jiaxiong Qiu , Xinjie Wang , Zhizhong Su

We present a method for single image 3D cuboid object detection and multi-view object SLAM in both static and dynamic environments, and demonstrate that the two parts can improve each other. Firstly for single image object detection, we…

Robotics · Computer Science 2019-04-08 Shichao Yang , Sebastian Scherer

In this paper, we tackle the problem of generating a novel image from an arbitrary viewpoint given a single frame as input. While existing methods operating in this setup aim at predicting the target view depth map to guide the synthesis,…

Computer Vision and Pattern Recognition · Computer Science 2023-08-29 Giovanni Minelli , Matteo Poggi , Samuele Salti

In this paper, an efficient closed-form solution for the state initialization in visual-inertial odometry (VIO) and simultaneous localization and mapping (SLAM) is presented. Unlike the state-of-the-art, we do not derive linear equations…

Computer Vision and Pattern Recognition · Computer Science 2021-01-29 Georgios Evangelidis , Branislav Micusik

We propose XVO, a semi-supervised learning method for training generalized monocular Visual Odometry (VO) models with robust off-the-self operation across diverse datasets and settings. In contrast to standard monocular VO approaches which…

Computer Vision and Pattern Recognition · Computer Science 2023-10-10 Lei Lai , Zhongkai Shangguan , Jimuyang Zhang , Eshed Ohn-Bar

Visual-inertial simultaneous localization and mapping (SLAM) is a key module of robotics and low-speed autonomous vehicles, which is usually limited by the high computation burden for practical applications. To this end, an innovative…

Robotics · Computer Science 2025-05-28 Bingxiang Kang , Jie Zou , Guofa Li , Pengwei Zhang , Jie Zeng , Kan Wang , Jie Li

In recent years, learning-based feature detection and matching have outperformed manually-designed methods in in-air cases. However, it is challenging to learn the features in the underwater scenario due to the absence of annotated…

Computer Vision and Pattern Recognition · Computer Science 2024-02-05 Jinghe Yang , Mingming Gong , Girish Nair , Jung Hoon Lee , Jason Monty , Ye Pu

Efficient three-dimensional reconstruction and real-time visualization are critical in surgical scenarios such as endoscopy. In recent years, 3D Gaussian Splatting (3DGS) has demonstrated remarkable performance in efficient 3D…

Computer Vision and Pattern Recognition · Computer Science 2025-07-08 Taoyu Wu , Yiyi Miao , Zhuoxiao Li , Haocheng Zhao , Kang Dang , Jionglong Su , Limin Yu , Haoang Li

The state of the art in human-centric computer vision achieves high accuracy and robustness across a diverse range of tasks. The most effective models in this domain have billions of parameters, thus requiring extremely large datasets,…

Computer Vision and Pattern Recognition · Computer Science 2025-07-22 Fatemeh Saleh , Sadegh Aliakbarian , Charlie Hewitt , Lohit Petikam , Xiao-Xian , Antonio Criminisi , Thomas J. Cashman , Tadas Baltrušaitis

Conventional SLAM techniques strongly rely on scene rigidity to solve data association, ignoring dynamic parts of the scene. In this work we present Semi-Direct DefSLAM (SD-DefSLAM), a novel monocular deformable SLAM method able to map…

Computer Vision and Pattern Recognition · Computer Science 2020-10-20 Juan J. Gómez Rodríguez , José Lamarca , Javier Morlana , Juan D. Tardós , José M. M. Montiel

This paper addresses the problem of learning to complete a scene's depth from sparse depth points and images of indoor scenes. Specifically, we study the case in which the sparse depth is computed from a visual-inertial simultaneous…

Computer Vision and Pattern Recognition · Computer Science 2020-08-18 Kourosh Sartipi , Tien Do , Tong Ke , Khiem Vuong , Stergios I. Roumeliotis

Traditional unsupervised optical flow methods are vulnerable to occlusions and motion boundaries due to lack of object-level information. Therefore, we propose UnSAMFlow, an unsupervised flow network that also leverages object information…

Computer Vision and Pattern Recognition · Computer Science 2024-05-07 Shuai Yuan , Lei Luo , Zhuo Hui , Can Pu , Xiaoyu Xiang , Rakesh Ranjan , Denis Demandolx

Monocular depth inference has gained tremendous attention from researchers in recent years and remains as a promising replacement for expensive time-of-flight sensors, but issues with scale acquisition and implementation overhead still…

Computer Vision and Pattern Recognition · Computer Science 2021-08-17 Kenny Chen , Alexandra Pogue , Brett T. Lopez , Ali-akbar Agha-mohammadi , Ankur Mehta
‹ Prev 1 8 9 10 Next ›