English
Related papers

Related papers: Learning Multiplicative Interactions with Bayesian…

200 papers

This paper fosters the idea that deep learning methods can be used to complement classical visual odometry pipelines to improve their accuracy and to associate uncertainty models to their estimations. We show that the biases inherent to the…

Computer Vision and Pattern Recognition · Computer Science 2020-07-30 Andrea De Maio , Simon Lacroix

In this work, we research and evaluate end-to-end learning of monocular semantic-metric occupancy grid mapping from weak binocular ground truth. The network learns to predict four classes, as well as a camera to bird's eye view mapping. At…

Robotics · Computer Science 2019-05-01 Chenyang Lu , Marinus Jacobus Gerardus van de Molengraft , Gijs Dubbelman

With monocular Visual-Inertial Odometry (VIO) system, 3D point cloud and camera motion can be estimated simultaneously. Because pure sparse 3D points provide a structureless representation of the environment, generating 3D mesh from sparse…

Computer Vision and Pattern Recognition · Computer Science 2021-01-15 Xin Li , Yijia He , Jinlong Lin , Xiao Liu

Visual Odometry (VO) is essential to downstream mobile robotics and augmented/virtual reality tasks. Despite recent advances, existing VO methods still rely on heuristic design choices that require several weeks of hyperparameter tuning by…

Computer Vision and Pattern Recognition · Computer Science 2024-07-23 Nico Messikommer , Giovanni Cioffi , Mathias Gehrig , Davide Scaramuzza

Integration of data from multiple omics techniques is becoming increasingly important in biomedical research. Due to non-uniformity and technical limitations in omics platforms, such integrative analyses on multiple omics, which we refer to…

Machine Learning · Computer Science 2021-02-11 Changhee Lee , Mihaela van der Schaar

Visual inertial odometry (VIO) is a process for fusing visual and kinematic data to understand a machine's state in a navigation task. Olfactory inertial odometry (OIO) is an analog to VIO that fuses signals from gas sensors with inertial…

Robotics · Computer Science 2025-06-06 Kordel K. France , Ovidiu Daescu , Anirban Paul , Shalini Prasad

Visual Odometry (VO) is fundamental to autonomous navigation, robotics, and augmented reality, with unsupervised approaches eliminating the need for expensive ground-truth labels. However, these methods struggle when dynamic objects violate…

Computer Vision and Pattern Recognition · Computer Science 2025-08-04 Jingchao Xie , Oussema Dhaouadi , Weirong Chen , Johannes Meier , Jacques Kaiser , Daniel Cremers

In recent years, transformer-based architectures become the de facto standard for sequence modeling in deep learning frameworks. Inspired by the successful examples, we propose a causal visual-inertial fusion transformer (VIFT) for pose…

Computer Vision and Pattern Recognition · Computer Science 2024-09-16 Yunus Bilge Kurt , Ahmet Akman , A. Aydın Alatan

Vision-and-language (VL) pre-training, which aims to learn a general representation of image-text pairs that can be transferred to various vision-and-language tasks. Compared with modeling uni-modal data, the main challenge of the VL model…

Computation and Language · Computer Science 2023-05-24 Hao Yang , Can Gao , Hao Líu , Xinyan Xiao , Yanyan Zhao , Bing Qin

Modern visual-inertial navigation systems (VINS) are faced with a critical challenge in real-world deployment: they need to operate reliably and robustly in highly dynamic environments. Current best solutions merely filter dynamic objects…

Robotics · Computer Science 2021-12-07 Karnik Ram , Chaitanya Kharyal , Sudarshan S. Harithas , K. Madhava Krishna

Visual-Inertial Odometry (VIO) utilizes an Inertial Measurement Unit (IMU) to overcome the limitations of Visual Odometry (VO). However, the VIO for vehicles in large-scale outdoor environments still has some difficulties in estimating…

Robotics · Computer Science 2017-08-15 Chang-Ryeol Lee , Kuk-Jin Yoon

Unsupervised Learning based monocular visual odometry (VO) has lately drawn significant attention for its potential in label-free leaning ability and robustness to camera parameters and environmental variations. However, partially due to…

Computer Vision and Pattern Recognition · Computer Science 2019-03-18 Yang Li , Yoshitaka Ushiku , Tatsuya Harada

We propose Deep Patch Visual Odometry (DPVO), a new deep learning system for monocular Visual Odometry (VO). DPVO uses a novel recurrent network architecture designed for tracking image patches across time. Recent approaches to VO have…

Computer Vision and Pattern Recognition · Computer Science 2023-05-24 Zachary Teed , Lahav Lipson , Jia Deng

Current approaches for visual-inertial odometry (VIO) are able to attain highly accurate state estimation via nonlinear optimization. However, real-time optimization quickly becomes infeasible as the trajectory grows over time, this problem…

Robotics · Computer Science 2016-11-01 Christian Forster , Luca Carlone , Frank Dellaert , Davide Scaramuzza

This paper introduces a fully deep learning approach to monocular SLAM, which can perform simultaneous localization using a neural network for learning visual odometry (L-VO) and dense 3D mapping. Dense 2D flow and a depth image are…

Robotics · Computer Science 2018-07-26 Cheng Zhao , Li Sun , Pulak Purkait , Tom Duckett , Rustam Stolkin

Despite learning-based visual odometry (VO) has shown impressive results in recent years, the pretrained networks may easily collapse in unseen environments. The large domain gap between training and testing data makes them difficult to…

Computer Vision and Pattern Recognition · Computer Science 2021-03-30 Shunkai Li , Xin Wu , Yingdian Cao , Hongbin Zha

Monocular visual-inertial odometry (VIO) is a low-cost solution to provide high-accuracy, low-drifting pose estimation. However, it has been meeting challenges in vehicular scenarios due to limited dynamics and lack of stable features. In…

Robotics · Computer Science 2023-06-21 Yuxuan Zhou , Xingxing Li , Shengyu Li , Xuanbin Wang , Zhiheng Shen

Most learning-based methods estimate ego-motion by utilizing visual sensors, which suffer from dramatic lighting variations and textureless scenarios. In this paper, we incorporate sparse but accurate depth measurements obtained from lidars…

Computer Vision and Pattern Recognition · Computer Science 2021-01-06 Bin Li , Mu Hu , Shuling Wang , Lianghao Wang , Xiaojin Gong

Multiple signal modalities, such as vision and sounds, are naturally present in real-world phenomena. Recently, there has been growing interest in learning generative models, in particular variational autoencoder (VAE), to for multimodal…

Machine Learning · Computer Science 2024-12-31 Peijie Qiu , Wenhui Zhu , Sayantan Kumar , Xiwen Chen , Xiaotong Sun , Jin Yang , Abolfazl Razi , Yalin Wang , Aristeidis Sotiras

Vision and voice are two vital keys for agents' interaction and learning. In this paper, we present a novel indoor navigation model called Memory Vision-Voice Indoor Navigation (MVV-IN), which receives voice commands and analyzes multimodal…

Computer Vision and Pattern Recognition · Computer Science 2020-09-02 Liqi Yan , Dongfang Liu , Yaoxian Song , Changbin Yu