English
Related papers

Related papers: Beyond Photometric Loss for Self-Supervised Ego-Mo…

200 papers

Egocentric 3D human pose estimation has been actively studied using cameras installed in front of a head-mounted device (HMD). While frontal placement is the optimal and the only option for some tasks, such as hand tracking, it remains…

Computer Vision and Pattern Recognition · Computer Science 2025-08-25 Hiroyasu Akada , Jian Wang , Vladislav Golyanik , Christian Theobalt

Estimating the camera's pose given images from a single camera is a traditional task in mobile robots and autonomous vehicles. This problem is called monocular visual odometry and often relies on geometric approaches that require…

Computer Vision and Pattern Recognition · Computer Science 2025-01-22 André O. Françani , Marcos R. O. A. Maximo

Visual localization on standard-definition (SD) maps has emerged as a promising low-cost and scalable solution for autonomous driving. However, existing regression-based approaches often overlook inherent geometric priors, resulting in…

Computer Vision and Pattern Recognition · Computer Science 2026-01-08 Xuchang Zhong , Xu Cao , Jinke Feng , Hao Fang

Visual odometry networks commonly use pretrained optical flow networks in order to derive the ego-motion between consecutive frames. The features extracted by these networks represent the motion of all the pixels between frames. However,…

Computer Vision and Pattern Recognition · Computer Science 2020-11-18 Hamed Damirchi , Rooholla Khorrambakht , Hamid D. Taghirad

Self-supervised monocular depth estimation has been widely studied, owing to its practical importance and recent promising improvements. However, most works suffer from limited supervision of photometric consistency, especially in weak…

Computer Vision and Pattern Recognition · Computer Science 2021-08-20 Hyunyoung Jung , Eunhyeok Park , Sungjoo Yoo

Recently, deep metric learning techniques received attention, as the learned distance representations are useful to capture the similarity relationship among samples and further improve the performance of various of supervised or…

Computer Vision and Pattern Recognition · Computer Science 2023-04-21 Zhiyuan Li , Anca Ralescu

In this paper we propose a robust visual odometry system for a wide-baseline camera rig with wide field-of-view (FOV) fisheye lenses, which provides full omnidirectional stereo observations of the environment. For more robust and accurate…

Computer Vision and Pattern Recognition · Computer Science 2019-03-04 Hochang Seok , Jongwoo Lim

We introduce OpenVO, a novel framework for Open-world Visual Odometry (VO) with temporal awareness under limited input conditions. OpenVO effectively estimates real-world-scale ego-motion from monocular dashcam footage with varying…

Computer Vision and Pattern Recognition · Computer Science 2026-04-28 Phuc D. A. Nguyen , Anh N. Nhu , Ming C. Lin

We introduce FocalPose++, a neural render-and-compare method for jointly estimating the camera-object 6D pose and camera focal length given a single RGB input image depicting a known object. The contributions of this work are threefold.…

Computer Vision and Pattern Recognition · Computer Science 2024-11-08 Martin Cífka , Georgy Ponimatkin , Yann Labbé , Bryan Russell , Mathieu Aubry , Vladimir Petrik , Josef Sivic

Depth estimation is a critical topic for robotics and vision-related tasks. In monocular depth estimation, in comparison with supervised learning that requires expensive ground truth labeling, self-supervised methods possess great potential…

Computer Vision and Pattern Recognition · Computer Science 2024-08-30 Jinchang Zhang , Praveen Kumar Reddy , Xue-Iuan Wong , Yiannis Aloimonos , Guoyu Lu

In this paper we propose an edge-direct visual odometry algorithm that efficiently utilizes edge pixels to find the relative pose that minimizes the photometric error between images. Prior work on exploiting edge pixels instead treats edges…

Computer Vision and Pattern Recognition · Computer Science 2019-06-13 Kevin Christensen , Martial Hebert

We propose D3VO as a novel framework for monocular visual odometry that exploits deep networks on three levels -- deep depth, pose and uncertainty estimation. We first propose a novel self-supervised monocular depth estimation network…

Computer Vision and Pattern Recognition · Computer Science 2020-03-31 Nan Yang , Lukas von Stumberg , Rui Wang , Daniel Cremers

Estimating 3D human motion from an egocentric video sequence plays a critical role in human behavior understanding and has various applications in VR/AR. However, naively learning a mapping between egocentric videos and human motions is…

Computer Vision and Pattern Recognition · Computer Science 2023-08-29 Jiaman Li , C. Karen Liu , Jiajun Wu

Visual-based localization has made significant progress, yet its performance often drops in large-scale, outdoor, and long-term settings due to factors like lighting changes, dynamic scenes, and low-texture areas. These challenges degrade…

Robotics · Computer Science 2025-09-11 Sai Puneeth Reddy Gottam , Haoming Zhang , Eivydas Keras

In monocular 3D human pose estimation a common setup is to first detect 2D positions and then lift the detection into 3D coordinates. Many algorithms suffer from overfitting to camera positions in the training set. We propose a siamese…

Computer Vision and Pattern Recognition · Computer Science 2019-02-19 Márton Véges , Viktor Varga , András Lőrincz

This paper tackles the challenges of self-supervised monocular depth estimation in indoor scenes caused by large rotation between frames and low texture. We ease the learning process by obtaining coarse camera poses from monocular sequences…

Computer Vision and Pattern Recognition · Computer Science 2023-09-29 Chaoqiang Zhao , Matteo Poggi , Fabio Tosi , Lei Zhou , Qiyu Sun , Yang Tang , Stefano Mattoccia

Building vehicles capable of operating without human supervision requires the determination of the agent's pose. Visual Odometry (VO) algorithms estimate the egomotion using only visual changes from the input images. The most recent VO…

Robotics · Computer Science 2021-07-08 Iury Cleveston , Esther L. Colombini

Making multi-camera visual SLAM systems easier to set up and more robust to the environment is attractive for vision robots. Existing monocular and binocular vision SLAM systems have narrow sensing Field-of-View (FoV), resulting in…

Robotics · Computer Science 2025-03-26 Huai Yu , Junhao Wang , Yao He , Wen Yang , Gui-Song Xia

Visual re-localization means using a single image as input to estimate the camera's location and orientation relative to a pre-recorded environment. The highest-scoring methods are "structure based," and need the query camera's intrinsics…

Computer Vision and Pattern Recognition · Computer Science 2021-04-13 Mehmet Ozgur Turkoglu , Eric Brachmann , Konrad Schindler , Gabriel Brostow , Aron Monszpart

This paper fosters the idea that deep learning methods can be used to complement classical visual odometry pipelines to improve their accuracy and to associate uncertainty models to their estimations. We show that the biases inherent to the…

Computer Vision and Pattern Recognition · Computer Science 2020-07-30 Andrea De Maio , Simon Lacroix
‹ Prev 1 8 9 10 Next ›