English
Related papers

Related papers: Learning Dense Flow Field for Highly-accurate Cros…

200 papers

State-of-the-art image segmentation algorithms generally consist of at least two successive and distinct computations: a boundary detection process that uses local image information to classify image locations as boundaries between objects,…

Computer Vision and Pattern Recognition · Computer Science 2016-11-03 Michał Januszewski , Jeremy Maitin-Shepard , Peter Li , Jörgen Kornfeld , Winfried Denk , Viren Jain

Optical Flow (OF) and depth are commonly used for visual odometry since they provide sufficient information about camera ego-motion in a rigid scene. We reformulate the problem of ego-motion estimation as a problem of motion estimation of a…

Computer Vision and Pattern Recognition · Computer Science 2019-07-18 Igor Slinko , Anna Vorontsova , Filipp Konokhov , Olga Barinova , Anton Konushin

We present a robust and real-time monocular six degree of freedom visual relocalization system. We use a Bayesian convolutional neural network to regress the 6-DOF camera pose from a single RGB image. It is trained in an end-to-end manner…

Computer Vision and Pattern Recognition · Computer Science 2016-02-19 Alex Kendall , Roberto Cipolla

This paper proposes a generalizable, end-to-end deep learning-based method for relative pose regression between two images. Given two images of the same scene captured from different viewpoints, our method predicts the relative rotation and…

Computer Vision and Pattern Recognition · Computer Science 2024-04-17 Fadi Khatib , Yuval Margalit , Meirav Galun , Ronen Basri

We propose DFPNet -- an unsupervised, joint learning system for monocular Depth, Optical Flow and egomotion (Camera Pose) estimation from monocular image sequences. Due to the nature of 3D scene geometry these three components are coupled.…

Computer Vision and Pattern Recognition · Computer Science 2022-10-12 Dipan Mandal , Abhilash Jain

This work proposes a novel deep network architecture to solve the camera Ego-Motion estimation problem. A motion estimation network generally learns features similar to Optical Flow (OF) fields starting from sequences of images. This OF can…

Computer Vision and Pattern Recognition · Computer Science 2018-02-16 Gabriele Costante , Thomas A. Ciarfuglia

Accurate registration of 2D imagery with point clouds is a key technology for image-LiDAR point cloud fusion, camera to laser scanner calibration and camera localization. Despite continuous improvements, automatic registration of 2D and 3D…

Computer Vision and Pattern Recognition · Computer Science 2019-12-13 Huai Yu , Weikun Zhen , Wen Yang , Sebastian Scherer

Relative pose estimation is crucial for various computer vision applications, including Robotic and Autonomous Driving. Current methods primarily depend on selecting and matching feature points prone to incorrect matches, leading to poor…

Computer Vision and Pattern Recognition · Computer Science 2024-10-08 Zherong Zhang , Chunyu Lin , Shujuan Huang , Shangrong Yang , Yao Zhao

Visual re-localization means using a single image as input to estimate the camera's location and orientation relative to a pre-recorded environment. The highest-scoring methods are "structure based," and need the query camera's intrinsics…

Computer Vision and Pattern Recognition · Computer Science 2021-04-13 Mehmet Ozgur Turkoglu , Eric Brachmann , Konrad Schindler , Gabriel Brostow , Aron Monszpart

Monocular visual odometry (VO) is a fundamental computer vision problem with applications in autonomous navigation, augmented reality and more. While deep learning-based methods have recently shown superior accuracy compared to traditional…

Computer Vision and Pattern Recognition · Computer Science 2026-04-27 Dominik Kuczkowski , Laura Ruotsalainen

We propose a viewpoint invariant model for 3D human pose estimation from a single depth image. To achieve this, our discriminative model embeds local regions into a learned viewpoint invariant feature space. Formulated as a multi-task…

Computer Vision and Pattern Recognition · Computer Science 2016-07-27 Albert Haque , Boya Peng , Zelun Luo , Alexandre Alahi , Serena Yeung , Li Fei-Fei

We propose a feed-forward method for dense Signed Distance Field (SDF) regression from unstructured image collections in less than three seconds, without camera calibration or post-hoc fusion. Our key insight is that the intermediate…

Computer Vision and Pattern Recognition · Computer Science 2026-03-30 Laura Fink , Linus Franke , George Kopanas , Marc Stamminger , Peter Hedman

This paper proposes a new image-based localization framework that explicitly localizes the camera/robot by fusing Convolutional Neural Network (CNN) and sequential images' geometric constraints. The camera is localized using a single or few…

Computer Vision and Pattern Recognition · Computer Science 2022-01-06 Jingwei Song , Mitesh Patel , Maani Ghaffari

While 3D Vision Foundation Models (3DVFMs) have demonstrated remarkable zero-shot capabilities in visual geometry estimation, their direct application to generalizable novel view synthesis (NVS) remains challenging. In this paper, we…

Computer Vision and Pattern Recognition · Computer Science 2026-03-27 Minh-Quan Viet Bui , Jaeho Moon , Munchurl Kim

We propose a new approach called LiDAR-Flow to robustly estimate a dense scene flow by fusing a sparse LiDAR with stereo images. We take the advantage of the high accuracy of LiDAR to resolve the lack of information in some regions of…

Computer Vision and Pattern Recognition · Computer Science 2019-12-16 Ramy Battrawy , René Schuster , Oliver Wasenmüller , Qing Rao , Didier Stricker

We address the problem of scene flow: given a pair of stereo or RGB-D video frames, estimate pixelwise 3D motion. We introduce RAFT-3D, a new deep architecture for scene flow. RAFT-3D is based on the RAFT model developed for optical flow…

Computer Vision and Pattern Recognition · Computer Science 2021-04-07 Zachary Teed , Jia Deng

The objective of this work is human pose estimation in videos, where multiple frames are available. We investigate a ConvNet architecture that is able to benefit from temporal context by combining information across the multiple frames…

Computer Vision and Pattern Recognition · Computer Science 2015-11-10 Tomas Pfister , James Charles , Andrew Zisserman

In autonomous driving, 3D object detection is essential for accurately identifying and tracking objects. Despite the continuous development of various technologies for this task, a significant drawback is observed in most of them-they…

Computer Vision and Pattern Recognition · Computer Science 2025-02-05 Hsin-Cheng Lu , Chung-Yi Lin , Winston H. Hsu

In this work, we propose a method for object recognition and pose estimation from depth images using convolutional neural networks. Previous methods addressing this problem rely on manifold learning to learn low dimensional viewpoint…

Computer Vision and Pattern Recognition · Computer Science 2019-04-19 Mai Bui , Sergey Zakharov , Shadi Albarqouni , Slobodan Ilic , Nassir Navab

We propose BEV-Patch-PF, a GPS-free sequential geo-localization system that integrates a particle filter with learned bird's-eye-view (BEV) and aerial feature maps. From onboard RGB and depth images, we construct a BEV feature map. For each…

‹ Prev 1 8 9 10 Next ›