English
Related papers

Related papers: ParticleSfM: Exploiting Dense Point Trajectories f…

200 papers

Transferring image-based object detectors to the domain of videos remains a challenging problem. Previous efforts mostly exploit optical flow to propagate features across frames, aiming to achieve a good trade-off between accuracy and…

Computer Vision and Pattern Recognition · Computer Science 2019-08-19 Chaoxu Guo , Bin Fan , Jie Gu , Qian Zhang , Shiming Xiang , Veronique Prinet , Chunhong Pan

Understanding the geometry and pose of objects in 2D images is a fundamental necessity for a wide range of real world applications. Driven by deep neural networks, recent methods have brought significant improvements to object pose…

Computer Vision and Pattern Recognition · Computer Science 2018-09-05 Jogendra Nath Kundu , Rahul M. V. , Aditya Ganeshan , R. Venkatesh Babu

We propose a learning-based method that solves monocular stereo and can be extended to fuse depth information from multiple target frames. Given two unconstrained images from a monocular camera with known intrinsic calibration, our network…

Computer Vision and Pattern Recognition · Computer Science 2019-09-13 Kaixuan Wang , Shaojie Shen

Deformable Monocular SLAM algorithms recover the localization of a camera in an unknown deformable environment. Current approaches use a template-based deformable tracking to recover the camera pose and the deformation of the map. These…

Computer Vision and Pattern Recognition · Computer Science 2021-09-16 Jose Lamarca , Juan J. Gomez Rodriguez , Juan D. Tardos , J. M. M. Montiel

We present a method to reconstruct the three-dimensional trajectory of a moving instance of a known object category using stereo video data. We track the two-dimensional shape of objects on pixel level exploiting instance-aware semantic…

Computer Vision and Pattern Recognition · Computer Science 2018-08-29 Sebastian Bullinger , Christoph Bodensteiner , Michael Arens , Rainer Stiefelhagen

Structure-from-motion (SfM) is a long-standing problem in the computer vision community, which aims to reconstruct the camera poses and 3D structure of a scene from a set of unconstrained 2D images. Classical frameworks solve this problem…

Computer Vision and Pattern Recognition · Computer Science 2023-12-08 Jianyuan Wang , Nikita Karaev , Christian Rupprecht , David Novotny

Conventional SLAM techniques strongly rely on scene rigidity to solve data association, ignoring dynamic parts of the scene. In this work we present Semi-Direct DefSLAM (SD-DefSLAM), a novel monocular deformable SLAM method able to map…

Computer Vision and Pattern Recognition · Computer Science 2020-10-20 Juan J. Gómez Rodríguez , José Lamarca , Javier Morlana , Juan D. Tardós , José M. M. Montiel

Deep Learning based techniques have been adopted with precision to solve a lot of standard computer vision problems, some of which are image classification, object detection and segmentation. Despite the widespread success of these…

Computer Vision and Pattern Recognition · Computer Science 2016-11-21 Vikram Mohanty , Shubh Agrawal , Shaswat Datta , Arna Ghosh , Vishnu Dutt Sharma , Debashish Chakravarty

Existing approaches for Structure from Motion (SfM) produce impressive 3-D reconstruction results especially when using imagery captured with large parallax. However, to create engaging video-content in movies and TV shows, the amount by…

Computer Vision and Pattern Recognition · Computer Science 2022-04-07 Sheng Liu , Xiaohan Nie , Raffay Hamid

This paper presents a robust monocular visual SLAM system that simultaneously utilizes point, line, and vanishing point features for accurate camera pose estimation and mapping. To address the critical challenge of achieving reliable…

Robotics · Computer Science 2025-03-13 Bingzheng Jiang , Jiayuan Wang , Han Ding , Lijun Zhu

This paper introduces FlowMap, an end-to-end differentiable method that solves for precise camera poses, camera intrinsics, and per-frame dense depth of a video sequence. Our method performs per-video gradient-descent minimization of a…

Computer Vision and Pattern Recognition · Computer Science 2024-07-24 Cameron Smith , David Charatan , Ayush Tewari , Vincent Sitzmann

We introduce a way to learn to estimate a scene representation from a single image by predicting a low-dimensional subspace of optical flow for each training example, which encompasses the variety of possible camera and object movement.…

Computer Vision and Pattern Recognition · Computer Science 2022-10-28 Richard Strong Bowen , Richard Tucker , Ramin Zabih , Noah Snavely

DensePose supersedes traditional landmark detectors by densely mapping image pixels to body surface coordinates. This power, however, comes at a greatly increased annotation time, as supervising the model requires to manually label hundreds…

Computer Vision and Pattern Recognition · Computer Science 2019-06-14 Natalia Neverova , James Thewlis , Rıza Alp Güler , Iasonas Kokkinos , Andrea Vedaldi

A reliable sense-and-avoid system is critical to enabling safe autonomous operation of unmanned aircraft. Existing sense-and-avoid methods often require specialized sensors that are too large or power intensive for use on small unmanned…

Computer Vision and Pattern Recognition · Computer Science 2021-11-04 John Mern , Kyle Julian , Rachael E. Tompa , Mykel J. Kochenderfer

Reconstructing high-quality 3D models from sparse 2D images has garnered significant attention in computer vision. Recently, 3D Gaussian Splatting (3DGS) has gained prominence due to its explicit representation with efficient training speed…

Computer Vision and Pattern Recognition · Computer Science 2024-12-31 Keng-Wei Chang , Zi-Ming Wang , Shang-Hong Lai

We propose a depth map inference system from monocular videos based on a novel dataset for navigation that mimics aerial footage from gimbal stabilized monocular camera in rigid scenes. Unlike most navigation datasets, the lack of rotation…

Computer Vision and Pattern Recognition · Computer Science 2018-09-13 Clément Pinard , Laure Chevalley , Antoine Manzanera , David Filliat

Recovering a dynamic 3D scene from a long monocular video is crucial for dense geometry, camera motion, and temporal correspondence to remain consistent in a shared coordinate system. Existing methods face two key challenges: (1)…

Computer Vision and Pattern Recognition · Computer Science 2026-05-19 Chenyi Xu , Yihao Wu , Liqi Yan , Chao Yang , Jianhui Zhang , Fangli Guan , Pan Li

Recent geometric methods need reliable estimates of 3D motion parameters to procure accurate dense depth map of a complex dynamic scene from monocular images \cite{kumar2017monocular, ranftl2016dense}. Generally, to estimate…

Computer Vision and Pattern Recognition · Computer Science 2019-03-26 Suryansh Kumar , Ram Srivatsav Ghorakavi , Yuchao Dai , Hongdong Li

Efficient and accurate camera pose estimation forms the foundational requirement for dense reconstruction in autonomous navigation, robotic perception, and virtual simulation systems. This paper addresses the challenge via cuSfM, a…

Computer Vision and Pattern Recognition · Computer Science 2025-10-20 Jingrui Yu , Jun Liu , Kefei Ren , Joydeep Biswas , Rurui Ye , Keqiang Wu , Chirag Majithia , Di Zeng

Self-supervised monocular depth estimation approaches either ignore independently moving objects in the scene or need a separate segmentation step to identify them. We propose MonoDepthSeg to jointly estimate depth and segment moving…

Computer Vision and Pattern Recognition · Computer Science 2021-10-22 Sadra Safadoust , Fatma Güney