English
Related papers

Related papers: Visual Geometry Grounded Deep Structure From Motio…

200 papers

A proper scene representation is central to the pursuit of spatial intelligence where agents can robustly reconstruct and efficiently understand 3D scenes. A scene representation is either metric, such as landmark maps in 3D reconstruction,…

Computer Vision and Pattern Recognition · Computer Science 2024-11-21 Juexiao Zhang , Gao Zhu , Sihang Li , Xinhao Liu , Haorui Song , Xinran Tang , Chen Feng

Vision Foundation Models (VFMs) have delivered remarkable performance in Domain Generalized Semantic Segmentation (DGSS). However, recent methods often overlook the fact that visual cues are susceptible, whereas the underlying geometry…

Computer Vision and Pattern Recognition · Computer Science 2025-07-16 Siyu Chen , Ting Han , Changshe Zhang , Xin Luo , Meiliu Wu , Guorong Cai , Jinhe Su

3D change detection from multi-view images is essential for urban monitoring, disaster assessment, and autonomous driving. However, existing methods predominantly operate in the 2D domain, where viewpoint variations are mistaken for…

Computer Vision and Pattern Recognition · Computer Science 2026-05-19 Wei Zhang , Songhua Li , Yihang Wu , Qiang Li , Qi Wang

We present a method to reconstruct the three-dimensional trajectory of a moving instance of a known object category using stereo video data. We track the two-dimensional shape of objects on pixel level exploiting instance-aware semantic…

Computer Vision and Pattern Recognition · Computer Science 2018-08-29 Sebastian Bullinger , Christoph Bodensteiner , Michael Arens , Rainer Stiefelhagen

Driven by the advancement of 3D devices, stereo vision tasks including stereo matching and stereo conversion have emerged as a critical research frontier. Contemporary stereo vision backbones typically rely on either monocular depth…

Computer Vision and Pattern Recognition · Computer Science 2026-04-01 Ziyang Chen , Yansong Qu , You Shen , Xuan Cheng , Liujuan Cao

This paper presents an accurate and robust Structure-from-Motion (SfM) pipeline named LiVisSfM, which is an SfM-based reconstruction system that fully combines LiDAR and visual cues. Unlike most existing LiDAR-inertial odometry (LIO) and…

Computer Vision and Pattern Recognition · Computer Science 2024-10-30 Hanqing Jiang , Liyang Zhou , Zhuang Zhang , Yihao Yu , Guofeng Zhang

Over the last decades, ample achievements have been made on Structure from motion (SfM). However, the vast majority of them basically work in an offline manner, i.e., images are firstly captured and then fed together into a SfM pipeline for…

Computer Vision and Pattern Recognition · Computer Science 2024-02-15 Zongqian Zhan , Rui Xia , Yifei Yu , Yibo Xu , Xin Wang

Visual localization is the problem of estimating a camera within a scene and a key component in computer vision applications such as self-driving cars and Mixed Reality. State-of-the-art approaches for accurate visual localization use…

Computer Vision and Pattern Recognition · Computer Science 2021-03-30 Qunjie Zhou , Torsten Sattler , Marc Pollefeys , Laura Leal-Taixe

Calibrating large-scale camera arrays, such as those in dome-based setups, is time-intensive and typically requires dedicated captures of known patterns. While extrinsics in such arrays are fixed due to the physical setup, intrinsics often…

Computer Vision and Pattern Recognition · Computer Science 2025-08-04 Jinjiang You , Hewei Wang , Yijie Li , Mingxiao Huo , Long Van Tran Ha , Mingyuan Ma , Jinfeng Xu , Jiayi Zhang , Puzhen Wu , Shubham Garg , Wei Pu

Accurate 3D reconstruction is frequently hindered by visual aliasing, where visually similar but distinct surfaces (aka, doppelgangers), are incorrectly matched. These spurious matches distort the structure-from-motion (SfM) process,…

Computer Vision and Pattern Recognition · Computer Science 2025-04-08 Yuanbo Xiangli , Ruojin Cai , Hanyu Chen , Jeffrey Byrne , Noah Snavely

3D shape reconstruction from a single image is a highly ill-posed problem. Modern deep learning based systems try to solve this problem by learning an end-to-end mapping from image to shape via a deep network. In this paper, we aim to solve…

Computer Vision and Pattern Recognition · Computer Science 2019-08-02 Kejie Li , Ravi Garg , Ming Cai , Ian Reid

Statistical Shape Modeling (SSM) effectively analyzes anatomical variations within populations but is limited by the need for manual localization and segmentation, which relies on scarce medical expertise. Recent advances in deep learning…

Computer Vision and Pattern Recognition · Computer Science 2024-07-10 Janmesh Ukey , Tushar Kataria , Shireen Y. Elhabian

Incremental Structure from Motion (ISfM) has been widely used for UAV image orientation. Its efficiency, however, decreases dramatically due to the sequential constraint. Although the divide-and-conquer strategy has been utilized for…

Computer Vision and Pattern Recognition · Computer Science 2023-02-08 San Jiang , Qingquan Li , Wanshou Jiang , Wu Chen

We propose SDFDiff, a novel approach for image-based shape optimization using differentiable rendering of 3D shapes represented by signed distance functions (SDFs). Compared to other representations, SDFs have the advantage that they can…

Computer Vision and Pattern Recognition · Computer Science 2022-02-23 Yue Jiang , Dantong Ji , Zhizhong Han , Matthias Zwicker

Non-Rigid Structure-from-Motion (NRSfM) reconstructs a deformable 3D object from the correspondences established between monocular 2D images. Current NRSfM methods lack statistical robustness, which is the ability to cope with…

Computer Vision and Pattern Recognition · Computer Science 2021-06-03 Shaifali Parashar , Adrien Bartoli , Daniel Pizarro

Statistical shape modeling (SSM) is an essential tool for analyzing variations in anatomical morphology. In a typical SSM pipeline, 3D anatomical images, gone through segmentation and rigid registration, are represented using…

Computer Vision and Pattern Recognition · Computer Science 2024-01-02 Hong Xu , Shireen Y. Elhabian

Fine-grained high-resolution remote sensing mapping typically relies on localized visual features, which restricts cross-domain generalizability and often leads to fragmented predictions of large-scale land covers. While global geospatial…

Computer Vision and Pattern Recognition · Computer Science 2026-04-23 Jienan Lyu , Miao Yang , Jinchen Cai , Yiwen Hu , Guanyi Lu , Junhao Qiu , Runmin Dong

This paper addresses the problem of recovering projective camera matrices from collections of fundamental matrices in multiview settings. We make two main contributions. First, given ${n \choose 2}$ fundamental matrices computed for $n$…

Computer Vision and Pattern Recognition · Computer Science 2021-06-08 Yoni Kasten , Amnon Geifman , Meirav Galun , Ronen Basri

Modern Vision-Language Models (VLMs) achieve strong semantic recognition, yet remain brittle on elementary spatial relations such as left of, on, behind, and between. One cause of this failure arises before language reasoning begins: the…

Computer Vision and Pattern Recognition · Computer Science 2026-05-19 Renjie Gu , Kaichen Zhou , Yan Luo , Mengyu Wang

The perspective camera and the isometric surface prior have recently gathered increased attention for Non-Rigid Structure-from-Motion (NRSfM). Despite the recent progress, several challenges remain, particularly the computational complexity…

Computer Vision and Pattern Recognition · Computer Science 2018-08-14 Thomas Probst , Danda Pani Paudel , Ajad Chhatkuli , Luc Van Gool