English
Related papers

Related papers: Incremental Non-Rigid Structure-from-Motion with U…

200 papers

We propose a Pose-Free Large Reconstruction Model (PF-LRM) for reconstructing a 3D object from a few unposed images even with little visual overlap, while simultaneously estimating the relative camera poses in ~1.3 seconds on a single A100…

Computer Vision and Pattern Recognition · Computer Science 2023-11-27 Peng Wang , Hao Tan , Sai Bi , Yinghao Xu , Fujun Luan , Kalyan Sunkavalli , Wenping Wang , Zexiang Xu , Kai Zhang

Calibrating large-scale camera arrays, such as those in dome-based setups, is time-intensive and typically requires dedicated captures of known patterns. While extrinsics in such arrays are fixed due to the physical setup, intrinsics often…

Computer Vision and Pattern Recognition · Computer Science 2025-08-04 Jinjiang You , Hewei Wang , Yijie Li , Mingxiao Huo , Long Van Tran Ha , Mingyuan Ma , Jinfeng Xu , Jiayi Zhang , Puzhen Wu , Shubham Garg , Wei Pu

We introduce a novel camera model for monocular 3D Morphable Model (3DMM) regression methods that effectively captures the perspective distortion effect commonly seen in close-up facial images. Fitting 3D morphable models to video is a key…

Computer Vision and Pattern Recognition · Computer Science 2026-03-06 Toby Chong , Ryota Nakajima

We present MatDecompSDF, a novel framework for recovering high-fidelity 3D shapes and decomposing their physically-based material properties from multi-view images. The core challenge of inverse rendering lies in the ill-posed…

Computer Vision and Pattern Recognition · Computer Science 2025-12-30 Chengyu Wang , Isabella Bennett , Henry Scott , Liang Zhang , Mei Chen , Hao Li , Rui Zhao

Recovering the 3D structure of the surrounding environment is an essential task in any vision-controlled Structure-from-Motion (SfM) scheme. This paper focuses on the theoretical properties of the SfM, known as the incremental active depth…

Robotics · Computer Science 2020-03-17 Romulo T. Rodrigues , Pedro Miraldo , Dimos V. Dimarogonas , A. Pedro Aguiar

In this work, we present a new method for 3D face reconstruction from sparse-view RGB images. Unlike previous methods which are built upon 3D morphable models (3DMMs) with limited details, we leverage an implicit representation to encode…

Computer Vision and Pattern Recognition · Computer Science 2022-10-04 Moran Li , Haibin Huang , Yi Zheng , Mengtian Li , Nong Sang , Chongyang Ma

We revisit scene-level 3D object detection as the output of an object-centric framework capable of both localization and mapping using 3D oriented boxes as the underlying geometric primitive. While existing 3D object detection approaches…

Computer Vision and Pattern Recognition · Computer Science 2025-05-30 Justin Lazarow , Kai Kang , Afshin Dehghan

In recent years, neural rendering methods such as NeRFs and 3D Gaussian Splatting (3DGS) have made significant progress in scene reconstruction and novel view synthesis. However, they heavily rely on preprocessed camera poses and 3D…

Graphics · Computer Science 2025-07-01 Chenhao Zhang , Yezhi Shen , Fengqing Zhu

We present an improved model for MRF-based depth upsampling, guided by image- as well as 3D surface normal features. By exploiting the underlying camera model we define a novel regularization term that implicitly evaluates the planarity of…

Computational Geometry · Computer Science 2017-06-20 Sascha Wirges , Björn Roxin , Eike Rehder , Tilman Kühner , Martin Lauer

Depth and ego-motion estimations are essential for the localization and navigation of autonomous robots and autonomous driving. Recent studies make it possible to learn the per-pixel depth and ego-motion from the unlabeled monocular video.…

Computer Vision and Pattern Recognition · Computer Science 2022-06-09 Guangming Wang , Jiquan Zhong , Shijie Zhao , Wenhua Wu , Zhe Liu , Hesheng Wang

Typical Structure-from-Motion (SfM) pipelines rely on finding correspondences across images, recovering the projective structure of the observed scene and upgrading it to a metric frame using camera self-calibration constraints. Solving…

Computer Vision and Pattern Recognition · Computer Science 2020-07-07 Rui Gong , Danda Pani Paudel , Ajad Chhatkuli , Luc Van Gool

Structure from motion algorithms have an inherent limitation that the reconstruction can only be determined up to the unknown scale factor. Modern mobile devices are equipped with an inertial measurement unit (IMU), which can be used for…

Computer Vision and Pattern Recognition · Computer Science 2017-08-14 Janne Mustaniemi , Juho Kannala , Simo Särkkä , Jiri Matas , Janne Heikkilä

Structure from motion (SfM) is an essential computer vision problem which has not been well handled by deep learning. One of the promising trends is to apply explicit structural constraint, e.g. 3D cost volume, into the network. However,…

Computer Vision and Pattern Recognition · Computer Science 2020-08-11 Xingkui Wei , Yinda Zhang , Zhuwen Li , Yanwei Fu , Xiangyang Xue

Shape-from-Template (SfT) methods estimate 3D surface deformations from a single monocular RGB camera while assuming a 3D state known in advance (a template). This is an important yet challenging problem due to the under-constrained nature…

Computer Vision and Pattern Recognition · Computer Science 2022-03-23 Navami Kairanda , Edith Tretschk , Mohamed Elgharib , Christian Theobalt , Vladislav Golyanik

Scene reconstruction from unorganized RGB images is an important task in many computer vision applications. Multi-view Stereo (MVS) is a common solution in photogrammetry applications for the dense reconstruction of a static scene. The…

Computer Vision and Pattern Recognition · Computer Science 2019-01-15 Matthias Innmann , Kihwan Kim , Jinwei Gu , Matthias Niessner , Charles Loop , Marc Stamminger , Jan Kautz

3D Gaussian Splatting (3DGS) has revolutionized neural rendering with its efficiency and quality, but like many novel view synthesis methods, it heavily depends on accurate camera poses from Structure-from-Motion (SfM) systems. Although…

Computer Vision and Pattern Recognition · Computer Science 2025-04-08 Zhisheng Huang , Peng Wang , Jingdong Zhang , Yuan Liu , Xin Li , Wenping Wang

We address the challenging problem of dense dynamic scene reconstruction and camera pose estimation from multiple freely moving cameras -- a setting that arises naturally when multiple observers capture a shared event. Prior approaches…

Computer Vision and Pattern Recognition · Computer Science 2026-03-17 Shuo Sun , Unal Artan , Malcolm Mielle , Achim J. Lilienthaland , Martin Magnusson

In multi-view human body capture systems, the recovered 3D geometry or even the acquired imagery data can be heavily corrupted due to occlusions, noise, limited field of- view, etc. Direct estimation of 3D pose, body shape or motion on…

Computer Vision and Pattern Recognition · Computer Science 2018-02-02 Zhong Li , Yu Ji , Wei Yang , Jinwei Ye , Jingyi Yu

Neural surface reconstruction relies heavily on accurate camera poses as input. Despite utilizing advanced pose estimators like COLMAP or ARKit, camera poses can still be noisy. Existing pose-NeRF joint optimization methods handle poses…

Computer Vision and Pattern Recognition · Computer Science 2024-11-22 Yi Gu , Dongjun Ye , Zhaorui Wang , Jiaxu Wang , Jiahang Cao , Renjing Xu

It is an exciting task to recover the scene's 3d-structure and camera pose from the video sequence. Most of the current solutions divide it into two parts, monocular depth recovery and camera pose estimation. The monocular depth recovery is…

Computer Vision and Pattern Recognition · Computer Science 2018-05-24 YanTong Wu , Yang Liu