English
Related papers

Related papers: Marginalized Bundle Adjustment: Multi-View Camera …

200 papers

We present a novel unsupervised learning framework for single view depth estimation using monocular videos. It is well known in 3D vision that enlarging the baseline can increase the depth estimation accuracy, and jointly optimizing a set…

Computer Vision and Pattern Recognition · Computer Science 2018-12-11 Lipu Zhou , Jiamin Ye , Montiel Abello , Shengze Wang , Michael Kaess

Monocular depth estimation (MDE) has attracted intense study due to its low cost and critical functions for robotic tasks such as localization, mapping and obstacle detection. Supervised approaches have led to great success with the advance…

Computer Vision and Pattern Recognition · Computer Science 2022-08-02 Shao-Yuan Lo , Wei Wang , Jim Thomas , Jingjing Zheng , Vishal M. Patel , Cheng-Hao Kuo

Most methods for Bundle Adjustment (BA) in computer vision are either centralized or operate incrementally. This leads to poor scaling and affects the quality of solution as the number of images grows in large scale structure from motion…

Computer Vision and Pattern Recognition · Computer Science 2017-08-29 Karthikeyan Natesan Ramamurthy , Chung-Ching Lin , Aleksandr Aravkin , Sharath Pankanti , Raphael Viguier

Structure from Motion (SfM) and visual localization in indoor texture-less scenes and industrial scenarios present prevalent yet challenging research topics. Existing SfM methods designed for natural scenes typically yield low accuracy or…

Robotics · Computer Science 2024-05-28 Yusen Xie , Zhenmin Huang , Kai Chen , Lei Zhu , Jun Ma

Pipe inspection is a critical task for many industries and infrastructure of a city. The 3D information of a pipe can be used for revealing the deformation of the pipe surface and position of the camera during the inspection. In this paper,…

Computer Vision and Pattern Recognition · Computer Science 2020-07-06 Sho kagami , Hajime Taira , Naoyuki Miyashita , Akihiko Torii , Masatoshi Okutomi

Self-supervised monocular depth estimation (MDE) has gained popularity for obtaining depth predictions directly from videos. However, these methods often produce scale invariant results, unless additional training signals are provided.…

Computer Vision and Pattern Recognition · Computer Science 2025-12-05 Gasser Elazab , Torben Gräber , Michael Unterreiner , Olaf Hellwich

Dense scene reconstruction for photo-realistic view synthesis has various applications, such as VR/AR, autonomous vehicles. However, most existing methods have difficulties in large-scale scenes due to three core challenges: \textit{(a)…

Computer Vision and Pattern Recognition · Computer Science 2025-12-24 Tianchen Deng , Nailin Wang , Chongdi Wang , Shenghai Yuan , Jingchuan Wang , Hesheng Wang , Danwei Wang , Weidong Chen

The structure from motion (SfM) problem in computer vision is the problem of recovering the three-dimensional ($3$D) structure of a stationary scene from a set of projective measurements, represented as a collection of two-dimensional…

Computer Vision and Pattern Recognition · Computer Science 2017-05-10 Onur Ozyesil , Vladislav Voroninski , Ronen Basri , Amit Singer

Usual Structure-from-Motion (SfM) techniques require at least trifocal overlaps to calibrate cameras and reconstruct a scene. We consider here scenarios of reduced image sets with little overlap, possibly as low as two images at most seeing…

Computer Vision and Pattern Recognition · Computer Science 2017-03-29 Yohann Salaun , Renaud Marlet , Pascal Monasse

Classical Bundle Adjustment (BA) is fundamentally limited by its reliance on precise metric initialization and prior camera intrinsics. While modern dense matchers offer high-fidelity correspondences, traditional Structure-from-Motion (SfM)…

Computer Vision and Pattern Recognition · Computer Science 2026-04-08 Jason Chui , Hector Andrade-Loarca , Daniel Cremers

Monocular Depth Estimation (MDE) aims to predict pixel-wise depth given a single RGB image. For both, the convolutional as well as the recent attention-based models, encoder-decoder-based architectures have been found to be useful due to…

Computer Vision and Pattern Recognition · Computer Science 2022-10-18 Ashutosh Agarwal , Chetan Arora

Shape from Focus (SFF) is a depth reconstruction technique that estimates scene structure from focus variations observed across a focal stack, that is, a sequence of images captured at different focus settings. A key limitation of SFF…

Computer Vision and Pattern Recognition · Computer Science 2026-04-03 Khurram Ashfaq , Muhammad Tariq Mahmood

Accurate 6D object pose estimation is a prerequisite for successfully completing robotic prehensile and non-prehensile manipulation tasks. At present, 6D pose estimation for robotic manipulation generally relies on depth sensors based on,…

Robotics · Computer Science 2025-06-23 Teng Guo , Baichuan Huang , Jingjin Yu

While initial approaches to Structure-from-Motion (SfM) revolved around both global and incremental methods, most recent applications rely on incremental systems to estimate camera poses due to their superior robustness. Though there has…

Computer Vision and Pattern Recognition · Computer Science 2023-12-01 Ayush Baid , John Lambert , Travis Driver , Akshay Krishnan , Hayk Stepanyan , Frank Dellaert

Calibrating large-scale camera arrays, such as those in dome-based setups, is time-intensive and typically requires dedicated captures of known patterns. While extrinsics in such arrays are fixed due to the physical setup, intrinsics often…

Computer Vision and Pattern Recognition · Computer Science 2025-08-04 Jinjiang You , Hewei Wang , Yijie Li , Mingxiao Huo , Long Van Tran Ha , Mingyuan Ma , Jinfeng Xu , Jiayi Zhang , Puzhen Wu , Shubham Garg , Wei Pu

Typical Structure-from-Motion (SfM) pipelines rely on finding correspondences across images, recovering the projective structure of the observed scene and upgrading it to a metric frame using camera self-calibration constraints. Solving…

Computer Vision and Pattern Recognition · Computer Science 2020-07-07 Rui Gong , Danda Pani Paudel , Ajad Chhatkuli , Luc Van Gool

Monocular metric depth estimation (MMDE) is a core challenge in computer vision, playing a pivotal role in real-world applications that demand accurate spatial understanding. Although prior works have shown promising zero-shot performance…

Computer Vision and Pattern Recognition · Computer Science 2026-04-09 Girish Chandar Ganesan , Yuliang Guo , Liu Ren , Xiaoming Liu

Monocular metric depth estimation (MMDE) is a crucial task to solve for indoor scene reconstruction on edge devices. Despite this importance, existing models are sensitive to factors such as boundary frequency of objects in the scene and…

Computer Vision and Pattern Recognition · Computer Science 2024-11-05 Sanghyun Byun , Jacob Song , Woo Seong Chung

In the last year, universal monocular metric depth estimation (universal MMDE) has gained considerable attention, serving as the foundation model for various multimedia tasks, such as video and image editing. Nonetheless, current approaches…

Computer Vision and Pattern Recognition · Computer Science 2024-08-16 Yihao Liu , Feng Xue , Anlong Ming , Mingshuai Zhao , Huadong Ma , Nicu Sebe

Accurate perception of the vehicle's 3D surroundings, including fine-scale road geometry, such as bumps, slopes, and surface irregularities, is essential for safe and comfortable vehicle control. However, conventional monocular depth…

Computer Vision and Pattern Recognition · Computer Science 2025-12-05 Gasser Elazab , Maximilian Jansen , Michael Unterreiner , Olaf Hellwich