English
Related papers

Related papers: Global Structure-from-Motion Meets Feedforward Rec…

200 papers

Transformers have revolutionized deep learning based computer vision with improved performance as well as robustness to natural corruptions and adversarial attacks. Transformers are used predominantly for 2D vision tasks, including image…

Computer Vision and Pattern Recognition · Computer Science 2023-12-19 Hemang Chawla , Arnav Varma , Elahe Arani , Bahram Zonooz

We present AMB3R, a multi-view feed-forward model for dense 3D reconstruction on a metric-scale that addresses diverse 3D vision tasks. The key idea is to leverage a sparse, yet compact, volumetric scene representation as our backend,…

Computer Vision and Pattern Recognition · Computer Science 2025-11-26 Hengyi Wang , Lourdes Agapito

Feed-forward 3D reconstruction models are efficient but rigid: once trained, they perform inference in a zero-shot manner and cannot adapt to the test scene. As a result, visually plausible reconstructions often contain errors, particularly…

Computer Vision and Pattern Recognition · Computer Science 2026-04-16 Yuhang Dai , Xingyi Yang

Active 3D measurement, especially structured light (SL) has been widely used in various fields for its robustness against textureless or equivalent surfaces by low light illumination. In addition, reconstruction of large scenes by moving…

Computer Vision and Pattern Recognition · Computer Science 2024-10-22 Kazuto Ichimaru , Diego Thomas , Takafumi Iwaguchi , Hiroshi Kawasaki

Structure-from-Motion (SfM) is a fundamental 3D vision task for recovering camera parameters and scene geometry from multi-view images. While recent deep learning advances enable accurate Monocular Depth Estimation (MDE) from single images…

Computer Vision and Pattern Recognition · Computer Science 2026-02-24 Shengjie Zhu , Ahmed Abdelkader , Mark J. Matthews , Xiaoming Liu , Wen-Sheng Chu

This paper introduces a general approach to dynamic scene reconstruction from multiple moving cameras without prior knowledge or limiting constraints on the scene structure, appearance, or illumination. Existing techniques for dynamic scene…

Computer Vision and Pattern Recognition · Computer Science 2015-10-01 Armin Mustafa , Hansung Kim , Jean-Yves Guillemaut , Adrian Hilton

3D recovery from multi-stereo and stereo images, as an important application of the image-based perspective geometry, serves many applications in computer vision, remote sensing and Geomatics. In this chapter, the authors utilize the…

Computer Vision and Pattern Recognition · Computer Science 2021-06-29 Rongjun Qin , Shuang Song , Xiao Ling , Mostafa Elhashash

View-graph is an essential input to large-scale structure from motion (SfM) pipelines. Accuracy and efficiency of large-scale SfM is crucially dependent on the input view-graph. Inconsistent or inaccurate edges can lead to inferior or wrong…

Computer Vision and Pattern Recognition · Computer Science 2017-12-05 Rajvi Shah , Visesh Chari , P J Narayanan

Image based modeling and laser scanning are two commonly used approaches in large-scale architectural scene reconstruction nowadays. In order to generate a complete scene reconstruction, an effective way is to completely cover the scene…

Computer Vision and Pattern Recognition · Computer Science 2019-09-19 Xiang Gao , Shuhan Shen , Lingjie Zhu , Tianxin Shi , Zhiheng Wang , Zhanyi Hu

Most of the previous 3D human pose estimation work relied on the powerful memory capability of the network to obtain suitable 2D-3D mappings from the training data. Few works have studied the modeling of human posture deformation in motion.…

Computer Vision and Pattern Recognition · Computer Science 2023-08-22 Haorui Ji , Hui Deng , Yuchao Dai , Hongdong Li

Sparse-view 3D reconstruction is essential for applications in which dense image acquisition is impractical, such as robotics, augmented/virtual reality (AR/VR), and autonomous systems. In these settings, minimal image overlap prevents…

Computer Vision and Pattern Recognition · Computer Science 2025-07-23 Tanveer Younis , Zhanglin Cheng

We introduce the Deformable Gaussian Splats Large Reconstruction Model (DGS-LRM), the first feed-forward method predicting deformable 3D Gaussian splats from a monocular posed video of any dynamic scene. Feed-forward scene reconstruction…

3D editing is a fundamental capability for scalable 3D content creation. While image editing has rapidly evolved toward large-scale feedforward generative paradigms, 3D AI generation remains dominated by training-free editing pipelines. A…

Computer Vision and Pattern Recognition · Computer Science 2026-05-28 Jiawei Weng , Saining Zhang , Zhenxin Diao , Peishuo Li , Henghaofan Zhang , Junhao Chen , Hao Zhao

3D Gaussian Splatting is a powerful visual representation, providing high-quality and efficient 3D scene reconstruction, but it is crucially dependent on accurate camera poses typically obtained from computationally intensive processes like…

Robotics · Computer Science 2026-04-15 Daniel Yang , Jungseok Hong , John J. Leonard , Yogesh Girdhar

Reconstructing high-quality point clouds from images remains challenging in computer vision. Existing generative-model-based approaches, particularly diffusion-model approaches that directly learn the posterior, may suffer from…

Computer Vision and Pattern Recognition · Computer Science 2025-11-11 Seunghyeok Shin , Dabin Kim , Hongki Lim

Accurate three-dimensional (3D) reconstruction of cardiac chamber motion from time-resolved medical imaging modalities is of growing interest in both the clinical and biomechanical fields. Despite recent advancement, the cardiac motion…

Medical Physics · Physics 2025-04-18 Francesco Capuano , Yue-Hin Loke , Ibrahim Yildiran , Laura Olivieri , Elias Balaras

Reassembling multiple axially symmetric pots from fragmentary sherds is crucial for cultural heritage preservation, yet it poses significant challenges due to thin and sharp fracture surfaces that generate numerous false positive matches…

Image and Video Processing · Electrical Eng. & Systems 2025-02-21 Seong Jong Yoo , Sisung Liu , Muhammad Zeeshan Arshad , Jinhyeok Kim , Young Min Kim , Yiannis Aloimonos , Cornelia Fermuller , Kyungdon Joo , Jinwook Kim , Je Hyeong Hong

It is well known that visual SLAM systems based on dense matching are locally accurate but are also susceptible to long-term drift and map corruption. In contrast, feature matching methods can achieve greater long-term consistency but can…

Computer Vision and Pattern Recognition · Computer Science 2022-08-09 Xingrui Yang , Yuhang Ming , Zhaopeng Cui , Andrew Calway

Recovering the 3D structure of the surrounding environment is an essential task in any vision-controlled Structure-from-Motion (SfM) scheme. This paper focuses on the theoretical properties of the SfM, known as the incremental active depth…

Robotics · Computer Science 2020-03-17 Romulo T. Rodrigues , Pedro Miraldo , Dimos V. Dimarogonas , A. Pedro Aguiar

We propose a framework that extends Blender to exploit Structure from Motion (SfM) and Multi-View Stereo (MVS) techniques for image-based modeling tasks such as sculpting or camera and motion tracking. Applying SfM allows us to determine…

Computer Vision and Pattern Recognition · Computer Science 2021-04-06 Sebastian Bullinger , Christoph Bodensteiner , Michael Arens
‹ Prev 1 8 9 10 Next ›