English
Related papers

Related papers: SAIL-Recon: Large SfM by Augmenting Scene Regressi…

200 papers

Estimating a dense depth map from a single view is geometrically ill-posed, and state-of-the-art methods rely on learning depth's relation with visual appearance using deep neural networks. On the other hand, Structure from Motion (SfM)…

Computer Vision and Pattern Recognition · Computer Science 2023-04-03 Sergio Izquierdo , Javier Civera

Solving image-to-3D from a single view is an ill-posed problem, and current neural reconstruction methods addressing it through diffusion models still rely on scene-specific optimization, constraining their generalization capability. To…

Computer Vision and Pattern Recognition · Computer Science 2024-01-09 Christian Simon , Sen He , Juan-Manuel Perez-Rua , Mengmeng Xu , Amine Benhalloum , Tao Xiang

Scene coordinate regression (SCR) methods have emerged as a promising area of research due to their potential for accurate visual localization. However, many existing SCR approaches train on samples from all image regions, including dynamic…

Computer Vision and Pattern Recognition · Computer Science 2024-09-09 Ting-Ru Liu , Hsuan-Kung Yang , Jou-Min Liu , Chun-Wei Huang , Tsung-Chih Chiang , Quan Kong , Norimasa Kobori , Chun-Yi Lee

Combining the signed distance function (SDF) and differentiable volume rendering has emerged as a powerful paradigm for surface reconstruction from multi-view images without 3D supervision. However, current methods are impeded by requiring…

Computer Vision and Pattern Recognition · Computer Science 2024-06-05 Rui Peng , Xiaodong Gu , Luyang Tang , Shihe Shen , Fanqi Yu , Ronggang Wang

Streaming 3D reconstruction aims to recover 3D information, such as camera poses and point clouds, from a video stream, which necessitates geometric accuracy, temporal consistency, and computational efficiency. Motivated by the principles…

Computer Vision and Pattern Recognition · Computer Science 2026-04-17 Lin-Zhuo Chen , Jian Gao , Yihang Chen , Ka Leong Cheng , Yipengjing Sun , Liangxiao Hu , Nan Xue , Xing Zhu , Yujun Shen , Yao Yao , Yinghao Xu

This paper introduces FlowMap, an end-to-end differentiable method that solves for precise camera poses, camera intrinsics, and per-frame dense depth of a video sequence. Our method performs per-video gradient-descent minimization of a…

Computer Vision and Pattern Recognition · Computer Science 2024-07-24 Cameron Smith , David Charatan , Ayush Tewari , Vincent Sitzmann

Visual re-localization means using a single image as input to estimate the camera's location and orientation relative to a pre-recorded environment. The highest-scoring methods are "structure based," and need the query camera's intrinsics…

Computer Vision and Pattern Recognition · Computer Science 2021-04-13 Mehmet Ozgur Turkoglu , Eric Brachmann , Konrad Schindler , Gabriel Brostow , Aron Monszpart

We present a neural-field-based large-scale reconstruction system that fuses lidar and vision data to generate high-quality reconstructions that are geometrically accurate and capture photo-realistic textures. This system adapts the…

Robotics · Computer Science 2025-02-18 Yifu Tao , Yash Bhalgat , Lanke Frank Tarimo Fu , Matias Mattamala , Nived Chebrolu , Maurice Fallon

Multi-camera systems are increasingly vital in the environmental perception of autonomous vehicles and robotics. Their physical configuration offers inherent fixed relative pose constraints that benefit Structure-from-Motion (SfM). However,…

Computer Vision and Pattern Recognition · Computer Science 2025-07-08 Peilin Tao , Hainan Cui , Diantao Tu , Shuhan Shen

Learning-based visual localization methods that use scene coordinate regression (SCR) offer the advantage of smaller map sizes. However, on datasets with complex illumination changes or image-level ambiguities, it remains a less robust…

Computer Vision and Pattern Recognition · Computer Science 2025-04-14 Xudong Jiang , Fangjinhua Wang , Silvano Galliani , Christoph Vogel , Marc Pollefeys

3D instance segmentation methods typically rely on high-quality point clouds or posed RGB-D scans, requiring complex multi-stage processing pipelines, and are highly sensitive to reconstruction noise. While recent feed-forward transformers…

Computer Vision and Pattern Recognition · Computer Science 2026-03-23 Jinyuan Qu , Hongyang Li , Lei Zhang

Simultaneous Localization and Mapping (SLAM) is pivotal in robotics, with photorealistic scene reconstruction emerging as a key challenge. To address this, we introduce Computational Alignment for Real-Time Gaussian Splatting SLAM (CaRtGS),…

Computer Vision and Pattern Recognition · Computer Science 2025-03-11 Dapeng Feng , Zhiqiang Chen , Yizhen Yin , Shipeng Zhong , Yuhua Qi , Hongbo Chen

Large scale 3D scene reconstruction is important for applications such as virtual reality and simulation. Existing neural rendering approaches (e.g., NeRF, 3DGS) have achieved realistic reconstructions on large scenes, but optimize per…

Computer Vision and Pattern Recognition · Computer Science 2024-10-01 Yun Chen , Jingkang Wang , Ze Yang , Sivabalan Manivasagam , Raquel Urtasun

In this paper, we introduce a novel image-goal navigation approach, named RFSG. Our focus lies in leveraging the fine-grained connections between goals, observations, and the environment within limited image data, all the while keeping the…

Robotics · Computer Science 2025-03-17 Zhicheng Feng , Xieyuanli Chen , Chenghao Shi , Lun Luo , Zhichao Chen , Yun-Hui Liu , Huimin Lu

We propose a new method for realistic real-time novel-view synthesis (NVS) of large scenes. Existing neural rendering methods generate realistic results, but primarily work for small scale scenes (<50 square meters) and have difficulty at…

Computer Vision and Pattern Recognition · Computer Science 2023-11-10 Jeffrey Yunfan Liu , Yun Chen , Ze Yang , Jingkang Wang , Sivabalan Manivasagam , Raquel Urtasun

We propose DrivingForward, a feed-forward Gaussian Splatting model that reconstructs driving scenes from flexible surround-view input. Driving scene images from vehicle-mounted cameras are typically sparse, with limited overlap, and the…

Computer Vision and Pattern Recognition · Computer Science 2024-12-24 Qijian Tian , Xin Tan , Yuan Xie , Lizhuang Ma

Scene-level novel view synthesis (NVS) is fundamental to many vision and graphics applications. Recently, pose-conditioned diffusion models have led to significant progress by extracting 3D information from 2D foundation models, but these…

Computer Vision and Pattern Recognition · Computer Science 2024-08-23 Joseph Tung , Gene Chou , Ruojin Cai , Guandao Yang , Kai Zhang , Gordon Wetzstein , Bharath Hariharan , Noah Snavely

Current non-rigid structure from motion (NRSfM) algorithms are mainly limited with respect to: (i) the number of images, and (ii) the type of shape variability they can handle. This has hampered the practical utility of NRSfM for many…

Computer Vision and Pattern Recognition · Computer Science 2019-08-13 Chen Kong , Simon Lucey

3D Gaussian Splatting (3DGS) has demonstrated impressive performance in 3D scene reconstruction. Beyond novel view synthesis, it shows great potential for multi-view surface reconstruction. Existing methods employ optimization-based…

Computer Vision and Pattern Recognition · Computer Science 2026-04-10 Chensheng Dai , Shengjun Zhang , Min Chen , Yueqi Duan

Scene coordinate regression (SCR) has established itself as a promising learning-based approach to visual relocalization. After mere minutes of scene-specific training, SCR models estimate camera poses of query images with high accuracy.…

Computer Vision and Pattern Recognition · Computer Science 2025-10-14 Leonard Bruns , Axel Barroso-Laguna , Tommaso Cavallari , Áron Monszpart , Sowmya Munukutla , Victor Adrian Prisacariu , Eric Brachmann