中文
相关论文

相关论文: AMB3R: Accurate Feed-forward Metric-scale 3D Recon…

200 篇论文

Simultaneous localization and mapping is essential for position tracking and scene understanding. 3D Gaussian-based map representations enable photorealistic reconstruction and real-time rendering of scenes using multiple posed cameras. We…

计算机视觉与模式识别 · 计算机科学 2024-04-02 Lisong C. Sun , Neel P. Bhatt , Jonathan C. Liu , Zhiwen Fan , Zhangyang Wang , Todd E. Humphreys , Ufuk Topcu

Feed-forward 3D reconstruction has advanced rapidly, but current models remain unreliable in UAV photogrammetric acquisition. We argue that this failure is caused not only by appearance-domain shift, but also by UAV-specific camera-geometry…

计算机视觉与模式识别 · 计算机科学 2026-05-20 Xiang Yang , Yongli Wang , HaiFeng Li , Yunsheng Zhang

Traditional high-quality 3D scanning and reconstruction typically relies on human labor to plan the scanning procedure. With the rapid development of embodied systems such as drones and robots, there is a growing demand of performing…

计算机视觉与模式识别 · 计算机科学 2025-12-05 Chentao Shen , Sizhe Zheng , Bingqian Wu , Yaohua Feng , Yuanchen Fei , Mingyu Mei , Hanwen Jiang , Xiangru Huang

Dense 3D tracking from monocular video is fundamental to dynamic scene understanding. While recent 3D foundation models provide reliable per-frame geometry, recovering object motion in this geometry remains challenging and benefits from…

计算机视觉与模式识别 · 计算机科学 2026-05-14 Jisu Nam , Jahyeok Koo , Soowon Son , Jaewoo Jung , Honggyu An , Junhwa Hur , Seungryong Kim

Synthetic aperture radar tomographic imaging reconstructs the three-dimensional reflectivity of a scene from a set of coherent acquisitions performed in an interferometric configuration. In forest areas, a large number of elements…

图像与视频处理 · 电气工程与系统科学 2024-02-09 Zoé Berenger , Loïc Denis , Florence Tupin , Laurent Ferro-Famil , Yue Huang

Volumetric models have become a popular representation for 3D scenes in recent years. One breakthrough leading to their popularity was KinectFusion, which focuses on 3D reconstruction using RGB-D sensors. However, monocular SLAM has since…

计算机视觉与模式识别 · 计算机科学 2017-08-03 Victor Adrian Prisacariu , Olaf Kähler , Stuart Golodetz , Michael Sapienza , Tommaso Cavallari , Philip H S Torr , David W Murray

We present a novel method to reconstruct 3D scenes from images by leveraging deep dense monocular SLAM and fast uncertainty propagation. The proposed approach is able to 3D reconstruct scenes densely, accurately, and in real-time while…

计算机视觉与模式识别 · 计算机科学 2022-10-18 Antoni Rosinol , John J. Leonard , Luca Carlone

Recent advances in 3D Gaussian Splatting (3DGS) have enabled generalizable, on-the-fly reconstruction of sequential input views. However, existing methods often predict per-pixel Gaussians and combine Gaussians from all views as the scene…

计算机视觉与模式识别 · 计算机科学 2025-10-20 Jiaxin Guo , Tongfan Guan , Wenzhen Dong , Wenzhao Zheng , Wenting Wang , Yue Wang , Yeung Yam , Yun-Hui Liu

We present a unified framework capable of solving a broad range of 3D tasks. Our approach features a stateful recurrent model that continuously updates its state representation with each new observation. Given a stream of images, this…

计算机视觉与模式识别 · 计算机科学 2025-01-22 Qianqian Wang , Yifei Zhang , Aleksander Holynski , Alexei A. Efros , Angjoo Kanazawa

We propose a straightforward method that simultaneously reconstructs the 3D facial structure and provides dense alignment. To achieve this, we design a 2D representation called UV position map which records the 3D shape of a complete face…

计算机视觉与模式识别 · 计算机科学 2018-03-22 Yao Feng , Fan Wu , Xiaohu Shao , Yanfeng Wang , Xi Zhou

3D scene reconstruction is a long-standing vision task. Existing approaches can be categorized into geometry-based and learning-based methods. The former leverages multi-view geometry but can face catastrophic failures due to the reliance…

计算机视觉与模式识别 · 计算机科学 2023-08-11 Guangkai Xu , Wei Yin , Hao Chen , Chunhua Shen , Kai Cheng , Feng Zhao

Multi-view stereo reconstruction (MVS) in the wild requires to first estimate the camera parameters e.g. intrinsic and extrinsic parameters. These are usually tedious and cumbersome to obtain, yet they are mandatory to triangulate…

计算机视觉与模式识别 · 计算机科学 2024-12-03 Shuzhe Wang , Vincent Leroy , Yohann Cabon , Boris Chidlovskii , Jerome Revaud

On-the-fly 3D reconstruction from monocular image sequences is a long-standing challenge in computer vision, critical for applications such as real-to-sim, AR/VR, and robotics. Existing methods face a major tradeoff: per-scene optimization…

计算机视觉与模式识别 · 计算机科学 2025-10-10 Guanghao Li , Kerui Ren , Linning Xu , Zhewen Zheng , Changjian Jiang , Xin Gao , Bo Dai , Jian Pu , Mulin Yu , Jiangmiao Pang

We introduce the first learning-based dense matching algorithm, termed Equirectangular Projection-Oriented Dense Kernelized Feature Matching (EDM), specifically designed for omnidirectional images. Equirectangular projection (ERP) images,…

计算机视觉与模式识别 · 计算机科学 2025-03-03 Dongki Jung , Jaehoon Choi , Yonghan Lee , Somi Jeong , Taejae Lee , Dinesh Manocha , Suyong Yeon

The ability to estimate rich geometry and camera motion from monocular imagery is fundamental to future interactive robotics and augmented reality applications. Different approaches have been proposed that vary in scene geometry…

计算机视觉与模式识别 · 计算机科学 2020-01-16 Jan Czarnowski , Tristan Laidlow , Ronald Clark , Andrew J. Davison

Accurate 3D reconstruction of vehicles is vital for applications such as vehicle inspection, predictive maintenance, and urban planning. Existing methods like Neural Radiance Fields and Gaussian Splatting have shown impressive results but…

计算机视觉与模式识别 · 计算机科学 2025-07-17 Davide Di Nucci , Matteo Tomei , Guido Borghi , Luca Ciuffreda , Roberto Vezzani , Rita Cucchiara

Automatic reconstruction of 3D models from images using multi-view Structure-from-Motion methods has been one of the most fruitful outcomes of computer vision. These advances combined with the growing popularity of Micro Aerial Vehicles as…

机器人学 · 计算机科学 2016-11-15 Shreyansh Daftry , Christof Hoppe , Horst Bischof

Joint camera pose and dense geometry estimation from a set of images or a monocular video remains a challenging problem due to its computational complexity and inherent visual ambiguities. Most dense incremental reconstruction systems…

计算机视觉与模式识别 · 计算机科学 2024-04-18 Kirill Mazur , Gwangbin Bae , Andrew J. Davison

Recent works on 3D reconstruction from posed images have demonstrated that direct inference of scene-level 3D geometry without test-time optimization is feasible using deep neural networks, showing remarkable promise and high efficiency.…

计算机视觉与模式识别 · 计算机科学 2023-08-22 Noah Stier , Anurag Ranjan , Alex Colburn , Yajie Yan , Liang Yang , Fangchang Ma , Baptiste Angles

Humans can infer 3D structure from 2D images of an object based on past experience and improve their 3D understanding as they see more images. Inspired by this behavior, we introduce SAP3D, a system for 3D reconstruction and novel view…

计算机视觉与模式识别 · 计算机科学 2024-04-05 Xinyang Han , Zelin Gao , Angjoo Kanazawa , Shubham Goel , Yossi Gandelsman