English
Related papers

Related papers: Revisiting PatchMatch Multi-View Stereo for Urban …

200 papers

Multi-task approaches to joint depth and segmentation prediction are well-studied for monocular images. Yet, predictions from a single-view are inherently limited, while multiple views are available in many robotics applications. On the…

Computer Vision and Pattern Recognition · Computer Science 2023-11-02 Mykhailo Shvets , Dongxu Zhao , Marc Niethammer , Roni Sengupta , Alexander C. Berg

While the keypoint-based maps created by sparse monocular simultaneous localisation and mapping (SLAM) systems are useful for camera tracking, dense 3D reconstructions may be desired for many robotic tasks. Solutions involving depth cameras…

Computer Vision and Pattern Recognition · Computer Science 2022-07-26 Tristan Laidlow , Jan Czarnowski , Stefan Leutenegger

In this paper, we propose a novel end-to-end deep neural network model for omnidirectional depth estimation from a wide-baseline multi-view stereo setup. The images captured with ultra wide field-of-view (FOV) cameras on an omnidirectional…

Computer Vision and Pattern Recognition · Computer Science 2019-08-20 Changhee Won , Jongbin Ryu , Jongwoo Lim

We present an end-to-end deep learning architecture for depth map inference from multi-view images. In the network, we first extract deep visual image features, and then build the 3D cost volume upon the reference camera frustum via the…

Computer Vision and Pattern Recognition · Computer Science 2018-07-18 Yao Yao , Zixin Luo , Shiwei Li , Tian Fang , Long Quan

Generalizable 3D Gaussian Splatting reconstruction showcases advanced Image-to-3D content creation but requires substantial computational resources and large datasets, posing challenges to training models from scratch. Current methods…

Computer Vision and Pattern Recognition · Computer Science 2026-01-05 Xiufeng Huang , Ka Chun Cheung , Runmin Cong , Simon See , Renjie Wan

Reconstructing accurate 3D surfaces for street-view scenarios is crucial for applications such as digital entertainment and autonomous driving simulation. However, existing street-view datasets, including KITTI, Waymo, and nuScenes, only…

Computer Vision and Pattern Recognition · Computer Science 2024-11-08 Yubin Hu , Kairui Wen , Heng Zhou , Xiaoyang Guo , Yong-Jin Liu

Multi-view stereo methods have achieved great success for depth estimation based on the coarse-to-fine depth learning frameworks, however, the existing methods perform poorly in recovering the depth of object boundaries and detail regions.…

Computer Vision and Pattern Recognition · Computer Science 2025-04-01 Haitao Tian , Junyang Li , Chenxing Wang , Helong Jiang

We present a novel 3D instance segmentation framework for Multi-View Stereo (MVS) buildings in urban scenes. Unlike existing works focusing on semantic segmentation of urban scenes, the emphasis of this work lies in detecting and segmenting…

Computer Vision and Pattern Recognition · Computer Science 2022-08-10 Jiazhou Chen , Yanghui Xu , Shufang Lu , Ronghua Liang , Liangliang Nan

Current methods for 3D scene reconstruction from sparse posed images employ intermediate 3D representations such as neural fields, voxel grids, or 3D Gaussians, to achieve multi-view consistent scene appearance and geometry. In this paper…

Computer Vision and Pattern Recognition · Computer Science 2025-02-03 Vitor Guizilini , Muhammad Zubair Irshad , Dian Chen , Greg Shakhnarovich , Rares Ambrus

Multi-View Stereo plays a pivotal role in civil engineering by facilitating 3D modeling, precise engineering surveying, quantitative analysis, as well as monitoring and maintenance. It serves as a valuable tool, offering high-precision and…

Computer Vision and Pattern Recognition · Computer Science 2024-10-25 Hongxin Peng , Yongjian Liao , Weijun Li , Chuanyu Fu , Guoxin Zhang , Ziquan Ding , Zijie Huang , Qiku Cao , Shuting Cai

Deep learning has made significant impacts on multi-view stereo systems. State-of-the-art approaches typically involve building a cost volume, followed by multiple 3D convolution operations to recover the input image's pixel-wise depth.…

Computer Vision and Pattern Recognition · Computer Science 2021-12-14 Zhenpei Yang , Zhile Ren , Qi Shan , Qixing Huang

In this paper, we present a system for incrementally reconstructing a dense 3D model of the geometry of an outdoor environment using a single monocular camera attached to a moving vehicle. Dense models provide a rich representation of the…

Computer Vision and Pattern Recognition · Computer Science 2021-08-18 Louis Gallagher , Varun Ravi Kumar , Senthil Yogamani , John B. McDonald

We present Multi-Baseline Gaussian Splatting (MuGS), a generalized feed-forward approach for novel view synthesis that effectively handles diverse baseline settings, including sparse input views with both small and large baselines.…

Computer Vision and Pattern Recognition · Computer Science 2025-10-27 Yaopeng Lou , Liao Shen , Tianqi Liu , Jiaqi Li , Zihao Huang , Huiqiang Sun , Zhiguo Cao

Multi-view stereopsis (MVS) tries to recover the 3D model from 2D images. As the observations become sparser, the significant 3D information loss makes the MVS problem more challenging. Instead of only focusing on densely sampled…

Computer Vision and Pattern Recognition · Computer Science 2020-05-28 Mengqi Ji , Jinzhi Zhang , Qionghai Dai , Lu Fang

Various datasets have been proposed for simultaneous localization and mapping (SLAM) and related problems. Existing datasets often include small environments, have incomplete ground truth, or lack important sensor data, such as depth and…

Computer Vision and Pattern Recognition · Computer Science 2023-01-04 Janne Mustaniemi , Juho Kannala , Esa Rahtu , Li Liu , Janne Heikkilä

This paper presents a learning-based method for multi-view depth estimation from posed images. Our core idea is a "learning-to-optimize" paradigm that iteratively indexes a plane-sweeping cost volume and regresses the depth map via a…

Computer Vision and Pattern Recognition · Computer Science 2024-11-06 Changjiang Cai , Pan Ji , Qingan Yan , Yi Xu

Integration of aerial and ground images has been proved as an efficient approach to enhance the surface reconstruction in urban environments. However, as the first step, the feature point matching between aerial and ground images is…

Computer Vision and Pattern Recognition · Computer Science 2020-11-24 Qing Zhu , Zhendong Wang , Han Hu , Linfu Xie , Xuming Ge , Yeting Zhang

We present a method for dynamic surface reconstruction of large-scale urban scenes from LiDAR. Depth-based reconstructions tend to focus on small-scale objects or large-scale SLAM reconstructions that treat moving objects as outliers. We…

Computer Vision and Pattern Recognition · Computer Science 2025-05-07 Nathaniel Chodosh , Anish Madan , Simon Lucey , Deva Ramanan

6D object pose estimation for unseen objects is essential in robotics but traditionally relies on trained models that require large datasets, high computational costs, and struggle to generalize. Zero-shot approaches eliminate the need for…

Computer Vision and Pattern Recognition · Computer Science 2025-03-11 Melvin Reka , Tessa Pulli , Markus Vincze

We present Uncertainty-aware Cascaded Stereo Network (UCS-Net) for 3D reconstruction from multiple RGB images. Multi-view stereo (MVS) aims to reconstruct fine-grained scene geometry from multi-view images. Previous learning-based MVS…

Computer Vision and Pattern Recognition · Computer Science 2020-04-21 Shuo Cheng , Zexiang Xu , Shilin Zhu , Zhuwen Li , Li Erran Li , Ravi Ramamoorthi , Hao Su