English
Related papers

Related papers: DVP-MVS++: Synergize Depth-Normal-Edge and Harmoni…

200 papers

We present a novel deep-learning-based method for Multi-View Stereo. Our method estimates high resolution and highly precise depth maps iteratively, by traversing the continuous space of feasible depth values at each pixel in a binary…

Computer Vision and Pattern Recognition · Computer Science 2021-11-30 Christian Sormann , Mattia Rossi , Andreas Kuhn , Friedrich Fraundorfer

We propose a learning-based network for depth map estimation from multi-view stereo (MVS) images. Our proposed network consists of three sub-networks: 1) a base network for initial depth map estimation from an unstructured stereo image…

Computer Vision and Pattern Recognition · Computer Science 2020-03-03 Sizhang Dai , Weibing Huang

We introduce MonSter++, a geometric foundation model for multi-view depth estimation, unifying rectified stereo matching and unrectified multi-view stereo. Both tasks fundamentally recover metric depth from correspondence search and…

Computer Vision and Pattern Recognition · Computer Science 2025-09-26 Junda Cheng , Wenjing Liao , Zhipeng Cai , Longliang Liu , Gangwei Xu , Xianqi Wang , Yuzhou Wang , Zikang Yuan , Yong Deng , Jinliang Zang , Yangyang Shi , Jinhui Tang , Xin Yang

This paper presents a neural architecture MVDiffusion++ for 3D object reconstruction that synthesizes dense and high-resolution views of an object given one or a few images without camera poses. MVDiffusion++ achieves superior flexibility…

Computer Vision and Pattern Recognition · Computer Science 2024-05-01 Shitao Tang , Jiacheng Chen , Dilin Wang , Chengzhou Tang , Fuyang Zhang , Yuchen Fan , Vikas Chandra , Yasutaka Furukawa , Rakesh Ranjan

We present a novel framework for enhancing the visual fidelity and consistency of text-guided 3D Gaussian Splatting (3DGS) editing. Existing editing approaches face two critical challenges: inconsistent geometric reconstructions across…

Computer Vision and Pattern Recognition · Computer Science 2025-03-17 Xuanqi Zhang , Jieun Lee , Chris Joslin , Wonsook Lee

Multi-view Stereo (MVS) with known camera parameters is essentially a 1D search problem within a valid depth range. Recent deep learning-based MVS methods typically densely sample depth hypotheses in the depth range, and then construct…

Computer Vision and Pattern Recognition · Computer Science 2021-12-07 Zhenxing Mi , Di Chang , Dan Xu

Multi-view stereo (MVS) models based on progressive depth hypothesis narrowing have made remarkable advancements. However, existing methods haven't fully utilized the potential that the depth coverage of individual instances is smaller than…

Computer Vision and Pattern Recognition · Computer Science 2025-05-20 Yinzhe Wang , Yiwen Xiao , Hu Wang , Yiping Xu , Yan Tian

Recent progress in vision-language pretraining has enabled significant improvements to many downstream computer vision applications, such as classification, retrieval, segmentation and depth prediction. However, a fundamental capability…

Stereo matching is a core component in many computer vision and robotics systems. Despite significant advances over the last decade, handling matching ambiguities in ill-posed regions and large disparities remains an open challenge. In this…

Computer Vision and Pattern Recognition · Computer Science 2025-05-13 Gangwei Xu , Xianqi Wang , Zhaoxing Zhang , Junda Cheng , Chunyuan Liao , Xin Yang

We present MVSGaussian, a new generalizable 3D Gaussian representation approach derived from Multi-View Stereo (MVS) that can efficiently reconstruct unseen scenes. Specifically, 1) we leverage MVS to encode geometry-aware Gaussian…

Computer Vision and Pattern Recognition · Computer Science 2024-07-16 Tianqi Liu , Guangcong Wang , Shoukang Hu , Liao Shen , Xinyi Ye , Yuhang Zang , Zhiguo Cao , Wei Li , Ziwei Liu

There has recently been great interest in neural rendering methods. Some approaches use 3D geometry reconstructed with Multi-View Stereo (MVS) but cannot recover from the errors of this process, while others directly learn a volumetric…

Computer Vision and Pattern Recognition · Computer Science 2021-09-09 Georgios Kopanas , Julien Philip , Thomas Leimkühler , George Drettakis

We introduce a novel multi-view stereo (MVS) method that can simultaneously recover not just per-pixel depth but also surface normals, together with the reflectance of textureless, complex non-Lambertian surfaces captured under known but…

Computer Vision and Pattern Recognition · Computer Science 2022-11-11 Kohei Yamashita , Yuto Enyo , Shohei Nobuhara , Ko Nishino

Recovering clear images from blurry ones with an unknown blur kernel is a challenging problem. Deep image prior (DIP) proposes to use the deep network as a regularizer for a single image rather than as a supervised model, which achieves…

Computer Vision and Pattern Recognition · Computer Science 2023-11-13 Tingting Wu , Zhiyan Du , Zhi Li , Feng-Lei Fan , Tieyong Zeng

Self-supervised monocular methods can efficiently learn depth information of weakly textured surfaces or reflective objects. However, the depth accuracy is limited due to the inherent ambiguity in monocular geometric modeling. In contrast,…

Computer Vision and Pattern Recognition · Computer Science 2022-08-22 Xiaofeng Wang , Zheng Zhu , Guan Huang , Xu Chi , Yun Ye , Ziwei Chen , Xingang Wang

Recent advances in Novel View Synthesis (NVS) and 3D generation have significantly improved editing tasks, with a primary emphasis on maintaining cross-view consistency throughout the generative process. Contemporary methods typically…

Graphics · Computer Science 2025-06-23 Pham Khai Nguyen Do , Bao Nguyen Tran , Nam Nguyen , Duc Dung Nguyen

3D Gaussian Splatting has achieved impressive performance in novel view synthesis with real-time rendering capabilities. However, reconstructing high-quality surfaces with fine details using 3D Gaussians remains a challenging task. In this…

Computer Vision and Pattern Recognition · Computer Science 2024-12-03 Jiepeng Wang , Yuan Liu , Peng Wang , Cheng Lin , Junhui Hou , Xin Li , Taku Komura , Wenping Wang

3D Gaussian splatting (3DGS) has demonstrated impressive performance in synthesizing high-fidelity novel views. Nonetheless, its effectiveness critically depends on the quality of the initialized point cloud. Specifically, achieving uniform…

Computer Vision and Pattern Recognition · Computer Science 2025-10-13 Yikang Zhang , Rui Fan

We propose D3VO as a novel framework for monocular visual odometry that exploits deep networks on three levels -- deep depth, pose and uncertainty estimation. We first propose a novel self-supervised monocular depth estimation network…

Computer Vision and Pattern Recognition · Computer Science 2020-03-31 Nan Yang , Lukas von Stumberg , Rui Wang , Daniel Cremers

Accurate metric depth is critical for autonomous driving perception and simulation, yet current approaches struggle to achieve high metric accuracy, multi-view and temporal consistency, and cross-domain generalization. To address these…

Computer Vision and Pattern Recognition · Computer Science 2026-03-05 Qihao Sun , Jiarun Liu , Ziqian Ni , Jianyun Xu , Tao Xie , Lijun Zhao , Ruifeng Li , Sheng Yang

3D Gaussian Splatting (3DGS) has shown significant advantages in novel view synthesis (NVS), particularly in achieving high rendering speeds and high-quality results. However, its geometric accuracy in 3D reconstruction remains limited due…

Graphics · Computer Science 2025-02-21 Qilin Zhang , Olaf Wysocki , Steffen Urban , Boris Jutzi