English
Related papers

Related papers: Revisiting Gradient-based Uncertainty for Monocula…

200 papers

Monocular depth estimation is the task of obtaining a measure of distance for each pixel using a single image. It is an important problem in computer vision and is usually solved using neural networks. Though recent works in this area have…

Computer Vision and Pattern Recognition · Computer Science 2022-09-27 Nikita Durasov , Mikhail Romanov , Valeriya Bubnova , Pavel Bogomolov , Anton Konushin

Recent work has shown that CNN-based depth and ego-motion estimators can be learned using unlabelled monocular videos. However, the performance is limited by unidentified moving objects that violate the underlying static scene assumption in…

Computer Vision and Pattern Recognition · Computer Science 2019-10-04 Jia-Wang Bian , Zhichao Li , Naiyan Wang , Huangying Zhan , Chunhua Shen , Ming-Ming Cheng , Ian Reid

Without ground truth supervision, self-supervised depth estimation can be trapped in a local minimum due to the gradient-locality issue of the photometric loss. In this paper, we present a framework to enhance depth by leveraging semantic…

Computer Vision and Pattern Recognition · Computer Science 2023-04-03 Shan Lin , Yuheng Zhi , Michael C. Yip

Monocular depth estimation using Convolutional Neural Networks (CNNs) has shown impressive performance in outdoor driving scenes. However, self-supervised learning of indoor depth from monocular sequences is quite challenging for…

Computer Vision and Pattern Recognition · Computer Science 2023-12-05 Chao Fan , Zhenyu Yin , Yue Li , Feiqing Zhang

We present a novel method for predicting accurate depths from monocular images with high efficiency. This optimal efficiency is achieved by exploiting wavelet decomposition, which is integrated in a fully differentiable encoder-decoder…

Computer Vision and Pattern Recognition · Computer Science 2021-08-17 Michaël Ramamonjisoa , Michael Firman , Jamie Watson , Vincent Lepetit , Daniyar Turmukhambetov

We propose a semantics-driven unsupervised learning approach for monocular depth and ego-motion estimation from videos in this paper. Recent unsupervised learning methods employ photometric errors between synthetic view and actual image as…

Computer Vision and Pattern Recognition · Computer Science 2020-06-09 Xiaobin Wei , Jianjiang Feng , Jie Zhou

Dense depth estimation is essential to scene-understanding for autonomous driving. However, recent self-supervised approaches on monocular videos suffer from scale-inconsistency across long sequences. Utilizing data from the ubiquitously…

Computer Vision and Pattern Recognition · Computer Science 2023-02-03 Hemang Chawla , Arnav Varma , Elahe Arani , Bahram Zonooz

Monocular depth estimation has improved significantly in recent years, driven by increasingly powerful models and large-scale training data. Predicted depth is increasingly used as an input signal for downstream tasks such as…

Computer Vision and Pattern Recognition · Computer Science 2026-05-20 Viktor Kocur , Sithu Aung , Gabrielle Flood , Yaqing Ding , Lukas Bujnak , Torsten Sattler , Zuzana Kukelova

Self-supervised monocular depth estimation approaches either ignore independently moving objects in the scene or need a separate segmentation step to identify them. We propose MonoDepthSeg to jointly estimate depth and segment moving…

Computer Vision and Pattern Recognition · Computer Science 2021-10-22 Sadra Safadoust , Fatma Güney

Event cameras offer distinct advantages over conventional frame-based sensors, including microsecond-level temporal resolution, high dynamic range, and low bandwidth. In this paper, we predict per-pixel depth distributions from monocular…

Computer Vision and Pattern Recognition · Computer Science 2026-05-12 Viktor Bergkvist , Felix Rydell , Per-Erik Forssén , David Gustafsson , Johan Rideg

Monocular depth estimation is critical for applications such as autonomous driving and scene reconstruction. While existing methods perform well under normal scenarios, their performance declines in adverse weather, due to challenging…

Computer Vision and Pattern Recognition · Computer Science 2025-05-20 Kui Jiang , Jing Cao , Zhaocheng Yu , Junjun Jiang , Jingchun Zhou

Accurate depth estimation from images is a fundamental task in many applications including scene understanding and reconstruction. Existing solutions for depth estimation often produce blurry approximations of low resolution. This paper…

Computer Vision and Pattern Recognition · Computer Science 2019-03-12 Ibraheem Alhashim , Peter Wonka

Monocular depth prediction plays a crucial role in understanding 3D scene geometry. Although recent methods have achieved impressive progress in evaluation metrics such as the pixel-wise relative error, most methods neglect the geometric…

Computer Vision and Pattern Recognition · Computer Science 2019-08-02 Wei Yin , Yifan Liu , Chunhua Shen , Youliang Yan

In self-supervised monocular depth estimation, the depth discontinuity and motion objects' artifacts are still challenging problems. Existing self-supervised methods usually utilize a single view to train the depth estimation network.…

Computer Vision and Pattern Recognition · Computer Science 2020-06-29 Jianrong Wang , Ge Zhang , Zhenyu Wu , XueWei Li , Li Liu

Self-supervised monocular depth estimation has emerged as a promising approach since it does not rely on labeled training data. Most methods combine convolution and Transformer to model long-distance dependencies to estimate depth…

Computer Vision and Pattern Recognition · Computer Science 2024-09-27 Xuezhi Xiang , Yao Wang , Lei Zhang , Denis Ombati , Himaloy Himu , Xiantong Zhen

In this article, we tackle the problem of depth estimation from single monocular images. Compared with depth estimation using multiple images such as stereo depth perception, depth from monocular images is much more challenging. Prior work…

Computer Vision and Pattern Recognition · Computer Science 2016-11-17 Fayao Liu , Chunhua Shen , Guosheng Lin , Ian Reid

Monocular depth reconstruction of complex and dynamic scenes is a highly challenging problem. While for rigid scenes learning-based methods have been offering promising results even in unsupervised cases, there exists little to no…

Computer Vision and Pattern Recognition · Computer Science 2021-10-29 Ayça Takmaz , Danda Pani Paudel , Thomas Probst , Ajad Chhatkuli , Martin R. Oswald , Luc Van Gool

Due to the domain differences and unbalanced disparity distribution across multiple datasets, current stereo matching approaches are commonly limited to a specific dataset and generalize poorly to others. Such domain shift issue is usually…

Computer Vision and Pattern Recognition · Computer Science 2023-08-01 Zhelun Shen , Xibin Song , Yuchao Dai , Dingfu Zhou , Zhibo Rao , Liangjun Zhang

Three-dimensional (3D) reconstruction from a single image is an ill-posed problem with inherent ambiguities, i.e. scale. Predicting a 3D scene from text description(s) is similarly ill-posed, i.e. spatial arrangements of objects described.…

Computer Vision and Pattern Recognition · Computer Science 2024-06-04 Ziyao Zeng , Daniel Wang , Fengyu Yang , Hyoungseob Park , Yangchao Wu , Stefano Soatto , Byung-Woo Hong , Dong Lao , Alex Wong

Existing monocular depth estimation methods have achieved excellent robustness in diverse scenes, but they can only retrieve affine-invariant depth, up to an unknown scale and shift. However, in some video-based scenarios such as video…

Computer Vision and Pattern Recognition · Computer Science 2023-04-07 Guangkai Xu , Wei Yin , Hao Chen , Chunhua Shen , Kai Cheng , Feng Wu , Feng Zhao