English
Related papers

Related papers: Advancing Depth Anything Model for Unsupervised Mo…

200 papers

In this paper we consider the task of image-guided depth completion where our system must infer the depth at every pixel of an input image based on the image content and a sparse set of depth measurements. We propose a novel approach that…

Computer Vision and Pattern Recognition · Computer Science 2019-12-24 Chao Qu , Ty Nguyen , Camillo J. Taylor

For a monocular 360 image, depth estimation is a challenging because the distortion increases along the latitude. To perceive the distortion, existing methods devote to designing a deep and complex network architecture. In this paper, we…

Computer Vision and Pattern Recognition · Computer Science 2022-08-04 Zhijie Shen , Chunyu Lin , Lang Nie , Kang Liao , Yao Zhao

Monocular depth estimation remains challenging, as foundation models such as Depth Anything V2 (DA-V2) struggle with real-world images that are far from the training distribution. We introduce Re-Depth Anything, a test-time self-supervision…

Computer Vision and Pattern Recognition · Computer Science 2026-03-10 Ananta R. Bhattarai , Helge Rhodin

Previous methods on estimating detailed human depth often require supervised training with `ground truth' depth data. This paper presents a self-supervised method that can be trained on YouTube videos without known depth, which makes…

Computer Vision and Pattern Recognition · Computer Science 2020-05-08 Feitong Tan , Hao Zhu , Zhaopeng Cui , Siyu Zhu , Marc Pollefeys , Ping Tan

Estimating 3D geometry from monocular colonoscopy images is challenging due to non-Lambertian surfaces, moving light sources, and large textureless regions. While recent 3D geometric foundation models eliminate the need for multi-stage…

Image and Video Processing · Electrical Eng. & Systems 2025-12-01 Zhiyi Jiang , Yifu Wang , Xuelian Cheng , Zongyuan Ge

Depth estimation from monocular images is an important task in localization and 3D reconstruction pipelines for bronchoscopic navigation. Various supervised and self-supervised deep learning-based approaches have proven themselves on this…

Image and Video Processing · Electrical Eng. & Systems 2021-09-27 Mert Asim Karaoglu , Nikolas Brasch , Marijn Stollenga , Wolfgang Wein , Nassir Navab , Federico Tombari , Alexander Ladikos

We present a novel method for predicting accurate depths from monocular images with high efficiency. This optimal efficiency is achieved by exploiting wavelet decomposition, which is integrated in a fully differentiable encoder-decoder…

Computer Vision and Pattern Recognition · Computer Science 2021-08-17 Michaël Ramamonjisoa , Michael Firman , Jamie Watson , Vincent Lepetit , Daniyar Turmukhambetov

Monocular (relative or metric) depth estimation is a critical task for various applications, such as autonomous vehicles, augmented reality and image editing. In recent years, with the increasing availability of mobile devices, accurate and…

Computer Vision and Pattern Recognition · Computer Science 2021-05-26 Mehmet Kerim Yucel , Valia Dimaridou , Anastasios Drosou , Albert Saà-Garriga

Computer-aided diagnosis for low-dose computed tomography (CT) based on deep learning has recently attracted attention as a first-line automatic testing tool because of its high accuracy and low radiation exposure. However, existing methods…

Image and Video Processing · Electrical Eng. & Systems 2022-06-28 Kyung-Su Kim , Seong Je Oh , Ju Hwan Lee , Myung Jin Chung

Recent advances in self-supervised learning havedemonstrated that it is possible to learn accurate monoculardepth reconstruction from raw video data, without using any 3Dground truth for supervision. However, in robotics…

Computer Vision and Pattern Recognition · Computer Science 2020-04-14 Robert McCraith , Lukas Neumann , Andrew Zisserman , Andrea Vedaldi

Purpose: Surgical scene understanding plays a critical role in the technology stack of tomorrow's intervention-assisting systems in endoscopic surgeries. For this, tracking the endoscope pose is a key component, but remains challenging due…

Computer Vision and Pattern Recognition · Computer Science 2023-04-18 Michel Hayoz , Christopher Hahne , Mathias Gallardo , Daniel Candinas , Thomas Kurmann , Maximilian Allan , Raphael Sznitman

Learning based methods have shown very promising results for the task of depth estimation in single images. However, most existing approaches treat depth prediction as a supervised regression problem and as a result, require vast quantities…

Computer Vision and Pattern Recognition · Computer Science 2017-04-14 Clément Godard , Oisin Mac Aodha , Gabriel J. Brostow

Monocular depth and pose estimation play an important role in the development of colonoscopy-assisted navigation, as they enable improved screening by reducing blind spots, minimizing the risk of missed or recurrent lesions, and lowering…

Computer Vision and Pattern Recognition · Computer Science 2026-02-23 Xinwei Ju , Rema Daher , Danail Stoyanov , Sophia Bano , Francisco Vasconcelos

Monocular depth estimation is known as an ill-posed task in which objects in a 2D image usually do not contain sufficient information to predict their depth. Thus, it acts differently from other tasks (e.g., classification and segmentation)…

Computer Vision and Pattern Recognition · Computer Science 2023-08-11 Wencheng Han , Junbo Yin , Jianbing Shen

We present Depth Anything at Any Condition (DepthAnything-AC), a foundation monocular depth estimation (MDE) model capable of handling diverse environmental conditions. Previous foundation MDE models achieve impressive performance across…

Computer Vision and Pattern Recognition · Computer Science 2025-07-03 Boyuan Sun , Modi Jin , Bowen Yin , Qibin Hou

One of the major challenges in Minimally Invasive Surgery (MIS) such as laparoscopy is the lack of depth perception. In recent years, laparoscopic scene tracking and surface reconstruction has been a focus of investigation to provide rich…

Computer Vision and Pattern Recognition · Computer Science 2017-03-06 Long Chen , Wen Tang , Nigel W. John , Tao Ruan Wan , Jian Jun Zhang

A significant weakness of most current deep Convolutional Neural Networks is the need to train them using vast amounts of manu- ally labelled data. In this work we propose a unsupervised framework to learn a deep convolutional neural…

Computer Vision and Pattern Recognition · Computer Science 2016-08-01 Ravi Garg , Vijay Kumar BG , Gustavo Carneiro , Ian Reid

While the keypoint-based maps created by sparse monocular simultaneous localisation and mapping (SLAM) systems are useful for camera tracking, dense 3D reconstructions may be desired for many robotic tasks. Solutions involving depth cameras…

Computer Vision and Pattern Recognition · Computer Science 2022-07-26 Tristan Laidlow , Jan Czarnowski , Stefan Leutenegger

Motivated by the astonishing capabilities of natural intelligent agents and inspired by theories from psychology, this paper explores the idea that perception gets coupled to 3D properties of the world via interaction with the environment.…

Computer Vision and Pattern Recognition · Computer Science 2020-07-17 Antonio Loquercio , Alexey Dosovitskiy , Davide Scaramuzza

Monocular depth estimation is a highly challenging problem that is often addressed with deep neural networks. While these are able to use recognition of image features to predict reasonably looking depth maps the result often has low metric…

Computer Vision and Pattern Recognition · Computer Science 2020-12-22 Patrik Persson , Linn Öström , Carl Olsson