English
Related papers

Related papers: Task-guided Domain Gap Reduction for Monocular Dep…

200 papers

Monocular depth estimation and ego-motion estimation are significant tasks for scene perception and navigation in stable, accurate and efficient robot-assisted endoscopy. To tackle lighting variations and sparse textures in endoscopic…

Computer Vision and Pattern Recognition · Computer Science 2025-06-23 Liangjing Shao , Linxin Bai , Chenkang Du , Xinrong Chen

Monocular depth estimation in endoscopy videos can enable assistive and robotic surgery to obtain better coverage of the organ and detection of various health issues. Despite promising progress on mainstream, natural image depth estimation,…

Computer Vision and Pattern Recognition · Computer Science 2024-08-22 Akshay Paruchuri , Samuel Ehrenstein , Shuxian Wang , Inbar Fried , Stephen M. Pizer , Marc Niethammer , Roni Sengupta

Monocular relative and metric depth estimation has seen a tremendous boost in the last few years due to the sharp advancements in foundation models and in particular transformer based networks. As we start to see applications to the domain…

Computer Vision and Pattern Recognition · Computer Science 2025-09-24 Nicolas Toussaint , Emanuele Colleoni , Ricardo Sanchez-Matilla , Joshua Sutcliffe , Vanessa Thompson , Muhammad Asad , Imanol Luengo , Danail Stoyanov

Colonoscopy is a vital tool for the early diagnosis of colorectal cancer, which is one of the main causes of cancer-related mortality globally; hence, it is deemed an essential technique for the prevention and early detection of colorectal…

Computer Vision and Pattern Recognition · Computer Science 2025-08-11 Ojonugwa Oluwafemi Ejiga Peter , Akingbola Oluwapemiisin , Amalahu Chetachi , Adeniran Opeyemi , Fahmi Khalifa , Md Mahmudur Rahman

We present a novel method for predicting accurate depths from monocular images with high efficiency. This optimal efficiency is achieved by exploiting wavelet decomposition, which is integrated in a fully differentiable encoder-decoder…

Computer Vision and Pattern Recognition · Computer Science 2021-08-17 Michaël Ramamonjisoa , Michael Firman , Jamie Watson , Vincent Lepetit , Daniyar Turmukhambetov

Endoscopy is the most widely used medical technique for cancer and polyp detection inside hollow organs. However, images acquired by an endoscope are frequently affected by illumination artefacts due to the enlightenment source orientation.…

Image and Video Processing · Electrical Eng. & Systems 2022-07-07 Axel Garcia-Vega , Ricardo Espinosa , Gilberto Ochoa-Ruiz , Thomas Bazin , Luis Eduardo Falcon-Morales , Dominique Lamarque , Christian Daul

Self-supervised monocular depth estimation has been widely investigated to estimate depth images and relative poses from RGB images. This framework is attractive for researchers because the depth and pose networks can be trained from just…

Computer Vision and Pattern Recognition · Computer Science 2022-02-21 Noriaki Hirose , Kosuke Tahara

Monocular depth estimators can be trained with various forms of self-supervision from binocular-stereo data to circumvent the need for high-quality laser scans or other ground-truth data. The disadvantage, however, is that the photometric…

Computer Vision and Pattern Recognition · Computer Science 2019-09-20 Jamie Watson , Michael Firman , Gabriel J. Brostow , Daniyar Turmukhambetov

Purpose: Monocular depth estimation (MDE) is vital for scene understanding in minimally invasive surgery (MIS). However, endoscopic video sequences are often contaminated by smoke, specular reflections, blur, and occlusions, limiting the…

Computer Vision and Pattern Recognition · Computer Science 2026-03-05 Muhammad Asad , Emanuele Colleoni , Pritesh Mehta , Nicolas Toussaint , Ricardo Sanchez-Matilla , Maria Robu , Faisal Bashir , Rahim Mohammadi , Imanol Luengo , Danail Stoyanov

Simulators can efficiently generate large amounts of labeled synthetic data with perfect supervision for hard-to-label tasks like semantic segmentation. However, they introduce a domain gap that severely hurts real-world performance. We…

Computer Vision and Pattern Recognition · Computer Science 2021-08-19 Vitor Guizilini , Jie Li , Rares Ambrus , Adrien Gaidon

Current 3D semi-supervised segmentation methods face significant challenges such as limited consideration of contextual information and the inability to generate reliable pseudo-labels for effective unsupervised data use. To address these…

Computer Vision and Pattern Recognition · Computer Science 2023-11-22 Sanaz Karimijafarbigloo , Reza Azad , Yury Velichko , Ulas Bagci , Dorit Merhof

Colorectal cancer (CRC) remains one of the leading causes of cancer-related morbidity and mortality worldwide, with gastrointestinal (GI) polyps serving as critical precursors according to the World Health Organization (WHO). Early and…

Computer Vision and Pattern Recognition · Computer Science 2026-01-15 Akwasi Asare , Thanh-Huy Nguyen , Ulas Bagci

Self-supervised monocular depth estimation approaches suffer not only from scale ambiguity but also infer temporally inconsistent depth maps w.r.t. scale. While disambiguating scale during training is not possible without some kind of…

Computer Vision and Pattern Recognition · Computer Science 2023-04-19 Zeeshan Khan Suri

Geometric estimation including depth estimation and scene reconstruction is a crucial technique for colonoscopy which can provide surgeons with 3D spatial perception and navigation. However, geometric ground truth in colonoscopy is…

Computer Vision and Pattern Recognition · Computer Science 2026-05-14 Liangjing Shao , Beilei Cui , Hongliang Ren

We present a method for jointly training the estimation of depth, ego-motion, and a dense 3D translation field of objects relative to the scene, with monocular photometric consistency being the sole source of supervision. We show that this…

Computer Vision and Pattern Recognition · Computer Science 2020-11-10 Hanhan Li , Ariel Gordon , Hang Zhao , Vincent Casser , Anelia Angelova

Obtaining semantic labels on a large scale radiology image database (215,786 key images from 61,845 unique patients) is a prerequisite yet bottleneck to train highly effective deep convolutional neural network (CNN) models for image…

Computer Vision and Pattern Recognition · Computer Science 2016-03-28 Xiaosong Wang , Le Lu , Hoo-chang Shin , Lauren Kim , Isabella Nogues , Jianhua Yao , Ronald Summers

Single-view depth estimation can be remarkably effective if there is enough ground-truth depth data for supervised training. However, there are scenarios, especially in medicine in the case of endoscopies, where such data cannot be…

Computer Vision and Pattern Recognition · Computer Science 2024-11-21 Javier Rodríguez-Puigvert , Víctor M. Batlle , J. M. M. Montiel , Ruben Martinez-Cantin , Pascal Fua , Juan D. Tardós , Javier Civera

Purpose: Surgical scene understanding plays a critical role in the technology stack of tomorrow's intervention-assisting systems in endoscopic surgeries. For this, tracking the endoscope pose is a key component, but remains challenging due…

Computer Vision and Pattern Recognition · Computer Science 2023-04-18 Michel Hayoz , Christopher Hahne , Mathias Gallardo , Daniel Candinas , Thomas Kurmann , Maximilian Allan , Raphael Sznitman

Semantic segmentation of polyps and depth estimation are two important research problems in endoscopic image analysis. One of the main obstacles to conduct research on these research problems is lack of annotated data. Endoscopic…

Computer Vision and Pattern Recognition · Computer Science 2022-04-08 Shrawan Kumar Thapa , Pranav Poudel , Binod Bhattarai , Danail Stoyanov

Recent advances in self-supervised learning havedemonstrated that it is possible to learn accurate monoculardepth reconstruction from raw video data, without using any 3Dground truth for supervision. However, in robotics…

Computer Vision and Pattern Recognition · Computer Science 2020-04-14 Robert McCraith , Lukas Neumann , Andrew Zisserman , Andrea Vedaldi