English
Related papers

Related papers: EndoDAC: Efficient Adapting Foundation Model for S…

200 papers

Self-supervised learning of depth has been a highly studied topic of research as it alleviates the requirement of having ground truth annotations for predicting depth. Depth is learnt as an intermediate solution to the task of view…

Computer Vision and Pattern Recognition · Computer Science 2021-03-02 Vinay Kaushik , Kartik Jindgar , Brejesh Lall

Reconstructing the scene of robotic surgery from the stereo endoscopic video is an important and promising topic in surgical data science, which potentially supports many applications such as surgical visual perception, robotic surgery…

Computer Vision and Pattern Recognition · Computer Science 2021-07-02 Yonghao Long , Zhaoshuo Li , Chi Hang Yee , Chi Fai Ng , Russell H. Taylor , Mathias Unberath , Qi Dou

Vision-Language-Action models have emerged as a promising paradigm for robotic manipulation by unifying perception, language grounding, and action generation. However, they often struggle in scenarios requiring precise spatial…

Computer Vision and Pattern Recognition · Computer Science 2026-05-15 Tao Lin , Yuxin Du , Jiting Liu , Nuobei Zhu , Yunhe Li , Yuqian Fu , Yinxinyu Chen , Hongyi Cai , Zewei Ye , Bing Cheng , Kai Ye , Yiran Mao , Yilei Zhong , MingKang Dong , Junchi Yan , Gen Li , Bo Zhao

State-of-the-art vessel segmentation methods typically require large-scale annotated datasets and suffer from severe performance degradation under domain shifts. In clinical practice, however, acquiring extensive annotations for every new…

Image and Video Processing · Electrical Eng. & Systems 2026-03-02 Kirato Yoshihara , Yohei Sugawara , Yuta Tokuoka , Lihang Hong

Visualizing colonoscopy is crucial for medical auxiliary diagnosis to prevent undetected polyps in areas that are not fully observed. Traditional feature-based and depth-based reconstruction approaches usually end up with undesirable…

Computer Vision and Pattern Recognition · Computer Science 2024-07-24 Zhenhua Wu , Yanlin Jin , Liangdong Qiu , Xiaoguang Han , Xiang Wan , Guanbin Li

The Segment Anything Model (SAM) stands as a foundational framework for image segmentation. While it exhibits remarkable zero-shot generalization in typical scenarios, its advantage diminishes when applied to specialized domains like…

Computer Vision and Pattern Recognition · Computer Science 2024-02-01 Zihan Zhong , Zhiqiang Tang , Tong He , Haoyang Fang , Chun Yuan

Accurate 3D reconstruction of deformable soft tissues is essential for surgical robotic perception. However, low-texture surfaces, specular highlights, and instrument occlusions often fragment geometric continuity, posing a challenge for…

Computer Vision and Pattern Recognition · Computer Science 2026-05-13 Falong Fan , Yi Xie , Arnis Lektauers , Bo Liu , Jerzy Rozenblit

Estimating camera motion and intrinsics from casual videos is a core challenge in computer vision. Traditional bundle-adjustment based methods, such as SfM and SLAM, struggle to perform reliably on arbitrary data. Although specialized SfM…

Computer Vision and Pattern Recognition · Computer Science 2025-04-01 Felix Wimbauer , Weirong Chen , Dominik Muhle , Christian Rupprecht , Daniel Cremers

Laparoscopic Field of View (FOV) control is one of the most fundamental and important components in Minimally Invasive Surgery (MIS), nevertheless, the traditional manual holding paradigm may easily bring fatigue to surgical assistants, and…

Robotics · Computer Science 2021-09-23 Bin Li , Bo Lu , Yiang Lu , Qi Dou , Yun-Hui Liu

The advent of autonomous driving and advanced driver assistance systems necessitates continuous developments in computer vision for 3D scene understanding. Self-supervised monocular depth estimation, a method for pixel-wise distance…

Computer Vision and Pattern Recognition · Computer Science 2023-02-03 Arnav Varma , Hemang Chawla , Bahram Zonooz , Elahe Arani

Accurately estimating depth in 360-degree imagery is crucial for virtual reality, autonomous navigation, and immersive media applications. Existing depth estimation methods designed for perspective-view imagery fail when applied to…

Computer Vision and Pattern Recognition · Computer Science 2024-10-31 Ning-Hsu Wang , Yu-Lun Liu

Depth adjustment aims to enhance the visual experience of stereoscopic 3D (S3D) images, which accompanied with improving visual comfort and depth perception. For a human expert, the depth adjustment procedure is a sequence of iterative…

Computer Vision and Pattern Recognition · Computer Science 2021-04-15 Hak Gu Kim , Minho Park , Sangmin Lee , Seongyeop Kim , Yong Man Ro

Pre-training on image-text colonoscopy records offers substantial potential for improving endoscopic image analysis, but faces challenges including non-informative background images, complex medical terminology, and ambiguous multi-lesion…

Computer Vision and Pattern Recognition · Computer Science 2025-05-15 Yili He , Yan Zhu , Peiyao Fu , Ruijie Yang , Tianyi Chen , Zhihua Wang , Quanlin Li , Pinghong Zhou , Xian Yang , Shuo Wang

Depth estimation from monocular images is an important task in localization and 3D reconstruction pipelines for bronchoscopic navigation. Various supervised and self-supervised deep learning-based approaches have proven themselves on this…

Image and Video Processing · Electrical Eng. & Systems 2021-09-27 Mert Asim Karaoglu , Nikolas Brasch , Marijn Stollenga , Wolfgang Wein , Nassir Navab , Federico Tombari , Alexander Ladikos

The complex nature of medical image segmentation calls for models that are specifically designed to capture detailed, domain-specific features. Large foundation models offer considerable flexibility, yet the cost of fine-tuning these models…

Image and Video Processing · Electrical Eng. & Systems 2025-09-03 Abdelrahman Elsayed , Sarim Hashmi , Mohammed Elseiagy , Hu Wang , Mohammad Yaqub , Ibrahim Almakky

Recently, self-supervised learning technology has been applied to calculate depth and ego-motion from monocular videos, achieving remarkable performance in autonomous driving scenarios. One widely adopted assumption of depth and ego-motion…

Computer Vision and Pattern Recognition · Computer Science 2021-12-16 Shuwei Shao , Zhongcai Pei , Weihai Chen , Wentao Zhu , Xingming Wu , Dianmin Sun , Baochang Zhang

Surgical tool segmentation in endoscopic images is an important problem: it is a crucial step towards full instrument pose estimation and it is used for integration of pre- and intra-operative images into the endoscopic view. While many…

Computer Vision and Pattern Recognition · Computer Science 2020-07-10 Daniil Pakhomov , Wei Shen , Nassir Navab

Depth sensing is a critical function for robotic tasks such as localization, mapping and obstacle detection. There has been a significant and growing interest in depth estimation from a single RGB image, due to the relatively low cost and…

Computer Vision and Pattern Recognition · Computer Science 2019-03-11 Diana Wofk , Fangchang Ma , Tien-Ju Yang , Sertac Karaman , Vivienne Sze

Class-incremental learning (CIL) for endoscopic image analysis is crucial for real-world clinical applications, where diagnostic models should continuously adapt to evolving clinical data while retaining performance on previously learned…

Computer Vision and Pattern Recognition · Computer Science 2025-10-21 Bingrong Liu , Jun Shi , Yushan Zheng

Self-supervised learning of depth map prediction and motion estimation from monocular video sequences is of vital importance -- since it realizes a broad range of tasks in robotics and autonomous vehicles. A large number of research efforts…

Computer Vision and Pattern Recognition · Computer Science 2021-03-24 Ue-Hwan Kim , Jong-Hwan Kim