English
Related papers

Related papers: Structure-preserving Image Translation for Depth E…

200 papers

The representation of geometry in real-time 3D perception systems continues to be a critical research issue. Dense maps capture complete surface shape and can be augmented with semantic labels, but their high dimensionality makes them…

Computer Vision and Pattern Recognition · Computer Science 2019-04-16 Michael Bloesch , Jan Czarnowski , Ronald Clark , Stefan Leutenegger , Andrew J. Davison

Low-overlap aerial imagery poses significant challenges to traditional photogrammetric methods, which rely heavily on high image overlap to produce accurate and complete mapping products. In this study, we propose a novel workflow based on…

Computer Vision and Pattern Recognition · Computer Science 2025-03-07 Jiageng Zhong , Qi Zhou , Ming Li , Armin Gruen , Xuan Liao

Domain adaptation, which bridges the distributions across different modalities, plays a crucial role in multimodal medical image analysis. In endoscopic imaging, combining pre-operative data with intra-operative imaging is important for…

Image and Video Processing · Electrical Eng. & Systems 2025-07-16 Junyang Wu , Fangfang Xie , Jiayuan Sun , Yun Gu , Guang-Zhong Yang

Contrastive pretraining can substantially increase model generalisation and downstream performance. However, the quality of the learned representations is highly dependent on the data augmentation strategy applied to generate positive…

Computer Vision and Pattern Recognition · Computer Science 2025-06-17 Mélanie Roschewitz , Fabio De Sousa Ribeiro , Tian Xia , Galvin Khara , Ben Glocker

Learning to predict scene depth from RGB inputs is a challenging task both for indoor and outdoor robot navigation. In this work we address unsupervised learning of scene depth and robot ego-motion where supervision is provided by monocular…

Computer Vision and Pattern Recognition · Computer Science 2018-11-16 Vincent Casser , Soeren Pirk , Reza Mahjourian , Anelia Angelova

Image translation across different domains has attracted much attention in both machine learning and computer vision communities. Taking the translation from source domain $\mathcal{D}_s$ to target domain $\mathcal{D}_t$ as an example,…

Computer Vision and Pattern Recognition · Computer Science 2019-05-30 Jianxin Lin , Yingce Xia , Yijun Wang , Tao Qin , Zhibo Chen

Reliable depth estimation under real optical conditions remains a core challenge for camera vision in systems such as autonomous robotics and augmented reality. Despite recent progress in depth estimation and depth-of-field rendering,…

Computer Vision and Pattern Recognition · Computer Science 2026-04-21 Nisarg K. Trivedi , Vinayak A. Belludi , Li-Yun Wang

Disconnectivity and distortion are the two problems which must be coped with when processing 360 degrees equirectangular images. In this paper, we propose a method of estimating the depth of monocular panoramic image with a teacher-student…

Computer Vision and Pattern Recognition · Computer Science 2024-07-18 Jingguo Liu , Yijun Xu , Shigang Li , Jianfeng Li

This research paper presents an innovative multi-task learning framework that allows concurrent depth estimation and semantic segmentation using a single camera. The proposed approach is based on a shared encoder-decoder architecture, which…

Computer Vision and Pattern Recognition · Computer Science 2024-03-19 Pardis Taghavi , Reza Langari , Gaurav Pandey

Automatic colorectal polyp detection in colonoscopy video is a fundamental task, which has received a lot of attention. Manually annotating polyp region in a large scale video dataset is time-consuming and expensive, which limits the…

Image and Video Processing · Electrical Eng. & Systems 2021-01-01 Zhi-Qin Zhan , Huazhu Fu , Yan-Yao Yang , Jingjing Chen , Jie Liu , Yu-Gang Jiang

We propose a depth map inference system from monocular videos based on a novel dataset for navigation that mimics aerial footage from gimbal stabilized monocular camera in rigid scenes. Unlike most navigation datasets, the lack of rotation…

Computer Vision and Pattern Recognition · Computer Science 2018-09-13 Clément Pinard , Laure Chevalley , Antoine Manzanera , David Filliat

Accurate polyp segmentation in colonoscopy is essential for early colorectal cancer detection, yet real-world clinical environments pose persistent challenges such as motion blur, specular reflections, and illumination instability. Most…

Computer Vision and Pattern Recognition · Computer Science 2026-05-19 Zhuoyu Wu , Wenhui Ou , Lexi Zhang , Pei-Sze Tan , Dongjun Wu , Junhe Zhao , Wenqi Fang , Raphaël C. -W. Phan

Image-to-image translation has drawn great attention during the past few years. It aims to translate an image in one domain to a given reference image in another domain. Due to its effectiveness and efficiency, many applications can be…

Computer Vision and Pattern Recognition · Computer Science 2019-11-05 Weihao Xia , Yujiu Yang , Jing-Hao Xue

Transparent object perception is indispensable for numerous robotic tasks. However, accurately segmenting and estimating the depth of transparent objects remain challenging due to complex optical properties. Existing methods primarily delve…

Computer Vision and Pattern Recognition · Computer Science 2025-03-04 Jiangyuan Liu , Hongxuan Ma , Yuxin Guo , Yuhao Zhao , Chi Zhang , Wei Sui , Wei Zou

Laparoscopic video tracking primarily focuses on two target types: surgical instruments and anatomy. The former could be used for skill assessment, while the latter is necessary for the projection of virtual overlays. Where instrument and…

Computer Vision and Pattern Recognition · Computer Science 2024-03-29 Beerend G. A. Gerats , Jelmer M. Wolterink , Seb P. Mol , Ivo A. M. J. Broeders

Purpose: A major barrier to the implementation of artificial intelligence for medical applications is the lack of explainability and high confidence for incorrect decisions, specifically with out-of-domain samples. We propose a…

Computer Vision and Pattern Recognition · Computer Science 2026-01-22 Mikyla K. Bowen , Jesse W. Wilson

Cross-domain mapping has been a very active topic in recent years. Given one image, its main purpose is to translate it to the desired target domain, or multiple domains in the case of multiple labels. This problem is highly challenging due…

Computer Vision and Pattern Recognition · Computer Science 2019-09-06 Andrés Romero , Pablo Arbeláez , Luc Van Gool , Radu Timofte

Depth estimation from monocular endoscopic images presents significant challenges due to the complexity of endoscopic surgery, such as irregular shapes of human soft tissues, as well as variations in lighting conditions. Existing methods…

Image and Video Processing · Electrical Eng. & Systems 2025-02-07 Dawei Lu , Deqiang Xiao , Danni Ai , Jingfan Fan , Tianyu Fu , Yucong Lin , Hong Song , Xujiong Ye , Lei Zhang , Jian Yang

Video depth estimation is crucial in various applications, such as scene reconstruction and augmented reality. In contrast to the naive method of estimating depths from images, a more sophisticated approach uses temporal information,…

Computer Vision and Pattern Recognition · Computer Science 2023-05-05 Elena Kosheleva , Sunil Jaiswal , Faranak Shamsafar , Noshaba Cheema , Klaus Illgner-Fehns , Philipp Slusallek

In the field of remote sensing, the scarcity of stereo-matched and particularly lack of accurate ground truth data often hinders the training of deep neural networks. The use of synthetically generated images as an alternative, alleviates…

Computer Vision and Pattern Recognition · Computer Science 2024-04-16 Vasudha Venkatesan , Daniel Panangian , Mario Fuentes Reyes , Ksenia Bittner
‹ Prev 1 8 9 10 Next ›