English
Related papers

Related papers: PPS-Ctrl: Controllable Sim-to-Real Translation for…

200 papers

Relative monocular depth, inferring depth up to shift and scale from a single image, is an active research topic. Recent deep learning models, trained on large and varied meta-datasets, now provide excellent performance in the domain of…

Computer Vision and Pattern Recognition · Computer Science 2024-10-22 Charlie Budd , Tom Vercauteren

We present a generic image-to-image translation framework, pixel2style2pixel (pSp). Our pSp framework is based on a novel encoder network that directly generates a series of style vectors which are fed into a pretrained StyleGAN generator,…

Computer Vision and Pattern Recognition · Computer Science 2021-04-22 Elad Richardson , Yuval Alaluf , Or Patashnik , Yotam Nitzan , Yaniv Azar , Stav Shapiro , Daniel Cohen-Or

The ability of accurate depth prediction by a convolutional neural network (CNN) is a major challenge for its wide use in practical visual simultaneous localization and mapping (SLAM) applications, such as enhanced camera tracking and dense…

Robotics · Computer Science 2022-02-02 Shing Yan Loo , Moein Shakeri , Sai Hong Tang , Syamsiah Mashohor , Hong Zhang

Recent multimodal models such as Contrastive Language-Image Pre-training (CLIP) have shown remarkable ability to align visual and linguistic representations. However, domains where small visual differences carry large semantic significance,…

Computer Vision and Pattern Recognition · Computer Science 2026-03-03 Hiroshi Sasaki

The majority of prior monocular depth estimation methods without groundtruth depth guidance focus on driving scenarios. We show that such methods generalize poorly to unseen complex indoor scenes, where objects are cluttered and arbitrarily…

Computer Vision and Pattern Recognition · Computer Science 2022-03-30 Cho-Ying Wu , Jialiang Wang , Michael Hall , Ulrich Neumann , Shuochen Su

Accurate monocular depth estimation is critical in colonoscopy for lesion localization and navigation. Foundation models trained on natural images fail to generalize directly to colonoscopy. We identify the core issue not as a semantic gap,…

Computer Vision and Pattern Recognition · Computer Science 2026-03-17 Xiaoxian Zhang , Minghai Shi , Lei Li

Positron Emission Tomography (PET) imaging is a vital tool in medical diagnostics, offering detailed insights into molecular processes within the human body. However, PET images often suffer from complicated noise, which can obscure…

Computer Vision and Pattern Recognition · Computer Science 2026-01-30 Xuehua Ye , Hongxu Yang , Adam J. Schwarz

Neural implicit representations have recently shown promising progress in dense Simultaneous Localization And Mapping (SLAM). However, existing works have shortcomings in terms of reconstruction quality and real-time performance, mainly due…

Computer Vision and Pattern Recognition · Computer Science 2025-01-14 Zhen Hong , Bowen Wang , Haoran Duan , Yawen Huang , Xiong Li , Zhenyu Wen , Xiang Wu , Wei Xiang , Yefeng Zheng

Given the recent advances in depth prediction from Convolutional Neural Networks (CNNs), this paper investigates how predicted depth maps from a deep neural network can be deployed for accurate and dense monocular reconstruction. We propose…

Computer Vision and Pattern Recognition · Computer Science 2017-04-13 Keisuke Tateno , Federico Tombari , Iro Laina , Nassir Navab

The usefulness of deep learning models in robotics is largely dependent on the availability of training data. Manual annotation of training data is often infeasible. Synthetic data is a viable alternative, but suffers from domain gap. We…

Computer Vision and Pattern Recognition · Computer Science 2022-11-18 Benedikt T. Imbusch , Max Schwarz , Sven Behnke

Compressed sensing (CS) is a valuable technique for reconstructing measurements in numerous domains. CS has not yet gained widespread adoption in scanning tunneling microscopy (STM), despite potentially offering the advantages of lower…

Mesoscale and Nanoscale Physics · Physics 2022-02-09 Brian E. Lerner , Anayeli Flores-Garibay , Benjamin J. Lawrie , Petro Maksymovych

Generative adversarial networks has emerged as a defacto standard for image translation problems. To successfully drive such models, one has to rely on additional networks e.g., discriminators and/or perceptual networks. Training these…

Computer Vision and Pattern Recognition · Computer Science 2019-08-02 M. Saquib Sarfraz , Constantin Seibold , Haroon Khalid , Rainer Stiefelhagen

In surgical computer vision applications, obtaining labeled training data is challenging due to data-privacy concerns and the need for expert annotation. Unpaired image-to-image translation techniques have been explored to automatically…

Computer Vision and Pattern Recognition · Computer Science 2024-02-22 Danush Kumar Venkatesh , Dominik Rivoir , Micha Pfeiffer , Fiona Kolbinger , Marius Distler , Jürgen Weitz , Stefanie Speidel

In surgical oncology, screening colonoscopy plays a pivotal role in providing diagnostic assistance, such as biopsy, and facilitating surgical navigation, particularly in polyp detection. Computer-assisted endoscopic surgery has recently…

Computer Vision and Pattern Recognition · Computer Science 2024-10-31 Baoru Huang , Yida Wang , Anh Nguyen , Daniel Elson , Francisco Vasconcelos , Danail Stoyanov

Remote Photoplethysmography (rPPG) enables convenient non-contact physiological measurement. Existing Self-Supervised Learning (SSL) methods commonly fall into a correlation trap: they tend to learn the most dominant periodic signals in the…

Computer Vision and Pattern Recognition · Computer Science 2026-05-05 Zhiyi Niu , Xiaoguang Tu , Bo Zhao , Junzhe Cao , Dan Guo , Zitong Yu

Existing deep calibrated photometric stereo networks basically aggregate observations under different lights based on the pre-defined operations such as linear projection and max pooling. While they are effective with the dense capture,…

Computer Vision and Pattern Recognition · Computer Science 2022-11-22 Satoshi Ikehata

State-of-the-art approaches to infer dense depth measurements from images rely on CNNs trained end-to-end on a vast amount of data. However, these approaches suffer a drastic drop in accuracy when dealing with environments much different in…

Computer Vision and Pattern Recognition · Computer Science 2019-09-10 Alessio Tonioni , Matteo Poggi , Stefano Mattoccia , Luigi Di Stefano

The photometric stereo (PS) problem consists in reconstructing the 3D-surface of an object, thanks to a set of photographs taken under different lighting directions. In this paper, we propose a multi-scale architecture for PS which,…

Computer Vision and Pattern Recognition · Computer Science 2023-10-05 Clément Hardy , Yvain Quéau , David Tschumperlé

The generation of realistic medical images from text descriptions has significant potential to address data scarcity challenges in healthcare AI while preserving patient privacy. This paper presents a comprehensive study of text-to-image…

Computer Vision and Pattern Recognition · Computer Science 2025-05-13 Mikhail Chaichuk , Sushant Gautam , Steven Hicks , Elena Tutubalina

Large-scale text-to-image generative models have been a revolutionary breakthrough in the evolution of generative AI, allowing us to synthesize diverse images that convey highly complex visual concepts. However, a pivotal challenge in…

Computer Vision and Pattern Recognition · Computer Science 2022-11-24 Narek Tumanyan , Michal Geyer , Shai Bagon , Tali Dekel
‹ Prev 1 3 4 5 6 7 10 Next ›