English
Related papers

Related papers: End-to-end Learning for Joint Depth and Image Reco…

200 papers

Existing methods for 3D face reconstruction from a few casually captured images employ deep learning based models along with a 3D Morphable Model(3DMM) as face geometry prior. Structure From Motion(SFM), followed by Multi-View Stereo (MVS),…

Computer Vision and Pattern Recognition · Computer Science 2023-08-29 Raja Kumar , Jiahao Luo , Alex Pang , James Davis

The dual-pixel (DP) hardware works by splitting each pixel in half and creating an image pair in a single snapshot. Several works estimate depth/inverse depth by treating the DP pair as a stereo pair. However, dual-pixel disparity only…

Computer Vision and Pattern Recognition · Computer Science 2020-12-02 Liyuan Pan , Shah Chowdhury , Richard Hartley , Miaomiao Liu , Hongguang Zhang , Hongdong Li

Though there exists a reasonable forward model for blur based on optical physics, recovering depth from a collection of defocused images remains a computationally challenging optimization problem. In this paper, we show that with…

Computer Vision and Pattern Recognition · Computer Science 2026-02-27 Holly Jackson , Caleb Adams , Ignacio Lopez-Francos , Benjamin Recht

For better photography, most recent commercial cameras including smartphones have either adopted large-aperture lens to collect more light or used a burst mode to take multiple images within short times. These interesting features lead us…

Computer Vision and Pattern Recognition · Computer Science 2022-10-03 Changyeon Won , Hae-Gon Jeon

Diffusion Probabilistic Methods are employed for state-of-the-art image generation. In this work, we present a method for extending such models for performing image segmentation. The method learns end-to-end, without relying on a…

Computer Vision and Pattern Recognition · Computer Science 2022-09-08 Tomer Amit , Tal Shaharbany , Eliya Nachmani , Lior Wolf

Depth-from-focus (DFF) is a technique that infers depth using the focus change of a camera. In this work, we propose a convolutional neural network (CNN) to find the best-focused pixels in a focal stack and infer depth from the focus…

Computer Vision and Pattern Recognition · Computer Science 2022-03-21 Fengting Yang , Xiaolei Huang , Zihan Zhou

Recovering detailed facial geometry from a set of calibrated multi-view images is valuable for its wide range of applications. Traditional multi-view stereo (MVS) methods adopt an optimization-based scheme to regularize the matching cost.…

Computer Vision and Pattern Recognition · Computer Science 2022-05-06 Yunze Xiao , Hao Zhu , Haotian Yang , Zhengyu Diao , Xiangju Lu , Xun Cao

The general aim of multi-focus image fusion is to gather focused regions of different images to generate a unique all-in-focus fused image. Deep learning based methods become the mainstream of image fusion by virtue of its powerful feature…

Computer Vision and Pattern Recognition · Computer Science 2022-05-26 Boyuan Ma , Xiang Yin , Di Wu , Xiaojuan Ban

Depth information is the foundation of perception, essential for autonomous driving, robotics, and other source-constrained applications. Promptly obtaining accurate and efficient depth information allows for a rapid response in dynamic…

Computer Vision and Pattern Recognition · Computer Science 2022-10-26 Xin Zhang , Rabab Abdelfattah , Yuqi Song , Samuel A. Dauchert , Xiaofeng wang

Fully-supervised category-level pose estimation aims to determine the 6-DoF poses of unseen instances from known categories, requiring expensive mannual labeling costs. Recently, various self-supervised category-level pose estimation…

Computer Vision and Pattern Recognition · Computer Science 2024-03-20 Jingtao Sun , Yaonan Wang , Mingtao Feng , Chao Ding , Mike Zheng Shou , Ajmal Saeed Mian

Limited by the cost and technology, the resolution of depth map collected by depth camera is often lower than that of its associated RGB camera. Although there have been many researches on RGB image super-resolution (SR), a major problem…

Computer Vision and Pattern Recognition · Computer Science 2020-11-25 Chuhua Xian , Kun Qian , Zitian Zhang , Charlie C. L. Wang

Point cloud analysis is an area of increasing interest due to the development of 3D sensors that are able to rapidly measure the depth of scenes accurately. Unfortunately, applying deep learning techniques to perform point cloud analysis is…

Computer Vision and Pattern Recognition · Computer Science 2021-01-05 Junming Zhang , Ming-Yuan Yu , Ram Vasudevan , Matthew Johnson-Roberson

Current non-rigid structure from motion (NRSfM) algorithms are mainly limited with respect to: (i) the number of images, and (ii) the type of shape variability they can handle. This has hampered the practical utility of NRSfM for many…

Computer Vision and Pattern Recognition · Computer Science 2019-08-13 Chen Kong , Simon Lucey

The reconstruction of a high resolution image given a low resolution observation is an ill-posed inverse problem in imaging. Deep learning methods rely on training data to learn an end-to-end mapping from a low-resolution input to a…

Image and Video Processing · Electrical Eng. & Systems 2023-07-19 Iman Marivani , Evaggelia Tsiligianni , Bruno Cornelis , Nikos Deligiannis

In this paper we focus on learning optimized partial differential equation (PDE) models for image filtering. In this approach, the grey-scaled images are represented by a vector field of two real-valued functions and the image restoration…

Numerical Analysis · Mathematics 2019-07-12 Sílvia Barbeiro , Diogo Lobo

The latest developments in Face Restoration have yielded significant advancements in visual quality through the utilization of diverse diffusion priors. Nevertheless, the uncertainty of face identity introduced by identity-obscure inputs…

Computer Vision and Pattern Recognition · Computer Science 2025-08-29 Yushun Fang , Lu Liu , Xiang Gao , Qiang Hu , Ning Cao , Jianghe Cui , Gang Chen , Xiaoyun Zhang

We present a virtual image refocusing method over an extended depth of field (DOF) enabled by cascaded neural networks and a double-helix point-spread function (DH-PSF). This network model, referred to as W-Net, is composed of two cascaded…

Image and Video Processing · Electrical Eng. & Systems 2021-06-22 Xilin Yang , Luzhe Huang , Yilin Luo , Yichen Wu , Hongda Wang , Yair Rivenson , Aydogan Ozcan

A well-known challenge in applying deep-learning methods to omnidirectional images is spherical distortion. In dense regression tasks such as depth estimation, where structural details are required, using a vanilla CNN layer on the…

Computer Vision and Pattern Recognition · Computer Science 2022-03-30 Yuyan Li , Yuliang Guo , Zhixin Yan , Xinyu Huang , Ye Duan , Liu Ren

Monocular depth estimation and ego-motion estimation are significant tasks for scene perception and navigation in stable, accurate and efficient robot-assisted endoscopy. To tackle lighting variations and sparse textures in endoscopic…

Computer Vision and Pattern Recognition · Computer Science 2025-06-23 Liangjing Shao , Linxin Bai , Chenkang Du , Xinrong Chen

Monocular depth estimation (MDE) aims to infer per-pixel depth from a single RGB image. While diffusion models have advanced MDE with impressive generalization, they often exhibit limitations in accurately reconstructing far-range regions.…

Computer Vision and Pattern Recognition · Computer Science 2025-11-18 Mingxia Zhan , Li Zhang , Yingjie Wang , Xiaomeng Chu , Beibei Wang , Yanyong Zhang