English
Related papers

Related papers: MODEST: Multi-Optics Depth-of-Field Stereo Dataset

200 papers

Despite several solutions and experiments have been conducted recently addressing image super-resolution (SR), boosted by deep learning (DL) techniques, they do not usually design evaluations with high scaling factors, capping it at 2x or…

Image and Video Processing · Electrical Eng. & Systems 2023-06-19 Valdivino Alexandre de Santiago Júnior

Stereo vision generally involves the computation of pixel correspondences and estimation of disparities between rectified image pairs. In many applications, including simultaneous localization and mapping (SLAM) and 3D object detection, the…

Computer Vision and Pattern Recognition · Computer Science 2020-11-11 WeiQin Chuah , Ruwan Tennakoon , Reza Hoseinnezhad , Alireza Bab-Hadiashar , David Suter

We propose a learning-based method that solves monocular stereo and can be extended to fuse depth information from multiple target frames. Given two unconstrained images from a monocular camera with known intrinsic calibration, our network…

Computer Vision and Pattern Recognition · Computer Science 2019-09-13 Kaixuan Wang , Shaojie Shen

Stereo image and video generation, stereo geometry estimation, and condition-controlled view synthesis require paired data in which the variables that determine binocular geometry -- camera baseline, intrinsics, scene depth, and camera…

Computer Vision and Pattern Recognition · Computer Science 2026-05-25 Yangzhi Cui , Feng Qiao , Nathan Jacobs

We revisit the problem of visual depth estimation in the context of autonomous vehicles. Despite the progress on monocular depth estimation in recent years, we show that the gap between monocular and stereo depth accuracy remains large$-$a…

Computer Vision and Pattern Recognition · Computer Science 2020-07-09 Nikolai Smolyanskiy , Alexey Kamenev , Stan Birchfield

The reliable fusion of depth maps from multiple viewpoints has become an important problem in many 3D reconstruction pipelines. In this work, we investigate its impact on robotic bin-picking tasks such as 6D object pose estimation. The…

Robotics · Computer Science 2021-03-23 Jun Yang , Dong Li , Steven L. Waslander

We present a method for depth estimation with monocular images, which can predict high-quality depth on diverse scenes up to an affine transformation, thus preserving accurate shapes of a scene. Previous methods that predict metric depth…

Computer Vision and Pattern Recognition · Computer Science 2020-03-31 Wei Yin , Xinlong Wang , Chunhua Shen , Yifan Liu , Zhi Tian , Songcen Xu , Changming Sun , Dou Renyin

An accurate depth map of the environment is critical to the safe operation of autonomous robots and vehicles. Currently, either light detection and ranging (LIDAR) or stereo matching algorithms are used to acquire such depth information.…

Self-supervised depth estimation from monocular cameras in diverse outdoor conditions, such as daytime, rain, and nighttime, is challenging due to the difficulty of learning universal representations and the severe lack of labeled…

Computer Vision and Pattern Recognition · Computer Science 2025-03-27 Weilong Yan , Ming Li , Haipeng Li , Shuwei Shao , Robby T. Tan

We present a large-scale stereo RGB image object pose estimation dataset named the $\textbf{StereOBJ-1M}$ dataset. The dataset is designed to address challenging cases such as object transparency, translucency, and specular reflection, in…

Computer Vision and Pattern Recognition · Computer Science 2022-03-16 Xingyu Liu , Shun Iwase , Kris M. Kitani

We present DurLAR, a high-fidelity 128-channel 3D LiDAR dataset with panoramic ambient (near infrared) and reflectivity imagery, as well as a sample benchmark task using depth estimation for autonomous driving applications. Our driving…

Computer Vision and Pattern Recognition · Computer Science 2024-06-17 Li Li , Khalid N. Ismail , Hubert P. H. Shum , Toby P. Breckon

Optical-SAR image matching is a fundamental task for image fusion and visual navigation. However, all large-scale open SAR dataset for methods development are collected from single platform, resulting in limited satellite types and spatial…

Computer Vision and Pattern Recognition · Computer Science 2025-10-14 Yibin Ye , Xichao Teng , Shuo Chen , Yijie Bian , Tao Tan , Zhang Li

Depth estimation from a single image is an active research topic in computer vision. The most accurate approaches are based on fully supervised learning models, which rely on a large amount of dense and high-resolution (HR) ground-truth…

Computer Vision and Pattern Recognition · Computer Science 2021-09-27 Jialei Xu , Yuanchao Bai , Xianming Liu , Junjun Jiang , Xiangyang Ji

Multi-focus image fusion (MFIF) addresses the depth-of-field (DOF) limitations of optical lenses, where only objects within a specific range appear sharp. Although traditional and deep learning methods have advanced the field, challenges…

Computer Vision and Pattern Recognition · Computer Science 2025-10-23 Luca Piano , Peng Huanwen , Radu Ciprian Bilcu

Multi-View Photometric Stereo (MVPS) is a popular method for fine-detailed 3D acquisition of an object from images. Despite its outstanding results on diverse material objects, a typical MVPS experimental setup requires a well-calibrated…

Computer Vision and Pattern Recognition · Computer Science 2025-11-05 Suryansh Kumar

We present a diverse dataset of industrial metal objects. These objects are symmetric, textureless and highly reflective, leading to challenging conditions not captured in existing datasets. Our dataset contains both real-world and…

Computer Vision and Pattern Recognition · Computer Science 2022-08-24 Peter De Roovere , Steven Moonen , Nick Michiels , Francis Wyffels

We present a novel real-time visual odometry framework for a stereo setup of a depth and high-resolution event camera. Our framework balances accuracy and robustness against computational efficiency towards strong performance in challenging…

Robotics · Computer Science 2022-02-08 Yi-Fan Zuo , Jiaqi Yang , Jiaben Chen , Xia Wang , Yifu Wang , Laurent Kneip

We present a new, publicly-available image dataset generated by the NVIDIA Deep Learning Data Synthesizer intended for use in object detection, pose estimation, and tracking applications. This dataset contains 144k stereo image pairs that…

Computer Vision and Pattern Recognition · Computer Science 2020-08-14 Mona Jalal , Josef Spjut , Ben Boudaoud , Margrit Betke

Instance detection (InsDet) is a long-lasting problem in robotics and computer vision, aiming to detect object instances (predefined by some visual examples) in a cluttered scene. Despite its practical significance, its advancement is…

Computer Vision and Pattern Recognition · Computer Science 2023-10-31 Qianqian Shen , Yunhan Zhao , Nahyun Kwon , Jeeeun Kim , Yanan Li , Shu Kong

We present a new multi-sensor dataset for multi-view 3D surface reconstruction. It includes registered RGB and depth data from sensors of different resolutions and modalities: smartphones, Intel RealSense, Microsoft Kinect, industrial…