中文
相关论文

相关论文: High-Resolution Synthetic RGB-D Datasets for Monoc…

200 篇论文

Monocular depth estimation (MDE) plays a pivotal role in various computer vision applications, such as robotics, augmented reality, and autonomous driving. Despite recent advancements, existing methods often fail to meet key requirements…

计算机视觉与模式识别 · 计算机科学 2025-09-29 Andrii Litvynchuk , Ivan Livinsky , Anand Ravi , Nima Kalantari , Andrii Tsarov

Accurate three-dimensional perception is a fundamental task in several computer vision applications. Recently, commercial RGB-depth (RGB-D) cameras have been widely adopted as single-view depth-sensing devices owing to their efficient…

计算机视觉与模式识别 · 计算机科学 2022-07-27 Jiwan Kim , Minchang Kim , Yeong-Gil Shin , Minyoung Chung

Guided depth super-resolution (GDSR) is an essential topic in multi-modal image processing, which reconstructs high-resolution (HR) depth maps from low-resolution ones collected with suboptimal conditions with the help of HR RGB images of…

计算机视觉与模式识别 · 计算机科学 2022-04-22 Zixiang Zhao , Jiangshe Zhang , Shuang Xu , Zudi Lin , Hanspeter Pfister

Depth information is essential for on-board perception in autonomous driving and driver assistance. Monocular depth estimation (MDE) is very appealing since it allows for appearance and depth being on direct pixelwise correspondence without…

计算机视觉与模式识别 · 计算机科学 2022-06-07 Akhil Gurram , Ahmet Faruk Tuna , Fengyi Shen , Onay Urfalioglu , Antonio M. López

An accurate depth map of the environment is critical to the safe operation of autonomous robots and vehicles. Currently, either light detection and ranging (LIDAR) or stereo matching algorithms are used to acquire such depth information.…

Robust three-dimensional scene understanding is now an ever-growing area of research highly relevant in many real-world applications such as autonomous driving and robotic navigation. In this paper, we propose a multi-task learning-based…

计算机视觉与模式识别 · 计算机科学 2019-08-16 Amir Atapour-Abarghouei , Toby P. Breckon

Autonomous field robots operating in unstructured environments require robust perception to ensure safe and reliable operations. Recent advances in monocular depth estimation have demonstrated the potential of low-cost cameras as depth…

机器人学 · 计算机科学 2026-05-21 Marco Job , Thomas Stastny , Eleni Kelasidi , Roland Siegwart , Michael Pantic

Depth cameras have found applications in diverse fields, such as computer vision, artificial intelligence, and video gaming. However, the high latency and low frame rate of existing commodity depth cameras impose limitations on their…

计算机视觉与模式识别 · 计算机科学 2023-05-25 Peyman Gholami , Robert Xiao

Ground-truth RGBD data are fundamental for a wide range of computer vision applications; however, those labeled samples are difficult to collect and time-consuming to produce. A common solution to overcome this lack of data is to employ…

计算机视觉与模式识别 · 计算机科学 2024-05-28 L. Papa , P. Russo , I. Amerini

In this paper, we explore how synthetically generated 3D face models can be used to construct a high accuracy ground truth for depth. This allows us to train the Convolutional Neural Networks (CNN) to solve facial depth estimation problems.…

图像与视频处理 · 电气工程与系统科学 2020-03-27 Faisal Khan , Shubhajit Basak , Hossein Javidnia , Michael Schukat , Peter Corcoran

The capabilities of monocular depth estimation (MDE) models are limited by the availability of sufficient and diverse datasets. In the case of MDE models for autonomous driving, this issue is exacerbated by the linearity of the captured…

计算机视觉与模式识别 · 计算机科学 2024-09-17 Casimir Feldmann , Niall Siegenheim , Nikolas Hars , Lovro Rabuzin , Mert Ertugrul , Luca Wolfart , Marc Pollefeys , Zuria Bauer , Martin R. Oswald

RGB-D data is essential for solving many problems in computer vision. Hundreds of public RGB-D datasets containing various scenes, such as indoor, outdoor, aerial, driving, and medical, have been proposed. These datasets are useful for…

计算机视觉与模式识别 · 计算机科学 2022-08-10 Alexandre Lopes , Roberto Souza , Helio Pedrini

We introduce SceneNet RGB-D, expanding the previous work of SceneNet to enable large scale photorealistic rendering of indoor scene trajectories. It provides pixel-perfect ground truth for scene understanding problems such as semantic…

计算机视觉与模式识别 · 计算机科学 2017-01-31 John McCormac , Ankur Handa , Stefan Leutenegger , Andrew J. Davison

In this paper, we present a novel tone mapping algorithm that can be used for displaying wide dynamic range (WDR) images on low dynamic range (LDR) devices. The proposed algorithm is mainly motivated by the logarithmic response and local…

计算机视觉与模式识别 · 计算机科学 2021-02-02 Jie Yang , Ziyi Liu , Ulian Shahnovich , Orly Yadid-Pecht

Image dehazing is one of the important and popular topics in computer vision and machine learning. A reliable real-time dehazing method with reliable performance is highly desired for many applications such as autonomous driving, security…

计算机视觉与模式识别 · 计算机科学 2022-12-16 Ruoteng Li , Xiaoyi Zhang , Shaodi You , Yu Li

Nowadays, robotics, AR, and 3D modeling applications attract considerable attention to single-view depth estimation (SVDE) as it allows estimating scene geometry from a single RGB image. Recent works have demonstrated that the accuracy of…

计算机视觉与模式识别 · 计算机科学 2023-06-06 Nikolay Patakin , Mikhail Romanov , Anna Vorontsova , Mikhail Artemyev , Anton Konushin

Limited by the cost and technology, the resolution of depth map collected by depth camera is often lower than that of its associated RGB camera. Although there have been many researches on RGB image super-resolution (SR), a major problem…

计算机视觉与模式识别 · 计算机科学 2020-11-25 Chuhua Xian , Kun Qian , Zitian Zhang , Charlie C. L. Wang

In this paper we present a compositing image synthesis method that generates RGB canvases with well aligned segmentation maps and sparse depth maps, coupled with an in-painting network that transforms the RGB canvases into high quality RGB…

计算机视觉与模式识别 · 计算机科学 2022-06-06 Valentina Musat , Daniele De Martini , Matthew Gadd , Paul Newman

Monocular Depth Estimation (MDE) is a foundational task for computer vision. Traditional methods are limited by data scarcity and quality, hindering their robustness. To overcome this, we propose BRIDGE, an RL-optimized depth-to-image (D2I)…

计算机视觉与模式识别 · 计算机科学 2025-10-01 Dingning Liu , Haoyu Guo , Jingyi Zhou , Tong He

For many fundamental scene understanding tasks, it is difficult or impossible to obtain per-pixel ground truth labels from real images. We address this challenge by introducing Hypersim, a photorealistic synthetic dataset for holistic…

计算机视觉与模式识别 · 计算机科学 2021-08-19 Mike Roberts , Jason Ramapuram , Anurag Ranjan , Atulit Kumar , Miguel Angel Bautista , Nathan Paczan , Russ Webb , Joshua M. Susskind