中文
相关论文

相关论文: Helvipad: A Real-World Dataset for Omnidirectional…

200 篇论文

Modern cameras are equipped with a wide array of sensors that enable recording the geospatial context of an image. Taking advantage of this, we explore depth estimation under the assumption that the camera is geocalibrated, a problem we…

计算机视觉与模式识别 · 计算机科学 2021-09-22 Scott Workman , Hunter Blanton

Current self-supervised methods for monocular depth estimation are largely based on deeply nested convolutional networks that leverage stereo image pairs or monocular sequences during a training phase. However, they often exhibit inaccurate…

计算机视觉与模式识别 · 计算机科学 2021-10-25 Jaehoon Cho , Dongbo Min , Youngjung Kim , Kwanghoon Sohn

The SCARED dataset is a widely used benchmark for endoscopic depth estimation, offering ground-truth 3D reconstructions captured with a structured light sensor. However, the depth maps for non-keyframe images rely on robot kinematics that…

计算机视觉与模式识别 · 计算机科学 2026-05-19 John J. Han , Adam Schmidt , Max Allan , Jie Ying Wu , Omid Mohareri

Text-to-image diffusion models have emerged as powerful priors for real-world image super-resolution (Real-ISR). However, existing methods may produce unintended results due to noisy text prompts and their lack of spatial information. In…

计算机视觉与模式识别 · 计算机科学 2024-12-02 Li-Yuan Tsao , Hao-Wei Chen , Hao-Wei Chung , Deqing Sun , Chun-Yi Lee , Kelvin C. K. Chan , Ming-Hsuan Yang

We propose Stereo Direct Sparse Odometry (Stereo DSO) as a novel method for highly accurate real-time visual odometry estimation of large-scale environments from stereo cameras. It jointly optimizes for all the model parameters within the…

计算机视觉与模式识别 · 计算机科学 2017-08-29 Rui Wang , Martin Schwörer , Daniel Cremers

We present a novel real-time visual odometry framework for a stereo setup of a depth and high-resolution event camera. Our framework balances accuracy and robustness against computational efficiency towards strong performance in challenging…

机器人学 · 计算机科学 2022-02-08 Yi-Fan Zuo , Jiaqi Yang , Jiaben Chen , Xia Wang , Yifu Wang , Laurent Kneip

While head-mounted devices are becoming more compact, they provide egocentric views with significant self-occlusions of the device user. Hence, existing methods often fail to accurately estimate complex 3D poses from egocentric views. In…

计算机视觉与模式识别 · 计算机科学 2024-05-16 Hiroyasu Akada , Jian Wang , Vladislav Golyanik , Christian Theobalt

Learning accurate depth is essential to multi-view 3D object detection. Recent approaches mainly learn depth from monocular images, which confront inherent difficulties due to the ill-posed nature of monocular depth learning. Instead of…

计算机视觉与模式识别 · 计算机科学 2022-08-23 Zengran Wang , Chen Min , Zheng Ge , Yinhao Li , Zeming Li , Hongyu Yang , Di Huang

Despite recent advances in stereo matching, the extension to intricate underwater settings remains unexplored, primarily owing to: 1) the reduced visibility, low contrast, and other adverse effects of underwater images; 2) the difficulty in…

计算机视觉与模式识别 · 计算机科学 2024-09-04 Qingxuan Lv , Junyu Dong , Yuezun Li , Sheng Chen , Hui Yu , Shu Zhang , Wenhan Wang

Training deep networks for semantic segmentation requires large amounts of labeled training data, which presents a major challenge in practice, as labeling segmentation masks is a highly labor-intensive process. To address this issue, we…

计算机视觉与模式识别 · 计算机科学 2021-04-06 Lukas Hoyer , Dengxin Dai , Yuhua Chen , Adrian Köring , Suman Saha , Luc Van Gool

This paper addresses the problem of single image depth estimation (SIDE), focusing on improving the quality of deep neural network predictions. In a supervised learning scenario, the quality of predictions is intrinsically related to the…

计算机视觉与模式识别 · 计算机科学 2020-02-10 Nícolas Rosa , Vitor Guizilini , Valdir Grassi

Multimodal large language models (MLLMs) have achieved impressive performance across various tasks such as image captioning and visual question answer(VQA); however, they often struggle to accurately interpret depth information inherent in…

计算机视觉与模式识别 · 计算机科学 2026-03-09 Hao Yang , Hongbo Zhang , Yanyan Zhao , Bing Qin

Depth acquisition with the active stereo camera is a challenging task for highly reflective objects. When setup permits, multi-view fusion can provide increased levels of depth completion. However, due to the slow acquisition speed of…

计算机视觉与模式识别 · 计算机科学 2022-03-01 Jun Yang , Steven L. Waslander

Self-supervised depth estimation draws a lot of attention recently as it can promote the 3D sensing capabilities of self-driving vehicles. However, it intrinsically relies upon the photometric consistency assumption, which hardly holds…

计算机视觉与模式识别 · 计算机科学 2023-02-03 Yupeng Zheng , Chengliang Zhong , Pengfei Li , Huan-ang Gao , Yuhang Zheng , Bu Jin , Ling Wang , Hao Zhao , Guyue Zhou , Qichao Zhang , Dongbin Zhao

Omnidirectional image (ODI) data is captured with a 360x180 field-of-view, which is much wider than the pinhole cameras and contains richer spatial information than the conventional planar images. Accordingly, omnidirectional vision has…

计算机视觉与模式识别 · 计算机科学 2022-05-25 Hao Ai , Zidong Cao , Jinjing Zhu , Haotian Bai , Yucheng Chen , Lin Wang

Multi-modal depth estimation is one of the key challenges for endowing autonomous machines with robust robotic perception capabilities. There have been outstanding advances in the development of uni-modal depth estimation techniques based…

机器人学 · 计算机科学 2023-07-21 Johan S. Obando-Ceron , Victor Romero-Cano , Sildomar Monteiro

Due to difficulties in acquiring ground truth depth of equirectangular (360) images, the quality and quantity of equirectangular depth data today is insufficient to represent the various scenes in the world. Therefore, 360 depth estimation…

计算机视觉与模式识别 · 计算机科学 2023-04-17 Ilwi Yun , Hyuk-Jae Lee , Chae Eun Rhee

The rapid advancement of deep learning has intensified the need for comprehensive data for use by autonomous driving algorithms. High-quality datasets are crucial for the development of effective data-driven autonomous driving solutions.…

计算机视觉与模式识别 · 计算机科学 2026-02-10 Lianqing Zheng , Long Yang , Qunshu Lin , Wenjin Ai , Minghao Liu , Shouyi Lu , Jianan Liu , Hongze Ren , Jingyue Mo , Xiaokai Bai , Jie Bai , Zhixiong Ma , Xichan Zhu

Recent supervised multi-view depth estimation networks have achieved promising results. Similar to all supervised approaches, these networks require ground-truth data during training. However, collecting a large amount of multi-view depth…

计算机视觉与模式识别 · 计算机科学 2021-04-08 Jiayu Yang , Jose M. Alvarez , Miaomiao Liu

Image restoration under adverse conditions, such as underwater, haze or fog, and low-light environments, remains a highly challenging problem due to complex physical degradations and severe information loss. Existing datasets are…

计算机视觉与模式识别 · 计算机科学 2026-04-15 Deqing Yang , Yingying Liu , Qicong Wang , Zhi Zeng , Dajiang Lu , Yibin Tian