中文
相关论文

相关论文: RePer-360: Releasing Perspective Priors for 360$^\…

200 篇论文

The increasing use of 360 images across various domains has emphasized the need for robust depth estimation techniques tailored for omnidirectional images. However, obtaining large-scale labeled datasets for 360 depth estimation remains a…

计算机视觉与模式识别 · 计算机科学 2025-09-30 Dongki Jung , Jaehoon Choi , Yonghan Lee , Dinesh Manocha

Due to difficulties in acquiring ground truth depth of equirectangular (360) images, the quality and quantity of equirectangular depth data today is insufficient to represent the various scenes in the world. Therefore, 360 depth estimation…

计算机视觉与模式识别 · 计算机科学 2023-04-17 Ilwi Yun , Hyuk-Jae Lee , Chae Eun Rhee

While instruction-based image editing is emerging, extending it to 360$^\circ$ panoramas introduces additional challenges. Existing methods often produce implausible results in both equirectangular projections (ERP) and perspective views.…

计算机视觉与模式识别 · 计算机科学 2025-12-24 Haoyi Zhong , Fang-Lue Zhang , Andrew Chalmers , Taehyun Rhee

360{\deg} images are widely available over the last few years. This paper proposes a new technique for single 360{\deg} image depth prediction under open environments. Depth prediction from a 360{\deg} single image is not easy for two…

计算机视觉与模式识别 · 计算机科学 2022-04-05 Yuya Hasegawa , Ikehata Satoshi , Kiyoharu Aizawa

This paper presents VGGT-360, a novel training-free framework for zero-shot, geometry-consistent panoramic depth estimation. Unlike prior view-independent training-free approaches, VGGT-360 reformulates the task as panoramic reprojection…

计算机视觉与模式识别 · 计算机科学 2026-05-15 Jiayi Yuan , Haobo Jiang , De Wen Soh , Na Zhao

We present the first self-supervised method to train panoramic room layout estimation models without any labeled data. Unlike per-pixel dense depth that provides abundant correspondence constraints, layout representation is sparse and…

计算机视觉与模式识别 · 计算机科学 2022-03-31 Hao-Wen Ting , Cheng Sun , Hwann-Tzong Chen

Depth estimation from a monocular 360 image is important to the perception of the entire 3D environment. However, the inherent distortion and large field of view (FoV) in 360 images pose great challenges for this task. To this end, existing…

计算机视觉与模式识别 · 计算机科学 2026-02-05 Zhijie Shen , Chunyu Lin , Lang Nie , Kang Liao , Weisi Lin , Yao Zhao

As 360{\deg} cameras become prevalent in many autonomous systems (e.g., self-driving cars and drones), efficient 360{\deg} perception becomes more and more important. We propose a novel self-supervised learning approach for predicting the…

计算机视觉与模式识别 · 计算机科学 2018-11-14 Fu-En Wang , Hou-Ning Hu , Hsien-Tzu Cheng , Juan-Ting Lin , Shang-Ta Yang , Meng-Li Shih , Hung-Kuo Chu , Min Sun

We present 360-MLC, a self-training method based on multi-view layout consistency for finetuning monocular room-layout models using unlabeled 360-images only. This can be valuable in practical scenarios where a pre-trained model needs to be…

计算机视觉与模式识别 · 计算机科学 2022-10-25 Bolivar Solarte , Chin-Hsuan Wu , Yueh-Cheng Liu , Yi-Hsuan Tsai , Min Sun

Recent automotive vision work has focused almost exclusively on processing forward-facing cameras. However, future autonomous vehicles will not be viable without a more comprehensive surround sensing, akin to a human driver, as can be…

计算机视觉与模式识别 · 计算机科学 2018-08-21 Grégoire Payen de La Garanderie , Amir Atapour Abarghouei , Toby P. Breckon

Self-supervised monocular depth estimation aims to infer depth information without relying on labeled data. However, the lack of labeled information poses a significant challenge to the model's representation, limiting its ability to…

计算机视觉与模式识别 · 计算机科学 2024-06-14 Guodong Sun , Junjie Liu , Mingxuan Liu , Moyun Liu , Yang Zhang

In this work, we propose DiT360, a DiT-based framework that performs hybrid training on perspective and panoramic data for panoramic image generation. For the issues of maintaining geometric fidelity and photorealism in generation quality,…

计算机视觉与模式识别 · 计算机科学 2025-10-14 Haoran Feng , Dizhe Zhang , Xiangtai Li , Bo Du , Lu Qi

Depth estimation is an essential task toward full scene understanding since it allows the projection of rich semantic information captured by cameras into 3D space. While the field has gained much attention recently, datasets for depth…

计算机视觉与模式识别 · 计算机科学 2024-11-19 Markus Schön , Jona Ruof , Thomas Wodtko , Michael Buchholz , Klaus Dietmayer

Referring Expression Comprehension (REC), which aims to ground a local visual region via natural language, is a task that heavily relies on multimodal alignment. Most existing methods utilize powerful pre-trained models to transfer…

计算机视觉与模式识别 · 计算机科学 2025-06-23 Ting Liu , Zunnan Xu , Yue Hu , Liangtao Shi , Zhiqiang Wang , Quanjun Yin

Recent work on depth estimation up to now has only focused on projective images ignoring 360 content which is now increasingly and more easily produced. We show that monocular depth estimation models trained on traditional images produce…

计算机视觉与模式识别 · 计算机科学 2018-07-26 Nikolaos Zioulis , Antonis Karakottas , Dimitrios Zarpalas , Petros Daras

Self-supervised monocular depth estimation has been widely investigated to estimate depth images and relative poses from RGB images. This framework is attractive for researchers because the depth and pose networks can be trained from just…

计算机视觉与模式识别 · 计算机科学 2022-02-21 Noriaki Hirose , Kosuke Tahara

Accurately estimating depth in 360-degree imagery is crucial for virtual reality, autonomous navigation, and immersive media applications. Existing depth estimation methods designed for perspective-view imagery fail when applied to…

计算机视觉与模式识别 · 计算机科学 2024-10-31 Ning-Hsu Wang , Yu-Lun Liu

360 depth estimation has recently received great attention for 3D reconstruction owing to its omnidirectional field of view (FoV). Recent approaches are predominantly focused on cross-projection fusion with geometry-based re-projection:…

计算机视觉与模式识别 · 计算机科学 2024-05-28 Hao Ai , Lin Wang

A well-known challenge in applying deep-learning methods to omnidirectional images is spherical distortion. In dense regression tasks such as depth estimation, where structural details are required, using a vanilla CNN layer on the…

计算机视觉与模式识别 · 计算机科学 2022-03-30 Yuyan Li , Yuliang Guo , Zhixin Yan , Xinyu Huang , Ye Duan , Liu Ren

Portraits or selfie images taken from a close distance typically suffer from perspective distortion. In this paper, we propose an end-to-end deep learning-based rectification pipeline to mitigate the effects of perspective distortion. We…

计算机视觉与模式识别 · 计算机科学 2025-09-16 Ahmed Alhawwary , Janne Mustaniemi , Phong Nguyen-Ha , Janne Heikkilä
‹ 上一页 1 2 3 10 下一页 ›