中文
相关论文

相关论文: World-Shaper: A Unified Framework for 360{\deg} Pa…

200 篇论文

We present the first self-supervised method to train panoramic room layout estimation models without any labeled data. Unlike per-pixel dense depth that provides abundant correspondence constraints, layout representation is sparse and…

计算机视觉与模式识别 · 计算机科学 2022-03-31 Hao-Wen Ting , Cheng Sun , Hwann-Tzong Chen

Given a single RGB panorama, the goal of 3D layout reconstruction is to estimate the room layout by predicting the corners, floor boundary, and ceiling boundary. A common approach has been to use standard convolutional networks to predict…

计算机视觉与模式识别 · 计算机科学 2021-04-20 Shivansh Rao , Vikas Kumar , Daniel Kifer , Lee Giles , Ankur Mali

We present a method to synthesize novel views from a single $360^\circ$ panorama image based on the neural radiance field (NeRF). Prior studies in a similar setting rely on the neighborhood interpolation capability of multi-layer…

计算机视觉与模式识别 · 计算机科学 2022-10-04 Shreyas Kulkarni , Peng Yin , Sebastian Scherer

In panorama understanding, the widely used equirectangular projection (ERP) entails boundary discontinuity and spatial distortion. It severely deteriorates the conventional CNNs and vision Transformers on panoramas. In this paper, we…

计算机视觉与模式识别 · 计算机科学 2023-08-29 Zhixin Ling , Zhen Xing , Xiangdong Zhou , Manliang Cao , Guichun Zhou

Generalizable neural implicit surface reconstruction aims to obtain an accurate underlying geometry given a limited number of multi-view images from unseen scenes. However, existing methods select only informative and relevant views using…

计算机视觉与模式识别 · 计算机科学 2024-05-20 Youngju Na , Woo Jae Kim , Kyu Beom Han , Suhyeon Ha , Sung-eui Yoon

Scene understanding and reasoning has been a fundamental problem in 3D computer vision, requiring models to identify objects, their properties, and spatial or comparative relationships among the objects. Existing approaches enable this by…

计算机视觉与模式识别 · 计算机科学 2026-02-03 Vivek Madhavaram , Vartika Sengar , Arkadipta De , Charu Sharma

This paper presents a new approach for synthesizing a novel street-view panorama given an overhead satellite image. Taking a small satellite image patch as input, our method generates a Google's omnidirectional street-view type panorama, as…

计算机视觉与模式识别 · 计算机科学 2022-03-29 Yujiao Shi , Dylan Campbell , Xin Yu , Hongdong Li

Viewpoint missing of objects is common in scene reconstruction, as camera paths typically prioritize capturing the overall scene structure rather than individual objects. This makes it highly challenging to achieve high-fidelity…

计算机视觉与模式识别 · 计算机科学 2025-09-23 Ziwei Chen , Ziling Liu , Zitong Huang , Mingqi Gao , Feng Zheng

Colour is one of the most perceptually salient yet least controllable attributes in image generation. Although recent diffusion models can modify object colours from user instructions, their results often deviate from the intended hue,…

计算机视觉与模式识别 · 计算机科学 2026-03-20 Yuqi Yang , Dongliang Chang , Yijia Ling , Ruoyi Du , Zhanyu Ma

Large-scale text-to-image models enable a wide range of image editing techniques, using text prompts or even spatial controls. However, applying these editing methods to multi-view images depicting a single scene leads to 3D-inconsistent…

计算机视觉与模式识别 · 计算机科学 2024-02-23 Or Patashnik , Rinon Gal , Daniel Cohen-Or , Jun-Yan Zhu , Fernando De la Torre

Understanding 3D scenes requires flexible combinations of visual reasoning tasks, including depth estimation, novel view synthesis, and object manipulation, all of which are essential for perception and interaction. Existing approaches have…

计算机视觉与模式识别 · 计算机科学 2026-05-26 Wanhee Lee , Klemen Kotar , Rahul Mysore Venkatesh , Jared Watrous , Honglin Chen , Khai Loong Aw , Daniel L. K. Yamins

Using convolutional neural networks for 360images can induce sub-optimal performance due to distortions entailed by a planar projection. The distortion gets deteriorated when a rotation is applied to the 360image. Thus, many researches…

计算机视觉与模式识别 · 计算机科学 2022-02-14 Sungmin Cho , Raehyuk Jung , Junseok Kwon

A single Panorama can be drawn perspectively without distortions in arbitrary viewing directions and field-of-views when the camera position is at the origin. This is a key advantage in VR and virtual tour applications because it enables…

图形学 · 计算机科学 2021-11-24 Chi-Han Peng , Jiayao Zhang

In this paper, we propose a novel procedure for 3D layout recovery of indoor scenes from single 360 degrees panoramic images. With such images, all scene is seen at once, allowing to recover closed geometries. Our method combines…

计算机视觉与模式识别 · 计算机科学 2018-06-22 Clara Fernandez-Labrador , Alejandro Perez-Yus , Gonzalo Lopez-Nicolas , Jose J. Guerrero

In this paper, we address panoramic semantic segmentation which is under-explored due to two critical challenges: (1) image distortions and object deformations on panoramas; (2) lack of semantic annotations in the 360{\deg} imagery. To…

计算机视觉与模式识别 · 计算机科学 2024-06-03 Jiaming Zhang , Kailun Yang , Hao Shi , Simon Reiß , Kunyu Peng , Chaoxiang Ma , Haodong Fu , Philip H. S. Torr , Kaiwei Wang , Rainer Stiefelhagen

Recent advances in image generation have led to remarkable improvements in synthesizing perspective images. However, these models still struggle with panoramic image generation due to unique challenges, including varying levels of geometric…

计算机视觉与模式识别 · 计算机科学 2025-06-30 Hakan Çapuk , Andrew Bond , Muhammed Burak Kızıl , Emir Göçen , Erkut Erdem , Aykut Erdem

Human perception for effective object tracking in 2D video streams arises from the implicit use of prior 3D knowledge and semantic reasoning. In contrast, most generic object tracking (GOT) methods primarily rely on 2D features of the…

计算机视觉与模式识别 · 计算机科学 2026-02-25 Shih-Fang Chen , Jun-Cheng Chen , I-Hong Jhuo , Yen-Yu Lin

World models have become indispensable tools for embodied intelligence, serving as powerful simulators capable of generating realistic robotic videos while addressing critical data scarcity challenges. However, current embodied world models…

计算机视觉与模式识别 · 计算机科学 2025-07-01 Yu Shang , Xin Zhang , Yinzhou Tang , Lei Jin , Chen Gao , Wei Wu , Yong Li

Face reenactment and portrait relighting are essential tasks in portrait editing, yet they are typically addressed independently, without much synergy. Most face reenactment methods prioritize motion control and multiview consistency, while…

计算机视觉与模式识别 · 计算机科学 2025-05-28 Yizhou Zhao , Chunjiang Liu , Haoyu Chen , Bhiksha Raj , Min Xu , Tadas Baltrusaitis , Mitch Rundle , HsiangTao Wu , Kamran Ghasedi

Neural fields are receiving increased attention as a geometric representation due to their ability to compactly store detailed and smooth shapes and easily undergo topological changes. Compared to classic geometry representations, however,…

计算机视觉与模式识别 · 计算机科学 2023-04-26 Arturs Berzins , Moritz Ibing , Leif Kobbelt