中文
相关论文

相关论文: Seg2Reg: Differentiable 2D Segmentation to 1D Regr…

200 篇论文

Although significant progress has been made in room layout estimation, most methods aim to reduce the loss in the 2D pixel coordinate rather than exploiting the room structure in the 3D space. Towards reconstructing the room layout in 3D,…

计算机视觉与模式识别 · 计算机科学 2021-04-06 Fu-En Wang , Yu-Hsuan Yeh , Min Sun , Wei-Chen Chiu , Yi-Hsuan Tsai

The task of room layout estimation is to locate the wall-floor, wall-ceiling, and wall-wall boundaries. Most recent methods solve this problem based on edge/keypoint detection or semantic segmentation. However, these approaches have shown…

计算机视觉与模式识别 · 计算机科学 2020-08-17 Weidong Zhang , Wei Zhang , Yinda Zhang

We present the first self-supervised method to train panoramic room layout estimation models without any labeled data. Unlike per-pixel dense depth that provides abundant correspondence constraints, layout representation is sparse and…

计算机视觉与模式识别 · 计算机科学 2022-03-31 Hao-Wen Ting , Cheng Sun , Hwann-Tzong Chen

Single-image room layout reconstruction aims to reconstruct the enclosed 3D structure of a room from a single image. Most previous work relies on the cuboid-shape prior. This paper considers a more general indoor assumption, i.e., the room…

计算机视觉与模式识别 · 计算机科学 2021-10-12 Cheng Yang , Jia Zheng , Xili Dai , Rui Tang , Yi Ma , Xiaojun Yuan

In this study, we address the challenge of 3D scene structure recovery from monocular depth estimation. While traditional depth estimation methods leverage labeled datasets to directly predict absolute depth, recent advancements advocate…

计算机视觉与模式识别 · 计算机科学 2023-09-19 Chi Zhang , Wei Yin , Gang Yu , Zhibin Wang , Tao Chen , Bin Fu , Joey Tianyi Zhou , Chunhua Shen

Reconstructing a structured vector-graphics representation from a rasterized floorplan image is typically an important prerequisite for computational tasks involving floorplans such as automated understanding or CAD workflows. However,…

计算机视觉与模式识别 · 计算机科学 2026-05-12 Hao Phung , Hadar Averbuch-Elor

Based on the Manhattan World assumption, most existing indoor layout estimation schemes focus on recovering layouts from vertically compressed 1D sequences. However, the compression procedure confuses the semantics of different planes,…

计算机视觉与模式识别 · 计算机科学 2023-03-07 Zhijie Shen , Zishuo Zheng , Chunyu Lin , Lang Nie , Kang Liao , Shuai Zheng , Yao Zhao

Robots operating in unstructured environments often require accurate and consistent object-level representations. This typically requires segmenting individual objects from the robot's surroundings. While recent large models such as Segment…

机器人学 · 计算机科学 2025-04-07 Haozhan Tang , Tianyi Zhang , Oliver Kroemer , Matthew Johnson-Roberson , Weiming Zhi

This paper proposes a new approach, Flat2Layout, for estimating general indoor room layout from a single-view RGB image whereas existing methods can only produce layout topologies captured from the box-shaped room. The proposed flat…

计算机视觉与模式识别 · 计算机科学 2019-05-30 Chi-Wei Hsiao , Cheng Sun , Min Sun , Hwann-Tzong Chen

Existing panoramic layout estimation solutions tend to recover room boundaries from a vertically compressed sequence, yielding imprecise results as the compression process often muddles the semantics between various planes. Besides, these…

计算机视觉与模式识别 · 计算机科学 2024-08-30 Zhijie Shen , Chunyu Lin , Junsong Zhang , Lang Nie , Kang Liao , Yao Zhao

We present a novel method to reconstruct the 3D layout of a room (walls, floors, ceilings) from a single perspective view in challenging conditions, by contrast with previous single-view methods restricted to cuboid-shaped layouts. This…

计算机视觉与模式识别 · 计算机科学 2020-07-22 Sinisa Stekovic , Shreyas Hampali , Mahdi Rad , Sayan Deb Sarkar , Friedrich Fraundorfer , Vincent Lepetit

Medical image segmentation models are typically optimised with voxel-wise losses that constrain predictions only in the output space. This leaves latent feature representations largely unconstrained, potentially limiting generalisation. We…

图像与视频处理 · 电气工程与系统科学 2026-03-02 Puru Vaish , Amin Ranem , Felix Meister , Tobias Heimann , Christoph Brune , Jelmer M. Wolterink

We present a new approach to the problem of estimating the 3D room layout from a single panoramic image. We represent room layout as three 1D vectors that encode, at each image column, the boundary positions of floor-wall and ceiling-wall,…

计算机视觉与模式识别 · 计算机科学 2019-04-09 Cheng Sun , Chi-Wei Hsiao , Min Sun , Hwann-Tzong Chen

Semantic segmentation is the pixel-wise labelling of an image. Since the problem is defined at the pixel level, determining image class labels only is not acceptable, but localising them at the original image pixel resolution is necessary.…

计算机视觉与模式识别 · 计算机科学 2022-03-17 Irem Ulku , Erdem Akagunduz

Recent advances in differentiable rendering, which allow calculating the gradients of 2D pixel values with respect to 3D object models, can be applied to estimation of the model parameters by gradient-based optimization with only 2D…

图形学 · 计算机科学 2022-12-07 Yiping Xie , Nils Bore , John Folkesson

Recent approaches for predicting layouts from 360 panoramas produce excellent results. These approaches build on a common framework consisting of three steps: a pre-processing step based on edge-based alignment, prediction of layout…

计算机视觉与模式识别 · 计算机科学 2020-12-29 Chuhang Zou , Jheng-Wei Su , Chi-Han Peng , Alex Colburn , Qi Shan , Peter Wonka , Hung-Kuo Chu , Derek Hoiem

In recent years, floor plan segmentation has gained significant attention due to its wide range of applications in floor plan reconstruction and robotics. In this paper, we propose a novel 2D floor plan segmentation technique based on a…

计算机视觉与模式识别 · 计算机科学 2023-03-27 Mohammadreza Sharif , Kiran Mohan , Sarath Suvarna

Visual cognition of the indoor environment can benefit from the spatial layout estimation, which is to represent an indoor scene with a 2D box on a monocular image. In this paper, we propose to fully exploit the edge and semantic…

计算机视觉与模式识别 · 计算机科学 2019-01-04 Weidong Zhang , Wei Zhang , Jason Gu

Depth estimation plays a pivotal role in advancing human-robot interactions, especially in indoor environments where accurate 3D scene reconstruction is essential for tasks like navigation and object handling. Monocular depth estimation,…

计算机视觉与模式识别 · 计算机科学 2025-02-18 Siddiqui Muhammad Yasir , Hyunsik Ahn

We present Seg-R1, a preliminary exploration of using reinforcement learning (RL) to enhance the pixel-level understanding and reasoning capabilities of large multimodal models (LMMs). Starting with foreground segmentation tasks,…

计算机视觉与模式识别 · 计算机科学 2025-07-01 Zuyao You , Zuxuan Wu
‹ 上一页 1 2 3 10 下一页 ›