English
Related papers

Related papers: PlaneTR: Structure-Guided Transformers for 3D Plan…

200 papers

Acquiring 3D geometry of real world objects has various applications in 3D digitization, such as navigation and content generation in virtual environments. Image remains one of the most popular media for such visual tasks due to its…

Computer Vision and Pattern Recognition · Computer Science 2017-01-26 Shuai Du , Youyi Zheng

In this work, we introduce the Global Planar Convolution module as a building-block for fully-convolutional networks that aggregates global information and, therefore, enhances the context perception capabilities of segmentation networks in…

Image and Video Processing · Electrical Eng. & Systems 2019-08-28 Santi Puch , Irina Sánchez , Aura Hernández , Gemma Piella , Vesna Prchkovska

We introduce a novel framework for learning vector representations of tree-structured geometric data focusing on 3D vascular networks. Our approach employs two sequentially trained Transformer-based autoencoders. In the first stage, the…

Image and Video Processing · Electrical Eng. & Systems 2025-06-16 James Batten , Michiel Schaap , Matthew Sinclair , Ying Bai , Ben Glocker

Most recent transformer-based models show impressive performance on vision tasks, even better than Convolution Neural Networks (CNN). In this work, we present a novel, flexible, and effective transformer-based model for high-quality…

Computer Vision and Pattern Recognition · Computer Science 2021-08-18 Ruohao Guo , Dantong Niu , Liao Qu , Zhenbo Li

We introduce a novel neural network architecture for encoding and synthesis of 3D shapes, particularly their structures. Our key insight is that 3D shapes are effectively characterized by their hierarchical organization of parts, which…

Graphics · Computer Science 2017-05-16 Jun Li , Kai Xu , Siddhartha Chaudhuri , Ersin Yumer , Hao Zhang , Leonidas Guibas

This paper presents Planar Gaussian Splatting (PGS), a novel neural rendering approach to learn the 3D geometry and parse the 3D planes of a scene, directly from multiple RGB images. The PGS leverages Gaussian primitives to model the scene…

Computer Vision and Pattern Recognition · Computer Science 2024-12-04 Farhad G. Zanjani , Hong Cai , Hanno Ackermann , Leila Mirvakhabova , Fatih Porikli

In geometry processing, symmetry is a universal type of high-level structural information of 3D models and benefits many geometry processing tasks including shape segmentation, alignment, matching, and completion. Thus it is an important…

Graphics · Computer Science 2021-09-15 Lin Gao , Ling-Xiao Zhang , Hsien-Yu Meng , Yi-Hui Ren , Yu-Kun Lai , Leif Kobbelt

The task of reconstructing particles from low-level detector response data to predict the set of final state particles in collision events represents a set-to-set prediction task requiring the use of multiple features and their correlations…

In this paper, a novel neural network architecture is proposed attempting to rectify text images with mild assumptions. A new dataset of text images is collected to verify our model and open to public. We explored the capability of deep…

Computer Vision and Pattern Recognition · Computer Science 2016-11-15 Chengzhe Yan , Jie Hu , Changshui Zhang

We study the inverse graphics problem of inferring a holistic representation for natural images. Given an input image, our goal is to induce a neuro-symbolic, program-like representation that jointly models camera poses, object locations,…

Computer Vision and Pattern Recognition · Computer Science 2020-06-29 Yikai Li , Jiayuan Mao , Xiuming Zhang , William T. Freeman , Joshua B. Tenenbaum , Jiajun Wu

Reconstruction based on the stereo camera has received considerable attention recently, but two particular challenges still remain. The first concerns the need to aggregate similar pixels in an effective approach, and the second is to…

Computer Vision and Pattern Recognition · Computer Science 2017-03-31 Lei Fan , Ziyu Pan , Long Chen , Kai Huang

3D reconstruction aims to reconstruct 3D objects from 2D views. Previous works for 3D reconstruction mainly focus on feature matching between views or using CNNs as backbones. Recently, Transformers have been shown effective in multiple…

Computer Vision and Pattern Recognition · Computer Science 2021-11-17 Zai Shi , Zhao Meng , Yiran Xing , Yunpu Ma , Roger Wattenhofer

High-definition (HD) map provides abundant and precise environmental information of the driving scene, serving as a fundamental and indispensable component for planning in autonomous driving system. We present MapTR, a structured end-to-end…

Computer Vision and Pattern Recognition · Computer Science 2023-01-31 Bencheng Liao , Shaoyu Chen , Xinggang Wang , Tianheng Cheng , Qian Zhang , Wenyu Liu , Chang Huang

3D plane reconstruction from a single image is a crucial yet challenging topic in 3D computer vision. Previous state-of-the-art (SOTA) methods have focused on training their system on a single dataset from either indoor or outdoor domain,…

Computer Vision and Pattern Recognition · Computer Science 2025-06-04 Jiachen Liu , Rui Yu , Sili Chen , Sharon X. Huang , Hengkai Guo

In this paper we address the problem of representing 3D visual data with parameterized volumetric shape primitives. Specifically, we present a (two-stage) approach built around convolutional neural networks (CNNs) capable of segmenting…

Computer Vision and Pattern Recognition · Computer Science 2020-01-29 Jaka Šircelj , Tim Oblak , Klemen Grm , Uroš Petković , Aleš Jaklič , Peter Peer , Vitomir Štruc , Franc Solina

We present extraction of tree structures, such as airways, from image data as a graph refinement task. To this end, we propose a graph auto-encoder model that uses an encoder based on graph neural networks (GNNs) to learn embeddings from…

Computer Vision and Pattern Recognition · Computer Science 2018-04-13 Raghavendra Selvan , Thomas Kipf , Max Welling , Jesper H. Pedersen , Jens Petersen , Marleen de Bruijne

Convolutional neural networks (CNNs) have been the de facto standard for nowadays 3D medical image segmentation. The convolutional operations used in these networks, however, inevitably have limitations in modeling the long-range dependency…

Computer Vision and Pattern Recognition · Computer Science 2021-03-05 Yutong Xie , Jianpeng Zhang , Chunhua Shen , Yong Xia

We present Neural Kernel Fields: a novel method for reconstructing implicit 3D shapes based on a learned kernel ridge regression. Our technique achieves state-of-the-art results when reconstructing 3D objects and large scenes from sparse…

Computer Vision and Pattern Recognition · Computer Science 2021-11-29 Francis Williams , Zan Gojcic , Sameh Khamis , Denis Zorin , Joan Bruna , Sanja Fidler , Or Litany

Learning-based 3D reconstruction using implicit neural representations has shown promising progress not only at the object level but also in more complicated scenes. In this paper, we propose Dynamic Plane Convolutional Occupancy Networks,…

Computer Vision and Pattern Recognition · Computer Science 2020-11-12 Stefan Lionar , Daniil Emtsev , Dusan Svilarkovic , Songyou Peng

Single-image room layout reconstruction aims to reconstruct the enclosed 3D structure of a room from a single image. Most previous work relies on the cuboid-shape prior. This paper considers a more general indoor assumption, i.e., the room…

Computer Vision and Pattern Recognition · Computer Science 2021-10-12 Cheng Yang , Jia Zheng , Xili Dai , Rui Tang , Yi Ma , Xiaojun Yuan