English
Related papers

Related papers: Transferable End-to-end Room Layout Estimation via…

200 papers

This paper focuses on training implicit models of infinite layers. Specifically, previous works employ implicit differentiation and solve the exact gradient for the backward propagation. However, is it necessary to compute such an exact but…

Machine Learning · Computer Science 2022-01-13 Zhengyang Geng , Xin-Yu Zhang , Shaojie Bai , Yisen Wang , Zhouchen Lin

This paper describes an approach to automatically extracting floor plans from the kinds of incomplete measurements that could be acquired by an autonomous mobile robot. The approach proceeds by reasoning about extended structural layout…

Computer Vision and Pattern Recognition · Computer Science 2018-11-20 Armon Shariati , Bernd Pfrommer , Camillo J. Taylor

Unsupervised image-to-image translation methods aim to map images from one domain into plausible examples from another domain while preserving structures shared across two domains. In the many-to-many setting, an additional guidance example…

Computer Vision and Pattern Recognition · Computer Science 2021-11-29 Ben Usman , Dina Bashkirova , Kate Saenko

Coarse room layout estimation provides important geometric cues for many downstream tasks. Current state-of-the-art methods are predominantly based on single views and often assume panoramic images. We introduce PixCuboid, an…

Computer Vision and Pattern Recognition · Computer Science 2025-08-07 Gustav Hanning , Kalle Åström , Viktor Larsson

In this work, we focus on outdoor lighting estimation by aggregating individual noisy estimates from images, exploiting the rich image information from wide-angle cameras and/or temporal image sequences. Photographs inherently encode…

Computer Vision and Pattern Recognition · Computer Science 2022-02-21 Haebom Lee , Christian Homeyer , Robert Herzog , Jan Rexilius , Carsten Rother

We present uLayout, a unified model for estimating room layout geometries from both perspective and panoramic images, whereas traditional solutions require different model designs for each image type. The key idea of our solution is to…

Computer Vision and Pattern Recognition · Computer Science 2025-03-28 Jonathan Lee , Bolivar Solarte , Chin-Hsuan Wu , Jin-Cheng Jhang , Fu-En Wang , Yi-Hsuan Tsai , Min Sun

This paper presents a novel training method for end-to-end scene text recognition. End-to-end scene text recognition offers high recognition accuracy, especially when using the encoder-decoder model based on Transformer. To train a highly…

Computer Vision and Pattern Recognition · Computer Science 2021-11-25 Shota Orihashi , Yoshihiro Yamazaki , Naoki Makishima , Mana Ihori , Akihiko Takashima , Tomohiro Tanaka , Ryo Masumura

We introduce a novel strategy for learning to extract semantically meaningful features from aerial imagery. Instead of manually labeling the aerial imagery, we propose to predict (noisy) semantic features automatically extracted from…

Computer Vision and Pattern Recognition · Computer Science 2016-12-09 Menghua Zhai , Zachary Bessinger , Scott Workman , Nathan Jacobs

Understating and controlling generative models' latent space is a complex task. In this paper, we propose a novel method for learning to control any desired attribute in a pre-trained GAN's latent space, for the purpose of editing…

Computer Vision and Pattern Recognition · Computer Science 2021-11-18 Nir Diamant , Nitsan Sandor , Alex M Bronstein

In traditional topology optimization, the computing time required to iteratively update the material distribution within a design domain strongly depends on the complexity or size of the problem, limiting its application in real engineering…

Computational Engineering, Finance, and Science · Computer Science 2024-05-14 Gabriel Garayalde , Matteo Torzoni , Matteo Bruggi , Alberto Corigliano

Compressive imaging is an emerging application of compressed sensing, devoted to acquisition, encoding and reconstruction of images using random projections as measurements. In this paper we propose a novel method to provide a scalable…

Information Theory · Computer Science 2013-10-07 Diego Valsesia , Enrico Magli

Current techniques in Visual Simultaneous Localization and Mapping (VSLAM) estimate camera displacement by comparing image features of consecutive scenes. These algorithms depend on scene continuity, hence requires frequent camera inputs.…

Robotics · Computer Science 2024-01-25 Mingyang Li , Yue Ma , Qinru Qiu

This paper addresses the problem of end-to-end (E2E) design of learning and communication in a task-oriented semantic communication system. In particular, we consider a multi-device cooperative edge inference system over a wireless…

Information Theory · Computer Science 2024-09-02 Chang Cai , Xiaojun Yuan , Ying-Jun Angela Zhang

Semantic communication is considered the future of mobile communication, which aims to transmit data beyond Shannon's theorem of communications by transmitting the semantic meaning of the data rather than the bit-by-bit reconstruction of…

Computer Vision and Pattern Recognition · Computer Science 2023-04-11 Maheshi Lokumarambage , Vishnu Gowrisetty , Hossein Rezaei , Thushan Sivalingam , Nandana Rajatheva , Anil Fernando

We present the first self-supervised method to train panoramic room layout estimation models without any labeled data. Unlike per-pixel dense depth that provides abundant correspondence constraints, layout representation is sparse and…

Computer Vision and Pattern Recognition · Computer Science 2022-03-31 Hao-Wen Ting , Cheng Sun , Hwann-Tzong Chen

As research on image inversion advances, the process is generally divided into two stages. The first step is Image Embedding, involves using an encoder or optimization procedure to embed an image and obtain its corresponding latent code.…

Computer Vision and Pattern Recognition · Computer Science 2025-11-21 Xuekun Zhao , Pu Cao , Xiaoya Yang , Mingjian Zhang , Lu Yang , Qing Song

Based on the Manhattan World assumption, most existing indoor layout estimation schemes focus on recovering layouts from vertically compressed 1D sequences. However, the compression procedure confuses the semantics of different planes,…

Computer Vision and Pattern Recognition · Computer Science 2023-03-07 Zhijie Shen , Zishuo Zheng , Chunyu Lin , Lang Nie , Kang Liao , Shuai Zheng , Yao Zhao

We propose an end-to-end deep convolutional network to simultaneously localize and rank relative visual attributes, given only weakly-supervised pairwise image comparisons. Unlike previous methods, our network jointly learns the attribute's…

Computer Vision and Pattern Recognition · Computer Science 2016-08-10 Krishna Kumar Singh , Yong Jae Lee

This paper presents a method of estimating the geometry of a room and the 3D pose of objects from a single 360-degree panorama image. Assuming Manhattan World geometry, we formulate the task as a Bayesian inference problem in which we…

Computer Vision and Pattern Recognition · Computer Science 2016-10-03 Jiu Xu , Bjorn Stenger , Tommi Kerola , Tony Tung

Object goal navigation aims to navigate an agent to locations of a given object category in unseen environments. Classical methods explicitly build maps of environments and require extensive engineering while lacking semantic information…

Computer Vision and Pattern Recognition · Computer Science 2023-08-11 Shizhe Chen , Thomas Chabal , Ivan Laptev , Cordelia Schmid
‹ Prev 1 3 4 5 6 7 10 Next ›