English
Related papers

Related papers: Towards Localizing Structural Elements: Merging Ge…

200 papers

This paper addresses the problem of building augmented metric representations of scenes with semantic information from RGB-D images. We propose a complete framework to create an enhanced map representation of the environment with…

Computer Vision and Pattern Recognition · Computer Science 2020-03-16 Renato Martins , Dhiego Bersan , Mario F. M. Campos , Erickson R. Nascimento

Visual simultaneous localization and mapping (SLAM) plays a critical role in autonomous robotic systems, especially where accurate and reliable measurements are essential for navigation and sensing. In feature-based SLAM, the quantityand…

Robotics · Computer Science 2025-09-03 Haolan Zhang , Chenghao Li , Thanh Nguyen Canh , Lijun Wang , Nak Young Chong

There has been a growing adoption of computer vision tools and technologies in architectural design workflows over the past decade. Notable use cases include point cloud generation, visual content analysis, and spatial awareness for robotic…

Computer Vision and Pattern Recognition · Computer Science 2023-12-21 Demircan Tas , Rohit Priyadarshi Sanatani

Large language models (LLMs) have recently been used as structured decoders for indoor understanding from 3D point-token inputs. However, point cloud encoders often under-represent thin structural elements such as doors and windows after…

Computer Vision and Pattern Recognition · Computer Science 2026-05-19 Shuliang Zhu , Tomiwa Adey , Jinjia Zhou

Simultaneous localization and mapping (SLAM) in highly dynamic environments is challenging due to the correlation complexity between moving objects and the camera pose. Many methods have been proposed to deal with this problem; however, the…

Robotics · Computer Science 2024-10-17 Tuan Dang , Khang Nguyen , Mandfred Huber

Real-time 3D object detection from point clouds is essential for dynamic scene understanding in applications such as augmented reality, robotics and navigation. We introduce a novel Spatial-prioritized and Rank-aware 3D object detection…

Computer Vision and Pattern Recognition · Computer Science 2025-11-21 Chenyu Zhao , Xianwei Zheng , Zimin Xia , Linwei Yue , Nan Xue

The emergence of modern RGB-D sensors had a significant impact in many application fields, including robotics, augmented reality (AR) and 3D scanning. They are low-cost, low-power and low-size alternatives to traditional range sensors such…

Computer Vision and Pattern Recognition · Computer Science 2020-01-22 Javier Civera , Seong Hun Lee

LiDAR sensors are a powerful tool for robot simultaneous localization and mapping (SLAM) in unknown environments, but the raw point clouds they produce are dense, computationally expensive to store, and unsuited for direct use by downstream…

Robotics · Computer Science 2022-10-03 Adam Dai , Greg Lund , Grace Gao

The rapid growth of vacation rental (VR) platforms has led to an increasing volume of property images, often uploaded without structured categorization. This lack of organization poses significant challenges for travelers attempting to…

Computer Vision and Pattern Recognition · Computer Science 2025-07-02 Vignesh Ram Nithin Kappagantula , Shayan Hassantabar

This work presents a novel dense RGB-D SLAM approach for dynamic planar environments that enables simultaneous multi-object tracking, camera localisation and background reconstruction. Previous dynamic SLAM methods either rely on semantic…

Robotics · Computer Science 2022-10-19 Ran Long , Christian Rauch , Tianwei Zhang , Vladimir Ivan , Tin Lun Lam , Sethu Vijayakumar

In this paper, we present a monocular Simultaneous Localization and Mapping (SLAM) algorithm using high-level object and plane landmarks. The built map is denser, more compact and semantic meaningful compared to feature point based SLAM. We…

Robotics · Computer Science 2019-07-01 Shichao Yang , Sebastian Scherer

Recently there has been a growing interest in category-level object pose and size estimation, and prevailing methods commonly rely on single view RGB-D images. However, one disadvantage of such methods is that they require accurate depth…

Computer Vision and Pattern Recognition · Computer Science 2024-03-25 Jiaqi Yang , Yucong Chen , Xiangting Meng , Chenxin Yan , Min Li , Ran Cheng , Lige Liu , Tao Sun , Laurent Kneip

Real-life man-made objects often exhibit strong and easily-identifiable structure, as a direct result of their design or their intended functionality. Structure typically appears in the form of individual parts and their arrangement.…

Computer Vision and Pattern Recognition · Computer Science 2018-09-06 Vignesh Ganapathi-Subramanian , Olga Diamanti , Soeren Pirk , Chengcheng Tang , Matthias Niessner , Leonidas J. Guibas

Research works on the two topics of Semantic Segmentation and SLAM (Simultaneous Localization and Mapping) have been following separate tracks. Here, we link them quite tightly by delineating a category label fusion technique that allows…

Computer Vision and Pattern Recognition · Computer Science 2015-11-16 Tommaso Cavallari , Luigi Di Stefano

We introduce SceneNet RGB-D, expanding the previous work of SceneNet to enable large scale photorealistic rendering of indoor scene trajectories. It provides pixel-perfect ground truth for scene understanding problems such as semantic…

Computer Vision and Pattern Recognition · Computer Science 2017-01-31 John McCormac , Ankur Handa , Stefan Leutenegger , Andrew J. Davison

Accurate localization and 3D maps are increasingly needed for various artificial intelligence based IoT applications such as augmented reality, intelligent transportation, crowd monitoring, robotics, etc. This article proposes a novel…

Robotics · Computer Science 2021-03-23 Max Jwo Lem Lee , Li-Ta Hsu

Exploring an unfamiliar indoor environment and avoiding obstacles is challenging for visually impaired people. Currently, several approaches achieve the avoidance of static obstacles based on the mapping of indoor scenes. To solve the issue…

Computer Vision and Pattern Recognition · Computer Science 2022-04-05 Wenyan Ou , Jiaming Zhang , Kunyu Peng , Kailun Yang , Gerhard Jaworek , Karin Müller , Rainer Stiefelhagen

Current open-vocabulary scene graph generation algorithms highly rely on both 3D scene point cloud data and posed RGB-D images and thus have limited applications in scenarios where RGB-D images or camera poses are not readily available. To…

Robotics · Computer Science 2024-09-17 Yifan Xu , Ziming Luo , Qianwei Wang , Vineet Kamat , Carol Menassa

Robotic tasks such as planning and navigation require a hierarchical semantic understanding of a scene, which could include multiple floors and rooms. Current methods primarily focus on object segmentation for 3D scene understanding.…

Computer Vision and Pattern Recognition · Computer Science 2024-12-13 Yash Mehan , Kumaraditya Gupta , Rohit Jayanti , Anirudh Govil , Sourav Garg , Madhava Krishna

One of the major challenges of a real-time autonomous robotic system for construction monitoring is to simultaneously localize, map, and navigate over the lifetime of the robot, with little or no human intervention. Past research on…

‹ Prev 1 4 5 6 7 8 10 Next ›