中文
相关论文

相关论文: Perspective from a Broader Context: Can Room Style…

200 篇论文

Since a building's floorplans are easily accessible, consistent over time, and inherently robust to changes in visual appearance, self-localization within the floorplan has attracted researchers' interest. However, since floorplans are…

计算机视觉与模式识别 · 计算机科学 2025-07-28 Bolei Chen , Jiaxu Kang , Haonan Yang , Ping Zhong , Jianxin Wang

Visual Floorplan Localization (FLoc) struggles with severe structural aliasing caused by repetitive minimalist layouts. This occurs because physically distant poses share highly similar visual-geometric features, which degrades spatial…

机器人学 · 计算机科学 2026-05-11 Ping Zhong , Shiyong Meng , Bolei Chen , Tao Zou , Chaoxu Mu , Jianxin Wang

We propose UnLoc, an efficient data-driven solution for sequential camera localization within floorplans. Floorplan data is readily available, long-term persistent, and robust to changes in visual appearance. We address key limitations of…

计算机视觉与模式识别 · 计算机科学 2026-03-17 Matthias Wüest , Francis Engelmann , Ondrej Miksik , Marc Pollefeys , Daniel Barath

In this paper we propose an efficient data-driven solution to self-localization within a floorplan. Floorplan data is readily available, long-term persistent and inherently robust to changes in the visual appearance. Our method does not…

计算机视觉与模式识别 · 计算机科学 2025-05-15 Changan Chen , Rui Wang , Christoph Vogel , Marc Pollefeys

This paper presents a new approach to recognize elements in floor plan layouts. Besides walls and rooms, we aim to recognize diverse floor plan elements, such as doors, windows and different types of rooms, in the floor layouts. To this…

计算机视觉与模式识别 · 计算机科学 2019-08-30 Zhiliang Zeng , Xianzhi Li , Ying Kin Yu , Chi-Wing Fu

Visual place recognition is essential for vision-based robot localization and SLAM. Despite the tremendous progress made in recent years, place recognition in changing environments remains challenging. A promising approach to cope with…

机器人学 · 计算机科学 2023-04-17 Reihaneh Mirjalili , Michael Krawez , Wolfram Burgard

Many public buildings provide floorplans with a "you are here" indicator to help visitors orient themselves. Floorplan localization seeks to computationally replicate this capability by determining where visual observations were captured…

计算机视觉与模式识别 · 计算机科学 2026-05-22 Junhyeong Cho , Ruojin Cai , Hadar Averbuch-Elor

Recognition of floor plans has been a challenging and popular task. Despite that many recent approaches have been proposed for this task, they typically fail to make the room-level unified prediction. Specifically, multiple semantic…

计算机视觉与模式识别 · 计算机科学 2022-11-01 Zhangyu Wang , Ningyuan Sun

Floorplans provide a compact representation of the building's structure, revealing not only layout information but also detailed semantics such as the locations of windows and doors. However, contemporary floorplan localization techniques…

计算机视觉与模式识别 · 计算机科学 2025-07-16 Yuval Grader , Hadar Averbuch-Elor

Place Recognition enables the estimation of a globally consistent map and trajectory by providing non-local constraints in Simultaneous Localisation and Mapping (SLAM). This paper presents Locus, a novel place recognition method using 3D…

机器人学 · 计算机科学 2022-09-27 Kavisha Vidanapathirana , Peyman Moghadam , Ben Harwood , Muming Zhao , Sridha Sridharan , Clinton Fookes

We present LaLaLoc to localise in environments without the need for prior visitation, and in a manner that is robust to large changes in scene appearance, such as a full rearrangement of furniture. Specifically, LaLaLoc performs…

计算机视觉与模式识别 · 计算机科学 2021-10-13 Henry Howard-Jenkins , Jose-Raul Ruiz-Sarmiento , Victor Adrian Prisacariu

Many recent machine learning approaches to floorplanning represent placement decisions using discrete canvas coordinates, which creates scalability bottlenecks as the action space grows. In this work, we study the effect of learning a…

机器学习 · 计算机科学 2026-03-25 Fin Amin , Nirjhor Rouf , Tse-Han Pan , Sounak Dutta , Md Kamal Ibn Shafi , Paul D. Franzon

Contrastive Language-Image Pre-training (CLIP) has been a celebrated method for training vision encoders to generate image/text representations facilitating various applications. Recently, CLIP has been widely adopted as the vision backbone…

计算机视觉与模式识别 · 计算机科学 2025-02-20 Hong-You Chen , Zhengfeng Lai , Haotian Zhang , Xinze Wang , Marcin Eichner , Keen You , Meng Cao , Bowen Zhang , Yinfei Yang , Zhe Gan

Our work addresses the problem of learning to localize objects in an open-world setting, i.e., given the bounding box information of a limited number of object classes during training, the goal is to localize all objects, belonging to both…

计算机视觉与模式识别 · 计算机科学 2025-04-25 Ashish Singh , Michael J. Jones , Kuan-Chuan Peng , Anoop Cherian , Moitreya Chatterjee , Erik Learned-Miller

Vision-Language Models (VLMs) have shown significant progress in open-set challenges. However, the limited availability of 3D datasets hinders their effective application in 3D scene understanding. We propose LOC, a general language-guided…

计算机视觉与模式识别 · 计算机科学 2025-10-28 Yuhang Gao , Xiang Xiang , Sheng Zhong , Guoyou Wang

Recognising in what type of environment one is located is an important perception task. For instance, for a robot operating in indoors it is helpful to be aware whether it is in a kitchen, a hallway or a bedroom. Existing approaches attempt…

计算机视觉与模式识别 · 计算机科学 2020-07-06 Shengyu Huang , Mikhail Usvyatsov , Konrad Schindler

Scene graphs enhance 3D mapping capabilities in robotics by understanding the relationships between different spatial elements, such as rooms and objects. Recent research extends scene graphs to hierarchical layers, adding and leveraging…

机器人学 · 计算机科学 2025-10-20 Jeewon Kim , Minho Oh , Hyun Myung

Recent studies in long video understanding have harnessed the advanced visual-language reasoning capabilities of Large Multimodal Models (LMMs), driving the evolution of video-LMMs specialized for processing extended video sequences.…

计算机视觉与模式识别 · 计算机科学 2026-03-06 Janghoon Cho , Jungsoo Lee , Munawar Hayat , Kyuwoong Hwang , Fatih Porikli , Sungha Choi

Placing is a necessary skill for a personal robot to have in order to perform tasks such as arranging objects in a disorganized room. The object placements should not only be stable but also be in their semantically preferred placing areas…

机器人学 · 计算机科学 2012-02-09 Yun Jiang , Marcus Lim , Changxi Zheng , Ashutosh Saxena

Image captioning has been shown as an effective pretraining method similar to contrastive pretraining. However, the incorporation of location-aware information into visual pretraining remains an area with limited research. In this paper, we…

计算机视觉与模式识别 · 计算机科学 2024-11-13 Bo Wan , Michael Tschannen , Yongqin Xian , Filip Pavetic , Ibrahim Alabdulmohsin , Xiao Wang , André Susano Pinto , Andreas Steiner , Lucas Beyer , Xiaohua Zhai
‹ 上一页 1 2 3 10 下一页 ›