中文
相关论文

相关论文: LiLoc: Lifelong Localization using Adaptive Submap…

200 篇论文

In this paper we propose an efficient data-driven solution to self-localization within a floorplan. Floorplan data is readily available, long-term persistent and inherently robust to changes in the visual appearance. Our method does not…

计算机视觉与模式识别 · 计算机科学 2025-05-15 Changan Chen , Rui Wang , Christoph Vogel , Marc Pollefeys

Localization is a critical technology for various applications ranging from navigation and surveillance to assisted living. Localization systems typically fuse information from sensors viewing the scene from different perspectives to…

计算机视觉与模式识别 · 计算机科学 2024-06-12 Jason Wu , Ziqi Wang , Xiaomin Ouyang , Ho Lyun Jeong , Colin Samplawski , Lance Kaplan , Benjamin Marlin , Mani Srivastava

Fingerprinting-based localization often suffers from poor cross-environment generalization, especially when only a few labeled samples are available in the target environment. Existing methods mitigate distribution shifts through domain…

信号处理 · 电气工程与系统科学 2026-05-20 Jun Gao , Zheng Xing , Wenliang Lin , Weibing Zhao , Xuhui Zhang , Junting Chen , Zhongliang Deng , Shuguang Cui

We propose an accurate and robust multi-modal sensor fusion framework, MetroLoc, towards one of the most extreme scenarios, the large-scale metro vehicle localization and mapping. MetroLoc is built atop an IMU-centric state estimator that…

机器人学 · 计算机科学 2021-11-02 Yusheng Wang , Weiwei Song , Yi Zhang , Fei Huang , Zhiyong Tu , Yidong Lou

While large-scale image-text pretrained models such as CLIP have been used for multiple video-level tasks on trimmed videos, their use for temporal localization in untrimmed videos is still a relatively unexplored task. We design a new…

计算机视觉与模式识别 · 计算机科学 2023-08-23 Shen Yan , Xuehan Xiong , Arsha Nagrani , Anurag Arnab , Zhonghao Wang , Weina Ge , David Ross , Cordelia Schmid

Lifelong mapping is crucial for the long-term deployment of robots in dynamic environments. In this paper, we present ELite, an ephemerality-aided LiDAR-based lifelong mapping framework which can seamlessly align multiple session data,…

机器人学 · 计算机科学 2025-03-04 Hyeonjae Gil , Dongjae Lee , Giseop Kim , Ayoung Kim

Methods that use Large Language Models (LLM) as planners for embodied instruction following tasks have become widespread. To successfully complete tasks, the LLM must be grounded in the environment in which the robot operates. One solution…

机器人学 · 计算机科学 2025-12-25 Anatoly O. Onishchenko , Alexey K. Kovalev , Aleksandr I. Panov

Long-term scene changes present challenges to localization systems using a pre-built map. This paper presents a LiDAR-based system that can provide robust localization against those challenges. Our method starts with activation of a mapping…

机器人学 · 计算机科学 2022-03-09 Bin Peng , Hongle Xie , Weidong Chen

Relocalization is a fundamental task in the field of robotics and computer vision. There is considerable work in the field of deep camera relocalization, which directly estimates poses from raw images. However, learning-based methods have…

机器人学 · 计算机科学 2021-03-23 Wei Wang , Pedro P. B. de Gusmo , Bo Yang , Andrew Markham , Niki Trigoni

In this paper, we propose a unified localization framework (called UNILocPro) that integrates model-based localization and channel charting (CC) for mixed line-of-sight (LoS)/non-line-of-sight (NLoS) scenarios. Specifically, based on…

信号处理 · 电气工程与系统科学 2025-11-03 Yuhao Zhang , Guangjin Pan , Musa Furkan Keskin , Ossi Kaltiokallio , Mikko Valkama , Henk Wymeersch

Vision-language models such as CLIP have shown impressive capabilities in aligning images and text, but they often struggle with lengthy and detailed text descriptions due to pre-training on short and concise captions. We present FAST-GOAL…

人工智能 · 计算机科学 2026-05-27 Hyungyu Choi , Young Kyun Jang , Chanho Eom

This article introduces a novel method for object-level relocalization of robotic systems. It determines the pose of a camera sensor by robustly associating the object detections in the current frame with 3D objects in a lightweight…

机器人学 · 计算机科学 2024-08-16 Yutong Wang , Chaoyang Jiang , Xieyuanli Chen

Multi-session map merging is crucial for extended autonomous operations in large-scale environments. In this paper, we present GMLD, a learning-based local descriptor framework for large-scale multi-session point cloud map merging that…

机器人学 · 计算机科学 2026-01-01 Yanlong Ma , Nakul S. Joshi , Christa S. Robison , Philip R. Osteen , Brett T. Lopez

Wireless fingerprint-based localization has become one of the most promising technologies for ubiquitous location-aware computing and intelligent location-based services. However, due to RF vulnerability to environmental dynamics over time,…

信号处理 · 电气工程与系统科学 2024-02-20 Lingyan Zhang , Junlin Huang , Tingting Zhang , Qinyu Zhang

The environment of most real-world scenarios such as malls and supermarkets changes at all times. A pre-built map that does not account for these changes becomes out-of-date easily. Therefore, it is necessary to have an up-to-date model of…

机器人学 · 计算机科学 2021-11-23 Min Zhao , Xin Guo , Le Song , Baoxing Qin , Xuesong Shi , Gim Hee Lee , Guanghui Sun

Visual localization is a fundamental task that regresses the 6 Degree Of Freedom (6DoF) poses with image features in order to serve the high precision localization requests in many robotics applications. Degenerate conditions like motion…

计算机视觉与模式识别 · 计算机科学 2022-10-19 Yuchen Yang , Xudong Zhang , Shuang Gao , Jixiang Wan , Yishan Ping , Yuyue Liu , Jijunnan Li , Yandong Guo

State-of-the-art hierarchical localisation pipelines (HLoc) employ image retrieval (IR) to establish 2D-3D correspondences by selecting the top-$k$ most similar images from a reference database. While increasing $k$ improves localisation…

计算机视觉与模式识别 · 计算机科学 2025-03-05 Changkun Liu , Jianhao Jiao , Huajian Huang , Zhengyang Ma , Dimitrios Kanoulas , Tristan Braud

Multimodal document retrieval aims to retrieve query-relevant components from documents composed of textual, tabular, and visual elements. An effective multimodal retriever needs to handle two main challenges: (1) mitigate the effect of…

信息检索 · 计算机科学 2026-02-05 Joohyung Yun , Doyup Lee , Wook-Shin Han

Traditional approaches to adapting multi-modal large language models (MLLMs) to new tasks have relied heavily on fine-tuning. This paper introduces Efficient Multi-Modal Long Context Learning (EMLoC), a novel training-free alternative that…

计算机视觉与模式识别 · 计算机科学 2025-05-27 Zehong Ma , Shiliang Zhang , Longhui Wei , Qi Tian

Recent studies in long video understanding have harnessed the advanced visual-language reasoning capabilities of Large Multimodal Models (LMMs), driving the evolution of video-LMMs specialized for processing extended video sequences.…

计算机视觉与模式识别 · 计算机科学 2026-03-06 Janghoon Cho , Jungsoo Lee , Munawar Hayat , Kyuwoong Hwang , Fatih Porikli , Sungha Choi