中文
相关论文

相关论文: GMMLoc: Structure Consistent Visual Localization w…

200 篇论文

Structure from Motion (SfM) and visual localization in indoor texture-less scenes and industrial scenarios present prevalent yet challenging research topics. Existing SfM methods designed for natural scenes typically yield low accuracy or…

机器人学 · 计算机科学 2024-05-28 Yusen Xie , Zhenmin Huang , Kai Chen , Lei Zhu , Jun Ma

We propose an image representation and matching approach that substantially improves visual-based location estimation for images. The main novelty of the approach, called distinctive visual element matching (DVEM), is its use of…

多媒体 · 计算机科学 2016-01-29 Xinchao Li , Martha A. Larson , Alan Hanjalic

Monocular 3D Gaussian Splatting SLAM suffers from critical limitations in time efficiency, geometric accuracy, and multi-view consistency. These issues stem from the time-consuming $\textit{Train-from-Scratch}$ optimization and the lack of…

机器人学 · 计算机科学 2026-04-06 Zicheng Zhang , Ke Wu , Xiangting Meng , Keyu Liu , Jieru Zhao , Wenchao Ding

Visual localization is crucial for Computer Vision and Augmented Reality (AR) applications, where determining the camera or device's position and orientation is essential to accurately interact with the physical environment. Traditional…

机器人学 · 计算机科学 2025-01-22 Albert Gassol Puigjaner , Irvin Aloise , Patrik Schmuck

Structure-from-Motion -- the process of simultaneously estimating camera poses and 3D scene structure from a collection of images -- remains a central challenge in computer vision, with many open problems yet to be solved. Recent advances…

计算机视觉与模式识别 · 计算机科学 2026-05-27 Linfei Pan , Johannes Schönberger , Marc Pollefeys

Image compression is a fundamental research field and many well-known compression standards have been developed for many decades. Recently, learned compression methods exhibit a fast development trend with promising results. However, there…

图像与视频处理 · 电气工程与系统科学 2020-03-31 Zhengxue Cheng , Heming Sun , Masaru Takeuchi , Jiro Katto

Visual localization plays an important role for intelligent robots and autonomous driving, especially when the accuracy of GNSS is unreliable. Recently, camera localization in LiDAR maps has attracted more and more attention for its low…

计算机视觉与模式识别 · 计算机科学 2024-10-28 Zhipeng Zhao , Huai Yu , Chenwei Lyv , Wen Yang , Sebastian Scherer

This paper presents Vision-Language Global Localization (VLG-Loc), a novel global localization method that uses human-readable labeled footprint maps containing only names and areas of distinctive visual landmarks in an environment. While…

机器人学 · 计算机科学 2025-12-19 Mizuho Aoki , Kohei Honda , Yasuhiro Yoshimura , Takeshi Ishita , Ryo Yonetani

The ability for a moving agent to localize itself in environment is the basic demand for emerging applications, such as autonomous driving, etc. Many existing methods based on multiple sensors still suffer from drift. We propose a scheme…

计算机视觉与模式识别 · 计算机科学 2022-09-09 Longrui Dong , Gang Zeng

Distribution learning focuses on learning the probability density function from a set of data samples. In contrast, clustering aims to group similar objects together in an unsupervised manner. Usually, these two tasks are considered…

机器学习 · 计算机科学 2023-08-31 Guanfang Dong , Chenqiu Zhao , Anup Basu

Vision-language models (VLMs) like CLIP have showcased a remarkable ability to extract transferable features for downstream tasks. Nonetheless, the training process of these models is usually based on a coarse-grained contrastive loss…

Recent years have seen a surge of interest in anomaly detection for tackling industrial defect detection, event detection, etc. However, existing unsupervised anomaly detectors, particularly those for the vision modality, face significant…

计算机视觉与模式识别 · 计算机科学 2023-10-05 Dong Chen , Kaihang Pan , Guoming Wang , Yueting Zhuang , Siliang Tang

Precise estimation of global orientation and location is critical to ensure a compelling outdoor Augmented Reality (AR) experience. We address the problem of geo-pose estimation by cross-view matching of query ground images to a…

计算机视觉与模式识别 · 计算机科学 2023-03-30 Niluthpol Chowdhury Mithun , Kshitij Minhas , Han-Pang Chiu , Taragay Oskiper , Mikhail Sizintsev , Supun Samarasekera , Rakesh Kumar

In this paper, we address the problem of landmark-based visual place recognition. In the state-of-the-art method, accurate object proposal algorithms are first leveraged for generating a set of local regions containing particular landmarks…

机器人学 · 计算机科学 2018-08-24 Bo Yang , Jun Li , Xiaosu Xu , Hong Zhang

We propose an image-based cross-view geolocalization method that estimates the global pose of a UAV with the aid of georeferenced satellite imagery. Our method consists of two Siamese neural networks that extract relevant features despite…

机器人学 · 计算机科学 2018-09-18 Akshay Shetty , Grace Xingxin Gao

In this work, we present a camera geopositioning system based on matching a query image against a database with panoramic images. For matching, our system uses memory vectors aggregated from global image descriptors based on convolutional…

计算机视觉与模式识别 · 计算机科学 2019-03-14 Raffaele Imbriaco , Clint Sebastian , Egor Bondarev , Peter de With

The goal of point cloud localization based on linguistic description is to identify a 3D position using textual description in large urban environments, which has potential applications in various fields, such as determining the location…

计算机视觉与模式识别 · 计算机科学 2025-03-21 Yanlong Xu , Haoxuan Qu , Jun Liu , Wenxiao Zhang , Xun Yang

Matching corresponding features between two images is a fundamental task to computer vision with numerous applications in object recognition, robotics, and 3D reconstruction. Current state of the art in image feature matching has focused on…

计算机视觉与模式识别 · 计算机科学 2019-09-05 Chen Zhao , Jiaqi Yang , Ke Xian , Zhiguo Cao , Xin Li

Localizing textual descriptions within large-scale 3D scenes presents inherent ambiguities, such as identifying all traffic lights in a city. Addressing this, we introduce a method to generate distributions of camera poses conditioned on…

计算机视觉与模式识别 · 计算机科学 2025-02-04 Qi Ma , Runyi Yang , Bin Ren , Nicu Sebe , Ender Konukoglu , Luc Van Gool , Danda Pani Paudel

We describe a method for modeling spatial context to enable video anomaly detection. The main idea is to discover regions that share similar object-level activities by clustering joint object attributes using Gaussian mixture models. We…

计算机视觉与模式识别 · 计算机科学 2025-01-16 Zhengye Yang , Richard J. Radke