English
Related papers

Related papers: Scaling Image Geo-Localization to Continent Level

200 papers

Cross-view geo-localization identifies the locations of street-view images by matching them with geo-tagged satellite images or OSM. However, most existing studies focus on image-to-image retrieval, with fewer addressing text-guided…

Computer Vision and Pattern Recognition · Computer Science 2025-04-02 Junyan Ye , Honglin Lin , Leyan Ou , Dairong Chen , Zihao Wang , Qi Zhu , Conghui He , Weijia Li

LiDAR-based place recognition is an essential and challenging task both in loop closure detection and global relocalization. We propose Deep Scan Context (DSC), a general and discriminative global descriptor that captures the relationship…

Computer Vision and Pattern Recognition · Computer Science 2021-11-30 Jiafeng Cui , Tengfei Huang , Yingfeng Cai , Junqiao Zhao , Lu Xiong , Zhuoping Yu

We address the task of cross-domain visual place recognition, where the goal is to geolocalize a given query image against a labeled gallery, in the case where the query and the gallery belong to different visual domains. To achieve this,…

Computer Vision and Pattern Recognition · Computer Science 2021-01-08 Gabriele Moreno Berton , Valerio Paolicelli , Carlo Masone , Barbara Caputo

In this work we present a novel approach to joint semantic localisation and scene understanding. Our work is motivated by the need for localisation algorithms which not only predict 6-DoF camera pose but also simultaneously recognise…

Computer Vision and Pattern Recognition · Computer Science 2019-09-24 Ignas Budvytis , Marvin Teichmann , Tomas Vojir , Roberto Cipolla

Global visual localization estimates the absolute pose of a camera using a single image, in a previously mapped area. Obtaining the pose from a single image enables many robotics and augmented/virtual reality applications. Inspired by…

Computer Vision and Pattern Recognition · Computer Science 2023-12-19 Mohammad Altillawi , Shile Li , Sai Manoj Prakhya , Ziyuan Liu , Joan Serrat

Image-to-image translation is a technique that focuses on transferring images from one domain to another while maintaining the essential content representations. In recent years, image-to-image translation has gained significant attention…

Image and Video Processing · Electrical Eng. & Systems 2024-04-02 Xixian Wu , Dian Chao , Yang Yang

"Text can appear anywhere". This property requires us to carefully process all the pixels in an image in order to accurately localize all text instances. In particular, for the more difficult task of localizing small text regions, many…

Computer Vision and Pattern Recognition · Computer Science 2019-07-30 Elad Richardson , Yaniv Azar , Or Avioz , Niv Geron , Tomer Ronen , Zach Avraham , Stav Shapiro

We present GSplatLoc, a camera localization method that leverages the differentiable rendering capabilities of 3D Gaussian splatting for ultra-precise pose estimation. By formulating pose estimation as a gradient-based optimization problem…

Computer Vision and Pattern Recognition · Computer Science 2025-05-20 Atticus J. Zeller , Haijuan Wu

Cross-view geo-localization (CVGL), which involves matching and retrieving satellite images to determine the geographic location of a ground image, is crucial in GNSS-constrained scenarios. However, this task faces significant challenges…

Computer Vision and Pattern Recognition · Computer Science 2024-11-20 Gaoshuang Huang , Yang Zhou , Luying Zhao , Wenjian Gan

Generalizable cross-view geo-localization aims to match the same location across views in unseen regions and conditions without GPS supervision. Its core difficulty lies in severe semantic inconsistency caused by viewpoint variation and…

Computer Vision and Pattern Recognition · Computer Science 2026-04-17 Hongyang Zhang , Yinhao Liu , Haitao Zhang , Zhongyi Wen , Zhenyu Kuang , Shuxian Liang , Xiansheng Hua

Recent developments in 3D Gaussian Splatting have made significant advances in surface reconstruction. However, scaling these methods to large-scale scenes remains challenging due to high computational demands and the complex dynamic…

Graphics · Computer Science 2025-06-24 Shihan Chen , Zhaojin Li , Zeyu Chen , Qingsong Yan , Gaoyang Shen , Ran Duan

A major challenge in place recognition for autonomous driving is to be robust against appearance changes due to short-term (e.g., weather, lighting) and long-term (seasons, vegetation growth, etc.) environmental variations. A promising…

Computer Vision and Pattern Recognition · Computer Science 2019-08-02 Anh-Dzung Doan , Yasir Latif , Tat-Jun Chin , Yu Liu , Thanh-Toan Do , Ian Reid

In this paper, we address the problem of landmark-based visual place recognition. In the state-of-the-art method, accurate object proposal algorithms are first leveraged for generating a set of local regions containing particular landmarks…

Robotics · Computer Science 2018-08-24 Bo Yang , Jun Li , Xiaosu Xu , Hong Zhang

We address the question of to what accuracy remote sensing images of the surface of planets can be matched, so that the possible displacement of features on the surface can be accurately measured. This is relevant in the context of the…

Astrophysics · Physics 2007-05-23 Giuseppe Vacanti , Ernst-Jan Buis

Capturing and labeling camera images in the real world is an expensive task, whereas synthesizing labeled images in a simulation environment is easy for collecting large-scale image data. However, learning from only synthetic images may not…

Computer Vision and Pattern Recognition · Computer Science 2018-07-06 Tadanobu Inoue , Subhajit Chaudhury , Giovanni De Magistris , Sakyasingha Dasgupta

Fine-grained categories are more difficulty distinguished than generic categories due to the similarity of inter-class and the diversity of intra-class. Therefore, the fine-grained visual categorization (FGVC) is considered as one of…

Computer Vision and Pattern Recognition · Computer Science 2015-05-12 Guo Lihua , Guo Chenggan

Cross-view image matching for geo-localisation is a challenging problem due to the significant visual difference between aerial and ground-level viewpoints. The method provides localisation capabilities from geo-referenced images,…

Computer Vision and Pattern Recognition · Computer Science 2024-09-25 Tavis Shore , Simon Hadfield , Oscar Mendez

CLIP has shown impressive results in aligning images and texts at scale. However, its ability to capture detailed visual features remains limited because CLIP matches images and texts at a global level. To address this issue, we propose…

Computer Vision and Pattern Recognition · Computer Science 2024-12-05 Rui Xiao , Sanghwan Kim , Mariana-Iuliana Georgescu , Zeynep Akata , Stephan Alaniz

Accurate and robust image-based geo-localization at a global scale is challenging due to diverse environments, visually ambiguous scenes, and the lack of distinctive landmarks in many regions. While contrastive learning methods show…

Computer Vision and Pattern Recognition · Computer Science 2025-09-29 Boyi Chen , Zhangyu Wang , Fabian Deuser , Johann Maximilian Zollner , Martin Werner

This article presents an efficient end-to-end method to perform instance-level recognition employed to the task of labeling and ranking landmark images. In a first step, we embed images in a high dimensional feature space using…

Computer Vision and Pattern Recognition · Computer Science 2020-10-06 Christof Henkel , Philipp Singer