中文
相关论文

相关论文: Wrivinder: Towards Spatial Intelligence for Geo-lo…

200 篇论文

We address the problem of ground-to-satellite image geo-localization, that is, estimating the camera latitude, longitude and orientation (azimuth angle) by matching a query image captured at the ground level against a large-scale database…

计算机视觉与模式识别 · 计算机科学 2022-03-29 Yujiao Shi , Xin Yu , Liu Liu , Dylan Campbell , Piotr Koniusz , Hongdong Li

Zero-shot 3D Visual Grounding (3DVG) is a critical capability for open-world embodied AI. However, existing methods are fundamentally bottlenecked by the poor quality of open-vocabulary 3D proposals, suffering from inaccurate categories and…

计算机视觉与模式识别 · 计算机科学 2026-04-30 Yufei Yin , Jie Zheng , Qianke Meng , Zhou Yu , Minghao Chen , Jiajun Ding , Min Tan , Yuling Xi , Zhiwen Chen , Chengfei Lv

Classifying geospatial imagery remains a major bottleneck for applications such as disaster response and land-use monitoring-particularly in regions where annotated data is scarce or unavailable. Existing tools (e.g., RS-CLIP) that claim…

计算机视觉与模式识别 · 计算机科学 2025-06-02 Gilles Quentin Hacheme , Girmaw Abebe Tadesse , Caleb Robinson , Akram Zaytar , Rahul Dodhia , Juan M. Lavista Ferres

The goal of cross-view image based geo-localization is to determine the location of a given street view image by matching it against a collection of geo-tagged satellite images. This task is notoriously challenging due to the drastic…

计算机视觉与模式识别 · 计算机科学 2021-03-12 Aysim Toker , Qunjie Zhou , Maxim Maximov , Laura Leal-Taixé

3D visual grounding (3DVG) aims to localize objects in a 3D scene based on natural language queries. In this work, we explore zero-shot 3DVG from multi-view images alone, without requiring any geometric supervision or object priors. We…

计算机视觉与模式识别 · 计算机科学 2026-02-04 Nikita Drozdov , Andrey Lemeshko , Nikita Gavrilov , Anton Konushin , Danila Rukhovich , Maksim Kolodiazhnyi

This paper addresses the problem of vehicle-mounted camera localization by matching a ground-level image with an overhead-view satellite map. Existing methods often treat this problem as cross-view image retrieval, and use learned deep…

计算机视觉与模式识别 · 计算机科学 2022-09-07 Yujiao Shi , Hongdong Li

We introduce Sky2Ground, a three-view dataset designed for varying altitude camera localization, correspondence learning, and reconstruction. The dataset combines structured synthetic imagery with real, in-the-wild images, providing both…

计算机视觉与模式识别 · 计算机科学 2026-04-21 Zengyan Wang , Sirshapan Mitra , Rajat Modi , Grace Lim , Yogesh Rawat

In this paper, we discuss and review how combined multi-view imagery from satellite to street-level can benefit scene analysis. Numerous works exist that merge information from remote sensing and images acquired from the ground for tasks…

计算机视觉与模式识别 · 计算机科学 2017-09-29 Sébastien Lefèvre , Devis Tuia , Jan Dirk Wegner , Timothée Produit , Ahmed Samy Nassar

The problem of localization on a geo-referenced satellite map given a query ground view image is useful yet remains challenging due to the drastic change in viewpoint. To this end, in this paper we work on the extension of our earlier work…

计算机视觉与模式识别 · 计算机科学 2019-06-04 Sixing Hu , Gim Hee Lee

Nowadays the accurate geo-localization of ground-view images has an important role across domains as diverse as journalism, forensics analysis, transports, and Earth Observation. This work addresses the problem of matching a query…

计算机视觉与模式识别 · 计算机科学 2024-05-24 Francesco Pro , Nikolaos Dionelis , Luca Maiano , Bertrand Le Saux , Irene Amerini

We explore the task of geometric reconstruction of images captured from a mixture of ground and aerial views. Current state-of-the-art learning-based approaches fail to handle the extreme viewpoint variation between aerial-ground image…

计算机视觉与模式识别 · 计算机科学 2025-04-18 Khiem Vuong , Anurag Ghosh , Deva Ramanan , Srinivasa Narasimhan , Shubham Tulsiani

Zero-shot 3D visual grounding requires localizing objects in unstructured environments from free-form natural language. Recent vision-language model (VLM) approaches achieve promising results but rely on view-dependent reasoning or implicit…

计算机视觉与模式识别 · 计算机科学 2026-05-22 Xuefei Sun , Xujia Zhang , Brendan Crowe , Doncey Albin , Christoffer Heckman

We introduce BEVRender, a novel learning based approach for the localization of ground vehicles in Global Navigation Satellite System(GNSS)-denied off-road scenarios. These environments are typically challenging for conventional…

机器人学 · 计算机科学 2024-12-11 Lihong Jin , Wei Dong , Wenshan Wang , Michael Kaess

This work addresses visual cross-view metric localization for outdoor robotics. Given a ground-level color image and a satellite patch that contains the local surroundings, the task is to identify the location of the ground camera within…

计算机视觉与模式识别 · 计算机科学 2022-08-19 Zimin Xia , Olaf Booij , Marco Manfredi , Julian F. P. Kooij

Cross-view geo-localization aims at establishing location correspondences between different viewpoints. Existing approaches typically learn cross-view correlations through direct feature similarity matching, often overlooking semantic…

计算机视觉与模式识别 · 计算机科学 2025-09-30 Hongyang Zhang , Yinhao Liu , Zhenyu Kuang

The European Space Agency's Copernicus Sentinel-1 (S-1) mission is a constellation of C-band synthetic aperture radar (SAR) satellites that provide unprecedented monitoring of the world's oceans. S-1's wave mode (WV) captures 20x20 km image…

Accurate surround-view depth estimation provides a competitive alternative to laser-based sensors and is essential for 3D scene understanding in autonomous driving. While empirical studies have proposed various approaches that primarily…

计算机视觉与模式识别 · 计算机科学 2026-01-21 Weimin Liu , Wenjun Wang , Joshua H. Meng

We examine the challenge of estimating the location of a single ground-level image in the absence of GPS or other location metadata. Currently, geolocation systems are evaluated by measuring the Great Circle Distance between the predicted…

计算机视觉与模式识别 · 计算机科学 2024-09-19 Michael J. Bianco , David Eigen , Michael Gormish

Reliable image correspondences form the foundation of vision-based spatial perception, enabling recovery of 3D structure and camera poses. However, unconstrained feature matching across domains such as aerial, indoor, and outdoor scenes…

计算机视觉与模式识别 · 计算机科学 2025-12-30 Zhimin Shao , Abhay Yadav , Rama Chellappa , Cheng Peng

With the ability of providing direct and accurate enough range measurements, light detection and ranging (LiDAR) is playing an essential role in localization and detection for autonomous vehicles. Since single LiDAR suffers from hardware…

机器人学 · 计算机科学 2022-01-14 Yusheng Wang , Yidong Lou , Weiwei Song , Huan Yu , Zhiyong Tu
‹ 上一页 1 2 3 10 下一页 ›