中文
相关论文

相关论文: Cross-view image geo-localization with Panorama-BE…

200 篇论文

Bird's-eye-view (BEV) representations derived from multi-camera input have become a central interface for online high-definition (HD) map construction. However, most approaches rely solely on ego-centric supervision, requiring large-scale…

计算机视觉与模式识别 · 计算机科学 2026-05-13 Daniel Lengerer , Mathias Pechinger , Klaus Bogenberger , Carsten Markgraf

Metric Cross-View Geo-Localization (MCVGL) aims to estimate the 3-DoF camera pose (position and heading) by matching ground and satellite images. In this work, instead of pinhole and satellite images, we study robust MCVGL using holistic…

计算机视觉与模式识别 · 计算机科学 2026-05-29 Junwei Zheng , Ruize Dai , Ruiping Liu , Zichao Zeng , Yufan Chen , Fangjinhua Wang , Kunyu Peng , Kailun Yang , Jiaming Zhang , Rainer Stiefelhagen

Cross-view geo-localization aims to spot images of the same location shot from two platforms, e.g., the drone platform and the satellite platform. Existing methods usually focus on optimizing the distance between one embedding with others…

计算机视觉与模式识别 · 计算机科学 2022-11-11 Tingyu Wang , Zhedong Zheng , Zunjie Zhu , Yuhan Gao , Yi Yang , Chenggang Yan

Satellite imagery differs fundamentally from natural images: its aerial viewpoint, very high resolution, diverse scale variations, and abundance of small objects demand both region-level spatial reasoning and holistic scene understanding.…

计算机视觉与模式识别 · 计算机科学 2025-12-15 Emanuel Sánchez Aimar , Gulnaz Zhambulova , Fahad Shahbaz Khan , Yonghao Xu , Michael Felsberg

We propose an accurate and interpretable fine-grained cross-view localization method that estimates the 3 Degrees of Freedom (DoF) pose of a ground-level image by matching its local features with a reference aerial image. Unlike prior…

计算机视觉与模式识别 · 计算机科学 2026-02-27 Zimin Xia , Chenghao Xu , Alexandre Alahi

Predicting realistic ground views from satellite imagery in urban scenes is a challenging task due to the significant view gaps between satellite and ground-view images. We propose a novel pipeline to tackle this challenge, by generating…

计算机视觉与模式识别 · 计算机科学 2024-09-16 Ningli Xu , Rongjun Qin

Accurate 3D lane detection from monocular images presents significant challenges due to depth ambiguity and imperfect ground modeling. Previous attempts to model the ground have often used a planar ground assumption with limited degrees of…

计算机视觉与模式识别 · 计算机科学 2025-01-27 Chaesong Park , Eunbin Seo , Jongwoo Lim

Nature disasters play a key role in shaping human-urban infrastructure interactions. Effective and efficient response to natural disasters is essential for building resilience and a sustainable urban environment. Two types of information…

计算机视觉与模式识别 · 计算机科学 2024-08-14 Hao Li , Fabian Deuser , Wenping Yina , Xuanshu Luo , Paul Walther , Gengchen Mai , Wei Huang , Martin Werner

Cross-view geo-localization (CVGL) matches query images ($\textit{e.g.}$, drone) to geographically corresponding opposite-view imagery ($\textit{e.g.}$, satellite). While supervised methods achieve strong performance, their reliance on…

计算机视觉与模式识别 · 计算机科学 2025-11-18 Cuiqun Chen , Qi Chen , Bin Yang , Xingyi Zhang

We propose a pipeline for combined multi-class object geolocation and height estimation from street level RGB imagery, which is considered as a single available input data modality. Our solution is formulated via Markov Random Field…

计算机视觉与模式识别 · 计算机科学 2023-05-16 Matej Ulicny , Vladimir A. Krylov , Julie Connelly , Rozenn Dahyot

Street view imagery has become an essential source for geospatial data collection and urban analytics, enabling the extraction of valuable insights that support informed decision-making. However, synthesizing street-view images from…

计算机视觉与模式识别 · 计算机科学 2025-09-30 Khawlah Bajbaa , Abbas Anwar , Muhammad Saqib , Hafeez Anwar , Nabin Sharma , Muhammad Usman

Aerial imagery analysis is critical for many research fields. However, obtaining frequent high-quality aerial images is not always accessible due to its high effort and cost requirements. One solution is to use the Ground-to-Aerial (G2A)…

计算机视觉与模式识别 · 计算机科学 2024-08-22 Ahmad Arrabi , Xiaohan Zhang , Waqas Sultani , Chen Chen , Safwan Wshah

Cross-View object geo-localization (CVOGL) aims to precisely determine the geographic coordinates of a query object from a ground or drone perspective by referencing a satellite map. Segmentation-based approaches offer high precision but…

计算机视觉与模式识别 · 计算机科学 2026-03-17 Chenlin Fu , Ao Gong , Yingying Zhu

In this paper, we introduce a novel approach to fine-grained cross-view geo-localization. Our method aligns a warped ground image with a corresponding GPS-tagged satellite image covering the same area using homography estimation. We first…

计算机视觉与模式识别 · 计算机科学 2023-09-01 Xiaolong Wang , Runsen Xu , Zuofan Cui , Zeyu Wan , Yu Zhang

Many applications such as autonomous navigation, urban planning and asset monitoring, rely on the availability of accurate information about objects and their geolocations. In this paper we propose to automatically detect and compute the…

计算机视觉与模式识别 · 计算机科学 2018-05-08 Vladimir A. Krylov , Eamonn Kenny , Rozenn Dahyot

Planet-scale photo geolocalization involves the intricate task of estimating the geographic location depicted in an image purely based on its visual features. While deep learning models, particularly convolutional neural networks (CNNs),…

计算机视觉与模式识别 · 计算机科学 2026-03-26 David Faget , José Luis Lisani , Miguel Colom

With the advancement of collaborative perception, the role of aerial-ground collaborative perception, a crucial component, is becoming increasingly important. The demand for collaborative perception across different perspectives to…

计算机视觉与模式识别 · 计算机科学 2024-06-10 Yuchao Wang , Peirui Cheng , Pengju Tian , Ziyang Yuan , Liangjin Zhao , Jing Tian , Wensheng Wang , Zhirui Wang , Xian Sun

Image geo-localization is the task of predicting the specific location of an image and requires complex reasoning across visual, geographical, and cultural contexts. While prior Vision Language Models (VLMs) have the best accuracy at this…

计算与语言 · 计算机科学 2025-02-21 Zheyuan Zhang , Runze Li , Tasnim Kabir , Jordan Boyd-Graber

Domain-generalized LiDAR semantic segmentation (LSS) seeks to train models on source-domain point clouds that generalize reliably to multiple unseen target domains, which is essential for real-world LiDAR applications. However, existing…

计算机视觉与模式识别 · 计算机科学 2026-02-17 Jindong Zhao , Yuan Gao , Yang Xia , Sheng Nie , Jun Yue , Weiwei Sun , Shaobo Xia

Bird's-Eye-View (BEV) representation has emerged as a mainstream paradigm for multi-view 3D object detection, demonstrating impressive perceptual capabilities. However, existing methods overlook the geometric quality of BEV representation,…

计算机视觉与模式识别 · 计算机科学 2024-12-24 Jinqing Zhang , Yanan Zhang , Yunlong Qi , Zehua Fu , Qingjie Liu , Yunhong Wang