中文
相关论文

相关论文: Geospatial Foundational Embedder: Top-1 Winning So…

200 篇论文

Publicly available satellite imagery, such as Sentinel- 2, often lacks the spatial resolution required for accurate analysis of remote sensing tasks including urban planning and disaster response. Current super-resolution techniques are…

计算机视觉与模式识别 · 计算机科学 2025-01-31 Daniel Panangian , Ksenia Bittner

The forward full-wave modeling of ground-penetrating radar (GPR) facilitates the understanding and interpretation of GPR data. Traditional forward solvers require excessive computational resources, especially when their repetitive…

图像与视频处理 · 电气工程与系统科学 2022-10-05 Qiqi Dai , Yee Hui Lee , Hai-Han Sun , Jiwei Qian , Genevieve Ow , Mohamed Lokman Mohd Yusof , Abdulkadir C. Yucel

There exists a correlation between geospatial activity temporal patterns and type of land use. A novel self-supervised approach is proposed to stratify landscape based on mobility activity time series. First, the time series signal is…

计算机视觉与模式识别 · 计算机科学 2024-01-18 Yi Cao , Swetava Ganguli , Vipul Pandey

Geospatial Foundation Models (GeoFMs) are transforming Earth Observation (EO), but evaluation lacks standardized protocols. GEO-Bench-2 addresses this with a comprehensive framework spanning classification, segmentation, regression, object…

Generating learning-friendly representations for points in a 2D space is a fundamental and long-standing problem in machine learning. Recently, multi-scale encoding schemes (such as Space2Vec) were proposed to directly encode any point in…

计算机视觉与模式识别 · 计算机科学 2022-01-26 Gengchen Mai , Yao Xuan , Wenyun Zuo , Krzysztof Janowicz , Ni Lao

The increasing availability of geospatial foundation models has the potential to transform remote sensing applications such as land cover classification, environmental monitoring, and change detection. Despite promising benchmark results,…

In this paper we address the task of visual place recognition (VPR), where the goal is to retrieve the correct GPS coordinates of a given query image against a huge geotagged gallery. While recent works have shown that building descriptors…

计算机视觉与模式识别 · 计算机科学 2022-01-26 Valerio Paolicelli , Antonio Tavera , Carlo Masone , Gabriele Berton , Barbara Caputo

We propose Equiangular Basis Vectors (EBVs) for classification tasks. In deep neural networks, models usually end with a k-way fully connected layer with softmax to handle different classification tasks. The learning objective of these…

计算机视觉与模式识别 · 计算机科学 2023-05-09 Yang Shen , Xuhao Sun , Xiu-Shen Wei

Aerial imagery and its direct application to visual localization is an essential problem for many Robotics and Computer Vision tasks. While Global Navigation Satellite Systems (GNSS) are the standard default solution for solving the aerial…

计算机视觉与模式识别 · 计算机科学 2024-11-01 Ivan Moskalenko , Anastasiia Kornilova , Gonzalo Ferrer

In this paper, we present a high-performing solution to the UAVM 2025 Challenge, which focuses on matching narrow FOV street-level images to corresponding satellite imagery using the University-1652 dataset. As panoramic Cross-View…

计算机视觉与模式识别 · 计算机科学 2025-07-23 Xiaohan Zhang , Tavis Shore , Chen Chen , Oscar Mendez , Simon Hadfield , Safwan Wshah

Embedding models have been crucial in enabling various downstream tasks such as semantic similarity, information retrieval, and clustering. Recently, there has been a surge of interest in developing universal text embedding models that can…

计算机视觉与模式识别 · 计算机科学 2025-01-03 Ziyan Jiang , Rui Meng , Xinyi Yang , Semih Yavuz , Yingbo Zhou , Wenhu Chen

This study investigates whether the geospatial and multimodal features encoded in \textit{Earth Embeddings} can effectively guide deep learning (DL) regression models for regional surface height mapping. In particular, we focused on…

计算机视觉与模式识别 · 计算机科学 2026-02-20 Alireza Hamoudzadeh , Valeria Belloni , Roberta Ravanelli

This report describes the winning solution to the WeatherProof Dataset Challenge (CVPR 2024 UG2+ Track 3). Details regarding the challenge are available at https://cvpr2024ug2challenge.github.io/track3.html. We propose an enhanced semantic…

计算机视觉与模式识别 · 计算机科学 2024-06-10 Nan Zhang , Xidan Zhang , Jianing Wei , Fangjun Wang , Zhiming Tan

In this paper, we propose a new image-based visual place recognition (VPR) framework by exploiting the structural cues in bird's-eye view (BEV) from a single monocular camera. The motivation arises from two key observations about place…

计算机视觉与模式识别 · 计算机科学 2024-07-24 Fudong Ge , Yiwei Zhang , Shuhan Shen , Yue Wang , Weiming Hu , Jin Gao

Generating learning-friendly representations for points in space is a fundamental and long-standing problem in ML. Recently, multi-scale encoding schemes (such as Space2Vec and NeRF) were proposed to directly encode any point in 2D/3D…

计算机视觉与模式识别 · 计算机科学 2023-07-04 Gengchen Mai , Yao Xuan , Wenyun Zuo , Yutong He , Jiaming Song , Stefano Ermon , Krzysztof Janowicz , Ni Lao

Video super-resolution (VSR) is a critical task for enhancing low-bitrate and low-resolution videos, particularly in streaming applications. While numerous solutions have been developed, they often suffer from high computational demands,…

图像与视频处理 · 电气工程与系统科学 2024-09-27 Marcos V Conde , Zhijun Lei , Wen Li , Christos Bampis , Ioannis Katsavounidis , Radu Timofte

Existing approaches to drone visual geo-localization predominantly adopt the image-based setting, where a single drone-view snapshot is matched with images from other platforms. Such task formulation, however, underutilizes the inherent…

计算机视觉与模式识别 · 计算机科学 2025-07-29 Hao Ju , Shaofei Huang , Si Liu , Zhedong Zheng

Satellite images are snapshots of the Earth surface. We propose to forecast them. We frame Earth surface forecasting as the task of predicting satellite imagery conditioned on future weather. EarthNet2021 is a large dataset suitable for…

机器学习 · 计算机科学 2021-04-21 Christian Requena-Mesa , Vitus Benson , Markus Reichstein , Jakob Runge , Joachim Denzler

Continuous Spatio-Temporal Video Super-Resolution (C-STVSR) aims to simultaneously enhance the spatial resolution and frame rate of videos by arbitrary scale factors, offering greater flexibility than fixed-scale methods that are…

计算机视觉与模式识别 · 计算机科学 2026-04-21 Mingyu Shi , Xin Di , Long Peng , Boxiang Cao , Anran Wu , Zhanfeng Feng , Jiaming Guo , Renjing Pei , Xueyang Fu , Yang Cao , Zhengjun Zha

Standard Vision Transformers flatten 2D images into 1D sequences, disrupting the natural spatial topology. While Rotary Positional Embedding (RoPE) excels in 1D, it inherits this limitation, often treating spatially distant patches (e.g.,…

计算机视觉与模式识别 · 计算机科学 2025-12-05 Yupu Yao , Bowen Yang