English
Related papers

Related papers: Zero-Shot Satellite Image Retrieval through Joint …

200 papers

Recent advances in deep-learning based methods for image matching have demonstrated their superiority over traditional algorithms, enabling correspondence estimation in challenging scenes with significant differences in viewing angles,…

Computer Vision and Pattern Recognition · Computer Science 2025-12-17 Rahul Deshmukh , Avinash Kak

We present GeoSynth, a model for synthesizing satellite images with global style and image-driven layout control. The global style control is via textual prompts or geographic location. These enable the specification of scene semantics or…

Computer Vision and Pattern Recognition · Computer Science 2024-04-11 Srikumar Sastry , Subash Khanal , Aayush Dhakal , Nathan Jacobs

Zero-shot 3D Anomaly Detection is an emerging task that aims to detect anomalies in a target dataset without any target training data, which is particularly important in scenarios constrained by sample scarcity and data privacy concerns.…

Computer Vision and Pattern Recognition · Computer Science 2026-05-27 Zehao Deng , An Liu , Yan Wang

We propose a novel Generalized Zero-Shot learning (GZSL) method that is agnostic to both unseen images and unseen semantic vectors during training. Prior works in this context propose to map high-dimensional visual features to the semantic…

Computer Vision and Pattern Recognition · Computer Science 2019-04-10 Pengkai Zhu , Hanxiao Wang , Venkatesh Saligrama

Generalized zero-shot learning (GZSL) aims to recognize objects from both seen and unseen classes, when only the labeled examples from seen classes are provided. Recent feature generation methods learn a generative model that can synthesize…

Computer Vision and Pattern Recognition · Computer Science 2021-03-31 Zongyan Han , Zhenyong Fu , Shuo Chen , Jian Yang

Geo-Foundational Models (GFMs) enable fast and reliable extraction of spatiotemporal information from satellite imagery, improving flood inundation mapping by leveraging location and time embeddings. Despite their potential, it remains…

Computer Vision and Pattern Recognition · Computer Science 2026-01-28 Saurabh Kaushik , Lalit Maurya , Elizabeth Tellman , ZhiJie Zhang

This study introduces a novel approach to online embedding of multi-scale CLIP (Contrastive Language-Image Pre-Training) features into 3D maps. By harnessing CLIP, this methodology surpasses the constraints of conventional…

Robotics · Computer Science 2024-03-28 Shun Taguchi , Hideki Deguchi

Floods are among the most damaging weather-related hazards, and in 2024, the warmest year on record, extreme flood events affected communities across five continents. Earth observation (EO) satellites provide critical, frequent coverage for…

Computer Vision and Pattern Recognition · Computer Science 2025-12-03 Mirela G. Tulbure , Julio Caineta , Mark Broich , Mollie D. Gaines , Philippe Rufin , Leon-Friedrich Thomas , Hamed Alemohammad , Jan Hemmerling , Patrick Hostert

Zero-shot scene understanding in real-world settings presents major challenges due to the complexity and variability of natural scenes, where models must recognize new objects, actions, and contexts without prior labeled examples. This work…

Computer Vision and Pattern Recognition · Computer Science 2025-10-30 Manjunath Prasad Holenarasipura Rajiv , B. M. Vidyavathi

Recent geospatial foundation models (GFMs) produce spatially extensive representations of the Earth's surface that capture rich physical and environmental patterns. Among them, the AlphaEarth Foundation (AE) represents a major step,…

Artificial Intelligence · Computer Science 2026-03-17 Junyuan Liu , Quan Qin , Guangsheng Dong , Xinglei Wang , Jiazhuang Feng , Zichao Zeng , Tao Cheng

Vision-language pre-training such as CLIP enables zero-shot transfer that can classify images according to the candidate class names. While CLIP demonstrates an impressive zero-shot performance on diverse downstream tasks, the distribution…

Computer Vision and Pattern Recognition · Computer Science 2024-08-27 Qi Qian , Juhua Hu

The recent widespread adoption of drones for studying marine animals provides opportunities for deriving biological information from aerial imagery. The large scale of imagery data acquired from drones is well suited for machine learning…

Computer Vision and Pattern Recognition · Computer Science 2025-01-13 Chinmay K Lalgudi , Mark E Leone , Jaden V Clark , Sergio Madrigal-Mora , Mario Espinoza

Zero-shot learning (ZSL) makes object recognition in images possible in absence of visual training data for a part of the classes from a dataset. When the number of classes is large, classes are usually represented by semantic class…

Computer Vision and Pattern Recognition · Computer Science 2020-08-10 Yannick Le Cacheux , Adrian Popescu , Hervé Le Borgne

Effective embodied exploration requires agents to accumulate and retain spatial knowledge over time. However, existing scene representations, such as discrete scene graphs or static view-based snapshots, lack \textit{post-hoc…

Computer Vision and Pattern Recognition · Computer Science 2026-03-20 Yiren Lu , Yi Du , Disheng Liu , Yunlai Zhou , Chen Wang , Yu Yin

Modern vision language pipelines are driven by RGB vision encoders trained on massive image text corpora. While these pipelines have enabled impressive zero-shot capabilities and strong transfer across tasks, they still inherit two…

Computer Vision and Pattern Recognition · Computer Science 2025-12-25 Yasmine Omri , Connor Ding , Tsachy Weissman , Thierry Tambe

Large scale vision and language models can achieve impressive zero-shot recognition performance by mapping class specific text queries to image content. Two distinct challenges that remain however, are high sensitivity to the choice of…

Computer Vision and Pattern Recognition · Computer Science 2023-04-05 Sarah Parisot , Yongxin Yang , Steven McDonagh

We present a novel deep zero-shot learning (ZSL) model for inferencing human-object-interaction with verb-object (VO) query. While the previous two-stream ZSL approaches only use the semantic/textual information to be fed into the query…

Computer Vision and Pattern Recognition · Computer Science 2019-09-17 Sungmin Eum , Heesung Kwon

Zero-shot learning (ZSL) can be formulated as a cross-domain matching problem: after being projected into a joint embedding space, a visual sample will match against all candidate class-level semantic descriptions and be assigned to the…

Computer Vision and Pattern Recognition · Computer Science 2018-12-12 Lei Zhang , Peng Wang , Lingqiao Liu , Chunhua Shen , Wei Wei , Yannning Zhang , Anton Van Den Hengel

Cross-view geo-localization aims at establishing location correspondences between different viewpoints. Existing approaches typically learn cross-view correlations through direct feature similarity matching, often overlooking semantic…

Computer Vision and Pattern Recognition · Computer Science 2025-09-30 Hongyang Zhang , Yinhao Liu , Zhenyu Kuang

Subnational location data of disaster events are critical for risk assessment and disaster risk reduction. Disaster databases such as EM-DAT often report locations in unstructured textual form, with inconsistent granularity or spelling,…

Artificial Intelligence · Computer Science 2025-11-20 Michele Ronco , Damien Delforge , Wiebke S. Jäger , Christina Corbane