English
Related papers

Related papers: SITS-DECO: A Generative Decoder Is All You Need Fo…

200 papers

Detection transformer (DETR) relies on one-to-one assignment, assigning one ground-truth object to one prediction, for end-to-end detection without NMS post-processing. It is known that one-to-many assignment, assigning one ground-truth…

Computer Vision and Pattern Recognition · Computer Science 2023-09-01 Qiang Chen , Xiaokang Chen , Jian Wang , Shan Zhang , Kun Yao , Haocheng Feng , Junyu Han , Errui Ding , Gang Zeng , Jingdong Wang

Efficient and accurate annotation of datasets remains a significant challenge for deploying object detection models such as You Only Look Once (YOLO) in real-world applications, particularly in agriculture where rapid decision-making is…

Computer Vision and Pattern Recognition · Computer Science 2025-12-02 Mohamed Abdallah Salem , Ahmed Harb Rabia

Accurate crop mapping fundamentally relies on modeling multi-scale spatiotemporal patterns, where spatial scales range from individual field textures to landscape-level context, and temporal scales capture both short-term phenological…

Computer Vision and Pattern Recognition · Computer Science 2026-01-16 Wenyuan Li , Shunlin Liang , Keyan Chen , Yongzhe Chen , Han Ma , Jianglei Xu , Yichuan Ma , Shikang Guan , Husheng Fang , Zhenwei Shi

Earth Observation (EO) analysis is inherently interactive: resolving uncertainty often requires expanding the region of interest, retrieving historical observations, and switching across sensors such as optical and Synthetic Aperture Radar.…

Artificial Intelligence · Computer Science 2026-05-05 Sai Ma , Zhuang Li , Sichao Li , Xinyue Xu , Ruibiao Zhu , Tony Boston , John A. Taylor

We introduce the Cisco Time Series Model, a univariate zero-shot forecaster. This time series foundation model is the result of a general architectural innovation to a time series model enabling it to accept multiresolution input, applied…

Recent vision foundation models can extract universal representations and show impressive abilities in various tasks. However, their application on object detection is largely overlooked, especially without fine-tuning them. In this work,…

Computer Vision and Pattern Recognition · Computer Science 2024-10-28 Shenghao Fu , Junkai Yan , Qize Yang , Xihan Wei , Xiaohua Xie , Wei-Shi Zheng

Foundation Models (FMs) are large-scale, pre-trained artificial intelligence (AI) systems that have revolutionized natural language processing and computer vision, and are now advancing geospatial analysis and Earth Observation (EO). They…

Computer Vision and Pattern Recognition · Computer Science 2025-11-04 Pedram Ghamisi , Weikang Yu , Xiaokang Zhang , Aldino Rizaldy , Jian Wang , Chufeng Zhou , Richard Gloaguen , Gustau Camps-Valls

This paper addresses the challenge of capturing global temporaldependencies in long video sequences for Video Object Segmentation (VOS). Existing architectures often fail to effectively model these dependencies acrossextended temporal…

Computer Vision and Pattern Recognition · Computer Science 2025-09-16 Fenglei Hao , Yuliang Yang , Ruiyuan Su , Zhengran Zhao , Yukun Qiao , Mengyu Zhu

Nowadays, modern Earth Observation systems continuously generate huge amounts of data. A notable example is represented by the Sentinel-2 mission, which provides images at high spatial resolution (up to 10m) with high temporal revisit…

Computer Vision and Pattern Recognition · Computer Science 2018-09-21 Roberto Interdonato , Dino Ienco , Raffaele Gaetano , Kenji Ose

Earth observation (EO) foundation models (FMs) are increasingly trained on multisensor data, spanning multispectral imagery (MSI), synthetic aperture radar (SAR), and derived geospatial layers, but hyperspectral imagery (HSI) remains…

Computer Vision and Pattern Recognition · Computer Science 2026-05-21 Nassim Ait Ali Braham , Aaron Banze , Conrad M. Albrecht , Julien Mairal , Jocelyn Chanussot , Xiao Xiang Zhu

When we are primarily interested in solving several problems jointly with a given prescribed high performance accuracy for each target application, then Foundation Models should for most cases be used rather than problem-specific models. We…

Computer Vision and Pattern Recognition · Computer Science 2024-06-27 Nikolaos Dionelis , Casper Fibaek , Luke Camilleri , Andreas Luyts , Jente Bosmans , Bertrand Le Saux

Deep learning models are increasingly data-hungry, requiring significant resources to collect and compile the datasets needed to train them, with Earth Observation (EO) models being no exception. However, the landscape of datasets in EO is…

Computer Vision and Pattern Recognition · Computer Science 2024-06-24 Alistair Francis , Mikolaj Czerkawski

Object detection plays a crucial role in the field of computer vision by autonomously locating and identifying objects of interest. The You Only Look Once (YOLO) model is an effective single-shot detector. However, YOLO faces challenges in…

Computer Vision and Pattern Recognition · Computer Science 2025-07-28 Yash Zambre , Ekdev Rajkitkul , Akshatha Mohan , Joshua Peeples

Unprecedented access to multi-temporal satellite imagery has opened new perspectives for a variety of Earth observation tasks. Among them, pixel-precise panoptic segmentation of agricultural parcels has major economic and environmental…

Computer Vision and Pattern Recognition · Computer Science 2022-06-28 Vivien Sainte Fare Garnot , Loic Landrieu

With the emergence of deep learning in the last years, new opportunities arose in Earth observation research. Nevertheless, they also brought with them new challenges. The data-hungry training processes of deep learning models demand large,…

Computer Vision and Pattern Recognition · Computer Science 2022-04-21 Thorsten Hoeser , Claudia Kuenzer

Recent improvements in positioning technology has led to a much wider availability of massive moving object data. A crucial task is to find the moving objects that travel together. Usually, these object sets are called spatio-temporal…

Databases · Computer Science 2016-11-26 Phan Nhat Hai , Pascal Poncelet , Maguelonne Teisseire

Visual generative models (e.g., diffusion models) typically operate in compressed latent spaces to balance training efficiency and sample quality. In parallel, there has been growing interest in leveraging high-quality pre-trained visual…

Computer Vision and Pattern Recognition · Computer Science 2025-12-17 Yuan Gao , Chen Chen , Tianrong Chen , Jiatao Gu

Ground-based remote sensing cloud image sequence extrapolation is a key research area in the development of photovoltaic power systems. However, existing approaches exhibit several limitations:(1)they primarily rely on static kernels to…

Computer Vision and Pattern Recognition · Computer Science 2026-02-27 Penghui Niu , Taotao Cai , Suqi Zhang , Junhua Gua , Ping Zhanga , Qiqi Liu , Jianxin Li

Archival aerial imagery is a source of worldwide very high resolution data for documenting paste 3-D changes. However, external information is required so that accurate 3-D models can be computed from archival aerial imagery. In this…

Computer Vision and Pattern Recognition · Computer Science 2022-07-05 Denis Feurer , Fabrice Vinatier

Simultaneous Machine Translation (SiMT) generates translation while reading source tokens, essentially producing the target prefix based on the source prefix. To achieve good performance, it leverages the relationship between source and…

Computation and Language · Computer Science 2024-06-07 Shoutao Guo , Shaolei Zhang , Yang Feng