English
Related papers

Related papers: GTPBD-MM: A Global Terraced Parcel and Boundary Da…

200 papers

We present TerraMind, the first any-to-any generative, multimodal foundation model for Earth observation (EO). Unlike other multimodal models, TerraMind is pretrained on dual-scale representations combining both token-level and pixel-level…

Accurate estimates of Above Ground Biomass (AGB) are essential in addressing two of humanity's biggest challenges: climate change and biodiversity loss. Existing datasets for AGB estimation from satellite imagery are limited. Either they…

Computer Vision and Pattern Recognition · Computer Science 2025-04-08 Ghjulia Sialelli , Torben Peters , Jan D. Wegner , Konrad Schindler

Missing data is a recurrent problem in remote sensing, mainly due to cloud coverage for multispectral images and acquisition problems. This can be a critical issue for crop monitoring, especially for applications relying on machine learning…

Off-road environments remain significant challenges for autonomous ground vehicles, due to the lack of structured roads and the presence of complex obstacles, such as uneven terrain, vegetation, and occlusions. Traditional perception…

Robotics · Computer Science 2025-08-07 Zitong Chen , Chao Sun , Shida Nie , Chen Min , Changjiu Ning , Haoyu Li , Bo Wang

White matter tract segmentation is crucial for studying brain structural connectivity and neurosurgical planning. However, segmentation remains challenging due to issues like class imbalance between major and minor tracts, structural…

Computer Vision and Pattern Recognition · Computer Science 2025-01-28 Anoushkrit Goel , Bipanjit Singh , Ankita Joshi , Ranjeet Ranjan Jha , Chirag Ahuja , Aditya Nigam , Arnav Bhavsar

Cross-resolution land cover mapping aims to produce high-resolution semantic predictions from coarse or low-resolution supervision, yet the severe resolution mismatch makes effective learning highly challenging. Existing weakly supervised…

Computer Vision and Pattern Recognition · Computer Science 2025-12-24 Peng Gao , Ke Li , Di Wang , Yongshan Zhu , Yiming Zhang , Xuemei Luo , Yifeng Wang

Both in terrestrial and extraterrestrial environments, the precise and informative model of the ground and the surface ahead is crucial for navigation and obstacle avoidance. The ground surface is not always flat and it may be sloped, bumpy…

Machine Learning · Computer Science 2022-10-20 Pouria Mehrabi , Hamid D. Taghirad

Multi-modal large language models have demonstrated impressive performance across various tasks in different modalities. However, existing multi-modal models primarily emphasize capturing global information within each modality while…

Computer Vision and Pattern Recognition · Computer Science 2024-03-06 Zhaowei Li , Qi Xu , Dong Zhang , Hang Song , Yiqing Cai , Qi Qi , Ran Zhou , Junting Pan , Zefeng Li , Van Tu Vu , Zhida Huang , Tao Wang

This paper proposes a simple, generic and robust method to extract the grains from experimental tridimensionnal images of granular materials obtained by X-ray tomography. This extraction has two steps: segmentation and splitting. For the…

Computer Vision and Pattern Recognition · Computer Science 2008-07-23 Vincent Tariel

Recent studies have shown the benefits of using additional elevation data (e.g., DSM) for enhancing the performance of the semantic segmentation of aerial images. However, previous methods mostly adopt 3D elevation information as additional…

Computer Vision and Pattern Recognition · Computer Science 2020-09-23 Xiang Li , Lingjing Wang , Yi Fang

Fine-grained high-resolution remote sensing mapping typically relies on localized visual features, which restricts cross-domain generalizability and often leads to fragmented predictions of large-scale land covers. While global geospatial…

Computer Vision and Pattern Recognition · Computer Science 2026-04-23 Jienan Lyu , Miao Yang , Jinchen Cai , Yiwen Hu , Guanyi Lu , Junhao Qiu , Runmin Dong

Despite the good results that have been achieved in unimodal segmentation, the inherent limitations of individual data increase the difficulty of achieving breakthroughs in performance. For that reason, multi-modal learning is increasingly…

Image and Video Processing · Electrical Eng. & Systems 2024-04-16 Yameng Wang , Yi Wan , Yongjun Zhang , Bin Zhang , Zhi Gao

Semantic segmentation aims to robustly predict coherent class labels for entire regions of an image. It is a scene understanding task that powers real-world applications (e.g., autonomous navigation). One important application, the use of…

Computer Vision and Pattern Recognition · Computer Science 2023-02-16 Yuxiang Zhang , Sachin Mehta , Anat Caspi

The traditional deep learning paradigm that solely relies on labeled data has limitations in representing the spatial relationships between farmland elements and the surrounding environment.It struggles to effectively model the dynamic…

Computer Vision and Pattern Recognition · Computer Science 2025-04-01 Chao Tao , Dandan Zhong , Weiliang Mu , Zhuofei Du , Haiyang Wu

High-resolution elevation data is essential for hydrological modeling, hazard assessment, and environmental monitoring; however, globally consistent, fine-scale Digital Elevation Models (DEMs) remain unavailable. Very high-resolution…

Computer Vision and Pattern Recognition · Computer Science 2026-04-03 Osher Rafaeli , Tal Svoray , Ariel Nahlieli

Accurate and cost-effective quantification of the agroecosystem carbon cycle at decision-relevant scales is essential for climate mitigation and sustainable agriculture. However, both transfer learning and the exploitation of spatial…

Machine Learning · Computer Science 2025-12-19 Ruolei Zeng , Arun Sharma , Shuai An , Mingzhou Yang , Shengya Zhang , Licheng Liu , David Mulla , Shashi Shekhar

Spectral information has long been recognized as a critical cue in remote sensing observations. Although numerous vision-language models have been developed for pixel-level interpretation, spectral information remains underutilized,…

Computer Vision and Pattern Recognition · Computer Science 2026-03-10 Dongchen Si , Di Wang , Erzhong Gao , Xiaolei Qin , Liu Zhao , Jing Zhang , Minqiang Xu , Jianbo Zhan , Jianshe Wang , Lin Liu , Bo Du , Liangpei Zhang

In recent years, encoder-decoder networks have focused on expanding receptive fields and incorporating multi-scale context to capture global features for objects of varying sizes. However, as networks deepen, they often discard fine spatial…

Image and Video Processing · Electrical Eng. & Systems 2024-09-20 Xiaogang Du , Dongxin Gu , Tao Lei , Yipeng Jiao , Yibin Zou

Computational surface modeling that underlies material recognition has transitioned from reflectance modeling using in-lab controlled radiometric measurements to image-based representations based on internet-mined single-view images…

Computer Vision and Pattern Recognition · Computer Science 2020-09-24 Jia Xue , Hang Zhang , Ko Nishino , Kristin J. Dana

This paper presents GeoDecoder, a dedicated multimodal model designed for processing geospatial information in maps. Built on the BeitGPT architecture, GeoDecoder incorporates specialized expert modules for image and text processing. On the…

Computer Vision and Pattern Recognition · Computer Science 2024-02-20 Feng Qi , Mian Dai , Zixian Zheng , Chao Wang