中文
相关论文

相关论文: Vision Foundation Models in Remote Sensing: A Surv…

200 篇论文

Building multisensory AI systems that learn from multiple sensory inputs such as text, speech, video, real-world sensors, wearable devices, and medical data holds great promise for impact in many scientific areas with practical benefits,…

机器学习 · 计算机科学 2024-05-01 Paul Pu Liang

The rise of large foundation models, trained on extensive datasets, is revolutionizing the field of AI. Models such as SAM, DALL-E2, and GPT-4 showcase their adaptability by extracting intricate patterns and performing effectively across…

计算机视觉与模式识别 · 计算机科学 2024-01-17 Xu Yan , Haiming Zhang , Yingjie Cai , Jingming Guo , Weichao Qiu , Bin Gao , Kaiqiang Zhou , Yue Zhao , Huan Jin , Jiantao Gao , Zhen Li , Lihui Jiang , Wei Zhang , Hongbo Zhang , Dengxin Dai , Bingbing Liu

While Wi-Fi sensing offers a compelling, privacy-preserving alternative to cameras, its practical utility has been fundamentally undermined by a lack of robustness across domains. Models trained in one setup fail to generalize to new…

计算机视觉与模式识别 · 计算机科学 2025-11-25 Cheng Jiang , Yihe Yan , Yanxiang Wang , Chun Tung Chou , Wen Hu

In recent years, black-box machine learning approaches have become a dominant modeling paradigm for knowledge extraction in remote sensing. Despite the potential benefits of uncovering the inner workings of these models with explainable AI,…

Spatio-temporal deep learning models aims to utilize useful patterns in such data to support tasks like prediction. However, previous deep learning models designed for specific tasks typically require separate training for each use case,…

Advances in foundation modeling have reshaped computational pathology. However, the increasing number of available models and lack of standardized benchmarks make it increasingly complex to assess their strengths, limitations, and potential…

计算机视觉与模式识别 · 计算机科学 2025-02-11 Andrew Zhang , Guillaume Jaume , Anurag Vaidya , Tong Ding , Faisal Mahmood

Artificial Intelligence (AI)-based radio fingerprinting (FP) outperforms classic localization methods in propagation environments with strong multipath effects. However, the model and data orchestration of FP are time-consuming and costly,…

信号处理 · 电气工程与系统科学 2024-10-02 Jonathan Ott , Jonas Pirkl , Maximilian Stahlke , Tobias Feigl , Christopher Mutschler

We envision the "virtual eye" as a next-generation, AI-powered platform that uses interconnected foundation models to simulate the eye's intricate structure and biological function across all scales. Advances in AI, imaging, and multiomics…

组织与器官 · 定量生物学 2025-05-12 Yue Wu , Yibo Guo , Yulong Yan , Jiancheng Yang , Xin Zhou , Ching-Yu Cheng , Danli Shi , Mingguang He

Vision foundation models are a new frontier in Geospatial Artificial Intelligence (GeoAI), an interdisciplinary research area that applies and extends AI for geospatial problem solving and geographic knowledge discovery, because of their…

计算机视觉与模式识别 · 计算机科学 2024-05-29 Wenwen Li , Hyunho Lee , Sizhe Wang , Chia-Yu Hsu , Samantha T. Arundel

The task of identifying and segmenting buildings within remote sensing imagery has perennially stood at the forefront of scholarly investigations. This manuscript accentuates the potency of harnessing diversified datasets in tandem with…

计算机视觉与模式识别 · 计算机科学 2023-10-27 Lei Li

The development of modern Artificial Intelligence (AI) models, particularly diffusion-based models employed in computer vision and image generation tasks, is undergoing a paradigmatic shift in development methodologies. Traditionally…

机器学习 · 计算机科学 2025-06-13 Sajjad Abdoli , Freeman Lewin , Gediminas Vasiliauskas , Fabian Schonholz

With advancements in GPS, remote sensing, and computational simulation, an enormous volume of spatiotemporal data is being collected at an increasing speed from various application domains, spanning Earth sciences, agriculture, smart…

机器学习 · 计算机科学 2023-11-01 Zhe Jiang

The advances in remote sensing technologies have boosted applications for Earth observation. These technologies provide multiple observations or views with different levels of information. They might contain static or temporary views with…

计算机视觉与模式识别 · 计算机科学 2024-02-06 Francisco Mena , Diego Arenas , Marlon Nuske , Andreas Dengel

Deep learning has largely reshaped remote sensing (RS) research for aerial image understanding and made a great success. Nevertheless, most of the existing deep models are initialized with the ImageNet pretrained weights. Since natural…

计算机视觉与模式识别 · 计算机科学 2023-07-19 Di Wang , Jing Zhang , Bo Du , Gui-Song Xia , Dacheng Tao

Robotic foundation models (RFMs) are emerging as a promising route towards flexible, instruction- and demonstration-driven robot control, however, a critical investigation of their industrial applicability is still lacking. This survey…

机器人学 · 计算机科学 2026-03-10 David Kube , Simon Hadwiger , Tobias Meisen

Deep learning has taken by storm all fields involved in data analysis, including remote sensing for Earth observation. However, despite significant advances in terms of performance, its lack of explainability and interpretability, inherent…

人工智能 · 计算机科学 2023-11-09 Gulsen Taskin , Erchan Aptoula , Alp Ertürk

With the rise in high resolution remote sensing technologies there has been an explosion in the amount of data available for forest monitoring, and an accompanying growth in artificial intelligence applications to automatically derive…

Effective foundation modeling in remote sensing requires spatially aligned heterogeneous modalities coupled with semantically grounded supervision, yet such resources remain limited at scale. We present GeoMeld, a large-scale multimodal…

Artificial intelligence (AI) is vital in ophthalmology, tackling tasks like diagnosis, classification, and visual question answering (VQA). However, existing AI models in this domain often require extensive annotation and are task-specific,…

计算机视觉与模式识别 · 计算机科学 2024-05-24 Danli Shi , Weiyi Zhang , Xiaolan Chen , Yexin Liu , Jiancheng Yang , Siyu Huang , Yih Chung Tham , Yingfeng Zheng , Mingguang He

We present VisionFM, a foundation model pre-trained with 3.4 million ophthalmic images from 560,457 individuals, covering a broad range of ophthalmic diseases, modalities, imaging devices, and demography. After pre-training, VisionFM…

‹ 上一页 1 8 9 10 下一页 ›