中文
相关论文

相关论文: RingMoE: Mixture-of-Modality-Experts Multi-Modal F…

200 篇论文

Reliable channel estimation (CE) is fundamental for robust communication in dynamic wireless environments, where models must generalize across varying conditions such as signal-to-noise ratios (SNRs), the number of resource blocks (RBs),…

信号处理 · 电气工程与系统科学 2025-09-22 Tianyu Li , Yan Xin , Jianzhong , Zhang

Large language models like ChatGPT have shown substantial progress in natural language understanding and generation, proving valuable across various disciplines, including the medical field. Despite advancements, challenges persist due to…

计算与语言 · 计算机科学 2024-04-16 Yusheng Liao , Shuyang Jiang , Yu Wang , Yanfeng Wang

Mixture-of-Experts (MoE) architectures achieve scalable learning by routing inputs to specialized subnetworks through conditional computation. However, conventional MoE designs assume homogeneous expert capability and domain-agnostic…

Deep learning methods have significantly advanced the development of intelligent rinterpretation in remote sensing (RS), with foundational model research based on large-scale pre-training paradigms rapidly reshaping various domains of Earth…

计算机视觉与模式识别 · 计算机科学 2025-07-02 Zhiwei Yi , Xin Cheng , Jingyu Ma , Ruifei Zhu , Junwei Tian , Yuanxiu Zhou , Xinge Zhao , Hongzhe Li

Supervised fine-tuning (SFT) is a milestone in aligning large language models with human instructions and adapting them to downstream tasks. In particular, Low-Rank Adaptation (LoRA) has gained widespread attention due to its parameter…

计算与语言 · 计算机科学 2025-11-05 Jia-Chen Zhang , Yu-Jie Xiong , Xi-He Qiu , Chun-Ming Xia , Fei Dai , Zheng Zhou

Cross-modal ship re-identification (ReID) between optical and synthetic aperture radar (SAR) imagery has recently emerged as a critical yet underexplored task in maritime intelligence and surveillance. However, the substantial modality gap…

计算机视觉与模式识别 · 计算机科学 2025-12-04 Yujian Zhao , Hankun Liu , Guanglin Niu

Large-scale foundation models (FMs) in remote sensing (RS) are developed based on the paradigms established in computer vision (CV) and have shown promise for various Earth observation applications. However, the direct transfer of scaling…

计算机视觉与模式识别 · 计算机科学 2026-02-02 Leonard Hackel , Tom Burgert , Begüm Demir

Mixture-of-Experts (MoE) Transformer, the backbone architecture of multiple phenomenal language models, leverages sparsity by activating only a fraction of model parameters for each input token. The sparse structure, while allowing constant…

Radiological analysis increasingly benefits from pretrained visual representations that can support heterogeneous downstream tasks across imaging modalities. In this work, we introduce OmniRad, a self-supervised radiological foundation…

计算机视觉与模式识别 · 计算机科学 2026-02-05 Luca Zedda , Andrea Loddo , Cecilia Di Ruberto

Cross-modal artificial intelligence, represented by visual language models, has achieved significant success in general image understanding. However, a fundamental cognitive inconsistency exists between general visual representation and…

计算机视觉与模式识别 · 计算机科学 2026-01-26 Yi Yang , Xiaokun Zhang , Qingchen Fang , Jing Liu , Ziqi Ye , Rui Li , Li Liu , Haipeng Wang

In the realm of geospatial analysis, the diversity of remote sensors, encompassing both optical and microwave technologies, offers a wealth of distinct observational capabilities. Recognizing this, we present msGFM, a multisensor geospatial…

计算机视觉与模式识别 · 计算机科学 2024-04-02 Boran Han , Shuai Zhang , Xingjian Shi , Markus Reichstein

Recent advancements in multimodal large language models (MLLMs) have demonstrated considerable potential for comprehensive 3D scene understanding. However, existing approaches typically utilize only one or a limited subset of 3D modalities,…

计算机视觉与模式识别 · 计算机科学 2025-05-28 Yue Zhang , Yingzhao Jian , Hehe Fan , Yi Yang , Roger Zimmermann

As satellite networks evolve to support increasingly diverse services and artificial general intelligence (AGI), large language models (LLMs) are emerging as a critical foundation for future space systems. However, deploying LLMs on…

网络与互联网体系结构 · 计算机科学 2026-05-19 Qian Chen , Xianhao Chen , Min Sheng , Kaibin Huang

Recent advances in Earth Observation have focused on large-scale foundation models. However, these models are computationally expensive, limiting their accessibility and reuse for downstream tasks. In this work, we investigate compact…

计算机视觉与模式识别 · 计算机科学 2025-09-16 Mohanad Albughdadi

We explore the scaling behaviors of artificial intelligence to establish practical techniques for training foundation models on high-resolution electro-optical (EO) datasets that exceed the current state-of-the-art scale by orders of…

计算机视觉与模式识别 · 计算机科学 2026-01-15 Charith Wickrema , Eliza Mace , Hunter Brown , Heidys Cabrera , Nick Krall , Matthew O'Neill , Shivangi Sarkar , Lowell Weissman , Eric Hughes , Guido Zarrella

Recent studies show that using diffusion models for time series signal reconstruction holds great promise. However, such approaches remain largely unexplored in the domain of medical time series. The unique characteristics of the…

机器学习 · 计算机科学 2026-01-13 Ci Zhang , Huayu Li , Changdi Yang , Jiangnan Xia , Yanzhi Wang , Xiaolong Ma , Jin Lu , Ao Li , Geng Yuan

Fine-tuning pre-trained large language models (LLMs) presents a dual challenge of balancing parameter efficiency and model capacity. Existing methods like low-rank adaptations (LoRA) are efficient but lack flexibility, while…

While Dense Retrieval Models (DRMs) have advanced Information Retrieval (IR), one limitation of these neural models is their narrow generalizability and robustness. To cope with this issue, one can leverage the Mixture-of-Experts (MoE)…

信息检索 · 计算机科学 2024-12-17 Effrosyni Sokli , Pranav Kasela , Georgios Peikos , Gabriella Pasi

Steered-Mixtures-of Experts (SMoE) present a unified framework for sparse representation and compression of image data with arbitrary dimensionality. Recent work has shown great improvements in the performance of such models for image and…

图像与视频处理 · 电气工程与系统科学 2022-09-14 Rolf Jongebloed , Erik Bochinski , Thomas Sikora

Anomaly detection is a critical task across numerous domains and modalities, yet existing methods are often highly specialized, limiting their generalizability. These specialized models, tailored for specific anomaly types like textural…

计算机视觉与模式识别 · 计算机科学 2025-08-11 Zhaopeng Gu , Bingke Zhu , Guibo Zhu , Yingying Chen , Wei Ge , Ming Tang , Jinqiao Wang