中文
相关论文

相关论文: Clin-JEPA: A Multi-Phase Co-Training Framework for…

200 篇论文

Geospatial foundation models compress multispectral observations into dense embeddings increasingly used in natural-language environmental reasoning systems. A single planetary-scale model, e.g. Google AlphaEarth, handles broad…

机器学习 · 计算机科学 2026-05-15 Mashrekur Rahman

A core strength of Model Predictive Control (MPC) for quadrupedal locomotion has been its ability to enforce constraints and provide interpretability of the sequence of commands over the horizon. However, despite being able to plan, MPC…

机器人学 · 计算机科学 2025-04-16 Aditya Shirwatkar , Naman Saxena , Kishore Chandra , Shishir Kolathaya

Deep learning models have shown promise in EEG-based outcome prediction for comatose patients after cardiac arrest, but their reliability is often compromised by subtle forms of data leakage. In particular, when long EEG recordings are…

机器学习 · 计算机科学 2026-03-30 Yixin Zhou , Zhixiang Liu , Vladimir I. Zadorozhny , Jonathan Elmer

We address the Continual Learning (CL) problem, wherein a model must learn a sequence of tasks from non-stationary distributions while preserving prior knowledge upon encountering new experiences. With the advancement of foundation models,…

机器学习 · 计算机科学 2024-07-08 Kyra Ahrens , Hans Hergen Lehmann , Jae Hee Lee , Stefan Wermter

Learning from multi-variate time-series with heterogeneous channel configurations remains a fundamental challenge for deep neural networks, particularly in clinical domains such as intracranial electroencephalography (iEEG), where channel…

机器学习 · 计算机科学 2025-08-26 Francesco Carzaniga , Michael Hersche , Abu Sebastian , Kaspar Schindler , Abbas Rahimi

Conventional object detectors rely on cross-entropy classification, which can be vulnerable to class imbalance and label noise. We propose CLIP-Joint-Detect, a simple and detector-agnostic framework that integrates CLIP-style contrastive…

计算机视觉与模式识别 · 计算机科学 2025-12-30 Behnam Raoufi , Hossein Sharify , Mohamad Mahdee Ramezanee , Khosrow Hajsadeghi , Saeed Bagheri Shouraki

Video-language foundation models have proven to be highly effective in zero-shot applications across a wide range of tasks. A particularly challenging area is the intraoperative surgical procedure domain, where labeled data is scarce, and…

计算机视觉与模式识别 · 计算机科学 2026-03-26 Florian Stilz , Vinkle Srivastav , Nassir Navab , Nicolas Padoy

We propose a novel multimodal deep learning framework for patient-level survival prediction, which integrates whole-slide histology features, RNA-seq expression profiles, and clinical variables. Our architecture combines an ABMIL…

定量方法 · 定量生物学 2026-05-15 Hassan Keshvarikhojasteh , Josien P. W. Pluim , Mitko Veta

Fine-grained medical image classification is challenged by subtle inter-class variations and visually ambiguous cases, where confidence estimates often exhibit uncertainty rather than being overconfident. In such scenarios, purely…

人工智能 · 计算机科学 2026-04-21 Zixuan Tang , Shen Zhao

Pre-trained vision-language (V-L) models such as CLIP have shown excellent performance in many downstream cross-modal tasks. However, most of them are only applicable to the English context. Subsequent research has focused on this problem…

计算机视觉与模式识别 · 计算机科学 2024-04-18 Wenbo Zhang , Yifan Zhang , Jianfeng Lin , Binqiang Huang , Jinlu Zhang , Wenhao Yu

A long-standing challenge in AI is to develop agents capable of solving a wide range of physical tasks and generalizing to new, unseen tasks and environments. A popular recent approach involves training a world model from state-action…

人工智能 · 计算机科学 2026-05-19 Basile Terver , Tsung-Yen Yang , Jean Ponce , Adrien Bardes , Yann LeCun

Tool-augmented multimodal reasoning enables visual language models (VLMs) to improve perception by interacting with external tools (e.g., cropping, depth estimation). However, such approaches incur substantial inference overhead, require…

机器学习 · 计算机科学 2026-04-10 Ashutosh Adhikari , Mirella Lapata

Spatio-temporal epidemic forecasting is critical for public health management, yet existing methods often struggle with insensitivity to weak epidemic signals, over-simplified spatial relations, and unstable parameter estimation. To address…

机器学习 · 计算机科学 2026-05-22 Sijie Ruan , Jinyu Li , Jia Wei , Zenghao Xu , Jie Bao , Junshi Xu , Junyang Qiu , Shuliang Wang , Xiaoxiao Wang , Hanning Yuan

Vision-language-action (VLA) models have significantly advanced robotic learning, enabling training on large-scale, cross-embodiment data and fine-tuning for specific robots. However, state-of-the-art autoregressive VLAs struggle with…

机器人学 · 计算机科学 2025-11-04 Chengmeng Li , Yaxin Peng

Autonomous drone fleets have immense potential in medical supply delivery during disaster incident response. However, coordinating multiple drones in such settings introduces compounding challenges: dynamic environmental hazards such as…

多智能体系统 · 计算机科学 2026-05-12 Aneesh Calyam , Subrahmanya Chandra Bhamidipati , Zack Murry , Sharan Srinivas

The increasing availability of big mobility data from ubiquitous portable devices enables human mobility prediction through deep learning approaches. However, the diverse complexity of human mobility data impedes model training, leading to…

机器学习 · 计算机科学 2026-03-10 Tianye Fang , Xuanshu Luo , Martin Werner

Electroencephalography (EEG) foundation models have shown promise for learning generalizable representations, yet they remain sensitive to channel heterogeneity, such as changes in channel composition or ordering. We propose channel-aware…

机器学习 · 计算机科学 2026-03-17 Hanseul Choi , Jinyeong Park , Seongwon Jin , Sungho Park , Jibum Kim

In real-world applications, not all instances in multi-view data are fully represented. To deal with incomplete data, Incomplete Multi-view Learning (IML) rises. In this paper, we propose the Joint Embedding Learning and Low-Rank…

机器学习 · 计算机科学 2019-12-17 Hong Tao , Chenping Hou , Dongyun Yi , Jubo Zhu , Dewen Hu

We present a pipeline for unbiased and robust multimodal registration of neuroimaging modalities with minimal pre-processing. While typical multimodal studies need to use multiple independent processing pipelines, with diverse options and…

计算机视觉与模式识别 · 计算机科学 2024-01-26 Adria Casamitjana , Juan Eugenio Iglesias , Raul Tudela , Aida Ninerola-Baizan , Roser Sala-Llonch

Epileptic seizure prediction from electroencephalographic (EEG) recordings remains challenging due to strong inter-patient variability and the complex temporal structure of neural signals. This paper presents a patient-adaptive transformer…

机器学习 · 计算机科学 2026-03-31 Mohamed Mahdi , Asma Baghdadi
‹ 上一页 1 8 9 10 下一页 ›