English
Related papers

Related papers: Clin-JEPA: A Multi-Phase Co-Training Framework for…

200 papers

Geospatial foundation models compress multispectral observations into dense embeddings increasingly used in natural-language environmental reasoning systems. A single planetary-scale model, e.g. Google AlphaEarth, handles broad…

Machine Learning · Computer Science 2026-05-15 Mashrekur Rahman

A core strength of Model Predictive Control (MPC) for quadrupedal locomotion has been its ability to enforce constraints and provide interpretability of the sequence of commands over the horizon. However, despite being able to plan, MPC…

Robotics · Computer Science 2025-04-16 Aditya Shirwatkar , Naman Saxena , Kishore Chandra , Shishir Kolathaya

Deep learning models have shown promise in EEG-based outcome prediction for comatose patients after cardiac arrest, but their reliability is often compromised by subtle forms of data leakage. In particular, when long EEG recordings are…

Machine Learning · Computer Science 2026-03-30 Yixin Zhou , Zhixiang Liu , Vladimir I. Zadorozhny , Jonathan Elmer

We address the Continual Learning (CL) problem, wherein a model must learn a sequence of tasks from non-stationary distributions while preserving prior knowledge upon encountering new experiences. With the advancement of foundation models,…

Machine Learning · Computer Science 2024-07-08 Kyra Ahrens , Hans Hergen Lehmann , Jae Hee Lee , Stefan Wermter

Learning from multi-variate time-series with heterogeneous channel configurations remains a fundamental challenge for deep neural networks, particularly in clinical domains such as intracranial electroencephalography (iEEG), where channel…

Machine Learning · Computer Science 2025-08-26 Francesco Carzaniga , Michael Hersche , Abu Sebastian , Kaspar Schindler , Abbas Rahimi

Conventional object detectors rely on cross-entropy classification, which can be vulnerable to class imbalance and label noise. We propose CLIP-Joint-Detect, a simple and detector-agnostic framework that integrates CLIP-style contrastive…

Computer Vision and Pattern Recognition · Computer Science 2025-12-30 Behnam Raoufi , Hossein Sharify , Mohamad Mahdee Ramezanee , Khosrow Hajsadeghi , Saeed Bagheri Shouraki

Video-language foundation models have proven to be highly effective in zero-shot applications across a wide range of tasks. A particularly challenging area is the intraoperative surgical procedure domain, where labeled data is scarce, and…

Computer Vision and Pattern Recognition · Computer Science 2026-03-26 Florian Stilz , Vinkle Srivastav , Nassir Navab , Nicolas Padoy

We propose a novel multimodal deep learning framework for patient-level survival prediction, which integrates whole-slide histology features, RNA-seq expression profiles, and clinical variables. Our architecture combines an ABMIL…

Quantitative Methods · Quantitative Biology 2026-05-15 Hassan Keshvarikhojasteh , Josien P. W. Pluim , Mitko Veta

Fine-grained medical image classification is challenged by subtle inter-class variations and visually ambiguous cases, where confidence estimates often exhibit uncertainty rather than being overconfident. In such scenarios, purely…

Artificial Intelligence · Computer Science 2026-04-21 Zixuan Tang , Shen Zhao

Pre-trained vision-language (V-L) models such as CLIP have shown excellent performance in many downstream cross-modal tasks. However, most of them are only applicable to the English context. Subsequent research has focused on this problem…

Computer Vision and Pattern Recognition · Computer Science 2024-04-18 Wenbo Zhang , Yifan Zhang , Jianfeng Lin , Binqiang Huang , Jinlu Zhang , Wenhao Yu

A long-standing challenge in AI is to develop agents capable of solving a wide range of physical tasks and generalizing to new, unseen tasks and environments. A popular recent approach involves training a world model from state-action…

Artificial Intelligence · Computer Science 2026-05-19 Basile Terver , Tsung-Yen Yang , Jean Ponce , Adrien Bardes , Yann LeCun

Tool-augmented multimodal reasoning enables visual language models (VLMs) to improve perception by interacting with external tools (e.g., cropping, depth estimation). However, such approaches incur substantial inference overhead, require…

Machine Learning · Computer Science 2026-04-10 Ashutosh Adhikari , Mirella Lapata

Spatio-temporal epidemic forecasting is critical for public health management, yet existing methods often struggle with insensitivity to weak epidemic signals, over-simplified spatial relations, and unstable parameter estimation. To address…

Machine Learning · Computer Science 2026-05-22 Sijie Ruan , Jinyu Li , Jia Wei , Zenghao Xu , Jie Bao , Junshi Xu , Junyang Qiu , Shuliang Wang , Xiaoxiao Wang , Hanning Yuan

Vision-language-action (VLA) models have significantly advanced robotic learning, enabling training on large-scale, cross-embodiment data and fine-tuning for specific robots. However, state-of-the-art autoregressive VLAs struggle with…

Robotics · Computer Science 2025-11-04 Chengmeng Li , Yaxin Peng

Autonomous drone fleets have immense potential in medical supply delivery during disaster incident response. However, coordinating multiple drones in such settings introduces compounding challenges: dynamic environmental hazards such as…

Multiagent Systems · Computer Science 2026-05-12 Aneesh Calyam , Subrahmanya Chandra Bhamidipati , Zack Murry , Sharan Srinivas

The increasing availability of big mobility data from ubiquitous portable devices enables human mobility prediction through deep learning approaches. However, the diverse complexity of human mobility data impedes model training, leading to…

Machine Learning · Computer Science 2026-03-10 Tianye Fang , Xuanshu Luo , Martin Werner

Electroencephalography (EEG) foundation models have shown promise for learning generalizable representations, yet they remain sensitive to channel heterogeneity, such as changes in channel composition or ordering. We propose channel-aware…

Machine Learning · Computer Science 2026-03-17 Hanseul Choi , Jinyeong Park , Seongwon Jin , Sungho Park , Jibum Kim

In real-world applications, not all instances in multi-view data are fully represented. To deal with incomplete data, Incomplete Multi-view Learning (IML) rises. In this paper, we propose the Joint Embedding Learning and Low-Rank…

Machine Learning · Computer Science 2019-12-17 Hong Tao , Chenping Hou , Dongyun Yi , Jubo Zhu , Dewen Hu

We present a pipeline for unbiased and robust multimodal registration of neuroimaging modalities with minimal pre-processing. While typical multimodal studies need to use multiple independent processing pipelines, with diverse options and…

Computer Vision and Pattern Recognition · Computer Science 2024-01-26 Adria Casamitjana , Juan Eugenio Iglesias , Raul Tudela , Aida Ninerola-Baizan , Roser Sala-Llonch

Epileptic seizure prediction from electroencephalographic (EEG) recordings remains challenging due to strong inter-patient variability and the complex temporal structure of neural signals. This paper presents a patient-adaptive transformer…

Machine Learning · Computer Science 2026-03-31 Mohamed Mahdi , Asma Baghdadi
‹ Prev 1 8 9 10 Next ›