中文
相关论文

相关论文: ADAPT: Multimodal Learning for Detecting Physiolog…

200 篇论文

In high-dimensional time series, the component processes are often assembled into a matrix to display their interrelationship. We focus on detecting mean shifts with unknown change point locations in these matrix time series. Series that…

统计方法学 · 统计学 2024-07-16 Xinyu Zhang , Kung-Sik Chan

Psychological distress is a significant and growing issue in society. Automatic detection, assessment, and analysis of such distress is an active area of research. Compared to modalities such as face, head, and vocal, research investigating…

计算机视觉与模式识别 · 计算机科学 2020-08-03 Weizhe Lin , Indigo Orton , Qingbiao Li , Gabriela Pavarini , Marwa Mahmoud

Recent progress in multi-modal conditioned face synthesis has enabled the creation of visually striking and accurately aligned facial images. Yet, current methods still face issues with scalability, limited flexibility, and a…

计算机视觉与模式识别 · 计算机科学 2024-03-22 Jingjing Ren , Cheng Xu , Haoyu Chen , Xinran Qin , Lei Zhu

This work addresses challenges in evaluating adaptive artificial intelligence (AI) models for medical devices, where iterative updates to both models and evaluation datasets complicate performance assessment. We introduce a novel approach…

人工智能 · 计算机科学 2026-04-07 Alexis Burgon , Berkman Sahiner , Nicholas A Petrick , Gene Pennello , Ravi K Samala

Many emerging applications of intelligent robots need to explore and understand new environments, where it is desirable to detect objects of novel classes on the fly with minimum online efforts. This is an object detection on demand (ODOD)…

计算机视觉与模式识别 · 计算机科学 2021-10-12 Xiangyun Zhao , Xu Zou , Ying Wu

Accurate recognition of human emotions is a crucial challenge in affective computing and human-robot interaction (HRI). Emotional states play a vital role in shaping behaviors, decisions, and social interactions. However, emotional…

机器人学 · 计算机科学 2024-09-19 Youssef Mohamed , Severin Lemaignan , Arzu Guneysu , Patric Jensfelt , Christian Smith

Accurate classification of medical device risk levels is essential for regulatory oversight and clinical safety. We present a Transformer-based multimodal framework that integrates textual descriptions and visual information to predict…

机器学习 · 计算机科学 2025-05-02 Yu Han , Aaron Ceross , Jeroen H. M. Bergmann

Federated learning is a decentralized training approach that keeps data under stakeholder control while achieving superior performance over isolated training. While inter-institutional feature discrepancies pose a challenge in all federated…

图像与视频处理 · 电气工程与系统科学 2025-07-01 Vasilis Siomos , Jonathan Passerat-Palmbach , Giacomo Tarroni

Assistive teleoperation enhances efficiency via shared control, yet inter-operator variability, stemming from diverse habits and expertise, induces highly heterogeneous trajectory distributions that undermine intent recognition stability.…

机器人学 · 计算机科学 2026-04-13 Yu Liu , Yihang Yin , Tianlv Huang , Fei Yan , Yuan Xu , Weinan Hong , Wei Han , Yue Cao , Xiangyu Chen , Zipei Fan , Xuan Song

Medical data poses a daunting challenge for AI algorithms: it exists in many different modalities, experiences frequent distribution shifts, and suffers from a scarcity of examples and labels. Recent advances, including transformers and…

Multimodal learning enables neural networks to integrate information from heterogeneous sources, but active learning in this setting faces distinct challenges. These include missing modalities, differences in modality difficulty, and…

机器学习 · 计算机科学 2026-04-01 Dustin Eisenhardt , Yunhee Jeong , Florian Buettner

In this paper, we introduce a novel Multi-Modal Contrastive Pre-training Framework that synergistically combines X-rays, electrocardiograms (ECGs), and radiology/cardiology reports. Our approach leverages transformers to encode these…

人工智能 · 计算机科学 2024-10-24 Samrajya Thapa , Koushik Howlader , Subhankar Bhattacharjee , Wei le

Modern multimodal systems deployed in industrial and safety-critical environments must remain reliable under partial sensor failures, signal degradation, or cross-modal inconsistencies. This work introduces a mathematically grounded…

机器学习 · 计算机科学 2026-03-27 Diyar Altinses , Andreas Schwung

Agile humanoid locomotion in complex 3D en- vironments requires balancing perceptual fidelity with com- putational efficiency, yet existing methods typically rely on rigid sensing configurations. We propose ADAPT (Adaptive dual-projection…

机器人学 · 计算机科学 2026-03-18 Shuo Shao , Tianchen Huang , Wei Gao , Shiwu Zhang

Recent multi-modal face anti-spoofing (FAS) methods have investigated the potential of leveraging multiple modalities to distinguish live and spoof faces. However, pre-adapted multi-modal FAS models often fail to detect unseen attacks from…

计算机视觉与模式识别 · 计算机科学 2025-09-30 Ming-Tsung Hsu , Fang-Yu Hsu , Yi-Ting Lin , Kai-Heng Chien , Jun-Ren Chen , Cheng-Hsiang Su , Yi-Chen Ou , Chiou-Ting Hsu , Pei-Kai Huang

Adaptive control for real-time manipulation requires quick estimation and prediction of object properties. While robot learning in this area primarily focuses on using vision, many tasks cannot rely on vision due to object occlusion. Here,…

机器人学 · 计算机科学 2021-10-12 Ahalya Prabhakar , Stanislas Furrer , Lorenzo Panchetti , Maxence Perret , Aude Billard

The ability to modify morphology in response to environmental changes represents a highly advantageous feature in biological organisms, facilitating their adaptation to diverse environmental conditions. While some robots have the capability…

机器人学 · 计算机科学 2023-09-20 Saurav Kumar Dutta , Yasemin Ozkan-Aydin

Unpaired medical image synthesis aims to provide complementary information for an accurate clinical diagnostics, and address challenges in obtaining aligned multi-modal medical scans. Transformer-based models excel in imaging translation…

计算机视觉与模式识别 · 计算机科学 2024-08-29 Vu Minh Hieu Phan , Yutong Xie , Bowen Zhang , Yuankai Qi , Zhibin Liao , Antonios Perperidis , Son Lam Phung , Johan W. Verjans , Minh-Son To

Multi-modal 3D object detection is pivotal for autonomous driving, integrating complementary sensors like LiDAR and cameras. However, its real-world reliability is challenged by transient data interruptions and missing, where modalities can…

计算机视觉与模式识别 · 计算机科学 2026-03-04 Shuangzhi Li , Lei Ma , Xingyu Li

Deploying multimodal systems in real-world environments often entails handling modality-missing scenarios, where one or more modalities are unavailable. While recent studies address this challenge for the general Multimodal Transformer (MT)…

计算机视觉与模式识别 · 计算机科学 2026-05-26 Jian Lang , Rongpei Hong , Ting Zhong , Fan Zhou