中文
相关论文

相关论文: DEDUCE: Multi-head attention decoupled contrastive…

200 篇论文

We propose AttentionMixer, a unified deep learning framework for multimodal detection of brain edema that combines structural head CT (HCT) with routine clinical metadata. While HCT provides rich spatial information, clinical variables such…

For personalized medicines, very crucial intrinsic information is present in high dimensional omics data which is difficult to capture due to the large number of molecular features and small number of available samples. Different types of…

机器学习 · 计算机科学 2022-02-04 Sayed Hashim , Muhammad Ali , Karthik Nandakumar , Mohammad Yaqub

Brain tumor is a common and fatal form of cancer which affects both adults and children. The classification of brain tumors into different types is hence a crucial task, as it greatly influences the treatment that physicians will prescribe.…

计算机视觉与模式识别 · 计算机科学 2021-08-10 Tian Yu Liu , Jiashi Feng

The human brain can easily focus on one speaker and suppress others in scenarios such as a cocktail party. Recently, researchers found that auditory attention can be decoded from the electroencephalogram (EEG) data. However, most existing…

声音 · 计算机科学 2023-08-09 Xiaoyu Chen , Changde Du , Qiongyi Zhou , Huiguang He

Building accurate and robust artificial intelligence systems for medical image assessment requires not only the research and design of advanced deep learning models but also the creation of large and curated sets of annotated training…

Data fusion enables powerful and generalizable analyses across multiple sources. However, different data collection capacities across different sources lead to blockwise missingness (BM), which poses challenges in practice. Meanwhile, the…

统计方法学 · 统计学 2025-12-09 Yiming Li , Ying Wei , Molei Liu

In recent years, there has been a notable increase in the use of supervised detection methods of major depressive disorder (MDD) based on electroencephalogram (EEG) signals. However, the process of labeling MDD remains challenging. As a…

机器学习 · 计算机科学 2025-12-17 Li-Xuan Zhao , Chen-Yang Xu , Wen-Qiang Li , Bo Wang , Rong-Xing Wei , Qing-Hao Menga

Traditional supervised learning with deep neural networks requires a tremendous amount of labelled data to converge to a good solution. For 3D medical images, it is often impractical to build a large homogeneous annotated dataset for a…

Accurately predicting the future trajectories of traffic agents is essential in autonomous driving. However, due to the inherent imbalance in trajectory distributions, tail data in natural datasets often represents more complex and…

计算机视觉与模式识别 · 计算机科学 2025-07-08 Bin Rao , Haicheng Liao , Yanchen Guan , Chengyue Wang , Bonan Wang , Jiaxun Zhang , Zhenning Li

Motivation: Cancer is heterogeneous, affecting the precise approach to personalized treatment. Accurate subtyping can lead to better survival rates for cancer patients. High-throughput technologies provide multiple omics data for cancer…

机器学习 · 计算机科学 2022-08-01 Hai Yang , Yuhang Sheng , Yi Jiang , Xiaoyang Fang , Dongdong Li , Jing Zhang , Zhe Wang

Multimodal survival prediction, a crucial yet challenging task, demands the integration of multimodal medical data (\eg Whole Slide Images (WSIs) and Genomic Profiles) to achieve accurate prognostic modeling. Given the inherent…

计算机视觉与模式识别 · 计算机科学 2026-05-21 Huayi Wang , Haochao Ying , Yuyang Xu , Qiyao Zheng , jun wang , Cheng Zhang , Ying Sun , Jian Wu

Accurately detecting Alzheimer's disease (AD) and predicting mini-mental state examination (MMSE) score are important tasks in elderly health by magnetic resonance imaging (MRI). Most of the previous methods on these two tasks are based on…

图像与视频处理 · 电气工程与系统科学 2023-07-10 Xu Tian , Jin Liu , Hulin Kuang , Yu Sheng , Jianxin Wang , The Alzheimer's Disease Neuroimaging Initiative

Contrastive Language-Image Pre-training (CLIP) has become a cornerstone in multimodal intelligence. However, recent studies discovered that CLIP can only encode one aspect of the feature space, leading to substantial information loss and…

计算机视觉与模式识别 · 计算机科学 2025-05-29 Jihai Zhang , Xiaoye Qu , Tong Zhu , Yu Cheng

Recent advancements in image classification have demonstrated that contrastive learning (CL) can aid in further learning tasks by acquiring good feature representation from a limited number of data samples. In this paper, we applied CL to…

机器学习 · 计算机科学 2024-10-22 Anchen Sun , Elizabeth J. Franzmann , Zhibin Chen , Xiaodong Cai

Acoustic Word Embeddings (AWEs) improve the efficiency of speech retrieval tasks such as Spoken Term Detection (STD) and Keyword Spotting (KWS). However, existing approaches suffer from limitations, including unimodal supervision, disjoint…

声音 · 计算机科学 2025-12-17 Ramesh Gundluru , Shubham Gupta , Sri Rama Murty K

Multi-view feature extraction is an efficient approach for alleviating the issue of dimensionality in highdimensional multi-view data. Contrastive learning (CL), which is a popular self-supervised learning method, has recently attracted…

计算机视觉与模式识别 · 计算机科学 2023-02-09 Hongjie Zhang

For early breast cancer detection, regular screening with mammography imaging is recommended. Routinary examinations result in datasets with a predominant amount of negative samples. A potential solution to such class-imbalance is joining…

计算机视觉与模式识别 · 计算机科学 2023-01-09 Amelia Jiménez-Sánchez , Mickael Tardy , Miguel A. González Ballester , Diana Mateus , Gemma Piella

Time series self-supervised learning (SSL) aims to exploit unlabeled data for pre-training to mitigate the reliance on labels. Despite the great success in recent years, there is limited discussion on the potential noise in the time series,…

机器学习 · 计算机科学 2024-06-10 Shuang Zhou , Daochen Zha , Xiao Shen , Xiao Huang , Rui Zhang , Fu-Lai Chung

Cross-corpus speech emotion recognition (SER) seeks to generalize the ability of inferring speech emotion from a well-labeled corpus to an unlabeled one, which is a rather challenging task due to the significant discrepancy between two…

声音 · 计算机科学 2023-08-07 Jiaxin Ye , Yujie Wei , Xin-Cheng Wen , Chenglong Ma , Zhizhong Huang , Kunhong Liu , Hongming Shan

State-of-the-art pre-trained image models predominantly adopt a two-stage approach: initial unsupervised pre-training on large-scale datasets followed by task-specific fine-tuning using Cross-Entropy loss~(CE). However, it has been…

计算机视觉与模式识别 · 计算机科学 2024-11-18 Zijun Long , George Killick , Lipeng Zhuang , Gerardo Aragon-Camarasa , Zaiqiao Meng , Richard Mccreadie