中文
相关论文

相关论文: Adaptive Clinical-Aware Latent Diffusion for Multi…

200 篇论文

Brain-to-image decoding has been recently propelled by the progress in generative AI models and the availability of large ultra-high field functional Magnetic Resonance Imaging (fMRI). However, current approaches depend on complicated…

计算机视觉与模式识别 · 计算机科学 2025-05-21 Marlène Careil , Yohann Benchetrit , Jean-Rémi King

Autism spectrum disorder (ASD) is a complex neurodevelopmental condition characterized by atypical functional brain connectivity and subtle structural alterations. rs-fMRI has been widely used to identify disruptions in large-scale brain…

计算机视觉与模式识别 · 计算机科学 2026-03-17 Ansar Rahman , Hassan Shojaee-Mend , Sepideh Hatamikia

This paper presents a general framework for obtaining interpretable multivariate discriminative models that allow efficient statistical inference for neuroimage analysis. The framework, termed generative discriminative machine (GDM),…

应用统计 · 统计学 2019-06-04 Erdem Varol , Aristeidis Sotiras , Ke Zeng , Christos Davatzikos

Automated diagnosis of Alzheimer Disease(AD) from brain imaging, such as magnetic resonance imaging (MRI), has become increasingly important and has attracted the community to contribute many deep learning methods. However, many of these…

图像与视频处理 · 电气工程与系统科学 2024-03-01 Yifeng Wang , Ke Chen , Haohan Wang

Multimodal clinical prediction faces three challenges: multiple foundation models (FMs) with complementary strengths per modality, pervasive missing modalities at training and test time, and sample-specific variation in modality…

机器学习 · 计算机科学 2026-05-19 Seungik Cho , Anqi Li , Wei Qiu

Industrial anomaly detection (IAD) increasingly benefits from integrating 2D and 3D data, but robust cross-modal fusion remains challenging. We propose a novel unsupervised framework, Multi-Modal Attention-Driven Fusion Restoration (MAFR),…

计算机视觉与模式识别 · 计算机科学 2025-10-28 Usman Ali , Ali Zia , Abdul Rehman , Umer Ramzan , Zohaib Hassan , Talha Sattar , Jing Wang , Wei Xiang

Alzheimer's disease is a progressive neurological disorder characterized by cognitive impairment and memory loss. With the increasing aging population, the incidence of AD is continuously rising, making early diagnosis and intervention an…

图像与视频处理 · 电气工程与系统科学 2023-11-23 Guian Fang , Mengsha Liu , Yi Zhong , Zhuolin Zhang , Jiehui Huang , Zhenchao Tang , Calvin Yu-Chian Chen

Recent advances in interactive text-to-image retrieval (I-TIR) use diffusion models to bridge the modality gap between the textual information need and the images to be searched, resulting in increased effectiveness. However, existing…

信息检索 · 计算机科学 2026-03-24 Zhuocheng Zhang , Xingwu Zhang , Kangheng Liang , Guanxuan Li , Richard Mccreadie , Zijun Long

Learning from multimodal datasets can leverage complementary information and improve performance in prediction tasks. A commonly used strategy to account for feature correlations in high-dimensional datasets is the latent variable approach.…

机器学习 · 计算机科学 2024-10-01 Lingchao Mao , Qi wang , Yi Su , Fleming Lure , Jing Li

Missing data problems, such as missing modalities in multi-modal brain MRI and missing slices in cardiac MRI, pose significant challenges in clinical practice. Existing methods rely on external guidance to supply detailed missing state for…

图像与视频处理 · 电气工程与系统科学 2026-03-11 Junkai Liu , Nay Aung , Theodoros N. Arvanitis , Joao A. C. Lima , Steffen E. Petersen , Le Zhang

Self-supervised learning has proved effective for skeleton-based human action understanding. However, previous works either rely on contrastive learning that suffers false negative problems or are based on reconstruction that learns too…

计算机视觉与模式识别 · 计算机科学 2024-09-17 Lehong Wu , Lilang Lin , Jiahang Zhang , Yiyang Ma , Jiaying Liu

Recent advances in generative modeling have positioned diffusion models as state-of-the-art tools for sampling from complex data distributions. While these models have shown remarkable success across single-modality domains such as images…

计算机视觉与模式识别 · 计算机科学 2025-10-28 Nimrod Berman , Omkar Joglekar , Eitan Kosman , Dotan Di Castro , Omri Azencot

Diffusion models are the go-to method for Text-to-Image generation, but their iterative denoising processes has high inference latency. Quantization reduces compute time by using lower bitwidths, but applies a fixed precision across all…

计算机视觉与模式识别 · 计算机科学 2026-03-17 Basile Lewandowski , Simon Kurz , Aditya Shankar , Robert Birke , Jian-Jia Chen , Lydia Y. Chen

Multimodal Emotion Recognition in Conversations (MERC) enhances emotional understanding through the fusion of multimodal signals. However, unpredictable modality absence in real-world scenarios significantly degrades the performance of…

计算机视觉与模式识别 · 计算机科学 2025-11-04 Xihang Qiu , Jiarong Cheng , Yuhao Fang , Wanpeng Zhang , Yao Lu , Ye Zhang , Chun Li

Unsupervised anomaly detection in brain images is crucial for identifying injuries and pathologies without access to labels. However, the accurate localization of anomalies in medical images remains challenging due to the inherent…

计算机视觉与模式识别 · 计算机科学 2025-07-25 Farzad Beizaee , Gregory Lodygensky , Christian Desrosiers , Jose Dolz

Although Convolutional Neural Networks (CNNs) have achieved promising results in image classification, they still are vulnerable to affine transformations including rotation, translation, flip and shuffle. The drawback motivates us to…

计算机视觉与模式识别 · 计算机科学 2023-12-14 Zijie Tan , Guanfang Dong , Chenqiu Zhao , Anup Basu

The scarcity of annotated surgical data poses a significant challenge for developing deep learning systems in computer-assisted interventions. While diffusion models can synthesize realistic images, they often suffer from data memorization,…

计算机视觉与模式识别 · 计算机科学 2025-09-24 Danush Kumar Venkatesh , Stefanie Speidel

Medical image segmentation is crucial for clinical diagnosis and treatment planning. Traditional methods typically produce a single segmentation mask, failing to capture inherent uncertainty. Recent generative models enable the creation of…

计算机视觉与模式识别 · 计算机科学 2026-02-27 Huynh Trinh Ngoc , Toan Nguyen Hai , Ba Luong Son , Long Tran Quoc

Neuroimaging data, particularly from techniques like MRI or PET, offer rich but complex information about brain structure and activity. To manage this complexity, latent representation models - such as Autoencoders, Generative Adversarial…

计算机视觉与模式识别 · 计算机科学 2024-12-31 C. Vázquez-García , F. J. Martínez-Murcia , F. Segovia Román , Juan M. Górriz

Multimodal medical image fusion helps in combining contrasting features from two or more input imaging modalities to represent fused information in a single image. One of the pivotal clinical applications of medical image fusion is the…

图像与视频处理 · 电气工程与系统科学 2019-09-20 Nishant Kumar , Nico Hoffmann , Martin Oelschlägel , Edmund Koch , Matthias Kirsch , Stefan Gumhold