English
Related papers

Related papers: MoXGATE: Modality-aware cross-attention for multi-…

200 papers

We propose AttentionMixer, a unified deep learning framework for multimodal detection of brain edema that combines structural head CT (HCT) with routine clinical metadata. While HCT provides rich spatial information, clinical variables such…

Purpose High dimensional, multimodal data can nowadays be analyzed by huge deep neural networks with little effort. Several fusion methods for bringing together different modalities have been developed. Given the prevalence of…

Computer Vision and Pattern Recognition · Computer Science 2025-10-03 Christian Gapp , Elias Tappeiner , Martin Welk , Karl Fritscher , Elke Ruth Gizewski , Rainer Schubert

Clustering is commonly performed as an initial analysis step for uncovering structure in 'omics datasets, e.g. to discover molecular subtypes of disease. The high-throughput, high-dimensional nature of these datasets means that they provide…

Methodology · Statistics 2023-03-02 Paul D. W. Kirk , Filippo Pagani , Sylvia Richardson

Mammography, an X-ray-based imaging technique, remains central to the early detection of breast cancer. Recent advances in artificial intelligence have enabled increasingly sophisticated computer-aided diagnostic methods, evolving from…

Image and Video Processing · Electrical Eng. & Systems 2025-10-09 Daniel G. P. Petrini , Hae Yong Kim

Understanding intricate and fast-paced movements of body parts is essential for the recognition and translation of sign language. The inclusion of additional information intended to identify and locate the moving body parts has been an…

Computer Vision and Pattern Recognition · Computer Science 2024-10-08 Zaber Ibn Abdul Hakim , Rasman Mubtasim Swargo , Muhammad Abdullah Adnan

Vision Transformers $(\texttt{ViT})$ have become the architecture of choice for many computer vision tasks, yet their performance in computer-aided diagnostics remains limited. Focusing on breast cancer detection from mammograms, we…

Computer Vision and Pattern Recognition · Computer Science 2026-04-22 Samyak Sanghvi , Piyush Miglani , Sarvesh Shashikumar , Kaustubh R Borgavi , Veenu Singla , Chetan Arora

Multimodal large language models (MLLMs) recently showed strong capacity in integrating data among multiple modalities, empowered by a generalizable attention architecture. Advanced methods predominantly focus on language-centric tuning…

Computer Vision and Pattern Recognition · Computer Science 2025-07-08 Zhicheng Zhang , Wuyou Xia , Chenxi Zhao , Zhou Yan , Xiaoqiang Liu , Yongjie Zhu , Wenyu Qin , Pengfei Wan , Di Zhang , Jufeng Yang

Mammography and ultrasound are extensively used by radiologists as complementary modalities to achieve better performance in breast cancer diagnosis. However, existing computer-aided diagnosis (CAD) systems for the breast are generally…

Image and Video Processing · Electrical Eng. & Systems 2020-09-24 Gavriel Habib , Nahum Kiryati , Miri Sklair-Levy , Anat Shalmon , Osnat Halshtok Neiman , Renata Faermann Weidenfeld , Yael Yagil , Eli Konen , Arnaldo Mayer

Multi-modal data comprising imaging (MRI, fMRI, PET, etc.) and non-imaging (clinical test, demographics, etc.) data can be collected together and used for disease prediction. Such diverse data gives complementary information about the…

Machine Learning · Computer Science 2018-12-27 Anees Kazi , S. Arvind krishna , Shayan Shekarforoush , Karsten Kortuem , Shadi Albarqouni , Nassir Navab

Due to the severe lack of labeled data, existing methods of medical visual question answering usually rely on transfer learning to obtain effective image feature representation and use cross-modal fusion of visual and linguistic features to…

Multimedia · Computer Science 2021-05-04 Haifan Gong , Guanqi Chen , Sishuo Liu , Yizhou Yu , Guanbin Li

Multimodal information retrieval (MIR) faces inherent challenges due to the heterogeneity of data sources and the complexity of cross-modal alignment. While previous studies have identified modal gaps in feature spaces, a systematic…

Computer Vision and Pattern Recognition · Computer Science 2025-05-28 Fanheng Kong , Jingyuan Zhang , Yahui Liu , Hongzhi Zhang , Shi Feng , Xiaocui Yang , Daling Wang , Yu Tian , Victoria W. , Fuzheng Zhang , Guorui Zhou

Multimodal sentiment analysis is an important research task to predict the sentiment score based on the different modality data from a specific opinion video. Many previous pieces of research have proved the significance of utilizing the…

Computation and Language · Computer Science 2022-08-26 Ming Jiang , Shaoxiong Ji

Cancer detection and classification from gigapixel whole slide images of stained tissue specimens has recently experienced enormous progress in computational histopathology. The limitation of available pixel-wise annotated scans shifted the…

Computer Vision and Pattern Recognition · Computer Science 2023-03-30 Mehdi Naouar , Gabriel Kalweit , Ignacio Mastroleo , Philipp Poxleitner , Marc Metzger , Joschka Boedecker , Maria Kalweit

Purpose: To determine whether deep learning models can distinguish between breast cancer molecular subtypes based on dynamic contrast-enhanced magnetic resonance imaging (DCE-MRI). Materials and methods: In this institutional review…

Computer Vision and Pattern Recognition · Computer Science 2017-12-01 Zhe Zhu , Ehab Albadawy , Ashirbani Saha , Jun Zhang , Michael R. Harowicz , Maciej A. Mazurowski

Optical coherence tomography (OCT) is one of the non-invasive and easy-to-acquire biomarkers (the thickness of the retinal layers, which is detectable within OCT scans) being investigated to diagnose Alzheimer's disease (AD). This work aims…

Image and Video Processing · Electrical Eng. & Systems 2022-06-14 Paria Jeihouni , Omid Dehzangi , Annahita Amireskandari , Ali Dabouei , Ali Rezai , Nasser M. Nasrabadi

Medical image segmentation of tumors and organs at risk is a time-consuming yet critical process in the clinic that utilizes multi-modality imaging (e.g, different acquisitions, data types, and sequences) to increase segmentation precision.…

Image and Video Processing · Electrical Eng. & Systems 2023-06-07 Qisheng He , Nicholas Summerfield , Ming Dong , Carri Glide-Hurst

Fluorescence microscopy allows for a detailed inspection of cells, cellular networks, and anatomical landmarks by staining with a variety of carefully-selected markers visualized as color channels. Quantitative characterization of…

Computer Vision and Pattern Recognition · Computer Science 2021-08-26 Alvaro Gomariz , Tiziano Portenier , Patrick M. Helbling , Stephan Isringhausen , Ute Suessbier , César Nombela-Arrieta , Orcun Goksel

Antibody binding site prediction plays a pivotal role in computational immunology and therapeutic antibody design. Existing sequence or structure methods rely on single-view features and fail to identify antibody-specific binding sites on…

Machine Learning · Computer Science 2025-09-12 Hongzong Li , Jiahao Ma , Zhanpeng Shi , Rui Xiao , Fanming Jin , Ye-Fan Hu , Hangjun Che , Jian-Dong Huang

Liver cancer is one of the most common cancers worldwide. Due to inconspicuous texture changes of liver tumor, contrast-enhanced computed tomography (CT) imaging is effective for the diagnosis of liver cancer. In this paper, we focus on…

Image and Video Processing · Electrical Eng. & Systems 2021-07-22 Yao Zhang , Jiawei Yang , Jiang Tian , Zhongchao Shi , Cheng Zhong , Yang Zhang , Zhiqiang He
‹ Prev 1 4 5 6 7 8 10 Next ›