中文
相关论文

相关论文: iMedImage Technical Report

200 篇论文

In this paper, an innovative multi-modal deep learning model is proposed to deeply integrate heterogeneous information from medical images and clinical reports. First, for medical images, convolutional neural networks were used to extract…

机器学习 · 计算机科学 2024-05-29 Ziyan Yao , Fei Lin , Sheng Chai , Weijie He , Lu Dai , Xinghui Fei

Purpose: Most studies evaluating artificial intelligence (AI) models that detect abnormalities in neuroimaging are either tested on unrepresentative patient cohorts or are insufficiently well-validated, leading to poor generalisability to…

图像与视频处理 · 电气工程与系统科学 2024-05-10 Siddharth Agarwal , David A. Wood , Mariusz Grzeda , Chandhini Suresh , Munaib Din , James Cole , Marc Modat , Thomas C Booth

Many clinical tasks require an understanding of specialized data, such as medical images and genomics, which is not typically found in general-purpose large multimodal models. Building upon Gemini's multimodal models, we develop several…

Recent advancements in foundation models have shown significant potential in medical image analysis. However, there is still a gap in models specifically designed for medical image localization. To address this, we introduce MedLAM, a 3D…

计算机视觉与模式识别 · 计算机科学 2024-10-10 Wenhui Lei , Xu Wei , Xiaofan Zhang , Kang Li , Shaoting Zhang

Foundation models (FMs) are transforming computational pathology by offering new ways to analyze histopathology images. However, FMs typically require weeks of training on large databases, making their creation a resource-intensive process.…

图像与视频处理 · 电气工程与系统科学 2026-01-27 Till Nicke , Daniela Schacherer , Jan Raphael Schäfer , Natalia Artysh , Antje Prasse , André Homeyer , Andrea Schenk , Henning Höfener , Johannes Lotz

The isocitrate dehydrogenase (IDH) gene mutation is an essential biomarker for the diagnosis and prognosis of glioma. It is promising to better predict glioma genotype by integrating focal tumor image and geometric features with brain…

计算机视觉与模式识别 · 计算机科学 2022-03-22 Yiran Wei , Xi Chen , Lei Zhu , Lipei Zhang , Carola-Bibiane Schönlieb , Stephen J. Price , Chao Li

Recent advances in multimodal large language models have enabled unified processing of visual and textual inputs, offering promising applications in general-purpose medical AI. However, their ability to generalize compositionally across…

计算机视觉与模式识别 · 计算机科学 2025-11-17 Pooja Singh , Siddhant Ujjain , Tapan Kumar Gandhi , Sandeep Kumar

Medical imaging is an essential tool in many areas of medical applications, used for both diagnosis and treatment. However, reading medical images and making diagnosis or treatment recommendations require specially trained medical…

计算机视觉与模式识别 · 计算机科学 2019-03-13 Wentao Zhu

While previous studies have demonstrated the potential of AI to diagnose diseases in imaging data, clinical implementation is still lagging behind. This is partly because AI models require training with large numbers of examples only…

Artificial intelligence (AI)-enabled diagnostics in maxillofacial pathology require structured, high-quality multimodal datasets. However, existing resources provide limited ameloblastoma coverage and lack the format consistency needed for…

人工智能 · 计算机科学 2026-02-06 Ajo Babu George , Anna Mariam John , Athul Anoop , Balu Bhasuran

Current vision-language models (VLMs) in medicine are primarily designed for categorical question answering (e.g., "Is this normal or abnormal?") or qualitative descriptive tasks. However, clinical decision-making often relies on…

计算机视觉与模式识别 · 计算机科学 2025-11-25 Yongcheng Yao , Yongshuo Zong , Raman Dutt , Yongxin Yang , Sotirios A Tsaftaris , Timothy Hospedales

Multi-modal medical image fusion (MMIF) is increasingly recognized as an essential technique for enhancing diagnostic precision and facilitating effective clinical decision-making within computer-aided diagnosis systems. MMIF combines data…

图像与视频处理 · 电气工程与系统科学 2025-05-22 Muhammad Zubair , Muzammil Hussai , Mousa Ahmad Al-Bashrawi , Malika Bendechache , Muhammad Owais

Anomaly detection is the problem of recognizing abnormal inputs based on the seen examples of normal data. Despite recent advances of deep learning in recognizing image anomalies, these methods still prove incapable of handling complex…

计算机视觉与模式识别 · 计算机科学 2021-09-14 Nina Shvetsova , Bart Bakker , Irina Fedulova , Heinrich Schulz , Dmitry V. Dylov

In medical imaging, chromosome straightening plays a significant role in the pathological study of chromosomes and in the development of cytogenetic maps. Whereas different approaches exist for the straightening task, typically geometric…

计算机视觉与模式识别 · 计算机科学 2021-10-20 Sifan Song , Daiyun Huang , Yalun Hu , Chunxiao Yang , Jia Meng , Fei Ma , Frans Coenen , Jiaming Zhang , Jionglong Su

Foundation models have demonstrated remarkable potential in medical domain. However, their application to complex cardiovascular diagnostics remains underexplored. In this paper, we present Cardiac-CLIP, a multi-modal foundation model…

Chromosome enumeration is an essential but tedious procedure in karyotyping analysis. To automate the enumeration process, we develop a chromosome enumeration framework, DeepACEv2, based on the region based object detection scheme. The…

计算机视觉与模式识别 · 计算机科学 2020-07-21 Li Xiao , Chunlong Luo , Tianqi Yu , Yufan Luo , Manqing Wang , Fuhai Yu , Yinhao Li , Chan Tian , Jie Qiao

Visual-language models have advanced the development of universal models, yet their application in medical imaging remains constrained by specific functional requirements and the limited data. Current general-purpose models are typically…

计算机视觉与模式识别 · 计算机科学 2024-10-08 Kaini Wang , Ling Yang , Siping Zhou , Guangquan Zhou , Wentao Zhang , Bin Cui , Shuo Li

The medical image analysis field has traditionally been focused on the development of organ-, and disease-specific methods. Recently, the interest in the development of more 20 comprehensive computational anatomical models has grown,…

While emerging 3D medical foundation models are envisioned as versatile tools with offer general-purpose capabilities, their validation remains largely confined to regional and structural imaging, leaving a significant modality discrepancy…

计算机视觉与模式识别 · 计算机科学 2026-02-10 Yichi Zhang , Feiyang Xiao , Le Xue , Wenbo Zhang , Gang Feng , Chenguang Zheng , Yuan Qi , Yuan Cheng , Zixin Hu