中文
相关论文

相关论文: Med-SORA: Symptom to Organ Reasoning in Abdomen CT…

200 篇论文

Multimodal large language models (MLLMs) have achieved significant success in the general field of image processing. Their emerging task generalization and freeform conversational capabilities can greatly facilitate medical diagnostic…

图像与视频处理 · 电气工程与系统科学 2024-09-17 Youzhu Jin , Yichen Zhang

Although self-supervised learning enables us to bootstrap the training by exploiting unlabeled data, the generic self-supervised methods for natural images do not sufficiently incorporate the context. For medical images, a desirable method…

图像与视频处理 · 电气工程与系统科学 2022-07-08 Li Sun , Ke Yu , Kayhan Batmanghelich

Deep segmentation networks achieve high performance when trained on specific datasets. However, in clinical practice, it is often desirable that pretrained segmentation models can be dynamically extended to enable segmenting new organs…

计算机视觉与模式识别 · 计算机科学 2024-10-08 Vince Zhu , Zhanghexuan Ji , Dazhou Guo , Puyang Wang , Yingda Xia , Le Lu , Xianghua Ye , Wei Zhu , Dakai Jin

Multimodal large language models have advanced rapidly, but their adoption in medicine is constrained by limited domain coverage, imperfect modality alignment, and insufficient grounded reasoning. We introduce MedMO, a medical multimodal…

计算机视觉与模式识别 · 计算机科学 2026-03-13 Ankan Deria , Komal Kumar , Adinath Madhavrao Dukre , Eran Segal , Salman Khan , Imran Razzak

Automatic radiology report generation can alleviate the workload for physicians and minimize regional disparities in medical resources, therefore becoming an important topic in the medical image analysis field. It is a challenging task, as…

计算机视觉与模式识别 · 计算机科学 2025-03-07 Xinyi Wang , Grazziela Figueredo , Ruizhe Li , Wei Emma Zhang , Weitong Chen , Xin Chen

Medical imaging plays a crucial role in diagnosis, with radiology reports serving as vital documentation. Automating report generation has emerged as a critical need to alleviate the workload of radiologists. While machine learning has…

图像与视频处理 · 电气工程与系统科学 2024-07-08 Ibrahim Ethem Hamamci , Sezgin Er , Bjoern Menze

There exists a large number of datasets for organ segmentation, which are partially annotated and sequentially constructed. A typical dataset is constructed at a certain time by curating medical images and annotating the organs of interest.…

图像与视频处理 · 电气工程与系统科学 2022-03-07 Pengbo Liu , Xia Wang , Mengsi Fan , Hongli Pan , Minmin Yin , Xiaohong Zhu , Dandan Du , Xiaoying Zhao , Li Xiao , Lian Ding , Xingwang Wu , S. Kevin Zhou

Physiological signals such as electrocardiograms (ECG) and electroencephalograms (EEG) provide complementary insights into human health and cognition, yet multi-modal integration is challenging due to limited multi-modal labeled data, and…

The rapid increase in the number of Computed Tomography (CT) scan examinations has created an urgent need for automated tools, such as organ segmentation, anomaly classification, and report generation, to assist radiologists with their…

计算机视觉与模式识别 · 计算机科学 2025-12-29 Theo Di Piazza , Carole Lazarus , Olivier Nempont , Loic Boussel

Transformers have demonstrated remarkable performance in natural language processing and computer vision. However, existing vision Transformers struggle to learn from limited medical data and are unable to generalize on diverse medical…

图像与视频处理 · 电气工程与系统科学 2023-04-06 Yunhe Gao , Mu Zhou , Di Liu , Zhennan Yan , Shaoting Zhang , Dimitris N. Metaxas

Image registration is an essential technique for the analysis of Computed Tomography (CT) images in clinical practice. However, existing methodologies are predominantly tailored to a specific organ of interest and often exhibit lower…

计算机视觉与模式识别 · 计算机科学 2025-03-31 Xuan Loc Pham , Mathias Prokop , Bram van Ginneken , Alessa Hering

Multi-organ segmentation in whole-body computed tomography (CT) is a constant pre-processing step which finds its application in organ-specific image retrieval, radiotherapy planning, and interventional image analysis. We address this…

计算机视觉与模式识别 · 计算机科学 2019-08-15 Fernando Navarro , Suprosanna Shit , Ivan Ezhov , Johannes Paetzold , Andrei Gafita , Jan Peeken , Stephanie Combs , Bjoern Menze

We introduce Med-CTX, a fully transformer based multimodal framework for explainable breast cancer ultrasound segmentation. We integrate clinical radiology reports to boost both performance and interpretability. Med-CTX achieves exact…

计算机视觉与模式识别 · 计算机科学 2025-08-20 Enobong Adahada , Isabel Sassoon , Kate Hone , Yongmin Li

Medical image segmentation is undergoing a paradigm shift from conventional visual pattern matching to cognitive reasoning analysis. Although Multimodal Large Language Models (MLLMs) have shown promise in integrating linguistic and visual…

计算机视觉与模式识别 · 计算机科学 2026-03-09 Yuxin Xie , Yuming Chen , Yishan Yang , Yi Zhou , Tao Zhou , Zhen Zhao , Jiacheng Liu , Huazhu Fu

Deformable registration of two-dimensional/three-dimensional (2D/3D) images of abdominal organs is a complicated task because the abdominal organs deform significantly and their contours are not detected in two-dimensional X-ray images. We…

图像与视频处理 · 电气工程与系统科学 2022-12-13 Ryuto Miura , Megumi Nakao , Mitsuhiro Nakamura , Tetsuya Matsuda

MLLMs MLLMs are beginning to appear in clinical workflows, but their ability to perform complex medical reasoning remains unclear. We present Med-CMR, a fine-grained Medical Complex Multimodal Reasoning benchmark. Med-CMR distinguishes from…

Clinical diagnosis is a highly specialized discipline requiring both domain expertise and strict adherence to rigorous guidelines. While current AI-driven medical research predominantly focuses on knowledge graphs or natural text…

机器学习 · 计算机科学 2025-12-12 Haolin Li , Tianjie Dai , Zhe Chen , Siyuan Du , Jiangchao Yao , Ya Zhang , Yanfeng Wang

Understanding how the brain represents visual information is a fundamental challenge in neuroscience and artificial intelligence. While AI-driven decoding of neural data has provided insights into the human visual system, integrating…

神经与进化计算 · 计算机科学 2025-10-07 Dongyang Li , Haoyang Qin , Mingyang Wu , Chen Wei , Quanying Liu

Autonomous surgical procedures, in particular minimal invasive surgeries, are the next frontier for Artificial Intelligence research. However, the existing challenges include precise identification of the human anatomy and the surgical…

计算机视觉与模式识别 · 计算机科学 2020-12-15 Salman Maqbool , Aqsa Riaz , Hasan Sajid , Osman Hasan