中文
相关论文

相关论文: Pillar-0: A New Frontier for Radiology Foundation …

200 篇论文

Pathology has played a crucial role in the diagnosis and evaluation of patient tissue samples obtained from surgeries and biopsies for many years. The advent of Whole Slide Scanners and the development of deep learning technologies have…

计算机视觉与模式识别 · 计算机科学 2024-08-07 Mieko Ochi , Daisuke Komura , Shumpei Ishikawa

Foundation models (FMs) are driving a prominent shift in biomedical imaging from task-specific models to unified backbone models for diverse tasks. This opens an avenue to integrate imaging, pathology, clinical records, and genomics data…

Despite the promise of foundation models in medical AI, current systems remain limited - they are modality-specific and lack transparent reasoning processes, hindering clinical adoption. To address this gap, we present EVLF-FM, a multimodal…

Recent advancements in foundation models, typically trained with self-supervised learning on large-scale and diverse datasets, have shown great potential in medical image analysis. However, due to the significant spatial heterogeneity of…

计算机视觉与模式识别 · 计算机科学 2024-01-25 Lingxiao Luo , Xuanzhong Chen , Bingda Tang , Xinsheng Chen , Rong Han , Chengpeng Hu , Yujiang Li , Ting Chen

Medical image captioning is a challenging task that requires generating clinically accurate and semantically meaningful descriptions of radiology images. While recent vision-language models (VLMs) such as BLIP, BLIP2, Gemini and ViT-GPT2…

图像与视频处理 · 电气工程与系统科学 2025-05-22 Manshi Limbu , Diwita Banerjee

Cardiac ultrasound diagnosis is critical for cardiovascular disease assessment, but acquiring standard views remains highly operator-dependent. Existing medical segmentation models often yield anatomically inconsistent results in images…

机器人学 · 计算机科学 2026-03-24 Zhiyan Cao , Zhengxi Wu , Yiwei Wang , Pei-Hsuan Lin , Li Zhang , Zhen Xie , Huan Zhao , Han Ding

Ultrasound foundation models have achieved strong performance on structured prediction tasks but remain exclusively vision-based, limiting zero-shot and few-shot transfer to novel tasks where task-specific annotation is scarce. We address…

计算机视觉与模式识别 · 计算机科学 2026-05-05 Zhuoyang Lyu , Yiyang Zhang , Tongxin Wang , Ruirui Lan

Foundation segmentation models such as SAM and SAM-2 perform well on natural images but struggle with brain MRIs where structures like the caudate and thalamus lack sharp boundaries and have low contrast. Rather than fine tune these models…

图像与视频处理 · 电气工程与系统科学 2025-11-26 Keith Moore

Cephalometric analysis is an important tool for orthodontic diagnosis. At present, most cephalometric analysis is performed with the help of image processing techniques. Hence, the resolution between millimeter and pixel is needed with high…

图像与视频处理 · 电气工程与系统科学 2019-12-10 Jia Guo , Shumeng Wang , Huiqi Li

Focal liver lesions (FLL) are common clinical findings during physical examination. Early diagnosis and intervention of liver malignancies are crucial to improving patient survival. Although the current 3D segmentation paradigm can…

图像与视频处理 · 电气工程与系统科学 2025-07-08 Jiacheng Hao , Xiaoming Zhang , Wei Liu , Xiaoli Yin , Yuan Gao , Chunli Li , Ling Zhang , Le Lu , Yu Shi , Xu Han , Ke Yan

We introduce CRIMSON, a clinically grounded evaluation framework for chest X-ray report generation that assesses reports based on diagnostic correctness, contextual relevance, and patient safety. Unlike prior metrics, CRIMSON incorporates…

In recent years, "U-shaped" neural networks featuring encoder and decoder structures have gained popularity in the field of medical image segmentation. Various variants of this model have been developed. Nevertheless, the evaluation of…

图像与视频处理 · 电气工程与系统科学 2023-06-02 Qi Ye , Lihua Guo

The advancement of artificial intelligence (AI) for organ segmentation and tumor detection is propelled by the growing availability of computed tomography (CT) datasets with detailed, per-voxel annotations. However, these AI models often…

图像与视频处理 · 电气工程与系统科学 2024-05-29 Jie Liu , Yixiao Zhang , Kang Wang , Mehmet Can Yavuz , Xiaoxi Chen , Yixuan Yuan , Haoliang Li , Yang Yang , Alan Yuille , Yucheng Tang , Zongwei Zhou

3D medical vision-language (VL) pretraining has shown potential in radiology by leveraging large-scale multimodal datasets with CT-report pairs. However, existing methods primarily rely on a global VL alignment directly adapted from 2D…

计算机视觉与模式识别 · 计算机科学 2025-12-03 Jingyang Lin , Yingda Xia , Jianpeng Zhang , Ke Yan , Kai Cao , Le Lu , Jiebo Luo , Ling Zhang

The field of computational pathology has recently seen rapid advances driven by the development of modern vision foundation models (FMs), typically trained on vast collections of pathology images. Recent studies demonstrate that increasing…

计算机视觉与模式识别 · 计算机科学 2025-04-08 Mikhail Karasikov , Joost van Doorn , Nicolas Känzig , Melis Erdal Cesur , Hugo Mark Horlings , Robert Berke , Fei Tang , Sebastian Otálora

Digital pathology is a tool of rapidly evolving importance within the discipline of pathology. Whole slide imaging promises numerous advantages; however, adoption is limited by challenges in ease of use and speed of high-quality image…

图形学 · 计算机科学 2025-04-24 Ryan Erik Landvater , Ulysses Balis

Ultrasound imaging is widely used in clinical diagnosis due to its non-invasive nature and real-time capabilities. However, traditional ultrasound diagnostics relies heavily on physician expertise and is often hampered by suboptimal image…

图像与视频处理 · 电气工程与系统科学 2025-12-18 Yuncheng Jiang , Chun-Mei Feng , Jinke Ren , Jun Wei , Zixun Zhang , Yiwen Hu , Yunbi Liu , Rui Sun , Xuemei Tang , Juan Du , Xiang Wan , Yong Xu , Bo Du , Xin Gao , Guangyu Wang , Shaohua Zhou , Shuguang Cui , Zhen Li

Foundation models (FMs) are large-scale deep learning models trained on massive datasets, often using self-supervised learning techniques. These models serve as a versatile base for a wide range of downstream tasks, including those in…

机器学习 · 计算机科学 2025-01-17 Wasif Khan , Seowung Leem , Kyle B. See , Joshua K. Wong , Shaoting Zhang , Ruogu Fang

Automated interpretation of CT images-particularly localizing and describing abnormal findings across multi-plane and whole-body scans-remains a significant challenge in clinical radiology. This work aims to address this challenge through…

图像与视频处理 · 电气工程与系统科学 2025-11-18 Ziheng Zhao , Lisong Dai , Ya Zhang , Yanfeng Wang , Weidi Xie

Robust preprocessing is rarely quantified in deep-learning pipelines for low-dose CT (LDCT) lung cancer screening. We develop and validate Virtual-Eyes, a clinically motivated 16-bit CT quality-control pipeline, and measure its differential…

计算机视觉与模式识别 · 计算机科学 2026-01-01 Md. Enamul Hoq , Linda Larson-Prior , Fred Prior