中文
相关论文

相关论文: Slide-Level Prompt Learning with Vision Language M…

200 篇论文

The interpretation of histopathology cases underlies many important diagnostic and treatment decisions in medicine. Notably, this process typically requires pathologists to integrate and summarize findings across multiple slides per case.…

In the application of Multiple Instance Learning (MIL) methods for Whole Slide Image (WSI) classification, attention mechanisms often focus on a subset of discriminative instances, which are closely linked to overfitting. To mitigate…

计算机视觉与模式识别 · 计算机科学 2024-07-08 Yunlong Zhang , Honglin Li , Yuxuan Sun , Sunyi Zheng , Chenglu Zhu , Lin Yang

In the field of computational pathology, the use of decision support systems powered by state-of-the-art deep learning solutions has been hampered by the lack of large labeled datasets. Until recently, studies relied on datasets in the…

计算机视觉与模式识别 · 计算机科学 2018-10-01 Gabriele Campanella , Vitor Werneck Krauss Silva , Thomas J. Fuchs

Despite the progress made by multimodal large language models (MLLMs) in computational pathology, they remain limited by a predominant focus on patch-level analysis, missing essential contextual information at the whole-slide level. The…

计算机视觉与模式识别 · 计算机科学 2025-03-20 Ying Chen , Guoan Wang , Yuanfeng Ji , Yanjun Li , Jin Ye , Tianbin Li , Ming Hu , Rongshan Yu , Yu Qiao , Junjun He

Learning good representation of giga-pixel level whole slide pathology images (WSI) for downstream tasks is critical. Previous studies employ multiple instance learning (MIL) to represent WSIs as bags of sampled patches because, for most…

计算机视觉与模式识别 · 计算机科学 2022-12-01 Chunyuan Li , Xinliang Zhu , Jiawen Yao , Junzhou Huang

Accurate prediction of placental diseases via whole slide images (WSIs) is critical for preventing severe maternal and fetal complications. However, WSI analysis presents significant computational challenges due to the massive data volume.…

计算机视觉与模式识别 · 计算机科学 2025-08-06 Hang Guo , Qing Zhang , Zixuan Gao , Siyuan Yang , Shulin Peng , Xiang Tao , Ting Yu , Yan Wang , Qingli Li

This paper explores training medical vision-language models (VLMs) -- where the visual and language inputs are embedded into a common space -- with a particular focus on scenarios where training data is limited, as is often the case in…

计算机视觉与模式识别 · 计算机科学 2023-04-03 Rhydian Windsor , Amir Jamaludin , Timor Kadir , Andrew Zisserman

In pre-clinical pathology, there is a paradox between the abundance of raw data (whole slide images from many organs of many individual animals) and the lack of pixel-level slide annotations done by pathologists. Due to time constraints and…

计算机视觉与模式识别 · 计算机科学 2023-02-06 Marco Bertolini , Van-Khoa Le , Jake Pencharz , Andreas Poehlmann , Djork-Arné Clevert , Santiago Villalba , Floriane Montanari

Recent advances in histopathology vision-language foundation models (VLFMs) have shown promise in addressing data scarcity for whole slide image (WSI) classification via zero-shot adaptation. However, these methods remain outperformed by…

计算机视觉与模式识别 · 计算机科学 2025-08-14 Tianqi Xiang , Yi Li , Qixiang Zhang , Xiaomeng Li

The expensive fine-grained annotation and data scarcity have become the primary obstacles for the widespread adoption of deep learning-based Whole Slide Images (WSI) classification algorithms in clinical practice. Unlike few-shot learning…

计算机视觉与模式识别 · 计算机科学 2024-10-01 Kexue Fu , Xiaoyuan Luo , Linhao Qu , Shuo Wang , Ying Xiong , Ilias Maglogiannis , Longxiang Gao , Manning Wang

Deep learning has shown strong potential in cancer classification from whole-slide images (WSIs), but the need for extensive expert annotations often limits its success. Annotation-free approaches, such as multiple instance learning (MIL)…

计算机视觉与模式识别 · 计算机科学 2026-03-12 Willmer Rafell Quinones Robles , Sakonporn Noree , Jongwoo Kim , Young Sin Ko , Bryan Wong , Mun Yong Yi

Digital pathology involves converting physical tissue slides into high-resolution Whole Slide Images (WSIs), which pathologists analyze for disease-affected tissues. However, large histology slides with numerous microscopic fields pose…

图像与视频处理 · 电气工程与系统科学 2024-02-19 Mobina Mansoori , Sajjad Shahabodini , Jamshid Abouei , Arash Mohammadi , Konstantinos N. Plataniotis

Fine-grained truck classification is critical for intelligent transportation systems (ITS), yet current LiDAR-based methods face scalability challenges due to their reliance on supervised deep learning and labor-intensive manual annotation.…

计算机视觉与模式识别 · 计算机科学 2026-02-11 Yiqiao Li , Bo Shang , Jie Wei

Weakly-supervised classification of histopathology slides is a computationally intensive task, with a typical whole slide image (WSI) containing billions of pixels to process. We propose Discriminative Region Active Sampling for Multiple…

图像与视频处理 · 电气工程与系统科学 2023-02-23 Jack Breen , Katie Allen , Kieran Zucker , Geoff Hall , Nicolas M. Orsi , Nishant Ravikumar

Whole slide images (WSIs) classification represents a fundamental challenge in computational pathology, where multiple instance learning (MIL) has emerged as the dominant paradigm. Current state-of-the-art (SOTA) MIL methods rely on…

计算机视觉与模式识别 · 计算机科学 2025-10-07 Chengying She , Chengwei Chen , Dongjie Fan , Lizhuang Liu , Chengwei Shao , Yun Bian , Ben Wang , Xinran Zhang

While multiple instance learning (MIL) has shown to be a promising approach for histopathological whole slide image (WSI) analysis, its reliance on permutation invariance significantly limits its capacity to effectively uncover semantic…

图像与视频处理 · 电气工程与系统科学 2025-07-14 Xiwen Chen , Peijie Qiu , Wenhui Zhu , Hao Wang , Huayu Li , Xuanzhao Dong , Xiaotong Sun , Xiaobing Yu , Yalin Wang , Abolfazl Razi , Aristeidis Sotiras

We developed a software pipeline for quality control (QC) of histopathology whole slide images (WSIs) that segments various regions, such as blurs of different levels, tissue regions, tissue folds, and pen marks. Given the necessity and…

图像与视频处理 · 电气工程与系统科学 2025-06-17 Abhijeet Patil , Garima Jain , Harsh Diwakar , Jay Sawant , Tripti Bameta , Swapnil Rane , Amit Sethi

Vision-language models (VLMs), such as CLIP, have shown strong generalization under zero-shot settings, yet adapting them to downstream tasks with limited supervision remains a significant challenge. Existing multi-modal prompt learning…

计算机视觉与模式识别 · 计算机科学 2025-12-01 Silin Cheng , Kai Han

Vision-language foundation models (VLMs) have shown great potential in feature transfer and generalization across a wide spectrum of medical-related downstream tasks. However, fine-tuning these models is resource-intensive due to their…

计算机视觉与模式识别 · 计算机科学 2025-11-18 Ye Du , Nanxi Yu , Shujun Wang

Digital whole slide images (WSIs) are generally captured at microscopic resolution and encompass extensive spatial data. Directly feeding these images to deep learning models is computationally intractable due to memory constraints, while…

图像与视频处理 · 电气工程与系统科学 2024-11-22 Manahil Raza , Ruqayya Awan , Raja Muhammad Saad Bashir , Talha Qaiser , Nasir M. Rajpoot
‹ 上一页 1 8 9 10 下一页 ›