中文
相关论文

相关论文: Slide-Level Prompt Learning with Vision Language M…

200 篇论文

Lifelong learning for whole slide images (WSIs) poses the challenge of training a unified model to perform multiple WSI-related tasks, such as cancer subtyping and tumor classification, in a distributed, continual fashion. This is a…

计算机视觉与模式识别 · 计算机科学 2025-04-23 Doanh C. Bui , Hoai Luan Pham , Vu Trung Duong Le , Tuan Hai Vu , Van Duy Tran , Yasuhiko Nakashima

While Multiple Instance Learning (MIL) has shown promising results in digital Pathology Whole Slide Image (WSI) classification, such a paradigm still faces performance and generalization problems due to challenges in high computational…

计算机视觉与模式识别 · 计算机科学 2023-03-16 Honglin Li , Chenglu Zhu , Yunlong Zhang , Yuxuan Sun , Zhongyi Shui , Wenwei Kuang , Sunyi Zheng , Lin Yang

Multiple instance learning (MIL) has been successfully applied for whole slide images (WSIs) analysis in computational pathology, enabling a wide range of prediction tasks from tumor subtyping to inferring genetic mutations and multi-omics…

计算机视觉与模式识别 · 计算机科学 2024-07-25 Junyu Li , Ye Zhang , Wen Shu , Xiaobing Feng , Yingchun Wang , Pengju Yan , Xiaolin Li , Chulin Sha , Min He

Cancer survival prediction is a challenging task that involves analyzing of the tumor microenvironment within Whole Slide Image (WSI). Previous methods cannot effectively capture the intricate interaction features among instances within the…

计算机视觉与模式识别 · 计算机科学 2024-10-23 Zekang Yang , Hong Liu , Xiangdong Wang

Pretrained biomedical vision-language models (VLMs) such as BioMedCLIP perform well on average but often degrade on challenging modalities where inter-class margins are small and acquisition-specific variations are pronounced, especially…

计算机视觉与模式识别 · 计算机科学 2026-04-21 Mainak Singha , Tanisha Gupta , Ankit Jha , Muhammad Haris Khan , Sayantani Ghosh , Biplab Banerjee

As data requirements continue to grow, efficient learning increasingly depends on the curation and distillation of high-value data rather than brute-force scaling of model sizes. In the case of a hyperspectral image (HSI), the challenge is…

计算机视觉与模式识别 · 计算机科学 2025-09-30 Abhiroop Chatterjee , Susmita Ghosh

Gigapixel image analysis, particularly for whole slide images (WSIs), often relies on multiple instance learning (MIL). Under the paradigm of MIL, patch image representations are extracted and then fixed during the training of the MIL…

计算机视觉与模式识别 · 计算机科学 2024-12-20 Kunming Tang , Zhiguo Jiang , Jun Shi , Wei Wang , Haibo Wu , Yushan Zheng

The survival analysis on histological whole-slide images (WSIs) is one of the most important means to estimate patient prognosis. Although many weakly-supervised deep learning models have been developed for gigapixel WSIs, their potential…

图像与视频处理 · 电气工程与系统科学 2023-11-06 Pei Liu , Luping Ji , Feng Ye , Bo Fu

Pretraining on large-scale, in-domain datasets grants histopathology foundation models (FM) the ability to learn task-agnostic data representations, enhancing transfer learning on downstream tasks. In computational pathology, automated…

计算机视觉与模式识别 · 计算机科学 2025-07-22 Pablo Meseguer , Rocío del Amor , Valery Naranjo

Improving the feature representation ability is the foundation of many whole slide pathological image (WSIs) tasks. Recent works have achieved great success in pathological-specific self-supervised learning (SSL). However, most of them only…

计算机视觉与模式识别 · 计算机科学 2023-07-21 Zhimiao Yu , Tiancheng Lin , Yi Xu

Pre-trained Vision-Language Models (VLMs), such as CLIP, have shown enhanced performance across a range of tasks that involve the integration of visual and linguistic modalities. When CLIP is used for depth estimation tasks, the patches,…

计算机视觉与模式识别 · 计算机科学 2023-11-03 Xueting Hu , Ce Zhang , Yi Zhang , Bowen Hai , Ke Yu , Zhihai He

Vision language models (VLM) have achieved success in both natural language comprehension and image recognition tasks. However, their use in pathology report generation for whole slide images (WSIs) is still limited due to the huge size of…

计算机视觉与模式识别 · 计算机科学 2024-09-25 Jing Wei Tan , SeungKyu Kim , Eunsu Kim , Sung Hak Lee , Sangjeong Ahn , Won-Ki Jeong

Whole Slide Images (WSIs) present a challenging computer vision task due to their gigapixel size and presence of numerous artefacts. Yet they are a valuable resource for patient diagnosis and stratification, often representing the gold…

计算机视觉与模式识别 · 计算机科学 2023-10-05 Amaya Gallagher-Syed , Luca Rossi , Felice Rivellese , Costantino Pitzalis , Myles Lewis , Michael Barnes , Gregory Slabaugh

Multi-Instance Learning (MIL) has shown impressive performance for histopathology whole slide image (WSI) analysis using bags or pseudo-bags. It involves instance sampling, feature representation, and decision-making. However, existing…

计算机视觉与模式识别 · 计算机科学 2024-03-14 Tingting Zheng , Kui Jiang , Hongxun Yao

Multiple instance learning (MIL) has enabled substantial progress in computational histopathology, where a large amount of patches from gigapixel whole slide images are aggregated into slide-level predictions. Heatmaps are widely used to…

Histomorphology is crucial in cancer diagnosis. However, existing whole slide image (WSI) classification methods struggle to effectively incorporate histomorphology information, limiting their ability to capture key pathological features.…

计算机视觉与模式识别 · 计算机科学 2025-12-01 Baizhi Wang , Rui Yan , Wenxin Ma , Xu Zhang , Yuhao Wang , Xiaolong Li , Yunjie Gu , Zihang Jiang , S. Kevin Zhou

Few-shot learning is a promising way for reducing the label cost in new categories adaptation with the guidance of a small, well labeled support set. But for few-shot semantic segmentation, the pixel-level annotations of support images are…

计算机视觉与模式识别 · 计算机科学 2023-11-27 Jing Wang , Yuang Liu , Qiang Zhou , Fan Wang

Whole Slide Images (WSI), obtained by high-resolution digital scanning of microscope slides at multiple scales, are the cornerstone of modern Digital Pathology. However, they represent a particular challenge to AI-based/AI-mediated analysis…

计算机视觉与模式识别 · 计算机科学 2024-04-12 Martim Afonso , Praphulla M. S. Bhawsar , Monjoy Saha , Jonas S. Almeida , Arlindo L. Oliveira

Large-scale pre-trained Vision-Language Models (VLMs), such as CLIP, establish the correlation between texts and images, achieving remarkable success on various downstream tasks with fine-tuning. In existing fine-tuning methods, the…

计算机视觉与模式识别 · 计算机科学 2023-07-31 Yi Zhang , Ce Zhang , Yushun Tang , Zhihai He

Recent advances in Vision-Language Models (VLMs) in histopathology, such as CONCH and QuiltNet, have demonstrated impressive zero-shot classification capabilities across various tasks. However, their general-purpose design may lead to…

计算机视觉与模式识别 · 计算机科学 2025-08-12 Jingna Qiu , Nishanth Jain , Jonas Ammeling , Marc Aubreville , Katharina Breininger