中文
相关论文

相关论文: Towards a Visual-Language Foundation Model for Com…

200 篇论文

While machine learning is currently transforming the field of histopathology, the domain lacks a comprehensive evaluation of state-of-the-art models based on essential but complementary quality requirements beyond a mere classification…

图像与视频处理 · 电气工程与系统科学 2023-05-11 Maximilian Springenberg , Annika Frommholz , Markus Wenzel , Eva Weicken , Jackie Ma , Nils Strodthoff

Foundation models increasingly offer potential to support interactive, agentic workflows that assist researchers during analysis and interpretation of image data. Such workflows often require coupling vision to language to provide a…

计算机视觉与模式识别 · 计算机科学 2026-02-27 Matthew Sutton , Katrin Amunts , Timo Dickscheid , Christian Schiffer

Recent breakthroughs in self-supervised learning have enabled the use of large unlabeled datasets to train visual foundation models that can generalize to a variety of downstream tasks. While this training paradigm is well suited for the…

Self-supervised pretraining attempts to enhance model performance by obtaining effective features from unlabeled data, and has demonstrated its effectiveness in the field of histopathology images. Despite its success, few works concentrate…

计算机视觉与模式识别 · 计算机科学 2023-09-22 Zhiyun Song , Penghui Du , Junpeng Yan , Kailu Li , Jianzhong Shou , Maode Lai , Yubo Fan , Yan Xu

Computational pathology foundation models (CPathFMs) have emerged as a powerful approach for analyzing histopathological data, leveraging self-supervised learning to extract robust feature representations from unlabeled whole-slide images.…

计算机视觉与模式识别 · 计算机科学 2025-02-27 Dong Li , Guihong Wan , Xintao Wu , Xinyu Wu , Ajit J. Nirmal , Christine G. Lian , Peter K. Sorger , Yevgeniy R. Semenov , Chen Zhao

Vision-language models can connect the text description of an object to its specific location in an image through visual grounding. This has potential applications in enhanced radiology reporting. However, these models require large…

计算机视觉与模式识别 · 计算机科学 2025-02-04 Zachary Huemann , Samuel Church , Joshua D. Warner , Daniel Tran , Xin Tie , Alan B McMillan , Junjie Hu , Steve Y. Cho , Meghan Lubner , Tyler J. Bradshaw

Unsupervised learning has made substantial progress over the last few years, especially by means of contrastive self-supervised learning. The dominating dataset for benchmarking self-supervised learning has been ImageNet, for which recent…

图像与视频处理 · 电气工程与系统科学 2022-08-17 Karin Stacke , Jonas Unger , Claes Lundström , Gabriel Eilertsen

The previous advancements in pathology image understanding primarily involved developing models tailored to specific tasks. Recent studies has demonstrated that the large vision-language model can enhance the performance of various…

人工智能 · 计算机科学 2024-08-20 Dawei Dai , Yuanhui Zhang , Long Xu , Qianlan Yang , Xiaojing Shen , Shuyin Xia , Guoyin Wang

Representation learning offers a conduit to elucidate distinctive features within the latent space and interpret the deep models. However, the randomness of lesion distribution and the complexity of low-quality factors in medical images…

计算机视觉与模式识别 · 计算机科学 2024-04-09 Qingshan Hou , Shuai Cheng , Peng Cao , Jinzhu Yang , Xiaoli Liu , Osmar R. Zaiane , Yih Chung Tham

Recently, deep neural networks have greatly advanced histopathology image segmentation but usually require abundant annotated data. However, due to the gigapixel scale of whole slide images and pathologists' heavy daily workload, obtaining…

计算机视觉与模式识别 · 计算机科学 2023-03-13 Wentao Pan , Jiangpeng Yan , Hanbo Chen , Jiawei Yang , Zhe Xu , Xiu Li , Jianhua Yao

Content-based medical image retrieval is an important diagnostic tool that improves the explainability of computer-aided diagnosis systems and provides decision making support to healthcare professionals. Medical imaging data, such as…

计算机视觉与模式识别 · 计算机科学 2022-11-23 Yunyan Xing , Benjamin J. Meyer , Mehrtash Harandi , Tom Drummond , Zongyuan Ge

Pathological image segmentation faces numerous challenges, particularly due to ambiguous semantic boundaries and the high cost of pixel-level annotations. Although recent semi-supervised methods based on consistency regularization (e.g.,…

计算机视觉与模式识别 · 计算机科学 2025-08-28 Mingxi Fu , Fanglei Fu , Xitong Ling , Huaitian Yuan , Tian Guan , Yonghong He , Lianghui Zhu

This work investigates descriptive captions as an additional source of supervision for biological multimodal foundation models. Images and captions can be viewed as complementary samples from the latent morphospace of a species, each…

Despite their successes in vision and language, foundation models have stumbled in pathology, revealing low accuracy, instability, and heavy computational demands. These shortcomings stem not from tuning problems but from deeper conceptual…

人工智能 · 计算机科学 2026-04-21 Hamid R. Tizhoosh

This study evaluates the generalisation capabilities of state-of-the-art histopathology foundation models on out-of-distribution multi-stain autoimmune Immunohistochemistry datasets. We compare 13 feature extractor models, including…

计算机视觉与模式识别 · 计算机科学 2024-10-30 Amaya Gallagher-Syed , Elena Pontarini , Myles J. Lewis , Michael R. Barnes , Gregory Slabaugh

The development of clinical-grade artificial intelligence in pathology is limited by the scarcity of diverse, high-quality annotated datasets. Generative models offer a potential solution but suffer from semantic instability and…

Multi-modal data abounds in biomedicine, such as radiology images and reports. Interpreting this data at scale is essential for improving clinical care and accelerating clinical research. Biomedical text with its complex semantics poses…

In computational pathology, understanding and generation have evolved along disparate paths: advanced understanding models already exhibit diagnostic-level competence, whereas generative models largely simulate pixels. Progress remains…

计算机视觉与模式识别 · 计算机科学 2026-02-27 Minghao Han , Yichen Liu , Yizhou Liu , Zizhi Chen , Jingqun Tang , Xuecheng Wu , Dingkang Yang , Lihua Zhang

Histopathological images contain rich phenotypic information that can be used to monitor underlying mechanisms contributing to diseases progression and patient survival outcomes. Recently, deep learning has become the mainstream…

图像与视频处理 · 电气工程与系统科学 2020-10-28 Chetan L. Srinidhi , Ozan Ciga , Anne L. Martel

The field of computational pathology has witnessed remarkable progress in the development of both task-specific predictive models and task-agnostic self-supervised vision encoders. However, despite the explosive growth of generative…

计算机视觉与模式识别 · 计算机科学 2023-12-14 Ming Y. Lu , Bowen Chen , Drew F. K. Williamson , Richard J. Chen , Kenji Ikamura , Georg Gerber , Ivy Liang , Long Phi Le , Tong Ding , Anil V Parwani , Faisal Mahmood