English
Related papers

Related papers: Uni-Hema: Unified Model for Digital Hematopatholog…

200 papers

Multimodal medical image segmentation faces significant challenges in the context of gastric cancer lesion analysis. This clinical context is defined by the scarcity of independent multimodal datasets and the imperative to amalgamate…

Image and Video Processing · Electrical Eng. & Systems 2025-05-28 Jiaming Liang , Lihuan Dai , Xiaoqi Sheng , Xiangguang Chen , Chun Yao , Guihua Tao , Qibin Leng , Hongmin Cai , Xi Zhong

Medical foundation models show promise to learn broadly generalizable features from large, diverse datasets. This could be the base for reliable cross-modality generalization and rapid adaptation to new, task-specific goals, with only a few…

In this work, we investigate multi-task learning as a way of pre-training models for classification tasks in digital pathology. It is motivated by the fact that many small and medium-size datasets have been released by the community over…

Image and Video Processing · Electrical Eng. & Systems 2020-05-19 Romain Mormont , Pierre Geurts , Raphaël Marée

Developing an AI-assisted gland segmentation method from histology images is critical for automatic cancer diagnosis and prognosis; however, the high cost of pixel-level annotations hinders its applications to broader diseases. Existing…

Computer Vision and Pattern Recognition · Computer Science 2022-06-28 Yi Li , Yiduo Yu , Yiwen Zou , Tianqi Xiang , Xiaomeng Li

The development of vision-language models (VLMs) is driven by large-scale and diverse multimodal datasets. However, progress toward generalist biomedical VLMs is limited by the lack of annotated, publicly accessible datasets across biology…

To train a robust deep learning model, one usually needs a balanced set of categories in the training data. The data acquired in a medical domain, however, frequently contains an abundance of healthy patients, versus a small variety of…

Image and Video Processing · Electrical Eng. & Systems 2020-03-09 Jevgenij Gamper , Brandon Chan , Yee Wah Tsang , David Snead , Nasir Rajpoot

Progress in a research field can be hard to assess, in particular when many concurrent methods are proposed in a short period of time. This is the case in digital pathology, where many foundation models have been released recently to serve…

Computer Vision and Pattern Recognition · Computer Science 2026-02-18 Pierre Marza , Leo Fillioux , Sofiène Boutaj , Kunal Mahatha , Christian Desrosiers , Pablo Piantanida , Jose Dolz , Stergios Christodoulidis , Maria Vakalopoulou

Vision-language models (VLMs) are increasingly important in medical applications; however, their evaluation in dermatology remains limited by datasets that focus primarily on image-level classification tasks such as lesion recognition.…

Computer Vision and Pattern Recognition · Computer Science 2026-01-21 Abdurrahim Yilmaz , Ozan Erdem , Ece Gokyayla , Ayda Acar , Burc Bugra Dagtas , Dilara Ilhan Erdil , Gulsum Gencoglan , Burak Temelkuran

Foundation models have substantially advanced computational pathology by learning transferable visual representations from large histological datasets, yet their performance varies widely across tasks due to differences in training data…

Computer Vision and Pattern Recognition · Computer Science 2026-02-16 Wenhui Lei , Yusheng Tan , Anqi Li , Hanyu Chen , Hengrui Tian , Ruiying Li , Zhengqun Jiang , Fang Yan , Xiaofan Zhang , Shaoting Zhang

Histopathology remains the gold standard for cancer diagnosis and prognosis. With the advent of transcriptome profiling, multi-modal learning combining transcriptomics with histology offers more comprehensive information. However, existing…

Image and Video Processing · Electrical Eng. & Systems 2026-03-03 Yupei Zhang , Xiaofei Wang , Anran Liu , Lequan Yu , Chao Li

Human perception of similarity across uni- and multimodal inputs is highly complex, making it challenging to develop automated metrics that accurately mimic it. General purpose vision-language models, such as CLIP and large multi-modal…

Computer Vision and Pattern Recognition · Computer Science 2024-12-17 Sara Ghazanfari , Siddharth Garg , Nicolas Flammarion , Prashanth Krishnamurthy , Farshad Khorrami , Francesco Croce

While deep learning models have become the predominant method for medical image segmentation, they are typically not capable of generalizing to unseen segmentation tasks involving new anatomies, image modalities, or labels. Given a new…

Computer Vision and Pattern Recognition · Computer Science 2023-04-14 Victor Ion Butoi , Jose Javier Gonzalez Ortiz , Tianyu Ma , Mert R. Sabuncu , John Guttag , Adrian V. Dalca

Earlier diagnosis of Leukemia can save thousands of lives annually. The prognosis of leukemia is challenging without the morphological information of White Blood Cells (WBC) and relies on the accessibility of expensive microscopes and the…

Image and Video Processing · Electrical Eng. & Systems 2024-05-20 Abdul Rehman , Talha Meraj , Aiman Mahmood Minhas , Ayisha Imran , Mohsen Ali , Waqas Sultani

Melanoma is the most lethal form of skin cancer, with an increasing incidence rate worldwide. Analyzing histological images of melanoma by localizing and classifying tissues and cell nuclei is considered the gold standard method for…

Computer Vision and Pattern Recognition · Computer Science 2025-09-03 Nima Torbati , Anastasia Meshcheryakova , Ramona Woitek , Sepideh Hatamikia , Diana Mechtcheriakova , Amirreza Mahbod

Melanoma is the most aggressive form of skin cancer, and early detection can significantly increase survival rates and prevent cancer spread. However, developing reliable automated detection techniques is difficult due to the lack of…

Computer Vision and Pattern Recognition · Computer Science 2024-03-25 SangHyuk Kim , Edward Gaibor , Daniel Haehn

Whole slide imaging (WSI) has transformed digital pathology by enabling computational analysis of gigapixel histopathology images. Recent foundation model advances have accelerated progress in computational pathology, facilitating joint…

Computer Vision and Pattern Recognition · Computer Science 2026-03-20 Peihang Wu , Zehong Chen , Lijian Xu

Recent advances in Large Multi-modal Models (LMMs) have demonstrated their remarkable success as general-purpose multi-modal assistants, with particular focuses on holistic image- and video-language understanding. Conversely, less attention…

Computer Vision and Pattern Recognition · Computer Science 2025-11-11 Ye Liu , Zongyang Ma , Junfu Pu , Zhongang Qi , Yang Wu , Ying Shan , Chang Wen Chen

Recent studies have demonstrated the feasibility of modeling single-cell data as natural languages and the potential of leveraging powerful large language models (LLMs) for understanding cell biology. However, a comprehensive evaluation of…

Quantitative Methods · Quantitative Biology 2025-05-14 Fan Zhang , Tianyu Liu , Zhihong Zhu , Hao Wu , Haixin Wang , Donghao Zhou , Yefeng Zheng , Kun Wang , Xian Wu , Pheng-Ann Heng

Foundation models (FMs) are transforming computational pathology by offering new ways to analyze histopathology images. However, FMs typically require weeks of training on large databases, making their creation a resource-intensive process.…

Image and Video Processing · Electrical Eng. & Systems 2026-01-27 Till Nicke , Daniela Schacherer , Jan Raphael Schäfer , Natalia Artysh , Antje Prasse , André Homeyer , Andrea Schenk , Henning Höfener , Johannes Lotz

Aggregating multi-site brain MRI data can enhance deep learning model training, but also introduces non-biological heterogeneity caused by site-specific variations (e.g., differences in scanner vendors, acquisition parameters, and imaging…

Computer Vision and Pattern Recognition · Computer Science 2026-01-14 Mengqi Wu , Yongheng Sun , Qianqian Wang , Pew-Thian Yap , Mingxia Liu