English
Related papers

Related papers: When Do Domain-Specific Foundation Models Justify …

200 papers

The rapid development of Vision Foundation Models (VFMs), particularly Vision Transformers (ViT) and Segment Anything Model (SAM), has sparked significant advances in the field of medical image analysis. These models have demonstrated…

Image and Video Processing · Electrical Eng. & Systems 2025-02-24 Pengchen Liang , Bin Pu , Haishan Huang , Yiwei Li , Hualiang Wang , Weibo Ma , Qing Chang

The scarcity of high-quality, labelled retinal imaging data, which presents a significant challenge in the development of machine learning models for ophthalmology, hinders progress in the field. Existing methods for synthesising Colour…

Image and Video Processing · Electrical Eng. & Systems 2025-07-18 Junzhi Ning , Cheng Tang , Kaijing Zhou , Diping Song , Lihao Liu , Ming Hu , Wei Li , Huihui Xu , Yanzhou Su , Tianbin Li , Jiyao Liu , Jin Ye , Sheng Zhang , Yuanfeng Ji , Junjun He

Purpose: This study provides the first comprehensive evaluation of foundation models in fetal ultrasound (US) imaging under low inter-class variability conditions. While recent vision foundation models such as DINOv3 have shown remarkable…

Computer Vision and Pattern Recognition · Computer Science 2025-11-05 Edoardo Conti , Riccardo Rosati , Lorenzo Federici , Adriano Mancini , Maria Chiara Fiorentin

Using deep learning models pre-trained on Imagenet is the traditional solution for medical image classification to deal with data scarcity. Nevertheless, relevant literature supports that this strategy may offer limited gains due to the…

Computer Vision and Pattern Recognition · Computer Science 2024-01-30 Julio Silva-Rodriguez , Jihed Chelbi , Waziha Kabir , Hadi Chakor , Jose Dolz , Ismail Ben Ayed , Riadh Kobbi

Diabetic Retinopathy (DR) constitutes 5% of global blindness cases. While numerous deep learning approaches have sought to enhance traditional DR grading methods, they often falter when confronted with new out-of-distribution data thereby…

Image and Video Processing · Electrical Eng. & Systems 2024-11-06 Sharon Chokuwa , Muhammad Haris Khan

This paper addresses the emerging task of recognizing multiple retinal diseases from wide-field (WF) and ultra-wide-field (UWF) fundus images. For an effective use of existing large amount of labeled color fundus photo (CFP) data and the…

Image and Video Processing · Electrical Eng. & Systems 2023-10-25 Qijie Wei , Jingyuan Yang , Bo Wang , Jinrui Wang , Jianchun Zhao , Xinyu Zhao , Sheng Yang , Niranchana Manivannan , Youxin Chen , Dayong Ding , Jing Zhou , Xirong Li

Although deep learning based diabetic retinopathy (DR) classification methods typically benefit from well-designed architectures of convolutional neural networks, the training setting also has a non-negligible impact on the prediction…

Image and Video Processing · Electrical Eng. & Systems 2022-10-19 Yijin Huang , Li Lin , Pujin Cheng , Junyan Lyu , Roger Tam , Xiaoying Tang

Vessel segmentation of retinal images is a key diagnostic capability in ophthalmology. This problem faces several challenges including low contrast, variable vessel size and thickness, and presence of interfering pathology such as…

Image and Video Processing · Electrical Eng. & Systems 2020-02-19 Venkateswararao Cherukuri , Vijay Kumar BG , Raja Bala , Vishal Monga

Despite the significant potential of Foundation Models (FMs) in medical imaging, their application to prognosis prediction remains challenging due to data scarcity, class imbalance, and task complexity, which limit their clinical adoption.…

Computer Vision and Pattern Recognition · Computer Science 2026-01-16 Filippo Ruffini , Elena Mulero Ayllon , Linlin Shen , Paolo Soda , Valerio Guarrasi

Background: The lack of explanations for the decisions made by algorithms such as deep learning has hampered their acceptance by the clinical community despite highly accurate results on multiple problems. Recently, attribution methods have…

Image and Video Processing · Electrical Eng. & Systems 2021-03-26 Amitojdeep Singh , J. Jothi Balaji , Mohammed Abdul Rasheed , Varadharajan Jayakumar , Rajiv Raman , Vasudevan Lakshminarayanan

Assessing the degree of disease severity in biomedical images is a task similar to standard classification but constrained by an underlying structure in the label space. Such a structure reflects the monotonic relationship between different…

Computer Vision and Pattern Recognition · Computer Science 2020-10-02 Adrian Galdran , José Dolz , Hadi Chakor , Hervé Lombaert , Ismail Ben Ayed

Our research focuses on the critical field of early diagnosis of disease by examining retinal blood vessels in fundus images. While automatic segmentation of retinal blood vessels holds promise for early detection, accurate analysis remains…

Image and Video Processing · Electrical Eng. & Systems 2024-05-14 Fatema Tuj Johora Faria , Mukaffi Bin Moin , Pronay Debnath , Asif Iftekher Fahim , Faisal Muhammad Shah

Foundational models are trained on extensive datasets to capture the general trends of a domain. However, in medical imaging, the scarcity of data makes pre-training for every domain, modality, or task challenging. Continual learning offers…

Image and Video Processing · Electrical Eng. & Systems 2025-08-20 Mohammad Areeb Qazi , Munachiso S Nwadike , Ibrahim Almakky , Mohammad Yaqub , Numan Saeed

Although vision foundation models (VFMs) are increasingly reused for biomedical image analysis, it remains unclear whether the latent representations they provide are general enough to support effective transfer and reuse across…

Computer Vision and Pattern Recognition · Computer Science 2026-02-10 Caterina Fuster-Barceló , Virginie Uhlmann

From diagnosing neovascular diseases to detecting white matter lesions, accurate tiny vessel segmentation in fundus images is critical. Promising results for accurate vessel segmentation have been known. However, their effectiveness in…

Computer Vision and Pattern Recognition · Computer Science 2021-04-20 Suraj Mishra , Danny Z. Chen , X. Sharon Hu

Recent advancements in pre-trained large foundation models (LFM) have yielded significant breakthroughs across various domains, including natural language processing and computer vision. These models have been particularly impactful in the…

Image and Video Processing · Electrical Eng. & Systems 2024-05-22 Ziqin Lin , Heng Li , Zinan Li , Huazhu Fu , Jiang Liu

Image super-resolution (SR) is a fast-moving field with novel architectures attracting the spotlight. However, most SR models were optimized with dated training strategies. In this work, we revisit the popular RCAN model and examine the…

Computer Vision and Pattern Recognition · Computer Science 2022-01-28 Zudi Lin , Prateek Garg , Atmadeep Banerjee , Salma Abdel Magid , Deqing Sun , Yulun Zhang , Luc Van Gool , Donglai Wei , Hanspeter Pfister

Convolutional neural networks (CNNs) show impressive performance for image classification and detection, extending heavily to the medical image domain. Nevertheless, medical experts are sceptical in these predictions as the nonlinear…

Computer Vision and Pattern Recognition · Computer Science 2017-06-30 Waleed M. Gondal , Jan M. Köhler , René Grzeszick , Gernot A. Fink , Michael Hirsch

Automated analysis of optical coherence tomography (OCT) and OCT angiography (OCTA) images is critical for robust ophthalmic diagnosis. Existing mainstream methods trained from scratch rely heavily on massive data and model scale, thereby…

Computer Vision and Pattern Recognition · Computer Science 2026-04-07 Xiaofei Su , Zengshuo Wang , Minghe Sun , Xin Zhao , Mingzhu Sun

We survey applications of pretrained foundation models in robotics. Traditional deep learning models in robotics are trained on small datasets tailored for specific tasks, which limits their adaptability across diverse applications. In…