English
Related papers

Related papers: Revisiting 2D Foundation Models for Scalable 3D Me…

200 papers

Face alignment, which fits a face model to an image and extracts the semantic meanings of facial pixels, has been an important topic in the computer vision community. However, most algorithms are designed for faces in small to medium poses…

Computer Vision and Pattern Recognition · Computer Science 2018-04-04 Xiangyu Zhu , Xiaoming Liu , Zhen Lei , Stan Z. Li

The recent integration of artificial intelligence into medical imaging has driven remarkable advances in automated organ segmentation. However, most existing 3D segmentation frameworks rely exclusively on visual learning from large…

Computer Vision and Pattern Recognition · Computer Science 2025-12-30 Hasan Faraz Khan , Noor Fatima , Muzammil Behzad

Recent advances in artificial intelligence (AI), in particular self-supervised learning of foundation models (FMs), are revolutionizing medical imaging and computational pathology (CPath). A constant challenge in the analysis of digital…

Vision foundation models (FMs) are accelerating the development of digital pathology algorithms and transforming biomedical research. These models learn, in a self-supervised manner, to represent histological features in highly…

This paper investigates an under-explored but important problem: given a collection of pre-trained neural networks, predicting their performance on each multi-modal task without fine-tuning them, such as image recognition, referring,…

Machine Learning · Computer Science 2023-08-14 Fanqing Meng , Wenqi Shao , Zhanglin Peng , Chonghe Jiang , Kaipeng Zhang , Yu Qiao , Ping Luo

Pathology foundation models (PFMs) have rapidly advanced and are becoming a common backbone for downstream clinical tasks, offering strong transferability across tissues and institutions. However, for dense prediction (e.g., segmentation),…

Image and Video Processing · Electrical Eng. & Systems 2026-02-05 Weiming Chen , Xitong Ling , Xidong Wang , Zhenyang Cai , Yijia Guo , Mingxi Fu , Ziyi Zeng , Minxi Ouyang , Jiawen Li , Yizhi Wang , Tian Guan , Benyou Wang , Yonghong He

Foundation models pre-trained on large-scale datasets demonstrate strong transfer learning capabilities; however, their adaptation to complex multi-label diagnostic tasks-such as comprehensive head CT finding detection-remains understudied.…

Foundation models are vital tools in various Computer Vision applications. They take as input a single RGB image and output a deep feature representation that is useful for various applications. However, in case we have multiple views of…

Computer Vision and Pattern Recognition · Computer Science 2025-12-18 Leo Segre , Or Hirschorn , Shai Avidan

The recent popularity of foundation models and the pre-train-and-adapt paradigm, where a large-scale model is transferred to downstream tasks, is gaining attention for volumetric medical image segmentation. However, current transfer…

Computer Vision and Pattern Recognition · Computer Science 2025-05-12 Julio Silva-Rodríguez , Jose Dolz , Ismail Ben Ayed

Three-dimensional medical image data and computer-aided decision making, particularly using deep learning, are becoming increasingly important in the medical field. To aid in these developments we introduce PR3DICTR: Platform for Research…

Computer Vision and Pattern Recognition · Computer Science 2026-04-06 Daniel C. MacRae , Luuk van der Hoek , Robert van der Wal , Suzanne P. M. de Vette , Hendrike Neh , Baoqiang Ma , Peter M. A. van Ooijen , Lisanne V. van Dijk

Medical image recognition serves as a key way to aid in clinical diagnosis, enabling more accurate and timely identification of diseases and abnormalities. Vision transformer-based approaches have proven effective in handling various…

Computer Vision and Pattern Recognition · Computer Science 2025-08-06 Zunhui Xia , Hongxing Li , Libin Lan

While 3D foundational models have shown promise for promptable segmentation of medical volumes, their robustness to imprecise prompts remains under-explored. In this work, we aim to address this gap by systematically studying the effect of…

Image and Video Processing · Electrical Eng. & Systems 2026-01-26 Soumitri Chattopadhyay , Basar Demir , Marc Niethammer

Foundation models for medical imaging demonstrate superior generalization capabilities across diverse anatomical structures and clinical applications. Their outstanding performance relies on substantial computational resources, limiting…

Image and Video Processing · Electrical Eng. & Systems 2026-04-15 Chen Ma , Jing Jiao , Shuyu Liang , Junhu Fu , Qin Wang , Zeju Li , Yuanyuan Wang , Yi Guo

In-context learning (ICL) offers a promising paradigm for universal medical image analysis, enabling models to perform diverse image processing tasks without retraining. However, current ICL models for medical imaging remain limited in two…

Computer Vision and Pattern Recognition · Computer Science 2025-11-21 Jiesi Hu , Jianfeng Cao , Yanwu Yang , Chenfei Ye , Yixuan Zhang , Hanyang Peng , Ting Ma

Medical data poses a daunting challenge for AI algorithms: it exists in many different modalities, experiences frequent distribution shifts, and suffers from a scarcity of examples and labels. Recent advances, including transformers and…

The rapid rise of large-scale foundation models has reshaped the landscape of image segmentation, with models such as Segment Anything achieving unprecedented versatility across diverse vision tasks. However, previous generations-including…

Computer Vision and Pattern Recognition · Computer Science 2025-11-25 Tianrun Chen , Runlong Cao , Xinda Yu , Lanyun Zhu , Chaotao Ding , Deyi Ji , Cheng Chen , Qi Zhu , Chunyan Xu , Papa Mao , Ying Zang

Pre-trained large-scale models have exhibited remarkable efficacy in computer vision, particularly for 2D image analysis. However, when it comes to 3D point clouds, the constrained accessibility of data, in contrast to the vast repositories…

Computer Vision and Pattern Recognition · Computer Science 2024-08-06 Mengke Li , Da Li , Guoqing Yang , Yiu-ming Cheung , Hui Huang

Automated diagnosis of Alzheimer Disease(AD) from brain imaging, such as magnetic resonance imaging (MRI), has become increasingly important and has attracted the community to contribute many deep learning methods. However, many of these…

Image and Video Processing · Electrical Eng. & Systems 2024-03-01 Yifeng Wang , Ke Chen , Haohan Wang

Radiological analysis increasingly benefits from pretrained visual representations that can support heterogeneous downstream tasks across imaging modalities. In this work, we introduce OmniRad, a self-supervised radiological foundation…

Computer Vision and Pattern Recognition · Computer Science 2026-02-05 Luca Zedda , Andrea Loddo , Cecilia Di Ruberto