English
Related papers

Related papers: Foundation versus Domain-specific Models: Performa…

200 papers

Foundation models are rapidly transforming Earth Observation data mining by enabling generalizable and scalable solutions for key tasks such as scene classification and semantic segmentation. While most efforts in the geospatial domain have…

Computer Vision and Pattern Recognition · Computer Science 2025-06-27 Man Duc Chuc

Cross-domain biometrics has been emerging as a new necessity, which poses several additional challenges, including harsh illumination changes, noise, pose variation, among others. In this paper, we explore approaches to cross-domain face…

Computer Vision and Pattern Recognition · Computer Science 2016-11-18 Guilherme Folego , Marcus A. Angeloni , José Augusto Stuchi , Alan Godoy , Anderson Rocha

Retinal foundation models aim to learn generalizable representations from diverse retinal images, facilitating label-efficient model adaptation across various ophthalmic tasks. Despite their success, current retinal foundation models are…

Computer Vision and Pattern Recognition · Computer Science 2024-08-13 Kai Yu , Yang Zhou , Yang Bai , Zhi Da Soh , Xinxing Xu , Rick Siow Mong Goh , Ching-Yu Cheng , Yong Liu

Foundation models hold promise for specialized medical imaging tasks, though their effectiveness in breast imaging remains underexplored. This study leverages BiomedCLIP as a foundation model to address challenges in model generalization.…

Understanding the semantics of individual regions or patches of unconstrained images, such as open-world object detection, remains a critical yet challenging task in computer vision. Building on the success of powerful image-level…

Computer Vision and Pattern Recognition · Computer Science 2025-04-22 Haosen Yang , Chuofan Ma , Bin Wen , Yi Jiang , Zehuan Yuan , Xiatian Zhu

Large-scale Vision-Language Foundation Models (VLFMs), such as CLIP, now underpin a wide range of computer vision research and applications. VLFMs are often adapted to various domain-specific tasks. However, VLFM performance on novel,…

Computer Vision and Pattern Recognition · Computer Science 2026-03-05 Chris Vorster , Mayug Maniparambil , Noel E. O'Connor , Noel Murphy , Derek Molloy

Recognizing and differentiating among both familiar and unfamiliar faces is a critical capability for face recognition systems and a key step toward artificial general intelligence (AGI). Motivated by this ability, this paper introduces…

Computer Vision and Pattern Recognition · Computer Science 2025-07-31 Yunseok Oh , Dong-Wan Choi

In this paper, we present a novel, scalable approach for constructing open set, instance-level 3D scene representations, advancing open world understanding of 3D environments. Existing methods require pre-constructed 3D scenes and face…

Computer Vision and Pattern Recognition · Computer Science 2024-09-17 Rafay Mohiuddin , Sai Manoj Prakhya , Fiona Collins , Ziyuan Liu , André Borrmann

Deep learning technology has enabled successful modeling of complex facial features when high quality images are available. Nonetheless, accurate modeling and recognition of human faces in real world scenarios `on the wild' or under adverse…

Computer Vision and Pattern Recognition · Computer Science 2020-11-30 S. W. Arachchilage , E. Izquierdo

Although deep neural networks offer better face detection results than shallow or handcrafted models, their complex architectures come with higher computational requirements and slower inference speeds than shallow neural networks. In this…

Computer Vision and Pattern Recognition · Computer Science 2018-11-29 Petru Soviany , Radu Tudor Ionescu

Facial image retrieval is a challenging task since faces have many similar features (areas), which makes it difficult for the retrieval systems to distinguish faces of different people. With the advent of deep learning, deep networks are…

Computer Vision and Pattern Recognition · Computer Science 2018-12-14 Ahmad S. Tarawneh , Ahmad B. A. Hassanat , Ceyhun Celik , Dmitry Chetverikov , M. Sohel Rahman , Chaman Verma

We address the problem of face anti-spoofing which aims to make the face verification systems robust in the real world settings. The context of detecting live vs. spoofed face images may differ significantly in the target domain, when…

Computer Vision and Pattern Recognition · Computer Science 2021-05-19 Ankush Panwar , Pratyush Singh , Suman Saha , Danda Pani Paudel , Luc Van Gool

Plastic surgery and disguise variations are two of the most challenging co-variates of face recognition. The state-of-art deep learning models are not sufficiently successful due to the availability of limited training samples. In this…

Computer Vision and Pattern Recognition · Computer Science 2018-11-20 Saksham Suri , Anush Sankaran , Mayank Vatsa , Richa Singh

Visual recognition in low-data regimes requires deep neural networks to learn generalized representations from limited training samples. Recently, CLIP-based methods have shown promising few-shot performance benefited from the contrastive…

Computer Vision and Pattern Recognition · Computer Science 2023-03-06 Renrui Zhang , Xiangfei Hu , Bohao Li , Siyuan Huang , Hanqiu Deng , Hongsheng Li , Yu Qiao , Peng Gao

Existing object recognition models have been shown to lack robustness in diverse geographical scenarios due to domain shifts in design and context. Class representations need to be adapted to more accurately reflect an object concept under…

Computer Vision and Pattern Recognition · Computer Science 2024-04-02 Kyle Buettner , Sina Malakouti , Xiang Lorraine Li , Adriana Kovashka

Current face anonymization techniques often depend on identity loss calculated by face recognition models, which can be inaccurate and unreliable. Additionally, many methods require supplementary data such as facial landmarks and masks to…

Computer Vision and Pattern Recognition · Computer Science 2024-11-04 Han-Wei Kung , Tuomas Varanka , Sanjay Saha , Terence Sim , Nicu Sebe

Foundation models have demonstrated a remarkable ability to learn rich, transferable representations across diverse modalities such as images, text, and audio. In modern machine learning pipelines, these representations often replace raw…

Computer Vision and Pattern Recognition · Computer Science 2026-03-10 Selene Cerna , Sara Si-Moussi , Wilfried Thuiller , Hadrien Hendrikx , Vincent Miele

2D face recognition encounters challenges in unconstrained environments due to varying illumination, occlusion, and pose. Recent studies focus on RGB-D face recognition to improve robustness by incorporating depth information. However,…

Computer Vision and Pattern Recognition · Computer Science 2024-03-19 Zijian Chen , Mei Wang , Weihong Deng , Hongzhi Shi , Dongchao Wen , Yingjie Zhang , Xingchen Cui , Jian Zhao

This work addresses a fundamental barrier in recommender systems: the inability to generalize across domains without extensive retraining. Traditional ID-based approaches fail entirely in cold-start and cross-domain scenarios where new…

Information Retrieval · Computer Science 2025-06-16 Yangqin Jiang , Xubin Ren , Lianghao Xia , Da Luo , Kangyi Lin , Chao Huang

To address the problem of data inconsistencies among different facial expression recognition (FER) datasets, many cross-domain FER methods (CD-FERs) have been extensively devised in recent years. Although each declares to achieve superior…

Computer Vision and Pattern Recognition · Computer Science 2021-12-01 Tianshui Chen , Tao Pu , Hefeng Wu , Yuan Xie , Lingbo Liu , Liang Lin