中文
相关论文

相关论文: Foundation versus Domain-specific Models: Performa…

200 篇论文

Foundation models are rapidly transforming Earth Observation data mining by enabling generalizable and scalable solutions for key tasks such as scene classification and semantic segmentation. While most efforts in the geospatial domain have…

计算机视觉与模式识别 · 计算机科学 2025-06-27 Man Duc Chuc

Cross-domain biometrics has been emerging as a new necessity, which poses several additional challenges, including harsh illumination changes, noise, pose variation, among others. In this paper, we explore approaches to cross-domain face…

计算机视觉与模式识别 · 计算机科学 2016-11-18 Guilherme Folego , Marcus A. Angeloni , José Augusto Stuchi , Alan Godoy , Anderson Rocha

Retinal foundation models aim to learn generalizable representations from diverse retinal images, facilitating label-efficient model adaptation across various ophthalmic tasks. Despite their success, current retinal foundation models are…

计算机视觉与模式识别 · 计算机科学 2024-08-13 Kai Yu , Yang Zhou , Yang Bai , Zhi Da Soh , Xinxing Xu , Rick Siow Mong Goh , Ching-Yu Cheng , Yong Liu

Foundation models hold promise for specialized medical imaging tasks, though their effectiveness in breast imaging remains underexplored. This study leverages BiomedCLIP as a foundation model to address challenges in model generalization.…

Understanding the semantics of individual regions or patches of unconstrained images, such as open-world object detection, remains a critical yet challenging task in computer vision. Building on the success of powerful image-level…

计算机视觉与模式识别 · 计算机科学 2025-04-22 Haosen Yang , Chuofan Ma , Bin Wen , Yi Jiang , Zehuan Yuan , Xiatian Zhu

Large-scale Vision-Language Foundation Models (VLFMs), such as CLIP, now underpin a wide range of computer vision research and applications. VLFMs are often adapted to various domain-specific tasks. However, VLFM performance on novel,…

计算机视觉与模式识别 · 计算机科学 2026-03-05 Chris Vorster , Mayug Maniparambil , Noel E. O'Connor , Noel Murphy , Derek Molloy

Recognizing and differentiating among both familiar and unfamiliar faces is a critical capability for face recognition systems and a key step toward artificial general intelligence (AGI). Motivated by this ability, this paper introduces…

计算机视觉与模式识别 · 计算机科学 2025-07-31 Yunseok Oh , Dong-Wan Choi

In this paper, we present a novel, scalable approach for constructing open set, instance-level 3D scene representations, advancing open world understanding of 3D environments. Existing methods require pre-constructed 3D scenes and face…

计算机视觉与模式识别 · 计算机科学 2024-09-17 Rafay Mohiuddin , Sai Manoj Prakhya , Fiona Collins , Ziyuan Liu , André Borrmann

Deep learning technology has enabled successful modeling of complex facial features when high quality images are available. Nonetheless, accurate modeling and recognition of human faces in real world scenarios `on the wild' or under adverse…

计算机视觉与模式识别 · 计算机科学 2020-11-30 S. W. Arachchilage , E. Izquierdo

Although deep neural networks offer better face detection results than shallow or handcrafted models, their complex architectures come with higher computational requirements and slower inference speeds than shallow neural networks. In this…

计算机视觉与模式识别 · 计算机科学 2018-11-29 Petru Soviany , Radu Tudor Ionescu

Facial image retrieval is a challenging task since faces have many similar features (areas), which makes it difficult for the retrieval systems to distinguish faces of different people. With the advent of deep learning, deep networks are…

计算机视觉与模式识别 · 计算机科学 2018-12-14 Ahmad S. Tarawneh , Ahmad B. A. Hassanat , Ceyhun Celik , Dmitry Chetverikov , M. Sohel Rahman , Chaman Verma

We address the problem of face anti-spoofing which aims to make the face verification systems robust in the real world settings. The context of detecting live vs. spoofed face images may differ significantly in the target domain, when…

计算机视觉与模式识别 · 计算机科学 2021-05-19 Ankush Panwar , Pratyush Singh , Suman Saha , Danda Pani Paudel , Luc Van Gool

Plastic surgery and disguise variations are two of the most challenging co-variates of face recognition. The state-of-art deep learning models are not sufficiently successful due to the availability of limited training samples. In this…

计算机视觉与模式识别 · 计算机科学 2018-11-20 Saksham Suri , Anush Sankaran , Mayank Vatsa , Richa Singh

Visual recognition in low-data regimes requires deep neural networks to learn generalized representations from limited training samples. Recently, CLIP-based methods have shown promising few-shot performance benefited from the contrastive…

计算机视觉与模式识别 · 计算机科学 2023-03-06 Renrui Zhang , Xiangfei Hu , Bohao Li , Siyuan Huang , Hanqiu Deng , Hongsheng Li , Yu Qiao , Peng Gao

Existing object recognition models have been shown to lack robustness in diverse geographical scenarios due to domain shifts in design and context. Class representations need to be adapted to more accurately reflect an object concept under…

计算机视觉与模式识别 · 计算机科学 2024-04-02 Kyle Buettner , Sina Malakouti , Xiang Lorraine Li , Adriana Kovashka

Current face anonymization techniques often depend on identity loss calculated by face recognition models, which can be inaccurate and unreliable. Additionally, many methods require supplementary data such as facial landmarks and masks to…

计算机视觉与模式识别 · 计算机科学 2024-11-04 Han-Wei Kung , Tuomas Varanka , Sanjay Saha , Terence Sim , Nicu Sebe

Foundation models have demonstrated a remarkable ability to learn rich, transferable representations across diverse modalities such as images, text, and audio. In modern machine learning pipelines, these representations often replace raw…

计算机视觉与模式识别 · 计算机科学 2026-03-10 Selene Cerna , Sara Si-Moussi , Wilfried Thuiller , Hadrien Hendrikx , Vincent Miele

2D face recognition encounters challenges in unconstrained environments due to varying illumination, occlusion, and pose. Recent studies focus on RGB-D face recognition to improve robustness by incorporating depth information. However,…

计算机视觉与模式识别 · 计算机科学 2024-03-19 Zijian Chen , Mei Wang , Weihong Deng , Hongzhi Shi , Dongchao Wen , Yingjie Zhang , Xingchen Cui , Jian Zhao

This work addresses a fundamental barrier in recommender systems: the inability to generalize across domains without extensive retraining. Traditional ID-based approaches fail entirely in cold-start and cross-domain scenarios where new…

信息检索 · 计算机科学 2025-06-16 Yangqin Jiang , Xubin Ren , Lianghao Xia , Da Luo , Kangyi Lin , Chao Huang

To address the problem of data inconsistencies among different facial expression recognition (FER) datasets, many cross-domain FER methods (CD-FERs) have been extensively devised in recent years. Although each declares to achieve superior…

计算机视觉与模式识别 · 计算机科学 2021-12-01 Tianshui Chen , Tao Pu , Hefeng Wu , Yuan Xie , Lingbo Liu , Liang Lin