English
Related papers

Related papers: FRoundation: Are Foundation Models Ready for Face …

200 papers

Recent years have seen an explosion of diverse general purpose pre-training methodologies for computer vision. However, the impact that these pre-training methodologies have on person identification tasks (re-id) remains under-explored. We…

Computer Vision and Pattern Recognition · Computer Science 2026-05-22 Thomas M. Metz , Matthew Q. Hill , Alice J. O'Toole

Foundation models leverage large-scale pretraining to capture extensive knowledge, demonstrating generalization in a wide range of language tasks. By comparison, vision foundation models (VFMs) often exhibit uneven improvements across…

Computer Vision and Pattern Recognition · Computer Science 2026-01-23 Shiqi Huang , Yipei Wang , Natasha Thorley , Alexander Ng , Shaheer Saeed , Mark Emberton , Shonit Punwani , Veeru Kasivisvanathan , Dean Barratt , Daniel Alexander , Yipeng Hu

The rapid evolution of generative models has enabled the creation of hyper-realistic facial deepfakes, exposing a critical vulnerability in modern digital forensics: the inability of detectors to generalize to unseen manipulation…

Computer Vision and Pattern Recognition · Computer Science 2026-05-26 Ibrahim Delibasoglu

This paper studies how to synthesize face images of non-existent persons, to create a dataset that allows effective training of face recognition (FR) models. Besides generating realistic face images, two other important goals are: 1) the…

Computer Vision and Pattern Recognition · Computer Science 2025-02-10 Haiyu Wu , Jaskirat Singh , Sicong Tian , Liang Zheng , Kevin W. Bowyer

The ability to accurately recognize an individual's face with respect to human aging factor holds significant importance for various private as well as government sectors such as customs and public security bureaus, passport office, and…

Computer Vision and Pattern Recognition · Computer Science 2024-10-23 Wang Yao , Muhammad Ali Farooq , Joseph Lemley , Peter Corcoran

Foundation models for computational pathology are expected to facilitate the development of high-performing, generalisable deep learning systems. However, in addition to biologically relevant features, current foundation models also capture…

Self-supervised learning (SSL) has enabled Vision Transformers (ViTs) to learn robust representations from large-scale natural image datasets, enhancing their generalization across domains. In retinal imaging, foundation models pretrained…

Image and Video Processing · Electrical Eng. & Systems 2025-05-23 Benjamin A. Cohen , Jonathan Fhima , Meishar Meisel , Baskin Meital , Luis Filipe Nakayama , Eran Berkowitz , Joachim A. Behar

Foundation models can be disruptive for future AI development by scaling up deep learning in terms of model size and training data's breadth and size. These models achieve state-of-the-art performance (often through further adaptation) on a…

Artificial Intelligence · Computer Science 2022-12-20 Johannes Schneider

The vulnerability of facial recognition systems to face morphing attacks is well known. Many different approaches for morphing attack detection have been proposed in the scientific literature. However, the morphing attack detection…

Cryptography and Security · Computer Science 2020-04-06 Ulrich Scherhag , Christian Rathgeb , Johannes Merkle , Christoph Busch

In recent years large model trained on huge amount of cross-modality data, which is usually be termed as foundation model, achieves conspicuous accomplishment in many fields, such as image recognition and generation. Though achieving great…

Computer Vision and Pattern Recognition · Computer Science 2023-08-02 Shiqi Yang , Atsushi Hashimoto , Yoshitaka Ushiku

Significant progress in the development of highly adaptable and reusable Artificial Intelligence (AI) models is expected to have a significant impact on Earth science and remote sensing. Foundation models are pre-trained on large unlabeled…

State-of-the-art face recognition (FR) approaches have shown remarkable results in predicting whether two faces belong to the same identity, yielding accuracies between 92% and 100% depending on the difficulty of the protocol. However, the…

Computer Vision and Pattern Recognition · Computer Science 2022-05-30 Stefan Hörmann , Tianlin Kong , Torben Teepe , Fabian Herzog , Martin Knoche , Gerhard Rigoll

Image-text training like CLIP has dominated the pretraining of vision foundation models in recent years. Subsequent efforts have been made to introduce region-level visual learning into CLIP's pretraining but face scalability challenges due…

Computer Vision and Pattern Recognition · Computer Science 2024-04-12 Xiaohu Jiang , Yixiao Ge , Yuying Ge , Dachuan Shi , Chun Yuan , Ying Shan

Recent deep face recognition models proposed in the literature utilized large-scale public datasets such as MS-Celeb-1M and VGGFace2 for training very deep neural networks, achieving state-of-the-art performance on mainstream benchmarks.…

Computer Vision and Pattern Recognition · Computer Science 2022-06-22 Fadi Boutros , Marco Huber , Patrick Siebke , Tim Rieber , Naser Damer

Although face recognition systems have seen a massive performance enhancement in recent years, they are still targeted by threats such as presentation attacks, leading to the need for generalizable presentation attack detection (PAD)…

Computer Vision and Pattern Recognition · Computer Science 2025-01-07 Guray Ozgur , Eduarda Caldeira , Tahar Chettaoui , Fadi Boutros , Raghavendra Ramachandra , Naser Damer

This paper presents Arc2Face, an identity-conditioned face foundation model, which, given the ArcFace embedding of a person, can generate diverse photo-realistic images with an unparalleled degree of face similarity than existing models.…

Computer Vision and Pattern Recognition · Computer Science 2024-08-26 Foivos Paraperas Papantoniou , Alexandros Lattas , Stylianos Moschoglou , Jiankang Deng , Bernhard Kainz , Stefanos Zafeiriou

Foundation Models (FMs) have shown impressive performance on various text and image processing tasks. They can generalize across domains and datasets in a zero-shot setting. This could make them suitable for automated quality inspection…

Computer Vision and Pattern Recognition · Computer Science 2025-09-26 Simon Baeuerle , Pratik Khanna , Nils Friederich , Angelo Jovin Yamachui Sitcheu , Damir Shakirov , Andreas Steimer , Ralf Mikut

Face recognition (FR) systems for video surveillance (VS) applications attempt to accurately detect the presence of target individuals over a distributed network of cameras. In video-based FR systems, facial models of target individuals are…

Computer Vision and Pattern Recognition · Computer Science 2018-06-28 Saman Bashbaghi , Eric Granger , Robert Sabourin , Mostafa Parchami

In this study, we explore the efficacy of advanced pre-trained architectures, such as Vision Transformers (ViT), ConvNeXt, and Swin Transformers in enhancing Federated Domain Generalization. These architectures capture global contextual…

Computer Vision and Pattern Recognition · Computer Science 2024-09-27 Avi Deb Raha , Apurba Adhikary , Mrityunjoy Gain , Yu Qiao , Choong Seon Hong

Foundation models are used to extract transferable representations from large amounts of unlabeled data, typically via self-supervised learning (SSL). However, many of these models rely on architectures that offer limited interpretability,…

Computer Vision and Pattern Recognition · Computer Science 2026-04-15 Samuel Ofosu Mensah , Camila Roa , Kerol Djoumessi , Philipp Berens