中文
相关论文

相关论文: FaceX-Zoo: A PyTorch Toolbox for Face Recognition

200 篇论文

Identifying species in biology among tens of thousands of visually similar taxa while discovering unknown species in open-world environments remains a fundamental challenge in biodiversity research. Current methods treat identification and…

计算机视觉与模式识别 · 计算机科学 2026-04-28 Jiawei Wang , Ming Lei , Yaning Yang , Xinyan Lin , Yuquan Le , Qiwei Ma , Zhiwei Xu , Zheqi Lv , Yuchen Ang , Zhe Quan , Tat-Seng Chua

Although multimodal fusion has made significant progress, its advancement is severely hindered by the lack of adequate evaluation benchmarks. Current fusion methods are typically evaluated on a small selection of public datasets, a limited…

机器学习 · 计算机科学 2026-05-07 Leyan Xue , Changqing Zhang , Kecheng Xue , Xiaohong Liu , Guangyu Wang , Zongbo Han

This is an official pytorch implementation of Deep High-Resolution Representation Learning for Human Pose Estimation. In this work, we are interested in the human pose estimation problem with a focus on learning reliable high-resolution…

计算机视觉与模式识别 · 计算机科学 2019-02-26 Ke Sun , Bin Xiao , Dong Liu , Jingdong Wang

We present MMFashion, a comprehensive, flexible and user-friendly open-source visual fashion analysis toolbox based on PyTorch. This toolbox supports a wide spectrum of fashion analysis tasks, including Fashion Attribute Prediction, Fashion…

计算机视觉与模式识别 · 计算机科学 2020-05-20 Xin Liu , Jiancheng Li , Jiaqi Wang , Ziwei Liu

Predicting attributes in the landmark free facial images is itself a challenging task which gets further complicated when the face gets occluded due to the usage of masks. Smart access control gates which utilize identity verification or…

计算机视觉与模式识别 · 计算机科学 2022-01-12 Prerana Mukherjee , Vinay Kaushik , Ronak Gupta , Ritika Jha , Daneshwari Kankanwadi , Brejesh Lall

Intelligent surveillance systems often handle perceptual tasks such as object detection, facial recognition, and emotion analysis independently, but they lack a unified, adaptive runtime scheduler that dynamically allocates computational…

Modeling 3D avatars benefits various application scenarios such as AR/VR, gaming, and filming. Character faces contribute significant diversity and vividity as a vital component of avatars. However, building 3D character face models usually…

计算机视觉与模式识别 · 计算机科学 2023-07-06 Zhongjin Luo , Dong Du , Heming Zhu , Yizhou Yu , Hongbo Fu , Xiaoguang Han

Modern face recognition systems (FRS) still fall short when the subjects are wearing facial masks, a common theme in the age of respiratory pandemics. An intuitive partial remedy is to add a mask detector to flag any masked faces so that…

计算机视觉与模式识别 · 计算机科学 2022-04-13 Jiayi Zhu , Qing Guo , Felix Juefei-Xu , Yihao Huang , Yang Liu , Geguang Pu

We present a method for highly efficient landmark detection that combines deep convolutional neural networks with well established model-based fitting algorithms. Motivated by established model-based fitting methods such as active shapes,…

计算机视觉与模式识别 · 计算机科学 2019-02-12 Marcin Kopaczka , Justus Schock , Dorit Merhof

While current dialogue systems like ChatGPT have made significant advancements in text-based interactions, they often overlook the potential of other modalities in enhancing the overall user experience. We present FaceChat, a web-based…

计算与语言 · 计算机科学 2023-03-14 Deema Alnuhait , Qingyang Wu , Zhou Yu

We propose a PiggyBack, a Visual Question Answering platform that allows users to apply the state-of-the-art visual-language pretrained models easily. The PiggyBack supports the full stack of visual question answering tasks, specifically…

计算机视觉与模式识别 · 计算机科学 2022-12-02 Zhihao Zhang , Siwen Luo , Junyi Chen , Sijia Lai , Siqu Long , Hyunsuk Chung , Soyeon Caren Han

Facial attribute analysis in the real world scenario is very challenging mainly because of complex face variations. Existing works of analyzing face attributes are mostly based on the cropped and aligned face images. However, this result in…

计算机视觉与模式识别 · 计算机科学 2017-07-28 Keke He , Yanwei Fu , Xiangyang Xue

With the comprehensive research conducted on various face analysis tasks, there is a growing interest among researchers to develop a unified approach to face perception. Existing methods mainly discuss unified representation and training,…

计算机视觉与模式识别 · 计算机科学 2024-03-15 Lixiong Qin , Mei Wang , Xuannan Liu , Yuhang Zhang , Wei Deng , Xiaoshuai Song , Weiran Xu , Weihong Deng

Vision Foundation Models (VFMs) have advanced representation learning through self-supervised methods. However, existing training pipelines are often inflexible, domain-specific, or computationally expensive, which limits their usability…

计算机视觉与模式识别 · 计算机科学 2025-11-04 Mahmut Selman Gokmen , Cody Bumgardner

As synthetic media, including video, audio, and text, become increasingly indistinguishable from real content, the risks of misinformation, identity fraud, and social manipulation escalate. This survey traces the evolution of deepfake…

计算机视觉与模式识别 · 计算机科学 2025-04-04 Ping Liu , Qiqi Tao , Joey Tianyi Zhou

Recent advances in deep learning have significantly increased the performance of face recognition systems. The performance and reliability of these models depend heavily on the amount and quality of the training data. However, the…

计算机视觉与模式识别 · 计算机科学 2018-02-19 Adam Kortylewski , Andreas Schneider , Thomas Gerig , Bernhard Egger , Andreas Morel-Forster , Thomas Vetter

Depth estimation is a fundamental task in 3D computer vision, crucial for applications such as 3D reconstruction, free-viewpoint rendering, robotics, autonomous driving, and AR/VR technologies. Traditional methods relying on hardware…

计算机视觉与模式识别 · 计算机科学 2025-10-23 Zhen Xu , Hongyu Zhou , Sida Peng , Haotong Lin , Haoyu Guo , Jiahao Shao , Peishan Yang , Qinglin Yang , Sheng Miao , Xingyi He , Yifan Wang , Yue Wang , Ruizhen Hu , Yiyi Liao , Xiaowei Zhou , Hujun Bao

The lack of a common platform and benchmark datasets for evaluating face obfuscation methods has been a challenge, with every method being tested using arbitrary experiments, datasets, and metrics. While prior work has demonstrated that…

计算机视觉与模式识别 · 计算机科学 2025-03-13 Seyyed Mohammad Sadegh Moosavi Khorzooghi , Poojitha Thota , Mohit Singhal , Abolfazl Asudeh , Gautam Das , Shirin Nilizadeh

Multimodal large language models (MLLMs) have shown remarkable performance in vision-language tasks. However, existing MLLMs are primarily trained on generic datasets, limiting their ability to reason on domain-specific visual cues such as…

计算机视觉与模式识别 · 计算机科学 2025-07-15 Hatef Otroshi Shahreza , Sébastien Marcel

Face Recognition has proven to be one of the most successful technology and has impacted heterogeneous domains. Deep learning has proven to be the most successful at computer vision tasks because of its convolution-based architecture. Since…

计算机视觉与模式识别 · 计算机科学 2022-01-11 Jash Dalvi , Sanket Bafna , Devansh Bagaria , Shyamal Virnodkar