中文
相关论文

相关论文: Faceptor: A Generalist Model for Face Perception

200 篇论文

Deepface generation has traditionally followed a task-driven paradigm, where distinct tasks (e.g., face transfer and hair transfer) are addressed by task-specific models. Nevertheless, this single-task setting severely limits model…

计算机视觉与模式识别 · 计算机科学 2026-03-23 Caiyi Sun , Yujing Sun , Xiangyu Li , Yuhang Zheng , Yiming Ren , Jiamin Wang , Yuexin Ma , Siu-Ming Yiu

In recent years, there has been increasing interest in automatic facial behavior analysis systems from computing communities such as vision, multimodal interaction, robotics, and affective computing. Building upon the widespread utility of…

计算机视觉与模式识别 · 计算机科学 2025-06-04 Jiewen Hu , Leena Mathur , Paul Pu Liang , Louis-Philippe Morency

Existing optical character recognition (OCR) methods rely on task-specific designs with divergent paradigms, architectures, and training strategies, which significantly increases the complexity of research and maintenance and hinders the…

计算机视觉与模式识别 · 计算机科学 2026-05-27 Dezhi Peng , Zhenhua Yang , Jiaxin Zhang , Chongyu Liu , Yongxin Shi , Kai Ding , Fengjun Guo , Lianwen Jin

Facial action unit recognition has many applications from market research to psychotherapy and from image captioning to entertainment. Despite its recent progress, deployment of these models has been impeded due to their limited…

计算机视觉与模式识别 · 计算机科学 2023-10-20 Javier Hernandez , Daniel McDuff , Ognjen , Rudovic , Alberto Fung , Mary Czerwinski

Recognizing the expressions of partially occluded faces is a challenging computer vision problem. Previous expression recognition methods, either overlooked this issue or resolved it using extreme assumptions. Motivated by the fact that the…

计算机视觉与模式识别 · 计算机科学 2020-05-14 Hui Ding , Peng Zhou , Rama Chellappa

In domains where computational resources and labeled data are limited, such as in robotics, deep networks with millions of weights might not be the optimal solution. In this paper, we introduce a connectivity scheme for pyramidal…

计算机视觉与模式识别 · 计算机科学 2021-03-24 Henrique Siqueira , Pablo Barros , Sven Magg , Cornelius Weber , Stefan Wermter

The human face plays a central role in social communication, necessitating the use of performant computer vision tools for human-centered applications. We propose Face-LLaVA, a multimodal large language model for face-centered, in-context…

计算机视觉与模式识别 · 计算机科学 2025-04-11 Ashutosh Chaubey , Xulang Guan , Mohammad Soleymani

Image processing is a fundamental task in computer vision, which aims at enhancing image quality and extracting essential features for subsequent vision applications. Traditionally, task-specific models are developed for individual tasks…

计算机视觉与模式识别 · 计算机科学 2024-02-22 Yihao Liu , Xiangyu Chen , Xianzheng Ma , Xintao Wang , Jiantao Zhou , Yu Qiao , Chao Dong

Adapting large-scale pretrained models to various downstream tasks via fine-tuning is a standard method in machine learning. Recently, parameter-efficient fine-tuning methods show promise in adapting a pretrained model to different tasks…

计算机视觉与模式识别 · 计算机科学 2022-10-10 Yen-Cheng Liu , Chih-Yao Ma , Junjiao Tian , Zijian He , Zsolt Kira

Tremendous progress has been made on face detection in recent years using convolutional neural networks. While many face detectors use designs designated for detecting faces, we treat face detection as a generic object detection task. We…

计算机视觉与模式识别 · 计算机科学 2022-01-28 Delong Qi , Weijun Tan , Qi Yao , Jingfeng Liu

Predicting individual aesthetic preferences holds significant practical applications and academic implications for human society. However, existing studies mainly focus on learning and predicting the commonality of facial attractiveness,…

计算机视觉与模式识别 · 计算机科学 2023-11-27 Luojun Lin , Zhifeng Shen , Jia-Li Yin , Qipeng Liu , Yuanlong Yu , Weijie Chen

Large-scale vision-language pre-trained models have shown promising transferability to various downstream tasks. As the size of these foundation models and the number of downstream tasks grow, the standard full fine-tuning paradigm becomes…

计算机视觉与模式识别 · 计算机科学 2023-05-23 Haoyu Lu , Yuqi Huo , Guoxing Yang , Zhiwu Lu , Wei Zhan , Masayoshi Tomizuka , Mingyu Ding

Pre-trained transformer-based language models are becoming increasingly popular due to their exceptional performance on various benchmarks. However, concerns persist regarding the presence of hidden biases within these models, which can…

计算与语言 · 计算机科学 2023-05-29 Bum Chul Kwon , Nandana Mihindukulasooriya

Automatic facial expression recognition (FER) has gained much attention due to its applications in human-computer interaction. Among the approaches to improve FER tasks, this paper focuses on deep architecture with the attention mechanism.…

计算机视觉与模式识别 · 计算机科学 2026-03-09 Luan Pham , The Huynh Vu , Tuan Anh Tran

We study the joint learning of image-to-text and text-to-image generations, which are naturally bi-directional tasks. Typical existing works design two separate task-specific models for each task, which impose expensive design efforts. In…

计算机视觉与模式识别 · 计算机科学 2021-10-20 Yupan Huang , Hongwei Xue , Bei Liu , Yutong Lu

We present User-predictable Face Editing (UP-FacE) -- a novel method for predictable face shape editing. In stark contrast to existing methods for face editing using trial and error, edits with UP-FacE are predictable by the human user.…

计算机视觉与模式识别 · 计算机科学 2024-07-12 Florian Strohm , Mihai Bâce , Andreas Bulling

Face detection and alignment in unconstrained environment are challenging due to various poses, illuminations and occlusions. Recent studies show that deep learning approaches can achieve impressive performance on these two tasks. In this…

计算机视觉与模式识别 · 计算机科学 2016-09-21 Kaipeng Zhang , Zhanpeng Zhang , Zhifeng Li , Yu Qiao

Face detection is a crucial first step in many facial recognition and face analysis systems. Early approaches for face detection were mainly based on classifiers built on top of hand-crafted features extracted from local image regions, such…

计算机视觉与模式识别 · 计算机科学 2021-04-15 Shervin Minaee , Ping Luo , Zhe Lin , Kevin Bowyer

A plethora of face forgery detectors exist to tackle facial deepfake risks. However, their practical application is hindered by the challenge of generalizing to forgeries unseen during the training stage. To this end, we introduce an…

计算机视觉与模式识别 · 计算机科学 2024-12-31 Xiaotian Si , Linghui Li , Liwei Zhang , Ziduo Guo , Kaiguo Yuan , Bingyu Li , Xiaoyong Li

In recent years, audio-driven 3D facial animation has gained significant attention, particularly in applications such as virtual reality, gaming, and video conferencing. However, accurately modeling the intricate and subtle dynamics of…

计算机视觉与模式识别 · 计算机科学 2023-11-14 Guinan Su , Yanwu Yang , Zhifeng Li