中文
相关论文

相关论文: Learning Gaze-aware Compositional GAN

200 篇论文

Developing gaze estimation models that generalize well to unseen domains and in-the-wild conditions remains a challenge with no known best solution. This is mostly due to the difficulty of acquiring ground truth data that cover the…

计算机视觉与模式识别 · 计算机科学 2023-12-15 Evangelos Ververas , Polydefkis Gkagkos , Jiankang Deng , Michail Christos Doukas , Jia Guo , Stefanos Zafeiriou

Current child face generators are restricted by the limited size of the available datasets. In addition, feature selection can prove to be a significant challenge, especially due to the large amount of features that need to be trained for.…

计算机视觉与模式识别 · 计算机科学 2023-01-24 Sofie Daniels , Jiugeng Sun , Jiaqing Xie

Non-invasive gaze estimation methods usually regress gaze directions directly from a single face or eye image. However, due to important variabilities in eye shapes and inner eye structures amongst individuals, universal models obtain…

计算机视觉与模式识别 · 计算机科学 2020-02-07 Gang Liu , Yu Yu , Kenneth A. Funes Mora , Jean-Marc Odobez

This paper is on face/head reenactment where the goal is to transfer the facial pose (3D head orientation and expression) of a target face to a source face. Previous methods focus on learning embedding networks for identity and pose…

计算机视觉与模式识别 · 计算机科学 2022-10-07 Stella Bounareli , Vasileios Argyriou , Georgios Tzimiropoulos

Generating realistic tissue images with annotations is a challenging task that is important in many computational histopathology applications. Synthetically generated images and annotations are valuable for training and evaluating…

图像与视频处理 · 电气工程与系统科学 2024-04-08 Srijay Deshpande , Fayyaz Minhas , Nasir Rajpoot

Powerful generative adversarial networks (GAN) have been developed to automatically synthesize realistic images from text. However, most existing tasks are limited to generating simple images such as flowers from captions. In this work, we…

机器学习 · 计算机科学 2019-11-27 Osaid Rehman Nasir , Shailesh Kumar Jha , Manraj Singh Grover , Yi Yu , Ajit Kumar , Rajiv Ratn Shah

The major challenge in today's computer vision scenario is the availability of good quality labeled data. In a field of study like image classification, where data is of utmost importance, we need to find more reliable methods which can…

计算机视觉与模式识别 · 计算机科学 2025-10-15 Aashish Dhawan , Divyanshu Mudgal

Unavailability of large training datasets is a bottleneck that needs to be overcome to realize the true potential of deep learning in histopathology applications. Although slide digitization via whole slide imaging scanners has increased…

图像与视频处理 · 电气工程与系统科学 2022-02-08 Komal Mariam , Osama Mohammed Afzal , Wajahat Hussain , Muhammad Umar Javed , Amber Kiyani , Nasir Rajpoot , Syed Ali Khurram , Hassan Aqeel Khan

Compared to the prosperity of pre-training models in natural image understanding, the research on large-scale pre-training models for facial knowledge learning is still limited. Current approaches mainly rely on manually assembled and…

计算机视觉与模式识别 · 计算机科学 2025-10-22 Yudong Li , Hao Li , Xianxu Hou , Linlin Shen

Leveraging datasets available to learn a model with high generalization ability to unseen domains is important for computer vision, especially when the unseen domain's annotated data are unavailable. We study a novel and practical problem…

计算机视觉与模式识别 · 计算机科学 2021-04-09 Yang Shu , Zhangjie Cao , Chenyu Wang , Jianmin Wang , Mingsheng Long

Fine-grained facial expression manipulation is a challenging problem, as fine-grained expression details are difficult to be captured. Most existing expression manipulation methods resort to discrete expression labels, which mainly edit…

计算机视觉与模式识别 · 计算机科学 2020-06-16 Junshu Tang , Zhiwen Shao , Lizhuang Ma

Facial landmarks refer to the localization of fundamental facial points on face images. There have been a tremendous amount of attempts to detect these points from facial images however, there has never been an attempt to synthesize a…

图像与视频处理 · 电气工程与系统科学 2018-02-02 Shabab Bazrafkan , Hossein Javidnia , Peter Corcoran

Current 3D gaze estimation methods struggle to generalize across diverse data domains, primarily due to i) the scarcity of annotated datasets, and ii) the insufficient diversity of labeled data. In this work, we present OmniGaze, a…

计算机视觉与模式识别 · 计算机科学 2025-10-17 Hongyu Qu , Jianan Wei , Xiangbo Shu , Yazhou Yao , Wenguan Wang , Jinhui Tang

Predominant techniques on talking head generation largely depend on 2D information, including facial appearances and motions from input face images. Nevertheless, dense 3D facial geometry, such as pixel-wise depth, plays a critical role in…

计算机视觉与模式识别 · 计算机科学 2023-12-12 Fa-Ting Hong , Li Shen , Dan Xu

Recent advances in Generative Adversarial Networks (GANs) have shown impressive results for task of facial expression synthesis. The most successful architecture is StarGAN, that conditions GANs generation process with images of a specific…

计算机视觉与模式识别 · 计算机科学 2018-08-30 Albert Pumarola , Antonio Agudo , Aleix M. Martinez , Alberto Sanfeliu , Francesc Moreno-Noguer

Recent years witness the tremendous success of generative adversarial networks (GANs) in synthesizing photo-realistic images. GAN generator learns to compose realistic images and reproduce the real data distribution. Through that, a…

计算机视觉与模式识别 · 计算机科学 2023-01-16 Yinghao Xu , Yujun Shen , Jiapeng Zhu , Ceyuan Yang , Bolei Zhou

We propose a novel method that trains a conditional Generative Adversarial Network (GAN) to generate visual interpretations of a Convolutional Neural Network (CNN). To comprehend a CNN, the GAN is trained with information on how the CNN…

计算机视觉与模式识别 · 计算机科学 2023-11-10 R T Akash Guna , Raul Benitez , O K Sikha

Generating face image with specific gaze information has attracted considerable attention. Existing approaches typically input gaze values directly for face generation, which is unnatural and requires annotated gaze datasets for training,…

计算机视觉与模式识别 · 计算机科学 2024-10-01 Hengfei Wang , Zhongqun Zhang , Yihua Cheng , Hyung Jin Chang

Previous Face Anti-spoofing (FAS) methods face the challenge of generalizing to unseen domains, mainly because most existing FAS datasets are relatively small and lack data diversity. Thanks to the development of face recognition in the…

计算机视觉与模式识别 · 计算机科学 2024-12-12 Xingming Long , Jie Zhang , Shiguang Shan

One of the biggest issues facing the use of machine learning in medical imaging is the lack of availability of large, labelled datasets. The annotation of medical images is not only expensive and time consuming but also highly dependent on…