中文
相关论文

相关论文: VidFace: A Full-Transformer Solver for Video FaceH…

200 篇论文

Video Face Swapping (VFS) requires seamlessly injecting a source identity into a target video while meticulously preserving the original pose, expression, lighting, background, and dynamic information. Existing methods struggle to maintain…

计算机视觉与模式识别 · 计算机科学 2026-01-06 Xu Guo , Fulong Ye , Xinghui Li , Pengqi Tu , Pengze Zhang , Qichao Sun , Songtao Zhao , Xiangwang Hou , Qian He

Vision Transformers (ViTs) have revolutionized large-scale visual modeling, yet remain underexplored in face recognition (FR) where CNNs still dominate. We identify a critical bottleneck: CNN-inspired training paradigms fail to unlock ViT's…

计算机视觉与模式识别 · 计算机科学 2025-08-18 Jinghan You , Shanglin Li , Yuanrui Sun , Jiangchuan Wei , Mingyu Guo , Chao Feng , Jiao Ran

Video face restoration faces a critical challenge in maintaining temporal consistency while recovering fine facial details from degraded inputs. This paper presents a novel approach that extends Vector-Quantized Variational Autoencoders…

计算机视觉与模式识别 · 计算机科学 2025-06-17 Yan Chen , Hanlin Shang , Ce Liu , Yuxuan Chen , Hui Li , Weihao Yuan , Hao Zhu , Zilong Dong , Siyu Zhu

Face hallucination, which is the task of generating a high-resolution face image from a low-resolution input image, is a well-studied problem that is useful in widespread application areas. Face hallucination is particularly challenging…

计算机视觉与模式识别 · 计算机科学 2016-04-28 Oncel Tuzel , Yuichi Taguchi , John R. Hershey

Face reenactment aims to generate realistic talking head videos by transferring motion from a driving video to a static source image while preserving the source identity. Although existing methods based on either implicit or explicit…

计算机视觉与模式识别 · 计算机科学 2025-07-23 Mingtao Guo , Guanyu Xing , Yanci Zhang , Yanli Liu

Face super-resolution is a challenging and highly ill-posed problem since a low-resolution (LR) face image may correspond to multiple high-resolution (HR) ones during the hallucination process and cause a dramatic identity change for the…

计算机视觉与模式识别 · 计算机科学 2021-10-01 Nitin Balachandran , Jun-Cheng Chen , Rama Chellappa

Most of the existing video face super-resolution (VFSR) methods are trained and evaluated on VoxCeleb1, which is designed specifically for speaker identification and the frames in this dataset are of low quality. As a consequence, the VFSR…

图像与视频处理 · 电气工程与系统科学 2022-05-10 Liangbin Xie. Xintao Wang , Honglun Zhang , Chao Dong , Ying Shan

Despite progress in human motion capture, existing multi-view methods often face challenges in estimating the 3D pose and shape of multiple closely interacting people. This difficulty arises from reliance on accurate 2D joint estimations,…

计算机视觉与模式识别 · 计算机科学 2024-08-21 Feichi Lu , Zijian Dong , Jie Song , Otmar Hilliges

We present a simple method to reconstruct a high-resolution video from a face-video, where the identity of a person is obscured by pixelization. This concealment method is popular because the viewer can still perceive a human face figure…

计算机视觉与模式识别 · 计算机科学 2020-09-30 Maayan Shuvi , Noa Fish , Kfir Aberman , Ariel Shamir , Daniel Cohen-Or

Multi-person total motion capture is extremely challenging when it comes to handle severe occlusions, different reconstruction granularities from body to face and hands, drastically changing observation scales and fast body movements. To…

计算机视觉与模式识别 · 计算机科学 2021-08-25 Yuxiang Zhang , Zhe Li , Liang An , Mengcheng Li , Tao Yu , Yebin Liu

This paper studies how to synthesize face images of non-existent persons, to create a dataset that allows effective training of face recognition (FR) models. Besides generating realistic face images, two other important goals are: 1) the…

计算机视觉与模式识别 · 计算机科学 2025-02-10 Haiyu Wu , Jaskirat Singh , Sicong Tian , Liang Zheng , Kevin W. Bowyer

In this paper, we study the task of hallucinating an authentic high-resolution (HR) face from an occluded thumbnail. We propose a multi-stage Progressive Upsampling and Inpainting Generative Adversarial Network, dubbed Pro-UIGAN, which…

计算机视觉与模式识别 · 计算机科学 2022-05-11 Yang Zhang , Xin Yu , Xiaobo Lu , Ping Liu

Multiple complex degradations are coupled in low-quality video faces in the real world. Therefore, blind video face restoration is a highly challenging ill-posed problem, requiring not only hallucinating high-fidelity details but also…

多媒体 · 计算机科学 2024-04-23 Kepeng Xu , Li Xu , Gang He , Wenxin Yu , Yunsong Li

Transformer-based video diffusion models rely on 3D attention over spatial and temporal tokens, which incurs quadratic time and memory complexity and makes end-to-end training for ultra-high-resolution videos prohibitively expensive. To…

计算机视觉与模式识别 · 计算机科学 2026-03-25 Yunfeng Wu , Hongying Cheng , Zihao He , Songhua Liu

Human image animation involves generating a video from a static image by following a specified pose sequence. Current approaches typically adopt a multi-stage pipeline that separately learns appearance and motion, which often leads to…

计算机视觉与模式识别 · 计算机科学 2024-05-29 Qilin Wang , Zhengkai Jiang , Chengming Xu , Jiangning Zhang , Yabiao Wang , Xinyi Zhang , Yun Cao , Weijian Cao , Chengjie Wang , Yanwei Fu

Face verification and recognition problems have seen rapid progress in recent years, however recognition from small size images remains a challenging task that is inherently intertwined with the task of face super-resolution. Tackling this…

计算机视觉与模式识别 · 计算机科学 2017-10-17 E. Ustinova , V. Lempitsky

Blind face restoration (BFR) on images has significantly progressed over the last several years, while real-world video face restoration (VFR), which is more challenging for more complex face motions such as moving gaze directions and…

计算机视觉与模式识别 · 计算机科学 2024-05-07 Ziyan Chen , Jingwen He , Xinqi Lin , Yu Qiao , Chao Dong

We propose FaceVR, a novel image-based method that enables video teleconferencing in VR based on self-reenactment. State-of-the-art face tracking methods in the VR context are focused on the animation of rigged 3d avatars. While they…

计算机视觉与模式识别 · 计算机科学 2018-03-23 Justus Thies , Michael Zollhöfer , Marc Stamminger , Christian Theobalt , Matthias Nießner

Existing methods of 3D dense face alignment mainly concentrate on accuracy, thus limiting the scope of their practical applications. In this paper, we propose a novel regression framework named 3DDFA-V2 which makes a balance among speed,…

计算机视觉与模式识别 · 计算机科学 2021-02-09 Jianzhu Guo , Xiangyu Zhu , Yang Yang , Fan Yang , Zhen Lei , Stan Z. Li

Face performance capture and reenactment techniques use multiple cameras and sensors, positioned at a distance from the face or mounted on heavy wearable devices. This limits their applications in mobile and outdoor environments. We present…

计算机视觉与模式识别 · 计算机科学 2019-05-28 Mohamed Elgharib , Mallikarjun BR , Ayush Tewari , Hyeongwoo Kim , Wentao Liu , Hans-Peter Seidel , Christian Theobalt