中文
相关论文

相关论文: LatentAvatar: Learning Latent Expression Code for …

200 篇论文

Face images are subject to many different factors of variation, especially in unconstrained in-the-wild scenarios. For most tasks involving such images, e.g. expression recognition from video streams, having enough labeled data is…

计算机视觉与模式识别 · 计算机科学 2020-08-19 Marah Halawa , Manuel Wöllhaf , Eduardo Vellasques , Urko Sánchez Sanz , Olaf Hellwich

Significant progress has been made in audio-driven human animation, while most existing methods focus mainly on facial movements, limiting their ability to create full-body animations with natural synchronization and fluidity. They also…

计算机视觉与模式识别 · 计算机科学 2025-06-24 Qijun Gan , Ruizi Yang , Jianke Zhu , Shaofei Xue , Steven Hoi

We present TimeWalker, a novel framework that models realistic, full-scale 3D head avatars of a person on lifelong scale. Unlike current human head avatar pipelines that capture identity at the momentary level(e.g., instant photography or…

计算机视觉与模式识别 · 计算机科学 2025-12-16 Dongwei Pan , Yang Li , Hongsheng Li , Kwan-Yee Lin

We present a deep learning framework for real-time speech-driven 3D facial animation from just raw waveforms. Our deep neural network directly maps an input sequence of speech audio to a series of micro facial action unit activations and…

计算机视觉与模式识别 · 计算机科学 2017-12-11 Hai X. Pham , Yuting Wang , Vladimir Pavlovic

Current 3D-aware pretraining methods for embodied perception and manipulation are largely built on differentiable rendering frameworks, producing either fully implicit neural fields or fully explicit geometric primitives. Implicit…

We present a novel framework for generating high-quality, animatable 4D avatar from a single image. While recent advances have shown promising results in 4D avatar creation, existing methods either require extensive multiview data or…

计算机视觉与模式识别 · 计算机科学 2025-04-22 Fei Yin , Mallikarjun B R , Chun-Han Yao , Rafał Mantiuk , Varun Jampani

In this paper, we propose a novel framework, Tracking-free Relightable Avatar (TRAvatar), for capturing and reconstructing high-fidelity 3D avatars. Compared to previous methods, TRAvatar works in a more practical and efficient setting.…

计算机视觉与模式识别 · 计算机科学 2023-09-11 Haotian Yang , Mingwu Zheng , Wanquan Feng , Haibin Huang , Yu-Kun Lai , Pengfei Wan , Zhongyuan Wang , Chongyang Ma

Authoring an appealing animation for a virtual character is a challenging task. In computer-aided keyframe animation artists define the key poses of a character by manipulating its underlying skeletons. To look plausible, a character pose…

图形学 · 计算机科学 2021-07-02 Léon Victor , Alexandre Meyer , Saïda Bouakaz

Latent space exploration is a technique that discovers interpretable latent directions and manipulates latent codes to edit various attributes in images generated by generative adversarial networks (GANs). However, in previous work, spatial…

计算机视觉与模式识别 · 计算机科学 2022-08-29 Yuki Endo

High-quality 3D avatar modeling faces a critical trade-off between fidelity and generalization. On the one hand, multi-view studio data enables high-fidelity modeling of humans with precise control over expressions and poses, but it…

Facial expression recognition has been an active research area over the past few decades, and it is still challenging due to the high intra-class variation. Traditional approaches for this problem rely on hand-crafted features such as SIFT,…

计算机视觉与模式识别 · 计算机科学 2019-02-05 Shervin Minaee , Amirali Abdolrashidi

Implicit Neural Networks (INRs) have emerged as powerful representations to encode all forms of data, including images, videos, audios, and scenes. With video, many INRs for video have been proposed for the compression task, and recent…

计算机视觉与模式识别 · 计算机科学 2024-08-06 Shishira R Maiya , Anubhav Gupta , Matthew Gwilliam , Max Ehrlich , Abhinav Shrivastava

We present a learning-based method for building driving-signal aware full-body avatars. Our model is a conditional variational autoencoder that can be animated with incomplete driving signals, such as human pose and facial keypoints, and…

计算机视觉与模式识别 · 计算机科学 2021-06-29 Timur Bagautdinov , Chenglei Wu , Tomas Simon , Fabian Prada , Takaaki Shiratori , Shih-En Wei , Weipeng Xu , Yaser Sheikh , Jason Saragih

In this work, we propose to model the interaction between visual and textual features for multi-modal neural machine translation (MMT) through a latent variable model. This latent variable can be seen as a multi-modal stochastic embedding…

计算与语言 · 计算机科学 2019-05-17 Iacer Calixto , Miguel Rios , Wilker Aziz

Talking head synthesis is an emerging technology with wide applications in film dubbing, virtual avatars and online education. Recent NeRF-based methods generate more natural talking videos, as they better capture the 3D structural…

计算机视觉与模式识别 · 计算机科学 2022-07-26 Shuai Shen , Wanhua Li , Zheng Zhu , Yueqi Duan , Jie Zhou , Jiwen Lu

Generating talking head videos through a face image and a piece of speech audio still contains many challenges. ie, unnatural head movement, distorted expression, and identity modification. We argue that these issues are mainly because of…

计算机视觉与模式识别 · 计算机科学 2023-03-14 Wenxuan Zhang , Xiaodong Cun , Xuan Wang , Yong Zhang , Xi Shen , Yu Guo , Ying Shan , Fei Wang

Lightweight creation of 3D digital avatars is a highly desirable but challenging task. With only sparse videos of a person under unknown illumination, we propose a method to create relightable and animatable neural avatars, which can be…

计算机视觉与模式识别 · 计算机科学 2023-12-21 Wenbin Lin , Chengwei Zheng , Jun-Hai Yong , Feng Xu

In this paper we address the problem of neural face reenactment, where, given a pair of a source and a target facial image, we need to transfer the target's pose (defined as the head pose and its facial expressions) to the source image, by…

计算机视觉与模式识别 · 计算机科学 2022-09-28 Stella Bounareli , Christos Tzelepis , Vasileios Argyriou , Ioannis Patras , Georgios Tzimiropoulos

This work addresses the problem of discovering, in an unsupervised manner, interpretable paths in the latent space of pretrained GANs, so as to provide an intuitive and easy way of controlling the underlying generative factors. In doing so,…

计算机视觉与模式识别 · 计算机科学 2021-09-29 Christos Tzelepis , Georgios Tzimiropoulos , Ioannis Patras

Creating high-fidelity 3D head avatars has always been a research hotspot, but it remains a great challenge under lightweight sparse view setups. In this paper, we propose HHAvatar represented by controllable 3D Gaussians for high-fidelity…

计算机视觉与模式识别 · 计算机科学 2024-11-21 Zhanfeng Liao , Yuelang Xu , Zhe Li , Qijing Li , Boyao Zhou , Ruifeng Bai , Di Xu , Hongwen Zhang , Yebin Liu