中文
相关论文

相关论文: VASA-3D: Lifelike Audio-Driven Gaussian Head Avata…

200 篇论文

The human face is central to communication. For immersive applications, the digital presence of a person should mirror the physical reality, capturing the users idiosyncrasies and detailed facial expressions. However, current 3D head avatar…

计算机视觉与模式识别 · 计算机科学 2026-04-16 Jalees Nehvi , Timo Bolkart , Thabo Beeler , Justus Thies

Creating high-fidelity 3D head avatars has always been a research hotspot, but it remains a great challenge under lightweight sparse view setups. In this paper, we propose HHAvatar represented by controllable 3D Gaussians for high-fidelity…

计算机视觉与模式识别 · 计算机科学 2024-11-21 Zhanfeng Liao , Yuelang Xu , Zhe Li , Qijing Li , Boyao Zhou , Ruifeng Bai , Di Xu , Hongwen Zhang , Yebin Liu

Recent advances in Gaussian Splatting have significantly boosted the reconstruction of head avatars, enabling high-quality facial modeling by representing an 3D avatar as a collection of 3D Gaussians. However, existing methods predominantly…

计算机视觉与模式识别 · 计算机科学 2025-08-29 Shiqi Xin , Xiaolin Zhang , Yanbin Liu , Peng Zhang , Caifeng Shan

The creation of 3D human avatars from multi-view videos is a significant yet challenging task in computer vision. However, existing techniques rely on high-quality, sharp images as input, which are often impractical to obtain in real-world…

计算机视觉与模式识别 · 计算机科学 2026-03-06 Muyao Niu , Yifan Zhan , Qingtian Zhu , Zhuoxiao Li , Wei Wang , Zhihang Zhong , Xiao Sun , Yinqiang Zheng

Creating digital avatars from textual prompts has long been a desirable yet challenging task. Despite the promising results achieved with 2D diffusion priors, current methods struggle to create high-quality and consistent animated avatars…

计算机视觉与模式识别 · 计算机科学 2024-12-24 Zhenglin Zhou , Fan Ma , Hehe Fan , Zongxin Yang , Yi Yang

Generating talking head videos through a face image and a piece of speech audio still contains many challenges. ie, unnatural head movement, distorted expression, and identity modification. We argue that these issues are mainly because of…

计算机视觉与模式识别 · 计算机科学 2023-03-14 Wenxuan Zhang , Xiaodong Cun , Xuan Wang , Yong Zhang , Xi Shen , Yu Guo , Ying Shan , Fei Wang

We propose FlashAvatar, a novel and lightweight 3D animatable avatar representation that could reconstruct a digital avatar from a short monocular video sequence in minutes and render high-fidelity photo-realistic images at 300FPS on a…

计算机视觉与模式识别 · 计算机科学 2024-04-01 Jun Xiang , Xuan Gao , Yudong Guo , Juyong Zhang

High-fidelity reconstruction of 3D human avatars has a wild application in visual reality. In this paper, we introduce FAGhead, a method that enables fully controllable human portraits from monocular videos. We explicit the traditional 3D…

计算机视觉与模式识别 · 计算机科学 2024-07-01 Yixin Xuan , Xinyang Li , Gongxin Yao , Shiwei Zhou , Donghui Sun , Xiaoxin Chen , Yu Pan

Creating high-fidelity, real-time drivable 3D head avatars is a core challenge in digital animation. While 3D Gaussian Splashing (3D-GS) offers unprecedented rendering speed and quality, current animation techniques often rely on a…

图形学 · 计算机科学 2026-01-22 Zhe Chang , Haodong Jin , Yan Song , Hui Yu

Reconstructing a high-quality, animatable 3D human avatar with expressive facial and hand motions from a single image has gained significant attention due to its broad application potential. 3D human avatar reconstruction typically requires…

计算机视觉与模式识别 · 计算机科学 2025-08-04 Dongbin Zhang , Yunfei Liu , Lijian Lin , Ye Zhu , Yang Li , Minghan Qin , Yu Li , Haoqian Wang

We present a new approach for video-driven animation of high-quality neural 3D head models, addressing the challenge of person-independent animation from video input. Typically, high-quality generative models are learned for specific…

计算机视觉与模式识别 · 计算机科学 2024-03-08 Wolfgang Paier , Paul Hinzer , Anna Hilsmann , Peter Eisert

We present MoGA, a novel method to reconstruct high-fidelity 3D Gaussian avatars from a single-view image. The main challenge lies in inferring unseen appearance and geometric details while ensuring 3D consistency and realism. Most previous…

计算机视觉与模式识别 · 计算机科学 2025-08-12 Zijian Dong , Longteng Duan , Jie Song , Michael J. Black , Andreas Geiger

The generation of high-fidelity, animatable 3D human avatars remains a core challenge in computer graphics and vision, with applications in VR, telepresence, and entertainment. Existing approaches based on implicit representations like…

计算机视觉与模式识别 · 计算机科学 2026-03-26 Ramazan Fazylov , Sergey Zagoruyko , Aleksandr Parkin , Stamatis Lefkimmiatis , Ivan Laptev

Despite recent progress in 3D Gaussian-based head avatar modeling, efficiently generating high fidelity avatars remains a challenge. Current methods typically rely on extensive multi-view capture setups or monocular videos with per-identity…

计算机视觉与模式识别 · 计算机科学 2026-02-02 Xinya Ji , Sebastian Weiss , Manuel Kansy , Jacek Naruniec , Xun Cao , Barbara Solenthaler , Derek Bradley

The creation of lifelike speech-driven 3D facial animation requires a natural and precise synchronization between audio input and facial expressions. However, existing works still fail to render shapes with flexible head poses and natural…

计算机视觉与模式识别 · 计算机科学 2023-11-01 Wei Zhao , Yijun Wang , Tianyu He , Lianying Yin , Jianxin Lin , Xin Jin

Codec Avatars are a recent class of learned, photorealistic face models that accurately represent the geometry and texture of a person in 3D (i.e., for virtual reality), and are almost indistinguishable from video. In this paper we describe…

计算机视觉与模式识别 · 计算机科学 2020-08-13 Alexander Richard , Colin Lea , Shugao Ma , Juergen Gall , Fernando de la Torre , Yaser Sheikh

Audio-driven talking head generation has drawn much attention in recent years, and many efforts have been made in lip-sync, expressive facial expressions, natural head pose generation, and high video quality. However, no model has yet led…

计算机视觉与模式识别 · 计算机科学 2023-12-08 Xusen Sun , Longhao Zhang , Hao Zhu , Peng Zhang , Bang Zhang , Xinya Ji , Kangneng Zhou , Daiheng Gao , Liefeng Bo , Xun Cao

We propose the first approach to automatically and jointly synthesize both the synchronous 3D conversational body and hand gestures, as well as 3D face and head animations, of a virtual character from speech input. Our algorithm uses a CNN…

计算机视觉与模式识别 · 计算机科学 2021-02-16 Ikhsanul Habibie , Weipeng Xu , Dushyant Mehta , Lingjie Liu , Hans-Peter Seidel , Gerard Pons-Moll , Mohamed Elgharib , Christian Theobalt

Current talking face generation methods mainly focus on speech-lip synchronization. However, insufficient investigation on the facial talking style leads to a lifeless and monotonous avatar. Most previous works fail to imitate expressive…

计算机视觉与模式识别 · 计算机科学 2024-11-21 Liyang Chen , Zhiyong Wu , Runnan Li , Weihong Bao , Jun Ling , Xu Tan , Sheng Zhao

Realistic 3D full-body talking avatars hold great potential in AR, with applications ranging from e-commerce live streaming to holographic communication. Despite advances in 3D Gaussian Splatting (3DGS) for lifelike avatar creation,…

计算机视觉与模式识别 · 计算机科学 2025-07-24 Jianchuan Chen , Jingchuan Hu , Gaige Wang , Zhonghua Jiang , Tiansong Zhou , Zhiwen Chen , Chengfei Lv