English
Related papers

Related papers: GPHM: Gaussian Parametric Head Model for Monocular…

200 papers

Digital humans and, especially, 3D facial avatars have raised a lot of attention in the past years, as they are the backbone of several applications like immersive telepresence in AR or VR. Despite the progress, facial avatars reconstructed…

Computer Vision and Pattern Recognition · Computer Science 2023-11-27 Berna Kabadayi , Wojciech Zielonka , Bharat Lal Bhatnagar , Gerard Pons-Moll , Justus Thies

In this paper, we explore a reconstruction and reenactment separated framework for 3D Gaussians head, which requires only a single portrait image as input to generate controllable avatar. Specifically, we developed a large-scale one-shot…

Computer Vision and Pattern Recognition · Computer Science 2025-09-18 Zhiling Ye , Cong Zhou , Xiubao Zhang , Haifeng Shen , Weihong Deng , Quan Lu

We present MoGA, a novel method to reconstruct high-fidelity 3D Gaussian avatars from a single-view image. The main challenge lies in inferring unseen appearance and geometric details while ensuring 3D consistency and realism. Most previous…

Computer Vision and Pattern Recognition · Computer Science 2025-08-12 Zijian Dong , Longteng Duan , Jie Song , Michael J. Black , Andreas Geiger

We introduce an approach that creates animatable human avatars from monocular videos using 3D Gaussian Splatting (3DGS). Existing methods based on neural radiance fields (NeRFs) achieve high-quality novel-view/novel-pose image synthesis but…

Computer Vision and Pattern Recognition · Computer Science 2024-04-05 Zhiyin Qian , Shaofei Wang , Marko Mihajlovic , Andreas Geiger , Siyu Tang

Real-time rendering of high-fidelity and animatable avatars from monocular videos remains a challenging problem in computer vision and graphics. Over the past few years, the Neural Radiance Field (NeRF) has made significant progress in…

Computer Vision and Pattern Recognition · Computer Science 2025-03-05 Qipeng Yan , Mingyang Sun , Lihua Zhang

The efficient reconstruction of high-quality and intuitively editable human avatars presents a pressing challenge in the field of computer vision. Recent advancements, such as 3DGS, have demonstrated impressive reconstruction efficiency and…

Graphics · Computer Science 2025-11-25 Mengtian Li , Shengxiang Yao , Yichen Pan , Haiyao Xiao , Zhongmei Li , Zhifeng Xie , Keyu Chen

We introduce GaussianSpeech, a novel approach that synthesizes high-fidelity animation sequences of photo-realistic, personalized 3D human head avatars from spoken audio. To capture the expressive, detailed nature of human heads, including…

Computer Vision and Pattern Recognition · Computer Science 2024-12-02 Shivangi Aneja , Artem Sevastopolsky , Tobias Kirschstein , Justus Thies , Angela Dai , Matthias Nießner

Creating realistic avatars from a single RGB image is an attractive yet challenging problem. Due to its ill-posed nature, recent works leverage powerful prior from 2D diffusion models pretrained on large datasets. Although 2D diffusion…

Computer Vision and Pattern Recognition · Computer Science 2024-12-17 Yuxuan Xue , Xianghui Xie , Riccardo Marin , Gerard Pons-Moll

High-fidelity head avatar reconstruction plays a crucial role in AR/VR, gaming, and multimedia content creation. Recent advances in 3D Gaussian Splatting (3DGS) have demonstrated effectiveness in modeling complex geometry with real-time…

Computer Vision and Pattern Recognition · Computer Science 2025-08-20 Shikun Zhang , Cunjian Chen , Yiqun Wang , Qiuhong Ke , Yong Li

We introduce GoMAvatar, a novel approach for real-time, memory-efficient, high-quality animatable human modeling. GoMAvatar takes as input a single monocular video to create a digital avatar capable of re-articulation in new poses and…

Computer Vision and Pattern Recognition · Computer Science 2024-04-12 Jing Wen , Xiaoming Zhao , Zhongzheng Ren , Alexander G. Schwing , Shenlong Wang

Neural radiance fields are capable of reconstructing high-quality drivable human avatars but are expensive to train and render and not suitable for multi-human scenes with complex shadows. To reduce consumption, we propose Animatable 3D…

Computer Vision and Pattern Recognition · Computer Science 2024-07-30 Yang Liu , Xiang Huang , Minghan Qin , Qinwei Lin , Haoqian Wang

Generating animatable human avatars from a single image is essential for various digital human modeling applications. Existing 3D reconstruction methods often struggle to capture fine details in animatable models, while generative…

Computer Vision and Pattern Recognition · Computer Science 2024-12-04 Lingteng Qiu , Shenhao Zhu , Qi Zuo , Xiaodong Gu , Yuan Dong , Junfei Zhang , Chao Xu , Zhe Li , Weihao Yuan , Liefeng Bo , Guanying Chen , Zilong Dong

Reconstructing photo-realistic and topology-aware animatable human avatars from monocular videos remains challenging in computer vision and graphics. Recently, methods using 3D Gaussians to represent the human body have emerged, offering…

Computer Vision and Pattern Recognition · Computer Science 2024-11-20 Haoyu Zhao , Chen Yang , Hao Wang , Xingyue Zhao , Wei Shen

Reconstructing a complete 3D head from a single portrait remains challenging because existing methods still face a sharp quality-speed trade-off: high-fidelity pipelines often rely on multi-stage processing and per-subject optimization,…

Computer Vision and Pattern Recognition · Computer Science 2026-04-16 Yujie Gao , Yao Xiao , Xiangnan Zhu , Ya Li , Yiyi Zhang , Liqing Zhang , Jianfu Zhang

In this work, we introduce Monocular and Generalizable Gaussian Talking Head Animation (MGGTalk), which requires monocular datasets and generalizes to unseen identities without personalized re-training. Compared with previous 3D Gaussian…

Computer Vision and Pattern Recognition · Computer Science 2025-04-02 Shengjie Gong , Haojie Li , Jiapeng Tang , Dongming Hu , Shuangping Huang , Hao Chen , Tianshui Chen , Zhuoman Liu

This work addresses the problem of real-time rendering of photorealistic human body avatars learned from multi-view videos. While the classical approaches to model and render virtual humans generally use a textured mesh, recent research has…

Computer Vision and Pattern Recognition · Computer Science 2024-03-29 Arthur Moreau , Jifei Song , Helisa Dhamo , Richard Shaw , Yiren Zhou , Eduardo Pérez-Pellitero

Modeling animatable human avatars from RGB videos is a long-standing and challenging problem. Recent works usually adopt MLP-based neural radiance fields (NeRF) to represent 3D humans, but it remains difficult for pure MLPs to regress…

Computer Vision and Pattern Recognition · Computer Science 2024-05-28 Zhe Li , Yipengjing Sun , Zerong Zheng , Lizhen Wang , Shengping Zhang , Yebin Liu

The creation of photorealistic dynamic hair remains a major challenge in digital human modeling because of the complex motions, occlusions, and light scattering. Existing methods often resort to static capture and physics-based models that…

Computer Vision and Pattern Recognition · Computer Science 2025-12-22 Junying Wang , Yuanlu Xu , Edith Tretschk , Ziyan Wang , Anastasia Ianina , Aljaz Bozic , Ulrich Neumann , Tony Tung

Reconstructing a high-quality, animatable 3D human avatar with expressive facial and hand motions from a single image has gained significant attention due to its broad application potential. 3D human avatar reconstruction typically requires…

Computer Vision and Pattern Recognition · Computer Science 2025-08-04 Dongbin Zhang , Yunfei Liu , Lijian Lin , Ye Zhu , Yang Li , Minghan Qin , Yu Li , Haoqian Wang

Nuanced expressiveness, particularly through fine-grained hand and facial expressions, is pivotal for enhancing the realism and vitality of digital human representations. In this work, we focus on investigating the expressiveness of human…

Computer Vision and Pattern Recognition · Computer Science 2024-07-04 Hezhen Hu , Zhiwen Fan , Tianhao Wu , Yihan Xi , Seoyoung Lee , Georgios Pavlakos , Zhangyang Wang