English
Related papers

Related papers: ELITE: Efficient Gaussian Head Avatar from a Monoc…

200 papers

Neural radiance fields are capable of reconstructing high-quality drivable human avatars but are expensive to train and render and not suitable for multi-human scenes with complex shadows. To reduce consumption, we propose Animatable 3D…

Computer Vision and Pattern Recognition · Computer Science 2024-07-30 Yang Liu , Xiang Huang , Minghan Qin , Qinwei Lin , Haoqian Wang

Reconstructing animatable and high-quality 3D head avatars from monocular videos, especially with realistic relighting, is a valuable task. However, the limited information from single-view input, combined with the complex head poses and…

Computer Vision and Pattern Recognition · Computer Science 2025-04-22 Dongbin Zhang , Yunfei Liu , Lijian Lin , Ye Zhu , Kangjie Chen , Minghan Qin , Yu Li , Haoqian Wang

Photorealistic avatars have become essential for immersive applications in virtual reality (VR) and augmented reality (AR), enabling lifelike interactions in areas such as training simulations, telemedicine, and virtual collaboration. These…

Graphics · Computer Science 2025-04-18 Rendong Zhang , Alexandra Watkins , Nilanjan Sarkar

We present FaceLift, a novel feed-forward approach for generalizable high-quality 360-degree 3D head reconstruction from a single image. Our pipeline first employs a multi-view latent diffusion model to generate consistent side and back…

Computer Vision and Pattern Recognition · Computer Science 2025-08-04 Weijie Lyu , Yi Zhou , Ming-Hsuan Yang , Zhixin Shu

We introduce GoMAvatar, a novel approach for real-time, memory-efficient, high-quality animatable human modeling. GoMAvatar takes as input a single monocular video to create a digital avatar capable of re-articulation in new poses and…

Computer Vision and Pattern Recognition · Computer Science 2024-04-12 Jing Wen , Xiaoming Zhao , Zhongzheng Ren , Alexander G. Schwing , Shenlong Wang

Building 3D animatable head avatars from a single image is an important yet challenging problem. Existing methods generally collapse under large camera pose variations, compromising the realism of 3D avatars. In this work, we propose a new…

Computer Vision and Pattern Recognition · Computer Science 2026-01-21 Shuling Zhao , Dan Xu

Personalized 3D avatars require an animatable representation of digital humans. Doing so instantly from monocular videos offers scalability to broad class of users and wide-scale applications. In this paper, we present a fast, simple, yet…

Computer Vision and Pattern Recognition · Computer Science 2024-07-17 Pramish Paudel , Anubhav Khanal , Ajad Chhatkuli , Danda Pani Paudel , Jyoti Tandukar

3D Gaussian Splatting (3DGS) provides an efficient method for high-quality scene reconstruction using anisotropic Gaussians. Recently, 3DGS-based methods have significantly improved the rendering quality of human avatars while enabling…

Computer Vision and Pattern Recognition · Computer Science 2026-05-26 Hongzhe Liao , Chuhua Xian , Hongmin Cai , Haiyang Liu , Fa-Ting Hong

Creating high-fidelity 3D head avatars has always been a research hotspot, but it remains a great challenge under lightweight sparse view setups. In this paper, we propose HHAvatar represented by controllable 3D Gaussians for high-fidelity…

Computer Vision and Pattern Recognition · Computer Science 2024-11-21 Zhanfeng Liao , Yuelang Xu , Zhe Li , Qijing Li , Boyao Zhou , Ruifeng Bai , Di Xu , Hongwen Zhang , Yebin Liu

Avatar reconstruction has traditionally relied on per-subject optimization that requires hours of computation or on expensive preprocessing that limits scalability. We introduce FFAvatar, a generalizable feed-forward framework that…

Graphics · Computer Science 2026-05-18 Thuan Hoang Nguyen , Jiahao Luo , Yinyu Nie , Hao Li , Gordon Guocheng Qian , Jian Wang

High-fidelity reconstruction of 3D human avatars has a wild application in visual reality. In this paper, we introduce FAGhead, a method that enables fully controllable human portraits from monocular videos. We explicit the traditional 3D…

Computer Vision and Pattern Recognition · Computer Science 2024-07-01 Yixin Xuan , Xinyang Li , Gongxin Yao , Shiwei Zhou , Donghui Sun , Xiaoxin Chen , Yu Pan

We present MATCH (Multi-view Avatars from Topologically Corresponding Heads), a multi-view Gaussian registration method for high-quality head avatar creation and editing. State-of-the-art multi-view head avatar methods require…

Computer Vision and Pattern Recognition · Computer Science 2026-03-18 Malte Prinzler , Paulo Gotardo , Siyu Tang , Timo Bolkart

Creating high-quality, generalizable speech-driven 3D talking heads remains a persistent challenge. Previous methods achieve satisfactory results for fixed viewpoints and small-scale audio variations, but they struggle with large head…

Computer Vision and Pattern Recognition · Computer Science 2025-07-11 Wentao Hu , Shunkai Li , Ziqiao Peng , Haoxian Zhang , Fan Shi , Xiaoqiang Liu , Pengfei Wan , Di Zhang , Hui Tian

The creation of 3D human avatars from multi-view videos is a significant yet challenging task in computer vision. However, existing techniques rely on high-quality, sharp images as input, which are often impractical to obtain in real-world…

Computer Vision and Pattern Recognition · Computer Science 2026-03-06 Muyao Niu , Yifan Zhan , Qingtian Zhu , Zhuoxiao Li , Wei Wang , Zhihang Zhong , Xiao Sun , Yinqiang Zheng

Nuanced expressiveness, particularly through fine-grained hand and facial expressions, is pivotal for enhancing the realism and vitality of digital human representations. In this work, we focus on investigating the expressiveness of human…

Computer Vision and Pattern Recognition · Computer Science 2024-07-04 Hezhen Hu , Zhiwen Fan , Tianhao Wu , Yihan Xi , Seoyoung Lee , Georgios Pavlakos , Zhangyang Wang

Leveraging pretrained 2D diffusion models and score distillation sampling (SDS), recent methods have shown promising results for text-to-3D avatar generation. However, generating high-quality 3D avatars capable of expressive animation…

Computer Vision and Pattern Recognition · Computer Science 2024-09-26 Yukun Huang , Jianan Wang , Ailing Zeng , Zheng-Jun Zha , Lei Zhang , Xihui Liu

Modeling animatable human avatars from videos is a long-standing and challenging problem. While conventional methods require per-instance optimization, recent feed-forward methods have been proposed to generate 3D Gaussians with a learnable…

Computer Vision and Pattern Recognition · Computer Science 2025-07-28 Yifan Liu , Shengjun Zhang , Chensheng Dai , Yang Chen , Hao Liu , Chen Li , Yueqi Duan

We present SynShot, a novel method for the few-shot inversion of a drivable head avatar based on a synthetic prior. We tackle three major challenges. First, training a controllable 3D generative network requires a large number of diverse…

Computer Vision and Pattern Recognition · Computer Science 2025-04-01 Wojciech Zielonka , Stephan J. Garbin , Alexandros Lattas , George Kopanas , Paulo Gotardo , Thabo Beeler , Justus Thies , Timo Bolkart

In this paper, we propose to create animatable avatars for interacting hands with 3D Gaussian Splatting (GS) and single-image inputs. Existing GS-based methods designed for single subjects often yield unsatisfactory results due to limited…

Computer Vision and Pattern Recognition · Computer Science 2024-10-14 Xuan Huang , Hanhui Li , Wanquan Liu , Xiaodan Liang , Yiqiang Yan , Yuhao Cheng , Chengqiang Gao

We present FastAvatar, a fast and robust algorithm for single-image 3D face reconstruction using 3D Gaussian Splatting (3DGS). Given a single input image from an arbitrary pose, FastAvatar recovers a high-quality, full-head 3DGS avatar in…

Computer Vision and Pattern Recognition · Computer Science 2025-11-27 Hao Liang , Zhixuan Ge , Soumendu Majee , Ashish Tiwari , G. M. Dilshan Godaliyadda , Ashok Veeraraghavan , Guha Balakrishnan