English
Related papers

Related papers: ELITE: Efficient Gaussian Head Avatar from a Monoc…

200 papers

We introduce an approach that creates animatable human avatars from monocular videos using 3D Gaussian Splatting (3DGS). Existing methods based on neural radiance fields (NeRFs) achieve high-quality novel-view/novel-pose image synthesis but…

Computer Vision and Pattern Recognition · Computer Science 2024-04-05 Zhiyin Qian , Shaofei Wang , Marko Mihajlovic , Andreas Geiger , Siyu Tang

Despite recent progress in 3D Gaussian-based head avatar modeling, efficiently generating high fidelity avatars remains a challenge. Current methods typically rely on extensive multi-view capture setups or monocular videos with per-identity…

Computer Vision and Pattern Recognition · Computer Science 2026-02-02 Xinya Ji , Sebastian Weiss , Manuel Kansy , Jacek Naruniec , Xun Cao , Barbara Solenthaler , Derek Bradley

Real-time rendering of human head avatars is a cornerstone of many computer graphics applications, such as augmented reality, video games, and films, to name a few. Recent approaches address this challenge with computationally efficient…

Computer Vision and Pattern Recognition · Computer Science 2024-09-19 Kartik Teotia , Hyeongwoo Kim , Pablo Garrido , Marc Habermann , Mohamed Elgharib , Christian Theobalt

We present a feed-forward framework for Gaussian full-head synthesis from a single unposed image. Unlike previous work that relies on time-consuming GAN inversion and test-time optimization, our framework can reconstruct the Gaussian…

Computer Vision and Pattern Recognition · Computer Science 2025-10-13 Peng Li , Yisheng He , Yingdong Hu , Yuan Dong , Weihao Yuan , Yuan Liu , Siyu Zhu , Gang Cheng , Zilong Dong , Yike Guo

This paper proposes an efficient 3D avatar coding framework that leverages compact human priors and canonical-to-target transformation to enable high-quality 3D human avatar video compression at ultra-low bit rates. The framework begins by…

Image and Video Processing · Electrical Eng. & Systems 2025-10-14 Shanzhi Yin , Bolin Chen , Xinju Wu , Ru-Ling Liao , Jie Chen , Shiqi Wang , Yan Ye

Gaussian splatting has emerged as a powerful 3D representation that harnesses the advantages of both explicit (mesh) and implicit (NeRF) 3D representations. In this paper, we seek to leverage Gaussian splatting to generate realistic…

Computer Vision and Pattern Recognition · Computer Science 2024-04-01 Ye Yuan , Xueting Li , Yangyi Huang , Shalini De Mello , Koki Nagano , Jan Kautz , Umar Iqbal

Animatable 3D reconstruction has significant applications across various fields, primarily relying on artists' handcraft creation. Recently, some studies have successfully constructed animatable 3D models from monocular videos. However,…

Computer Vision and Pattern Recognition · Computer Science 2024-03-19 Tingyang Zhang , Qingzhe Gao , Weiyu Li , Libin Liu , Baoquan Chen

Despite significant progress in 3D avatar reconstruction, it still faces challenges such as high time complexity, sensitivity to data quality, and low data utilization. We propose FastAvatar, a feedforward 3D avatar framework capable of…

Computer Vision and Pattern Recognition · Computer Science 2026-03-03 Yue Wu , Xuanhong Chen , Yufan Wu , Wen Li , Yuxi Lu , Kairui Feng

We propose FlashAvatar, a novel and lightweight 3D animatable avatar representation that could reconstruct a digital avatar from a short monocular video sequence in minutes and render high-fidelity photo-realistic images at 300FPS on a…

Computer Vision and Pattern Recognition · Computer Science 2024-04-01 Jun Xiang , Xuan Gao , Yudong Guo , Juyong Zhang

Creating relightable and animatable avatars from multi-view or monocular videos is a challenging task for digital human creation and virtual reality applications. Previous methods rely on neural radiance fields or ray tracing, resulting in…

Computer Vision and Pattern Recognition · Computer Science 2025-05-21 Youyi Zhan , Tianjia Shao , He Wang , Yin Yang , Kun Zhou

High-fidelity facial avatar reconstruction from a monocular video is a significant research problem in computer graphics and computer vision. Recently, Neural Radiance Field (NeRF) has shown impressive novel view rendering results and has…

Computer Vision and Pattern Recognition · Computer Science 2023-03-07 Yunpeng Bai , Yanbo Fan , Xuan Wang , Yong Zhang , Jingxiang Sun , Chun Yuan , Ying Shan

We introduce a novel approach to creating ultra-realistic head avatars and rendering them in real-time (>30fps at $2048 \times 1334$ resolution). First, we propose a hybrid explicit representation that combines the advantages of two…

Graphics · Computer Science 2025-02-20 Hongrui Cai , Yuting Xiao , Xuan Wang , Jiafei Li , Yudong Guo , Yanbo Fan , Shenghua Gao , Juyong Zhang

Gaussian-based human avatars have achieved an unprecedented level of visual fidelity. However, existing approaches based on high-capacity neural networks typically require a desktop GPU to achieve real-time performance for a single avatar,…

Computer Vision and Pattern Recognition · Computer Science 2025-07-01 Forrest Iandola , Stanislav Pidhorskyi , Igor Santesteban , Divam Gupta , Anuj Pahuja , Nemanja Bartolovic , Frank Yu , Emanuel Garbin , Tomas Simon , Shunsuke Saito

Reconstructing high-fidelity 3D head avatars is crucial in various applications such as virtual reality. The pioneering methods reconstruct realistic head avatars with Neural Radiance Fields (NeRF), which have been limited by training and…

Computer Vision and Pattern Recognition · Computer Science 2025-11-03 Peng Chen , Xiaobao Wei , Qingpo Wuwu , Xinyi Wang , Xingyu Xiao , Ming Lu

We introduce 3D Gaussian blendshapes for modeling photorealistic head avatars. Taking a monocular video as input, we learn a base head model of neutral expression, along with a group of expression blendshapes, each of which corresponds to a…

Graphics · Computer Science 2024-05-03 Shengjie Ma , Yanlin Weng , Tianjia Shao , Kun Zhou

Current personalized neural head avatars face a trade-off: lightweight models lack detail and realism, while high-quality, animatable avatars require significant computational resources, making them unsuitable for commodity devices. To…

Computer Vision and Pattern Recognition · Computer Science 2025-04-01 Wojciech Zielonka , Timo Bolkart , Thabo Beeler , Justus Thies

We present Vid2Avatar-Pro, a method to create photorealistic and animatable 3D human avatars from monocular in-the-wild videos. Building a high-quality avatar that supports animation with diverse poses from a monocular video is challenging…

Computer Vision and Pattern Recognition · Computer Science 2025-03-04 Chen Guo , Junxuan Li , Yash Kant , Yaser Sheikh , Shunsuke Saito , Chen Cao

In this paper, we present a novel method that facilitates the creation of vivid 3D Gaussian avatars from monocular video inputs (GVA). Our innovation lies in addressing the intricate challenges of delivering high-fidelity human body…

Computer Vision and Pattern Recognition · Computer Science 2024-03-20 Xinqi Liu , Chenming Wu , Jialun Liu , Xing Liu , Jinbo Wu , Chen Zhao , Haocheng Feng , Errui Ding , Jingdong Wang

Real-time rendering of high-fidelity and animatable avatars from monocular videos remains a challenging problem in computer vision and graphics. Over the past few years, the Neural Radiance Field (NeRF) has made significant progress in…

Computer Vision and Pattern Recognition · Computer Science 2025-03-05 Qipeng Yan , Mingyang Sun , Lihua Zhang

Creating digital avatars from textual prompts has long been a desirable yet challenging task. Despite the promising results achieved with 2D diffusion priors, current methods struggle to create high-quality and consistent animated avatars…

Computer Vision and Pattern Recognition · Computer Science 2024-12-24 Zhenglin Zhou , Fan Ma , Hehe Fan , Zongxin Yang , Yi Yang