English
Related papers

Related papers: LUCAS: Layered Universal Codec Avatars

200 papers

Recent advances in 3D vision have led to specialized models for either 3D understanding (e.g., shape classification, segmentation, reconstruction) or 3D generation (e.g., synthesis, completion, and editing). However, these tasks are often…

Computer Vision and Pattern Recognition · Computer Science 2026-04-21 Peng Huang , Yifeng Chen , Zeyu Zhang , Hao Tang

Reconstructing high-fidelity 3D head avatars is crucial in various applications such as virtual reality. The pioneering methods reconstruct realistic head avatars with Neural Radiance Fields (NeRF), which have been limited by training and…

Computer Vision and Pattern Recognition · Computer Science 2025-11-03 Peng Chen , Xiaobao Wei , Qingpo Wuwu , Xinyi Wang , Xingyu Xiao , Ming Lu

Reconstructing photorealistic and animatable 4D head avatars from a single portrait image remains a fundamental challenge in computer vision. While diffusion models have enabled remarkable progress in image and video generation for avatar…

Computer Vision and Pattern Recognition · Computer Science 2026-03-13 Chao Xu , Xiaochen Zhao , Xiang Deng , Jingxiang Sun , Donglin Di , Zhuo Su , Yebin Liu

DiffusionAvatars synthesizes a high-fidelity 3D head avatar of a person, offering intuitive control over both pose and expression. We propose a diffusion-based neural renderer that leverages generic 2D priors to produce compelling images of…

Computer Vision and Pattern Recognition · Computer Science 2024-04-18 Tobias Kirschstein , Simon Giebenhain , Matthias Nießner

Recent techniques on implicit geometry representation learning and neural rendering have shown promising results for 3D clothed human reconstruction from sparse video inputs. However, it is still challenging to reconstruct detailed surface…

Computer Vision and Pattern Recognition · Computer Science 2024-04-23 Hao Wang , Qingshan Xu , Hongyuan Chen , Rui Ma

Photorealistic avatars have become essential for immersive applications in virtual reality (VR) and augmented reality (AR), enabling lifelike interactions in areas such as training simulations, telemedicine, and virtual collaboration. These…

Graphics · Computer Science 2025-04-18 Rendong Zhang , Alexandra Watkins , Nilanjan Sarkar

Full-duplex speech interaction, as the most natural and intuitive mode of human communication, is driving artificial intelligence toward more human-like conversational systems. Traditional cascaded speech processing pipelines suffer from…

Artificial Intelligence · Computer Science 2026-05-01 Yadong Li , Guoxin Wu , Haiping Hou , Biye Li

Constructing drivable and photorealistic 3D head avatars has become a central task in AR/XR, enabling immersive and expressive user experiences. With the emergence of high-fidelity and efficient representations such as 3D Gaussians, recent…

Graphics · Computer Science 2025-12-25 Jaeseong Lee , Junyeong Ahn , Taewoong Kang , Jaegul Choo

In this paper, we present an end-to-end learning framework for detailed 3D face reconstruction from a single image. Our approach uses a 3DMM-based coarse model and a displacement map in UV-space to represent a 3D face. Unlike previous work…

Computer Vision and Pattern Recognition · Computer Science 2020-09-03 Yajing Chen , Fanzi Wu , Zeyu Wang , Yibing Song , Yonggen Ling , Linchao Bao

Text-driven avatar generation has gained significant attention owing to its convenience. However, existing methods typically model the human body with all garments as a single 3D model, limiting its usability, such as clothing replacement,…

Computer Vision and Pattern Recognition · Computer Science 2025-03-19 Jingyu Zhuang , Di Kang , Linchao Bao , Liang Lin , Guanbin Li

Creating high-fidelity, real-time drivable 3D head avatars is a core challenge in digital animation. While 3D Gaussian Splashing (3D-GS) offers unprecedented rendering speed and quality, current animation techniques often rely on a…

Graphics · Computer Science 2026-01-22 Zhe Chang , Haodong Jin , Yan Song , Hui Yu

We present GenLCA, a diffusion-based generative model for generating and editing photorealistic full-body avatars from text and image inputs. The generated avatars are faithful to the inputs, while supporting high-fidelity facial and…

Computer Vision and Pattern Recognition · Computer Science 2026-04-10 Yiqian Wu , Rawal Khirodkar , Egor Zakharov , Timur Bagautdinov , Lei Xiao , Zhaoen Su , Shunsuke Saito , Xiaogang Jin , Junxuan Li

Animatable clothing transfer, aiming at dressing and animating garments across characters, is a challenging problem. Most human avatar works entangle the representations of the human body and clothing together, which leads to difficulties…

Computer Vision and Pattern Recognition · Computer Science 2024-05-14 Siyou Lin , Zhe Li , Zhaoqi Su , Zerong Zheng , Hongwen Zhang , Yebin Liu

An important challenge for autonomous agents such as robots is to maintain a spatially and temporally consistent model of the world. It must be maintained through occlusions, previously-unseen views, and long time horizons (e.g., loop…

Computer Vision and Pattern Recognition · Computer Science 2023-10-03 Dominik A. Kloepfer , Dylan Campbell , João F. Henriques

Creating realistic avatars from a single RGB image is an attractive yet challenging problem. Due to its ill-posed nature, recent works leverage powerful prior from 2D diffusion models pretrained on large datasets. Although 2D diffusion…

Computer Vision and Pattern Recognition · Computer Science 2024-12-17 Yuxuan Xue , Xianghui Xie , Riccardo Marin , Gerard Pons-Moll

Single-view 3D hair reconstruction is challenging, due to the wide range of shape variations among diverse hairstyles. Current state-of-the-art methods are specialized in recovering un-braided 3D hairs and often take braided styles as their…

Computer Vision and Pattern Recognition · Computer Science 2024-09-26 Yujian Zheng , Yuda Qiu , Leyang Jin , Chongyang Ma , Haibin Huang , Di Zhang , Pengfei Wan , Xiaoguang Han

We propose a compositional method for constructing a complete 3D head avatar from a single image. Prior one-shot holistic approaches frequently fail to produce realistic hair dynamics during animation, largely due to inadequate decoupling…

Computer Vision and Pattern Recognition · Computer Science 2026-04-17 Yuan Sun , Xuan Wang , WeiLi Zhang , Wenxuan Zhang , Yu Guo , Fei Wang

We present Drivable 3D Gaussian Avatars (D3GA), a multi-layered 3D controllable model for human bodies that utilizes 3D Gaussian primitives embedded into tetrahedral cages. The advantage of using cages compared to commonly employed linear…

Computer Vision and Pattern Recognition · Computer Science 2025-02-12 Wojciech Zielonka , Timur Bagautdinov , Shunsuke Saito , Michael Zollhöfer , Justus Thies , Javier Romero

We present ARCH++, an image-based method to reconstruct 3D avatars with arbitrary clothing styles. Our reconstructed avatars are animation-ready and highly realistic, in both the visible regions from input views and the unseen regions.…

Computer Vision and Pattern Recognition · Computer Science 2022-03-01 Tong He , Yuanlu Xu , Shunsuke Saito , Stefano Soatto , Tony Tung

Diffusion-based video generation techniques have significantly improved zero-shot talking-head avatar generation, enhancing the naturalness of both head motion and facial expressions. However, existing methods suffer from poor…

Graphics · Computer Science 2025-04-24 Lingzhou Mu , Baiji Liu , Ruonan Zhang , Guiming Mo , Jiawei Jin , Kai Zhang , Haozhi Huang
‹ Prev 1 3 4 5 6 7 10 Next ›