中文
相关论文

相关论文: HeadStudio: Text to Animatable Head Avatars with 3…

200 篇论文

We present MATCH (Multi-view Avatars from Topologically Corresponding Heads), a multi-view Gaussian registration method for high-quality head avatar creation and editing. State-of-the-art multi-view head avatar methods require…

计算机视觉与模式识别 · 计算机科学 2026-03-18 Malte Prinzler , Paulo Gotardo , Siyu Tang , Timo Bolkart

Gaussian-based human avatars have achieved an unprecedented level of visual fidelity. However, existing approaches based on high-capacity neural networks typically require a desktop GPU to achieve real-time performance for a single avatar,…

The fidelity of relighting is bounded by both geometry and appearance representations. For geometry, both mesh and volumetric approaches have difficulty modeling intricate structures like 3D hair geometry. For appearance, existing…

图形学 · 计算机科学 2024-05-29 Shunsuke Saito , Gabriel Schwartz , Tomas Simon , Junxuan Li , Giljoo Nam

We introduce ELITE, an Efficient Gaussian head avatar synthesis from a monocular video via Learned Initialization and TEst-time generative adaptation. Prior works rely either on a 3D data prior or a 2D generative prior to compensate for…

计算机视觉与模式识别 · 计算机科学 2026-01-16 Kim Youwang , Lee Hyoseok , Subin Park , Gerard Pons-Moll , Tae-Hyun Oh

We present READ Avatars, a 3D-based approach for generating 2D avatars that are driven by audio input with direct and granular control over the emotion. Previous methods are unable to achieve realistic animation due to the many-to-many…

计算机视觉与模式识别 · 计算机科学 2023-03-02 Jack Saunders , Vinay Namboodiri

Recent advances in neural rendering have improved both training and rendering times by orders of magnitude. While these methods demonstrate state-of-the-art quality and speed, they are designed for photogrammetry of static scenes and do not…

计算机视觉与模式识别 · 计算机科学 2023-11-30 Muhammed Kocabas , Jen-Hao Rick Chang , James Gabriel , Oncel Tuzel , Anurag Ranjan

The creation of high-fidelity, digital versions of human heads is an important stepping stone in the process of further integrating virtual components into our everyday lives. Constructing such avatars is a challenging research problem, due…

计算机视觉与模式识别 · 计算机科学 2024-09-16 Simon Giebenhain , Tobias Kirschstein , Martin Rünz , Lourdes Agapito , Matthias Nießner

We introduce a novel framework for modeling high-fidelity, animatable 3D human avatars from motion-blurred monocular video inputs. Motion blur is prevalent in real-world dynamic video capture, especially due to human movements in 3D human…

计算机视觉与模式识别 · 计算机科学 2025-06-17 Xianrui Luo , Juewen Peng , Zhongang Cai , Lei Yang , Fan Yang , Zhiguo Cao , Guosheng Lin

With the advancement of virtual reality, the demand for 3D human avatars is increasing. The emergence of Gaussian Splatting technology has enabled the rendering of Gaussian avatars with superior visual quality and reduced computational…

图形学 · 计算机科学 2025-12-01 Xiaonuo Dongye , Hanzhi Guo , Le Luo , Haiyan Jiang , Yihua Bao , Jie Guo , Zeyu Tian , Dongdong Weng

Neural radiance fields are capable of reconstructing high-quality drivable human avatars but are expensive to train and render and not suitable for multi-human scenes with complex shadows. To reduce consumption, we propose Animatable 3D…

计算机视觉与模式识别 · 计算机科学 2024-07-30 Yang Liu , Xiang Huang , Minghan Qin , Qinwei Lin , Haoqian Wang

We present LAM, an innovative Large Avatar Model for animatable Gaussian head reconstruction from a single image. Unlike previous methods that require extensive training on captured video sequences or rely on auxiliary neural networks for…

计算机视觉与模式识别 · 计算机科学 2025-04-07 Yisheng He , Xiaodong Gu , Xiaodan Ye , Chao Xu , Zhengyi Zhao , Yuan Dong , Weihao Yuan , Zilong Dong , Liefeng Bo

We leverage increasingly popular three-dimensional neural representations in order to construct a unified and consistent explanation of a collection of uncalibrated images of the human face. Our approach utilizes Gaussian Splatting, since…

计算机视觉与模式识别 · 计算机科学 2025-12-19 Haodi He , Jihun Yu , Ronald Fedkiw

Recent advancements in automatic 3D avatar generation guided by text have made significant progress. However, existing methods have limitations such as oversaturation and low-quality output. To address these challenges, we propose X-Oscar,…

计算机视觉与模式识别 · 计算机科学 2024-05-03 Yiwei Ma , Zhekai Lin , Jiayi Ji , Yijun Fan , Xiaoshuai Sun , Rongrong Ji

Creating high-quality, generalizable speech-driven 3D talking heads remains a persistent challenge. Previous methods achieve satisfactory results for fixed viewpoints and small-scale audio variations, but they struggle with large head…

计算机视觉与模式识别 · 计算机科学 2025-07-11 Wentao Hu , Shunkai Li , Ziqiao Peng , Haoxian Zhang , Fan Shi , Xiaoqiang Liu , Pengfei Wan , Di Zhang , Hui Tian

Generating 3D human models directly from text helps reduce the cost and time of character modeling. However, achieving multi-attribute controllable and realistic 3D human avatar generation is still challenging due to feature coupling and…

计算机视觉与模式识别 · 计算机科学 2024-03-21 Chaoqun Gong , Yuqin Dai , Ronghui Li , Achun Bao , Jun Li , Jian Yang , Yachao Zhang , Xiu Li

We introduce AvatarPointillist, a novel framework for generating dynamic 4D Gaussian avatars from a single portrait image. At the core of our method is a decoder-only Transformer that autoregressively generates a point cloud for 3D Gaussian…

计算机视觉与模式识别 · 计算机科学 2026-04-21 Hongyu Liu , Xuan Wang , Zijian Wu , Yating Wang , Ziyu Wan , Yue Ma , Runtao Liu , Boyao Zhou , Yujun Shen , Qifeng Chen

This paper presents EGSTalker, a real-time audio-driven talking head generation framework based on 3D Gaussian Splatting (3DGS). Designed to enhance both speed and visual fidelity, EGSTalker requires only 3-5 minutes of training video to…

声音 · 计算机科学 2025-10-13 Tianheng Zhu , Yinfeng Yu , Liejun Wang , Fuchun Sun , Wendong Zheng

Creating high-quality animatable 3D human avatars from a single image remains a significant challenge in computer vision due to the inherent difficulty of reconstructing complete 3D information from a single viewpoint. Current approaches…

计算机视觉与模式识别 · 计算机科学 2025-05-09 Yonwoo Choi

Recent advances in diffusion models have made significant progress in digital human generation. However, most existing models still struggle to maintain 3D consistency, temporal coherence, and motion accuracy. A key reason for these…

图形学 · 计算机科学 2025-03-21 Xuan Gao , Jingtao Zhou , Dongyu Liu , Yuqi Zhou , Juyong Zhang

Building photorealistic, animatable full-body digital humans remains a longstanding challenge in computer graphics and vision. Recent advances in animatable avatar modeling have largely progressed along two directions: improving the…

计算机视觉与模式识别 · 计算机科学 2026-04-21 Heming Zhu , Guoxing Sun , Marc Habermann