中文
相关论文

相关论文: GaussianHead: High-fidelity Head Avatars with Lear…

200 篇论文

We present SplattingAvatar, a hybrid 3D representation of photorealistic human avatars with Gaussian Splatting embedded on a triangle mesh, which renders over 300 FPS on a modern GPU and 30 FPS on a mobile device. We disentangle the motion…

图形学 · 计算机科学 2024-03-11 Zhijing Shao , Zhaolong Wang , Zhuang Li , Duotun Wang , Xiangru Lin , Yu Zhang , Mingming Fan , Zeyu Wang

With the rising interest from the community in digital avatars coupled with the importance of expressions and gestures in communication, modeling natural avatar behavior remains an important challenge across many industries such as…

计算机视觉与模式识别 · 计算机科学 2025-04-11 Kefan Chen , Sergiu Oprea , Justin Theiss , Sreyas Mohan , Srinath Sridhar , Aayush Prakash

Leveraging pretrained 2D diffusion models and score distillation sampling (SDS), recent methods have shown promising results for text-to-3D avatar generation. However, generating high-quality 3D avatars capable of expressive animation…

计算机视觉与模式识别 · 计算机科学 2024-09-26 Yukun Huang , Jianan Wang , Ailing Zeng , Zheng-Jun Zha , Lei Zhang , Xihui Liu

Forecasting future scenarios in dynamic environments is essential for intelligent decision-making and navigation, a challenge yet to be fully realized in computer vision and robotics. Traditional approaches like video prediction and…

计算机视觉与模式识别 · 计算机科学 2024-05-31 Boming Zhao , Yuan Li , Ziyu Sun , Lin Zeng , Yujun Shen , Rui Ma , Yinda Zhang , Hujun Bao , Zhaopeng Cui

This paper introduces a novel clothed human model that can be learned from multiview RGB videos, with a particular emphasis on recovering physically accurate body and cloth movements. Our method, Position Based Dynamic Gaussians (PBDyG),…

计算机视觉与模式识别 · 计算机科学 2024-12-10 Shota Sasaki , Jane Wu , Ko Nishino

Animatable 3D reconstruction has significant applications across various fields, primarily relying on artists' handcraft creation. Recently, some studies have successfully constructed animatable 3D models from monocular videos. However,…

计算机视觉与模式识别 · 计算机科学 2024-03-19 Tingyang Zhang , Qingzhe Gao , Weiyu Li , Libin Liu , Baoquan Chen

We present Gaussian See, Gaussian Do, a novel approach for semantic 3D motion transfer from multiview video. Our method enables rig-free, cross-category motion transfer between objects with semantically meaningful correspondence. Building…

计算机视觉与模式识别 · 计算机科学 2025-11-20 Yarin Bekor , Gal Michael Harari , Or Perel , Or Litany

We present 3DGH, an unconditional generative model for 3D human heads with composable hair and face components. Unlike previous work that entangles the modeling of hair and face, we propose to separate them using a novel data representation…

3D Gaussian Splatting (3DGS) has demonstrated impressive novel view synthesis performance. While conventional methods require per-scene optimization, more recently several feed-forward methods have been proposed to generate pixel-aligned…

计算机视觉与模式识别 · 计算机科学 2025-03-21 Shengjun Zhang , Xin Fei , Fangfu Liu , Haixu Song , Yueqi Duan

We introduce GoMAvatar, a novel approach for real-time, memory-efficient, high-quality animatable human modeling. GoMAvatar takes as input a single monocular video to create a digital avatar capable of re-articulation in new poses and…

计算机视觉与模式识别 · 计算机科学 2024-04-12 Jing Wen , Xiaoming Zhao , Zhongzheng Ren , Alexander G. Schwing , Shenlong Wang

Although 2D generative models have made great progress in face image generation and animation, they often suffer from undesirable artifacts such as 3D inconsistency when rendering images from different camera viewpoints. This prevents them…

计算机视觉与模式识别 · 计算机科学 2022-10-13 Yue Wu , Yu Deng , Jiaolong Yang , Fangyun Wei , Qifeng Chen , Xin Tong

High-fidelity 3D garment synthesis from text is desirable yet challenging for digital avatar creation. Recent diffusion-based approaches via Score Distillation Sampling (SDS) have enabled new possibilities but either intricately couple with…

计算机视觉与模式识别 · 计算机科学 2024-06-25 Yufei Liu , Junshu Tang , Chu Zheng , Shijie Zhang , Jinkun Hao , Junwei Zhu , Dongjin Huang

We propose PixelGaussian, an efficient feed-forward framework for learning generalizable 3D Gaussian reconstruction from arbitrary views. Most existing methods rely on uniform pixel-wise Gaussian representations, which learn a fixed number…

计算机视觉与模式识别 · 计算机科学 2024-10-25 Xin Fei , Wenzhao Zheng , Yueqi Duan , Wei Zhan , Masayoshi Tomizuka , Kurt Keutzer , Jiwen Lu

Nuanced expressiveness, particularly through fine-grained hand and facial expressions, is pivotal for enhancing the realism and vitality of digital human representations. In this work, we focus on investigating the expressiveness of human…

计算机视觉与模式识别 · 计算机科学 2024-07-04 Hezhen Hu , Zhiwen Fan , Tianhao Wu , Yihan Xi , Seoyoung Lee , Georgios Pavlakos , Zhangyang Wang

Constructing a 3D scene capable of accommodating open-ended language queries, is a pivotal pursuit, particularly within the domain of robotics. Such technology facilitates robots in executing object manipulations based on human language…

Existing approaches for human avatar generation--both NeRF-based and 3D Gaussian Splatting (3DGS) based--struggle with maintaining 3D consistency and exhibit degraded detail reconstruction, particularly when training with sparse inputs. To…

计算机视觉与模式识别 · 计算机科学 2024-11-20 Haoyu Zhao , Hao Wang , Chen Yang , Wei Shen

3D Gaussian splatting (3DGS) has recently emerged as an alternative representation that leverages a 3D Gaussian-based representation and introduces an approximated volumetric rendering, achieving very fast rendering speed and promising…

计算机视觉与模式识别 · 计算机科学 2024-08-08 Joo Chan Lee , Daniel Rho , Xiangyu Sun , Jong Hwan Ko , Eunbyung Park

We present a feed-forward framework for Gaussian full-head synthesis from a single unposed image. Unlike previous work that relies on time-consuming GAN inversion and test-time optimization, our framework can reconstruct the Gaussian…

计算机视觉与模式识别 · 计算机科学 2025-10-13 Peng Li , Yisheng He , Yingdong Hu , Yuan Dong , Weihao Yuan , Yuan Liu , Siyu Zhu , Gang Cheng , Zilong Dong , Yike Guo

Accurate, real-time 3D reconstruction of human heads from monocular images and videos underlies numerous visual applications. As 3D ground truth data is hard to come by at scale, previous methods have sought to learn from abundant 2D videos…

计算机视觉与模式识别 · 计算机科学 2025-10-20 Liam Schoneveld , Zhe Chen , Davide Davoli , Jiapeng Tang , Saimon Terazawa , Ko Nishino , Matthias Nießner

We propose ActiveSplat, an autonomous high-fidelity reconstruction system leveraging Gaussian splatting. Taking advantage of efficient and realistic rendering, the system establishes a unified framework for online mapping, viewpoint…

机器人学 · 计算机科学 2025-06-17 Yuetao Li , Zijia Kuang , Ting Li , Qun Hao , Zike Yan , Guyue Zhou , Shaohui Zhang