中文
相关论文

相关论文: Sequential Gaussian Avatars with Hierarchical Moti…

200 篇论文

We introduce a novel framework for 3D human avatar generation and personalization, leveraging text prompts to enhance user engagement and customization. Central to our approach are key innovations aimed at overcoming the challenges in…

3D Gaussian Splatting (3DGS) provides an efficient method for high-quality scene reconstruction using anisotropic Gaussians. Recently, 3DGS-based methods have significantly improved the rendering quality of human avatars while enabling…

计算机视觉与模式识别 · 计算机科学 2026-05-26 Hongzhe Liao , Chuhua Xian , Hongmin Cai , Haiyang Liu , Fa-Ting Hong

The trend in sign language generation is centered around data-driven generative methods that require vast amounts of precise 2D and 3D human pose data to achieve an acceptable generation quality. However, currently, most sign language…

计算机视觉与模式识别 · 计算机科学 2025-12-25 Kaustubh Kundu , Hrishav Bakul Barua , Lucy Robertson-Bell , Zhixi Cai , Kalin Stefanov

We present FastAvatar, a fast and robust algorithm for single-image 3D face reconstruction using 3D Gaussian Splatting (3DGS). Given a single input image from an arbitrary pose, FastAvatar recovers a high-quality, full-head 3DGS avatar in…

计算机视觉与模式识别 · 计算机科学 2025-11-27 Hao Liang , Zhixuan Ge , Soumendu Majee , Ashish Tiwari , G. M. Dilshan Godaliyadda , Ashok Veeraraghavan , Guha Balakrishnan

State-of-the-art approaches for conditional human body rendering via Gaussian splatting typically focus on simple body motions captured from many views. This is often in the context of dancing or walking. However, for more complex use…

计算机视觉与模式识别 · 计算机科学 2025-05-06 Maksym Ivashechkin , Oscar Mendez , Richard Bowden

Leveraging pretrained 2D diffusion models and score distillation sampling (SDS), recent methods have shown promising results for text-to-3D avatar generation. However, generating high-quality 3D avatars capable of expressive animation…

计算机视觉与模式识别 · 计算机科学 2024-09-26 Yukun Huang , Jianan Wang , Ailing Zeng , Zheng-Jun Zha , Lei Zhang , Xihui Liu

Reconstructing dynamic humans interacting with real-world environments from monocular videos is an important and challenging task. Despite considerable progress in 4D neural rendering, existing approaches either model dynamic scenes…

计算机视觉与模式识别 · 计算机科学 2025-11-14 Wenqing Wang , Haosen Yang , Josef Kittler , Xiatian Zhu

We present PrismAvatar: a 3D head avatar model which is designed specifically to enable real-time animation and rendering on resource-constrained edge devices, while still enjoying the benefits of neural volumetric rendering at training…

计算机视觉与模式识别 · 计算机科学 2025-02-12 Prashant Raina , Felix Taubner , Mathieu Tuli , Eu Wern Teh , Kevin Ferreira

Current personalized neural head avatars face a trade-off: lightweight models lack detail and realism, while high-quality, animatable avatars require significant computational resources, making them unsuitable for commodity devices. To…

计算机视觉与模式识别 · 计算机科学 2025-04-01 Wojciech Zielonka , Timo Bolkart , Thabo Beeler , Justus Thies

Novel view synthesis has seen major advances in recent years, with 3D Gaussian splatting offering an excellent level of visual quality, fast training and real-time rendering. However, the resources needed for training and rendering…

计算机视觉与模式识别 · 计算机科学 2024-06-19 Bernhard Kerbl , Andréas Meuleman , Georgios Kopanas , Michael Wimmer , Alexandre Lanvin , George Drettakis

With the rising interest from the community in digital avatars coupled with the importance of expressions and gestures in communication, modeling natural avatar behavior remains an important challenge across many industries such as…

计算机视觉与模式识别 · 计算机科学 2025-04-11 Kefan Chen , Sergiu Oprea , Justin Theiss , Sreyas Mohan , Srinath Sridhar , Aayush Prakash

We introduce RMAvatar, a novel human avatar representation with Gaussian splatting embedded on mesh to learn clothed avatar from a monocular video. We utilize the explicit mesh geometry to represent motion and shape of a virtual human and…

计算机视觉与模式识别 · 计算机科学 2025-01-14 Sen Peng , Weixing Xie , Zilong Wang , Xiaohu Guo , Zhonggui Chen , Baorong Yang , Xiao Dong

We present an efficient neural 3D scene representation for novel-view synthesis (NVS) in large-scale, dynamic urban areas. Existing works are not well suited for applications like mixed-reality or closed-loop simulation due to their limited…

计算机视觉与模式识别 · 计算机科学 2024-11-26 Tobias Fischer , Jonas Kulhanek , Samuel Rota Bulò , Lorenzo Porzi , Marc Pollefeys , Peter Kontschieder

Existing methods like Neural Radiation Fields (NeRF) and 3D Gaussian Splatting (3DGS) have made significant strides in facial attribute control such as facial animation and components editing, yet they struggle with fine-grained…

计算机视觉与模式识别 · 计算机科学 2024-08-21 Pinxin Liu , Luchuan Song , Daoan Zhang , Hang Hua , Yunlong Tang , Huaijin Tu , Jiebo Luo , Chenliang Xu

DiffusionAvatars synthesizes a high-fidelity 3D head avatar of a person, offering intuitive control over both pose and expression. We propose a diffusion-based neural renderer that leverages generic 2D priors to produce compelling images of…

计算机视觉与模式识别 · 计算机科学 2024-04-18 Tobias Kirschstein , Simon Giebenhain , Matthias Nießner

We present a novel approach for tracking multiple people in video. Unlike past approaches which employ 2D representations, we focus on using 3D representations of people, located in three-dimensional space. To this end, we develop a method,…

计算机视觉与模式识别 · 计算机科学 2021-11-16 Jathushan Rajasegaran , Georgios Pavlakos , Angjoo Kanazawa , Jitendra Malik

Generating animatable human avatars from a single image is essential for various digital human modeling applications. Existing 3D reconstruction methods often struggle to capture fine details in animatable models, while generative…

计算机视觉与模式识别 · 计算机科学 2024-12-04 Lingteng Qiu , Shenhao Zhu , Qi Zuo , Xiaodong Gu , Yuan Dong , Junfei Zhang , Chao Xu , Zhe Li , Weihao Yuan , Liefeng Bo , Guanying Chen , Zilong Dong

We present TimeWalker, a novel framework that models realistic, full-scale 3D head avatars of a person on lifelong scale. Unlike current human head avatar pipelines that capture identity at the momentary level(e.g., instant photography or…

计算机视觉与模式识别 · 计算机科学 2025-12-16 Dongwei Pan , Yang Li , Hongsheng Li , Kwan-Yee Lin

Latent scene representation plays a significant role in training reinforcement learning (RL) agents. To obtain good latent vectors describing the scenes, recent works incorporate the 3D-aware latent-conditioned NeRF pipeline into scene…

机器人学 · 计算机科学 2024-09-30 Jiaxu Wang , Ziyi Zhang , Qiang Zhang , Jia Li , Jingkai Sun , Mingyuan Sun , Junhao He , Renjing Xu

Digital humans and, especially, 3D facial avatars have raised a lot of attention in the past years, as they are the backbone of several applications like immersive telepresence in AR or VR. Despite the progress, facial avatars reconstructed…

计算机视觉与模式识别 · 计算机科学 2023-11-27 Berna Kabadayi , Wojciech Zielonka , Bharat Lal Bhatnagar , Gerard Pons-Moll , Justus Thies