中文
相关论文

相关论文: Dream, Lift, Animate: From Single Images to Animat…

200 篇论文

Creating and customizing a 3D clothed avatar from textual descriptions is a critical and challenging task. Traditional methods often treat the human body and clothing as inseparable, limiting users' ability to freely mix and match garments.…

图形学 · 计算机科学 2025-03-18 Jia Gong , Shenyu Ji , Lin Geng Foo , Kang Chen , Hossein Rahmani , Jun Liu

Existing 3D clothed avatar reconstruction methods achieve high visual fidelity but ignore geometric structure and physical plausibility. They either model clothed humans as a single deformable surface or attempt garment disentanglement…

计算机视觉与模式识别 · 计算机科学 2026-05-21 Daniel Eskandar , Berna Kabadayi , Garvita Tiwari , Gerard Pons-Moll

Building 3D animatable head avatars from a single image is an important yet challenging problem. Existing methods generally collapse under large camera pose variations, compromising the realism of 3D avatars. In this work, we propose a new…

计算机视觉与模式识别 · 计算机科学 2026-01-21 Shuling Zhao , Dan Xu

Animatable clothing transfer, aiming at dressing and animating garments across characters, is a challenging problem. Most human avatar works entangle the representations of the human body and clothing together, which leads to difficulties…

计算机视觉与模式识别 · 计算机科学 2024-05-14 Siyou Lin , Zhe Li , Zhaoqi Su , Zerong Zheng , Hongwen Zhang , Yebin Liu

Creating photorealistic 3D head avatars from limited input has become increasingly important for applications in virtual reality, telepresence, and digital entertainment. While recent advances like neural rendering and 3D Gaussian splatting…

图形学 · 计算机科学 2026-03-12 Chen Guo , Zhuo Su , Liao Wang , Jian Wang , Shuang Li , Xu Chang , Zhaohu Li , Yang Zhao , Guidong Wang , Yebin Liu , Ruqi Huang

Sparse volumetric reconstruction and rendering via 3D Gaussian splatting have recently enabled animatable 3D head avatars that are rendered under arbitrary viewpoints with impressive photorealism. Today, such photoreal avatars are seen as a…

计算机视觉与模式识别 · 计算机科学 2025-05-12 Gengyan Li , Paulo Gotardo , Timo Bolkart , Stephan Garbin , Kripasindhu Sarkar , Abhimitra Meka , Alexandros Lattas , Thabo Beeler

Reconstructing a high-quality, animatable 3D human avatar with expressive facial and hand motions from a single image has gained significant attention due to its broad application potential. 3D human avatar reconstruction typically requires…

计算机视觉与模式识别 · 计算机科学 2025-08-04 Dongbin Zhang , Yunfei Liu , Lijian Lin , Ye Zhu , Yang Li , Minghan Qin , Yu Li , Haoqian Wang

We present LAM, an innovative Large Avatar Model for animatable Gaussian head reconstruction from a single image. Unlike previous methods that require extensive training on captured video sequences or rely on auxiliary neural networks for…

计算机视觉与模式识别 · 计算机科学 2025-04-07 Yisheng He , Xiaodong Gu , Xiaodan Ye , Chao Xu , Zhengyi Zhao , Yuan Dong , Weihao Yuan , Zilong Dong , Liefeng Bo

We present GaussianAvatar, an efficient approach to creating realistic human avatars with dynamic 3D appearances from a single video. We start by introducing animatable 3D Gaussians to explicitly represent humans in various poses and…

计算机视觉与模式识别 · 计算机科学 2024-04-02 Liangxiao Hu , Hongwen Zhang , Yuxiang Zhang , Boyao Zhou , Boning Liu , Shengping Zhang , Liqiang Nie

In this paper, we propose Generalizable and Animatable Gaussian head Avatar (GAGAvatar) for one-shot animatable head avatar reconstruction. Existing methods rely on neural radiance fields, leading to heavy rendering consumption and low…

计算机视觉与模式识别 · 计算机科学 2024-10-11 Xuangeng Chu , Tatsuya Harada

Vision-language-action (VLA) policies have advanced language-conditioned robotic manipulation by transferring semantic priors from pretrained vision-language models to action generation. However, standard action-imitation learning often…

机器人学 · 计算机科学 2026-05-29 Zijian Zhang , Yuqing Jiang , Qian Cheng , Xiaofan Li , Si Liu , Ding Zhao , Ping Luo , Weitao Zhou , Haibao Yu

We present Drivable 3D Gaussian Avatars (D3GA), a multi-layered 3D controllable model for human bodies that utilizes 3D Gaussian primitives embedded into tetrahedral cages. The advantage of using cages compared to commonly employed linear…

计算机视觉与模式识别 · 计算机科学 2025-02-12 Wojciech Zielonka , Timur Bagautdinov , Shunsuke Saito , Michael Zollhöfer , Justus Thies , Javier Romero

Generating animatable human avatars from a single image is essential for various digital human modeling applications. Existing 3D reconstruction methods often struggle to capture fine details in animatable models, while generative…

计算机视觉与模式识别 · 计算机科学 2024-12-04 Lingteng Qiu , Shenhao Zhu , Qi Zuo , Xiaodong Gu , Yuan Dong , Junfei Zhang , Chao Xu , Zhe Li , Weihao Yuan , Liefeng Bo , Guanying Chen , Zilong Dong

We present MoGA, a novel method to reconstruct high-fidelity 3D Gaussian avatars from a single-view image. The main challenge lies in inferring unseen appearance and geometric details while ensuring 3D consistency and realism. Most previous…

计算机视觉与模式识别 · 计算机科学 2025-08-12 Zijian Dong , Longteng Duan , Jie Song , Michael J. Black , Andreas Geiger

Modeling animatable human avatars from RGB videos is a long-standing and challenging problem. Recent works usually adopt MLP-based neural radiance fields (NeRF) to represent 3D humans, but it remains difficult for pure MLPs to regress…

计算机视觉与模式识别 · 计算机科学 2024-05-28 Zhe Li , Yipengjing Sun , Zerong Zheng , Lizhen Wang , Shengping Zhang , Yebin Liu

We present GALA, a framework that takes as input a single-layer clothed 3D human mesh and decomposes it into complete multi-layered 3D assets. The outputs can then be combined with other assets to create novel clothed human avatars with any…

计算机视觉与模式识别 · 计算机科学 2024-01-24 Taeksoo Kim , Byungjun Kim , Shunsuke Saito , Hanbyul Joo

We present a method that reconstructs and animates a 3D head avatar from a single-view portrait image. Existing methods either involve time-consuming optimization for a specific person with multiple images, or they struggle to synthesize…

计算机视觉与模式识别 · 计算机科学 2023-06-16 Xueting Li , Shalini De Mello , Sifei Liu , Koki Nagano , Umar Iqbal , Jan Kautz

We present LiftAvatar, a new paradigm that completes sparse monocular observations in kinematic space (e.g., facial expressions and head pose) and uses the completed signals to drive high-fidelity avatar animation. LiftAvatar is a…

计算机视觉与模式识别 · 计算机科学 2026-03-03 Hualiang Wei , Shunran Jia , Jialun Liu , Wenhui Li

We propose VASA-3D, an audio-driven, single-shot 3D head avatar generator. This research tackles two major challenges: capturing the subtle expression details present in real human faces, and reconstructing an intricate 3D head avatar from…

计算机视觉与模式识别 · 计算机科学 2025-12-17 Sicheng Xu , Guojun Chen , Jiaolong Yang , Yizhong Zhang , Yu Deng , Steve Lin , Baining Guo

We propose a novel framework for decomposing arbitrarily posed humans into animatable multi-layered 3D human avatars, separating the body and garments. Conventional single-layer reconstruction methods lock clothing to one identity, while…

计算机视觉与模式识别 · 计算机科学 2026-01-12 Yinghan Xu , John Dingliana
‹ 上一页 1 2 3 10 下一页 ›