中文
相关论文

相关论文: Stratified Avatar Generation from Sparse Observati…

200 篇论文

Deep learning greatly improved the realism of animatable human models by learning geometry and appearance from collections of 3D scans, template meshes, and multi-view imagery. High-resolution models enable photo-realistic avatars but at…

计算机视觉与模式识别 · 计算机科学 2022-10-13 Shih-Yang Su , Timur Bagautdinov , Helge Rhodin

Reconstructing photorealistic and animatable 4D head avatars from a single portrait image remains a fundamental challenge in computer vision. While diffusion models have enabled remarkable progress in image and video generation for avatar…

计算机视觉与模式识别 · 计算机科学 2026-03-13 Chao Xu , Xiaochen Zhao , Xiang Deng , Jingxiang Sun , Donglin Di , Zhuo Su , Yebin Liu

We introduce a framework that automates the transformation of static anime illustrations into manipulatable 2.5D models. Current professional workflows require tedious manual segmentation and the artistic ``hallucination'' of occluded…

计算机视觉与模式识别 · 计算机科学 2026-02-04 Jian Lin , Chengze Li , Haoyun Qin , Kwun Wang Chan , Yanghua Jin , Hanyuan Liu , Stephen Chun Wang Choy , Xueting Liu

Rigging and skinning clothed human avatars is a challenging task and traditionally requires a lot of manual work and expertise. Recent methods addressing it either generalize across different characters or focus on capturing the dynamics of…

计算机视觉与模式识别 · 计算机科学 2023-07-04 Zhouyingcheng Liao , Vladislav Golyanik , Marc Habermann , Christian Theobalt

We propose SparseFusion, a sparse view 3D reconstruction approach that unifies recent advances in neural rendering and probabilistic image generation. Existing approaches typically build on neural rendering with re-projected features but…

计算机视觉与模式识别 · 计算机科学 2023-02-17 Zhizhuo Zhou , Shubham Tulsiani

It is now possible to reconstruct dynamic human motion and shape from a sparse set of cameras using Neural Radiance Fields (NeRF) driven by an underlying skeleton. However, a challenge remains to model the deformation of cloth and skin in…

计算机视觉与模式识别 · 计算机科学 2023-10-02 Chunjin Song , Bastian Wandt , Helge Rhodin

The creation of 4D avatars (i.e., animated 3D avatars) from text description typically uses text-to-image (T2I) diffusion models to synthesize 3D avatars in the canonical space and subsequently applies animation with target motions.…

计算机视觉与模式识别 · 计算机科学 2024-06-10 Zenghao Chai , Chen Tang , Yongkang Wong , Mohan Kankanhalli

In multi-view 3D human pose estimation, models typically rely on images captured simultaneously from different camera views to predict a pose at a specific moment. While providing accurate spatial information, this traditional approach…

计算机视觉与模式识别 · 计算机科学 2026-05-15 Ling Li , Changjie Chen , Yuyan Wang , Jiaqing Lyu , Kenglun Chang , Yiyun Chen , Zhidong Deng

We present a methodology to model articulated objects using a sparse set of images with unknown poses. Current methods require dense multi-view observations and ground-truth camera poses. Our approach operates with as few as four views per…

计算机视觉与模式识别 · 计算机科学 2026-04-06 Jianning Deng , Kartic Subr , Hakan Bilen

Existing 3D clothed avatar reconstruction methods achieve high visual fidelity but ignore geometric structure and physical plausibility. They either model clothed humans as a single deformable surface or attempt garment disentanglement…

计算机视觉与模式识别 · 计算机科学 2026-05-21 Daniel Eskandar , Berna Kabadayi , Garvita Tiwari , Gerard Pons-Moll

We propose a novel approach for reconstructing animatable 3D Gaussian avatars from monocular videos captured by commodity devices like smartphones. Photorealistic 3D head avatar reconstruction from such recordings is challenging due to…

计算机视觉与模式识别 · 计算机科学 2025-04-15 Jiapeng Tang , Davide Davoli , Tobias Kirschstein , Liam Schoneveld , Matthias Niessner

Diffusion-based inpainting is a powerful tool for the reconstruction of images from sparse data. Its quality strongly depends on the choice of known data. Optimising their spatial location -- the inpainting mask -- is challenging. A…

图像与视频处理 · 电气工程与系统科学 2022-05-17 Tobias Alt , Pascal Peter , Joachim Weickert

Modeling animatable human avatars from monocular or multi-view videos has been widely studied, with recent approaches leveraging neural radiance fields (NeRFs) or 3D Gaussian Splatting (3DGS) achieving impressive results in novel-view and…

计算机视觉与模式识别 · 计算机科学 2025-04-03 Yahui Li , Zhi Zeng , Liming Pang , Guixuan Zhang , Shuwu Zhang

It is extremely challenging to create an animatable clothed human avatar from RGB videos, especially for loose clothes due to the difficulties in motion modeling. To address this problem, we introduce a novel representation on the basis of…

计算机视觉与模式识别 · 计算机科学 2022-03-29 Zerong Zheng , Han Huang , Tao Yu , Hongwen Zhang , Yandong Guo , Yebin Liu

The ability to create realistic, animatable and relightable head avatars from casual video sequences would open up wide ranging applications in communication and entertainment. Current methods either build on explicit 3D morphable meshes…

计算机视觉与模式识别 · 计算机科学 2023-03-01 Yufeng Zheng , Wang Yifan , Gordon Wetzstein , Michael J. Black , Otmar Hilliges

We present a novel approach for generating animatable 3D-aware art avatars from a single image, with controllable facial expressions, head poses, and shoulder movements. Unlike previous reenactment methods, our approach utilizes a…

计算机视觉与模式识别 · 计算机科学 2024-03-27 Shaoxu Li

The generation of high-quality, animatable 3D head avatars from text has enormous potential in content creation applications such as games, movies, and embodied virtual assistants. Current text-to-3D generation methods typically combine…

计算机视觉与模式识别 · 计算机科学 2025-04-23 Yiqian Wu , Malte Prinzler , Xiaogang Jin , Siyu Tang

The ability to generate diverse 3D articulated head avatars is vital to a plethora of applications, including augmented reality, cinematography, and education. Recent work on text-guided 3D object generation has shown great promise in…

计算机视觉与模式识别 · 计算机科学 2023-07-12 Alexander W. Bergman , Wang Yifan , Gordon Wetzstein

Recently, data-driven single-view reconstruction methods have shown great progress in modeling 3D dressed humans. However, such methods suffer heavily from depth ambiguities and occlusions inherent to single view inputs. In this paper, we…

计算机视觉与模式识别 · 计算机科学 2021-12-07 Pierre Zins , Yuanlu Xu , Edmond Boyer , Stefanie Wuhrer , Tony Tung

We propose PFAvatar (Pose-Fusion Avatar), a new method that reconstructs high-quality 3D avatars from Outfit of the Day(OOTD) photos, which exhibit diverse poses, occlusions, and complex backgrounds. Our method consists of two stages: (1)…

计算机视觉与模式识别 · 计算机科学 2025-11-19 Dianbing Xi , Guoyuan An , Jingsen Zhu , Zhijian Liu , Yuan Liu , Ruiyuan Zhang , Jiayuan Lu , Yuchi Huo , Rui Wang