中文
相关论文

相关论文: PERSE: Personalized 3D Generative Avatars from A S…

200 篇论文

We present a new approach for synthesizing novel views of people in new poses. Our novel differentiable renderer enables the synthesis of highly realistic images from any viewpoint. Rather than operating over mesh-based structures, our…

计算机视觉与模式识别 · 计算机科学 2022-02-22 Guillaume Rochette , Chris Russell , Richard Bowden

This paper proposes a new generative adversarial network for pose transfer, i.e., transferring the pose of a given person to a target pose. The generator of the network comprises a sequence of Pose-Attentional Transfer Blocks that each…

计算机视觉与模式识别 · 计算机科学 2019-05-14 Zhen Zhu , Tengteng Huang , Baoguang Shi , Miao Yu , Bofei Wang , Xiang Bai

Generation of high-quality person images is challenging, due to the sophisticated entanglements among image factors, e.g., appearance, pose, foreground, background, local details, global structures, etc. In this paper, we present a novel…

计算机视觉与模式识别 · 计算机科学 2020-07-20 Siyu Huang , Haoyi Xiong , Zhi-Qi Cheng , Qingzhong Wang , Xingran Zhou , Bihan Wen , Jun Huan , Dejing Dou

The creation of high-fidelity, digital versions of human heads is an important stepping stone in the process of further integrating virtual components into our everyday lives. Constructing such avatars is a challenging research problem, due…

计算机视觉与模式识别 · 计算机科学 2024-09-16 Simon Giebenhain , Tobias Kirschstein , Martin Rünz , Lourdes Agapito , Matthias Nießner

We present DreamAvatar, a text-and-shape guided framework for generating high-quality 3D human avatars with controllable poses. While encouraging results have been reported by recent methods on text-guided 3D common object generation,…

计算机视觉与模式识别 · 计算机科学 2023-12-01 Yukang Cao , Yan-Pei Cao , Kai Han , Ying Shan , Kwan-Yee K. Wong

Inferring full-body poses from Head Mounted Devices, which capture only 3-joint observations from the head and wrists, is a challenging task with wide AR/VR applications. Previous attempts focus on learning one-stage motion mapping and thus…

计算机视觉与模式识别 · 计算机科学 2025-05-13 Fangyu Du , Yang Yang , Xuehao Gao , Hongye Hou

In this paper, we introduce a novel text-to-avatar generation method that separately generates the human body and the clothes and allows high-quality animation on the generated avatar. While recent advancements in text-to-avatar generation…

计算机视觉与模式识别 · 计算机科学 2024-09-27 Jionghao Wang , Yuan Liu , Zhiyang Dou , Zhengming Yu , Yongqing Liang , Cheng Lin , Xin Li , Wenping Wang , Rong Xie , Li Song

We propose a novel framework for decomposing arbitrarily posed humans into animatable multi-layered 3D human avatars, separating the body and garments. Conventional single-layer reconstruction methods lock clothing to one identity, while…

计算机视觉与模式识别 · 计算机科学 2026-01-12 Yinghan Xu , John Dingliana

This report introduces Make-A-Character 2, an advanced system for generating high-quality 3D characters from single portrait photographs, ideal for game development and digital human applications. Make-A-Character 2 builds upon its…

计算机视觉与模式识别 · 计算机科学 2025-01-16 Lin Liu , Yutong Wang , Jiahao Chen , Jianfang Li , Tangli Xue , Longlong Li , Jianqiang Ren , Liefeng Bo

We introduce PortraitGen, a powerful portrait video editing method that achieves consistent and expressive stylization with multimodal prompts. Traditional portrait video editing methods often struggle with 3D and temporal consistency, and…

计算机视觉与模式识别 · 计算机科学 2024-09-23 Xuan Gao , Haiyao Xiao , Chenglai Zhong , Shimin Hu , Yudong Guo , Juyong Zhang

The estimation of 3D human body pose and shape from a single image has been extensively studied in recent years. However, the texture generation problem has not been fully discussed. In this paper, we propose an end-to-end learning strategy…

计算机视觉与模式识别 · 计算机科学 2019-04-09 Jian Wang , Yunshan Zhong , Yachun Li , Chi Zhang , Yichen Wei

To make 3D human avatars widely available, we must be able to generate a variety of 3D virtual humans with varied identities and shapes in arbitrary poses. This task is challenging due to the diversity of clothed body shapes, their complex…

计算机视觉与模式识别 · 计算机科学 2022-04-14 Xu Chen , Tianjian Jiang , Jie Song , Jinlong Yang , Michael J. Black , Andreas Geiger , Otmar Hilliges

The text-to-image (T2I) personalization diffusion model can generate images of the novel concept based on the user input text caption. However, existing T2I personalized methods either require test-time fine-tuning or fail to generate…

计算机视觉与模式识别 · 计算机科学 2024-12-25 Xiao Guo , Manh Tran , Jiaxin Cheng , Xiaoming Liu

We propose an approach to generate images of people given a desired appearance and pose. Disentangled representations of pose and appearance are necessary to handle the compound variability in the resulting generated images. Hence, we…

计算机视觉与模式识别 · 计算机科学 2021-04-27 Mengyao Zhai , Ruizhi Deng , Jiacheng Chen , Lei Chen , Zhiwei Deng , Greg Mori

Generating high-fidelity images of humans with fine-grained control over attributes such as hairstyle and clothing remains a core challenge in personalized text-to-image synthesis. While prior methods emphasize identity preservation from a…

计算机视觉与模式识别 · 计算机科学 2025-10-17 Guocheng Gordon Qian , Daniil Ostashev , Egor Nemchinov , Avihay Assouline , Sergey Tulyakov , Kuan-Chieh Jackson Wang , Kfir Aberman

We propose an alternative generator architecture for generative adversarial networks, borrowing from style transfer literature. The new architecture leads to an automatically learned, unsupervised separation of high-level attributes (e.g.,…

神经与进化计算 · 计算机科学 2019-04-01 Tero Karras , Samuli Laine , Timo Aila

Personalizing 3D scenes from a single reference image enables intuitive user-guided editing, which requires achieving both multi-view consistency across perspectives and referential consistency with the input image. However, these goals are…

计算机视觉与模式识别 · 计算机科学 2025-12-29 Yuxuan Wang , Xuanyu Yi , Qingshan Xu , Yuan Zhou , Long Chen , Hanwang Zhang

Existing full-body Gaussian avatar methods primarily optimize global reconstruction quality and often fail to preserve fine-grained facial geometry and expression details. This challenge arises from limited facial representational capacity…

计算机视觉与模式识别 · 计算机科学 2026-04-14 Willem Menu , Erkut Akdag , Pedro Quesado , Yasaman Kashefbahrami , Egor Bondarev

Creating a controllable and relightable digital avatar from multi-view video with fixed illumination is a very challenging problem since humans are highly articulated, creating pose-dependent appearance effects, and skin as well as clothing…

计算机视觉与模式识别 · 计算机科学 2024-07-29 Diogo Luvizon , Vladislav Golyanik , Adam Kortylewski , Marc Habermann , Christian Theobalt

We propose VLOGGER, a method for audio-driven human video generation from a single input image of a person, which builds on the success of recent generative diffusion models. Our method consists of 1) a stochastic human-to-3d-motion…

计算机视觉与模式识别 · 计算机科学 2024-03-14 Enric Corona , Andrei Zanfir , Eduard Gabriel Bazavan , Nikos Kolotouros , Thiemo Alldieck , Cristian Sminchisescu