中文
相关论文

相关论文: Text-guided 3D Human Generation from 2D Collection…

200 篇论文

3D digital garment generation and editing play a pivotal role in fashion design, virtual try-on, and gaming. Traditional methods struggle to meet the growing demand due to technical complexity and high resource costs. Learning-based…

图形学 · 计算机科学 2025-09-23 Ruiyan Wang , Zhengxue Cheng , Zonghao Lin , Jun Ling , Yuzhou Liu , Yanru An , Rong Xie , Li Song

In this paper, we present an end-to-end approach to generate high-resolution person images conditioned on texts only. State-of-the-art text-to-image generation models are mainly designed for center-object generation, e.g., flowers and…

计算机视觉与模式识别 · 计算机科学 2023-07-04 Deyin Liu , Lin Yuanbo Wu , Bo Li , Zongyuan Ge

While recent 3D generative models can produce high-quality texture images, they often fail to capture human preferences or meet task-specific requirements. Moreover, a core challenge in the 3D texture generation domain is that most existing…

计算机视觉与模式识别 · 计算机科学 2025-12-10 AmirHossein Zamani , Tianhao Xie , Amir G. Aghdam , Tiberiu Popa , Eugene Belilovsky

Achieving realistic animated human avatars requires accurate modeling of pose-dependent clothing deformations. Existing learning-based methods heavily rely on the Linear Blend Skinning (LBS) of minimally-clothed human models like SMPL to…

计算机视觉与模式识别 · 计算机科学 2025-04-10 Hang Ye , Xiaoxuan Ma , Hai Ci , Wentao Zhu , Yizhou Wang

Recent remarkable advances in large-scale text-to-image diffusion models have inspired a significant breakthrough in text-to-3D generation, pursuing 3D content creation solely from a given text prompt. However, existing text-to-3D…

计算机视觉与模式识别 · 计算机科学 2023-11-10 Yang Chen , Yingwei Pan , Yehao Li , Ting Yao , Tao Mei

Image extrapolation aims at expanding the narrow field of view of a given image patch. Existing models mainly deal with natural scene images of homogeneous regions and have no control of the content generation process. In this work, we…

计算机视觉与模式识别 · 计算机科学 2019-12-30 Yijun Li , Lu Jiang , Ming-Hsuan Yang

Traditional 3D garment creation is labor-intensive, involving sketching, modeling, UV mapping, and texturing, which are time-consuming and costly. Recent advances in diffusion-based generative models have enabled new possibilities for 3D…

计算机视觉与模式识别 · 计算机科学 2024-05-22 Boqian Li , Xuan Li , Ying Jiang , Tianyi Xie , Feng Gao , Huamin Wang , Yin Yang , Chenfanfu Jiang

This work focuses on generating realistic, physically-based human behaviors from multi-modal inputs, which may only partially specify the desired motion. For example, the input may come from a VR controller providing arm motion and body…

机器人学 · 计算机科学 2025-02-11 Aayam Shrestha , Pan Liu , German Ros , Kai Yuan , Alan Fern

We present Layout-Your-3D, a framework that allows controllable and compositional 3D generation from text prompts. Existing text-to-3D methods often struggle to generate assets with plausible object interactions or require tedious…

计算机视觉与模式识别 · 计算机科学 2025-04-08 Junwei Zhou , Xueting Li , Lu Qi , Ming-Hsuan Yang

Text-to-motion generation, which translates textual descriptions into human motions, faces the challenge that users often struggle to precisely convey their intended motions through text alone. To address this issue, this paper introduces…

计算机视觉与模式识别 · 计算机科学 2026-05-21 Tao Wang , Lei Jin , Zhihua Wu , Qiaozhi He , Jiaming Chu , Yu Cheng , Junliang Xing , Jian Zhao , Shuicheng Yan , Li Wang

Text-driven content creation has evolved to be a transformative technique that revolutionizes creativity. Here we study the task of text-driven human video generation, where a video sequence is synthesized from texts describing the…

计算机视觉与模式识别 · 计算机科学 2023-04-18 Yuming Jiang , Shuai Yang , Tong Liang Koh , Wayne Wu , Chen Change Loy , Ziwei Liu

For designing a wide range of everyday objects, the design process should be aware of both the human body and the underlying semantics of the design specification. However, these two objectives present significant challenges to the current…

计算机视觉与模式识别 · 计算机科学 2025-03-31 Michelle Guo , Mia Tang , Hannah Cha , Ruohan Zhang , C. Karen Liu , Jiajun Wu

We introduce MoMask, a novel masked modeling framework for text-driven 3D human motion generation. In MoMask, a hierarchical quantization scheme is employed to represent human motion as multi-layer discrete motion tokens with high-fidelity…

计算机视觉与模式识别 · 计算机科学 2023-12-04 Chuan Guo , Yuxuan Mu , Muhammad Gohar Javed , Sen Wang , Li Cheng

Realistic dynamic garments on animated characters have many AR/VR applications. While authoring such dynamic garment geometry is still a challenging task, data-driven simulation provides an attractive alternative, especially if it can be…

计算机视觉与模式识别 · 计算机科学 2022-09-26 Meng Zhang , Duygu Ceylan , Niloy J. Mitra

Pedestrian motion, due to its causal nature, is strongly influenced by domain gaps arising from discrepancies between training and testing data distributions. Focusing on 3D human pose estimation, this work presents a controllable human…

计算机视觉与模式识别 · 计算机科学 2026-05-13 Xinhao Hu , Yiyi Zhang , Liqing Zhang , Jianfu Zhang

If a picture is worth thousand words, a moving 3d shape must be worth a million. We build upon the success of recent generative methods that create images fitting the semantics of a text prompt, and extend it to the controlled generation of…

机器学习 · 计算机科学 2021-09-28 Nikolay Jetchev

High-fidelity 3D garment synthesis from text is desirable yet challenging for digital avatar creation. Recent diffusion-based approaches via Score Distillation Sampling (SDS) have enabled new possibilities but either intricately couple with…

计算机视觉与模式识别 · 计算机科学 2024-06-25 Yufei Liu , Junshu Tang , Chu Zheng , Shijie Zhang , Jinkun Hao , Junwei Zhu , Dongjin Huang

With the onset of diffusion-based generative models and their ability to generate text-conditioned images, content generation has received a massive invigoration. Recently, these models have been shown to provide useful guidance for the…

计算机视觉与模式识别 · 计算机科学 2023-11-30 Alexander Vilesov , Pradyumna Chari , Achuta Kadambi

Digital human motion synthesis is a vibrant research field with applications in movies, AR/VR, and video games. Whereas methods were proposed to generate natural and realistic human motions, most only focus on modeling humans and largely…

计算机视觉与模式识别 · 计算机科学 2023-11-07 Quanzhou Li , Jingbo Wang , Chen Change Loy , Bo Dai

We present DreamBooth3D, an approach to personalize text-to-3D generative models from as few as 3-6 casually captured images of a subject. Our approach combines recent advances in personalizing text-to-image models (DreamBooth) with…