English
Related papers

Related papers: C3DAG: Controlled 3D Animal Generation using 3D po…

200 papers

This paper proposes an end-to-end framework for generating 3D human pose datasets using Neural Radiance Fields (NeRF). Public datasets generally have limited diversity in terms of human poses and camera viewpoints, largely due to the…

Computer Vision and Pattern Recognition · Computer Science 2023-12-25 Mohsen Gholami , Rabab Ward , Z. Jane Wang

This paper presents InceptionHuman, a prompt-to-NeRF framework that allows easy control via a combination of prompts in different modalities (e.g., text, poses, edge, segmentation map, etc) as inputs to generate photorealistic 3D humans.…

Computer Vision and Pattern Recognition · Computer Science 2024-08-07 Shiu-hong Kao , Xinhang Liu , Yu-Wing Tai , Chi-Keung Tang

We present Layout-Your-3D, a framework that allows controllable and compositional 3D generation from text prompts. Existing text-to-3D methods often struggle to generate assets with plausible object interactions or require tedious…

Computer Vision and Pattern Recognition · Computer Science 2025-04-08 Junwei Zhou , Xueting Li , Lu Qi , Ming-Hsuan Yang

Recent attempts to solve the problem of head reenactment using a single reference image have shown promising results. However, most of them either perform poorly in terms of photo-realism, or fail to meet the identity preservation problem,…

Computer Vision and Pattern Recognition · Computer Science 2021-08-24 Michail Christos Doukas , Stefanos Zafeiriou , Viktoriia Sharmanska

Despite the unprecedented progress in the field of 3D generation, current systems still often fail to produce high-quality 3D assets that are visually appealing and geometrically and semantically consistent across multiple viewpoints. To…

Computer Vision and Pattern Recognition · Computer Science 2025-04-28 Shivam Duggal , Yushi Hu , Oscar Michel , Aniruddha Kembhavi , William T. Freeman , Noah A. Smith , Ranjay Krishna , Antonio Torralba , Ali Farhadi , Wei-Chiu Ma

Accurate and scalable quantification of animal pose and appearance is crucial for studying behavior. Current 3D pose estimation techniques, such as keypoint- and mesh-based techniques, often face challenges including limited…

Computer Vision and Pattern Recognition · Computer Science 2026-02-03 Jack Goffinet , Youngjo Min , Carlo Tomasi , David E. Carlson

We address the challenge of generating 3D articulated objects in a controllable fashion. Currently, modeling articulated 3D objects is either achieved through laborious manual authoring, or using methods from prior work that are hard to…

Computer Vision and Pattern Recognition · Computer Science 2024-03-21 Jiayi Liu , Hou In Ivan Tam , Ali Mahdavi-Amiri , Manolis Savva

We train a feed-forward text-to-3D diffusion generator for human characters using only single-view 2D data for supervision. Existing 3D generative models cannot yet match the fidelity of image or video generative models. State-of-the-art 3D…

Computer Vision and Pattern Recognition · Computer Science 2024-12-24 Souhaib Attaiki , Paul Guerrero , Duygu Ceylan , Niloy J. Mitra , Maks Ovsjanikov

Talking head generation is to generate video based on a given source identity and target motion. However, current methods face several challenges that limit the quality and controllability of the generated videos. First, the generated face…

Computer Vision and Pattern Recognition · Computer Science 2023-11-03 Yue Gao , Yuan Zhou , Jinglu Wang , Xiao Li , Xiang Ming , Yan Lu

We introduce Garment3DGen a new method to synthesize 3D garment assets from a base mesh given a single input image as guidance. Our proposed approach allows users to generate 3D textured clothes based on both real and synthetic images, such…

Computer Vision and Pattern Recognition · Computer Science 2025-05-01 Nikolaos Sarafianos , Tuur Stuyck , Xiaoyu Xiang , Yilei Li , Jovan Popovic , Rakesh Ranjan

Recent advancements in text-to-image generation have enabled significant progress in zero-shot 3D shape generation. This is achieved by score distillation, a methodology that uses pre-trained text-to-image diffusion models to optimize the…

Computer Vision and Pattern Recognition · Computer Science 2023-05-29 Zhenzhen Weng , Zeyu Wang , Serena Yeung

Creating detailed 3D human avatars with fitted garments traditionally requires specialized expertise and labor-intensive workflows. While recent advances in generative AI have enabled text-to-3D human and clothing synthesis, existing…

Computer Vision and Pattern Recognition · Computer Science 2026-04-01 Zhiyao Sun , Yu-Hui Wen , Ho-Jui Fang , Sheng Ye , Matthieu Lin , Tian Lv , Yong-Jin Liu

We introduce PAT3D, the first physics-augmented text-to-3D scene generation framework that integrates vision-language models with physics-based simulation to produce physically plausible, simulation-ready, and intersection-free 3D scenes.…

Computer Vision and Pattern Recognition · Computer Science 2026-04-24 Guying Lin , Kemeng Huang , Michael Liu , Ruihan Gao , Hanke Chen , Lyuhao Chen , Beijia Lu , Taku Komura , Yuan Liu , Jun-Yan Zhu , Minchen Li

Drag-based editing has become popular in 2D content creation, driven by the capabilities of image generative models. However, extending this technique to 3D remains a challenge. Existing 3D drag-based editing methods, whether employing…

Computer Vision and Pattern Recognition · Computer Science 2024-10-22 Honghua Chen , Yushi Lan , Yongwei Chen , Yifan Zhou , Xingang Pan

While NeRF-based 3D-aware image generation methods enable viewpoint control, limitations still remain to be adopted to various 3D applications. Due to their view-dependent and light-entangled volume representation, the 3D geometry presents…

Computer Vision and Pattern Recognition · Computer Science 2022-03-15 Minsoo Lee , Chaeyeon Chung , Hojun Cho , Minjung Kim , Sanghun Jung , Jaegul Choo , Minhyuk Sung

Inverse graphics aims to recover 3D models from 2D observations. Utilizing differentiable rendering, recent 3D-aware generative models have shown impressive results of rigid object generation using 2D images. However, it remains challenging…

Computer Vision and Pattern Recognition · Computer Science 2022-10-11 Fangzhou Hong , Zhaoxi Chen , Yushi Lan , Liang Pan , Ziwei Liu

We introduce AvatarForge, a framework for generating animatable 3D human avatars from text or image inputs using AI-driven procedural generation. While diffusion-based methods have made strides in general 3D object generation, they struggle…

Computer Vision and Pattern Recognition · Computer Science 2025-03-12 Xinhang Liu , Yu-Wing Tai , Chi-Keung Tang

This paper presents a novel method for generating diverse 3D human poses in scenes with semantic control. Existing methods heavily rely on the human-scene interaction dataset, resulting in a limited diversity of the generated human poses.…

Computer Vision and Pattern Recognition · Computer Science 2024-06-11 Bowen Dang , Xi Zhao

To represent people in mixed reality applications for collaboration and communication, we need to generate realistic and faithful avatar poses. However, the signal streams that can be applied for this task from head-mounted devices (HMDs)…

Computer Vision and Pattern Recognition · Computer Science 2022-03-14 Sadegh Aliakbarian , Pashmina Cameron , Federica Bogo , Andrew Fitzgibbon , Thomas J. Cashman

Generating realistic, room-level indoor scenes with semantically plausible and detailed appearances from in-the-wild images is crucial for various applications in VR, AR, and robotics. The success of NeRF-based generative methods indicates…

Computer Vision and Pattern Recognition · Computer Science 2025-04-04 Ming-Jia Yang , Yu-Xiao Guo , Yang Liu , Bin Zhou , Xin Tong