English
Related papers

Related papers: D$^3$-Human: Dynamic Disentangled Digital Human fr…

200 papers

Obtaining personalized 3D animatable avatars from a monocular camera has several real world applications in gaming, virtual try-on, animation, and VR/XR, etc. However, it is very challenging to model dynamic and fine-grained clothing…

Computer Vision and Pattern Recognition · Computer Science 2023-10-31 Yuxuan Xue , Bharat Lal Bhatnagar , Riccardo Marin , Nikolaos Sarafianos , Yuanlu Xu , Gerard Pons-Moll , Tony Tung

Reconstructing 3D clothed humans from monocular images and videos is a fundamental problem with applications in virtual try-on, avatar creation, and mixed reality. Despite significant progress in human body recovery, accurately…

Computer Vision and Pattern Recognition · Computer Science 2026-03-02 Yingxuan You , Ren Li , Corentin Dumery , Cong Cao , Hao Li , Pascal Fua

We present a novel method to improve the accuracy of the 3D reconstruction of clothed human shape from a single image. Recent work has introduced volumetric, implicit and model-based shape learning frameworks for reconstruction of objects…

Computer Vision and Pattern Recognition · Computer Science 2020-09-30 Akin Caliskan , Armin Mustafa , Evren Imre , Adrian Hilton

Reconstructing clothed humans from a single image is a fundamental task in computer vision with wide-ranging applications. Although existing monocular clothed human reconstruction solutions have shown promising results, they often rely on…

Computer Vision and Pattern Recognition · Computer Science 2025-10-20 Arindam Dutta , Meng Zheng , Zhongpai Gao , Benjamin Planche , Anwesha Choudhuri , Terrence Chen , Amit K. Roy-Chowdhury , Ziyan Wu

We propose CrossHuman, a novel method that learns cross-guidance from parametric human model and multi-frame RGB images to achieve high-quality 3D human reconstruction. To recover geometry details and texture even in invisible regions, we…

Computer Vision and Pattern Recognition · Computer Science 2022-07-21 Liliang Chen , Jiaqi Li , Han Huang , Yandong Guo

We present a novel deep learning-based approach to the 3D reconstruction of clothed humans using weak supervision via 2D normal maps. Given a single RGB image or multiview images, our network infers a signed distance function (SDF)…

Computer Vision and Pattern Recognition · Computer Science 2023-11-28 Jane Wu , Diego Thomas , Ronald Fedkiw

We present Vid2Avatar, a method to learn human avatars from monocular in-the-wild videos. Reconstructing humans that move naturally from monocular in-the-wild videos is difficult. Solving it requires accurately separating humans from…

Computer Vision and Pattern Recognition · Computer Science 2023-02-23 Chen Guo , Tianjian Jiang , Xu Chen , Jie Song , Otmar Hilliges

This paper presents an approach that reconstructs a hand-held object from a monocular video. In contrast to many recent methods that directly predict object geometry by a trained network, the proposed approach does not require any learned…

Computer Vision and Pattern Recognition · Computer Science 2022-12-01 Di Huang , Xiaopeng Ji , Xingyi He , Jiaming Sun , Tong He , Qing Shuai , Wanli Ouyang , Xiaowei Zhou

Recovering the 3D shape of a person from its 2D appearance is ill-posed due to ambiguities. Nevertheless, with the help of convolutional neural networks (CNN) and prior knowledge on the 3D human body, it is possible to overcome such…

Computer Vision and Pattern Recognition · Computer Science 2020-04-23 Hayato Onizuka , Zehra Hayirci , Diego Thomas , Akihiro Sugimoto , Hideaki Uchiyama , Rin-ichiro Taniguchi

Monocular 3D clothed human reconstruction aims to generate a complete and realistic textured 3D avatar from a single image. Existing methods are commonly trained under multi-view supervision with annotated geometric priors, and during…

Computer Vision and Pattern Recognition · Computer Science 2026-03-06 Nanjie Yao , Gangjian Zhang , Wenhao Shen , Jian Shu , Yu Feng , Hao Wang

For visual manipulation tasks, we aim to represent image content with semantically meaningful features. However, learning implicit representations from images often lacks interpretability, especially when attributes are intertwined. We…

Computer Vision and Pattern Recognition · Computer Science 2022-08-08 Xue Hu , Xinghui Li , Benjamin Busam , Yiren Zhou , Ales Leonardis , Shanxin Yuan

3D human body reconstruction from monocular images is an interesting and ill-posed problem in computer vision with wider applications in multiple domains. In this paper, we propose SHARP, a novel end-to-end trainable network that accurately…

Computer Vision and Pattern Recognition · Computer Science 2021-11-24 Sai Sagar Jinka , Rohan Chacko , Astitva Srivastava , Avinash Sharma , P. J. Narayanan

Fast 3D clothed human reconstruction from monocular video remains a significant challenge in computer vision, particularly in balancing computational efficiency with reconstruction quality. Current approaches are either focused on static…

Computer Vision and Pattern Recognition · Computer Science 2025-05-13 Matthew Marchellus , Nadhira Noor , In Kyu Park

We present a new end-to-end learning framework to obtain detailed and spatially coherent reconstructions of multiple people from a single image. Existing multi-person methods suffer from two main drawbacks: they are often model-based and…

Computer Vision and Pattern Recognition · Computer Science 2021-04-20 Armin Mustafa , Akin Caliskan , Lourdes Agapito , Adrian Hilton

This paper proposes a new end-to-end neural rendering architecture to transfer appearance and reenact human actors. Our method leverages a carefully designed graph convolutional network (GCN) to model the human body manifold structure,…

Computer Vision and Pattern Recognition · Computer Science 2021-10-25 Thiago L. Gomes , Thiago M. Coutinho , Rafael Azevedo , Renato Martins , Erickson R. Nascimento

Recently, implicit neural representation has been widely used to generate animatable human avatars. However, the materials and geometry of those representations are coupled in the neural network and hard to edit, which hinders their…

Computer Vision and Pattern Recognition · Computer Science 2024-05-21 Qifeng Chen , Rengan Xie , Kai Huang , Qi Wang , Wenting Zheng , Rong Li , Yuchi Huo

Monocular dynamic video reconstruction faces significant challenges in dynamic human scenes due to geometric inconsistencies and resolution degradation issues. Existing methods lack 3D human structural understanding, producing geometrically…

Computer Vision and Pattern Recognition · Computer Science 2025-12-10 Weitao Xiong , Zhiyuan Yuan , Jiahao Lu , Chengfeng Zhao , Peng Li , Yuan Liu

Recovering world-coordinate human motion from monocular videos with humanoid robot retargeting is significant for embodied intelligence and robotics. To avoid complex SLAM pipelines or heavy temporal models, we propose a lightweight,…

Robotics · Computer Science 2025-12-29 Zhangzheng Tu , Kailun Su , Shaolong Zhu , Yukun Zheng

Our work aims to reconstruct a 3D object that is held and rotated by a hand in front of a static RGB camera. Previous methods that use implicit neural representations to recover the geometry of a generic hand-held object from multi-view…

Computer Vision and Pattern Recognition · Computer Science 2023-12-29 Shijian Jiang , Qi Ye , Rengan Xie , Yuchi Huo , Xiang Li , Yang Zhou , Jiming Chen

In this paper, we propose StereoPIFu, which integrates the geometric constraints of stereo vision with implicit function representation of PIFu, to recover the 3D shape of the clothed human from a pair of low-cost rectified images. First,…

Computer Vision and Pattern Recognition · Computer Science 2021-04-14 Yang Hong , Juyong Zhang , Boyi Jiang , Yudong Guo , Ligang Liu , Hujun Bao