English
Related papers

Related papers: Learning Disentangled Avatars with Hybrid 3D Repre…

200 papers

Recovering photorealistic and drivable full-body avatars is crucial for numerous applications, including virtual reality, 3D games, and tele-presence. Most methods, whether reconstruction or generation, require large numbers of human motion…

Computer Vision and Pattern Recognition · Computer Science 2024-05-31 Yujiao Jiang , Qingmin Liao , Zhaolong Wang , Xiangru Lin , Zongqing Lu , Yuxi Zhao , Hanqing Wei , Jingrui Ye , Yu Zhang , Zhijing Shao

Hand pose estimation from the monocular 2D image is challenging due to the variation in lighting, appearance, and background. While some success has been achieved using deep neural networks, they typically require collecting a large dataset…

Computer Vision and Pattern Recognition · Computer Science 2019-04-17 Yikang Li , Chris Twigg , Yuting Ye , Lingling Tao , Xiaogang Wang

Deep generative modelling for human body analysis is an emerging problem with many interesting applications. However, the latent space learned by such approaches is typically not interpretable, resulting in less flexibility. In this work,…

Computer Vision and Pattern Recognition · Computer Science 2020-02-18 Rodrigo de Bem , Arnab Ghosh , Thalaiyasingam Ajanthan , Ondrej Miksik , Adnane Boukhayma , N. Siddharth , Philip Torr

In this paper, we propose a framework for disentangling the appearance and geometry representations in the face recognition task. To provide supervision for this aim, we generate geometrically identical faces by incorporating spatial…

Computer Vision and Pattern Recognition · Computer Science 2020-01-15 Ali Dabouei , Fariborz Taherkhani , Sobhan Soleymani , Jeremy Dawson , Nasser M. Nasrabadi

Learning robotic manipulation from human videos is a promising solution to the data bottleneck in robotics, but the distribution shift between humans and robots remains a critical challenge. Existing approaches often produce entangled…

Robotics · Computer Science 2026-05-06 Zhiyuan Li , Wenyan Yang , Wenshuai Zhao , Yue Ma , Yuanpeng Tu , Pekka Marttinen , Joni Pajarinen

It is extremely challenging to create an animatable clothed human avatar from RGB videos, especially for loose clothes due to the difficulties in motion modeling. To address this problem, we introduce a novel representation on the basis of…

Computer Vision and Pattern Recognition · Computer Science 2022-03-29 Zerong Zheng , Han Huang , Tao Yu , Hongwen Zhang , Yandong Guo , Yebin Liu

This paper proposes a technique for efficiently modeling dynamic humans by explicifying the implicit neural fields via a Neural Explicit Surface (NES). Implicit neural fields have advantages over traditional explicit representations in…

Computer Vision and Pattern Recognition · Computer Science 2023-08-11 Ruiqi Zhang , Jie Chen , Qiang Wang

How can intelligent agents solve a diverse set of tasks in a data-efficient manner? The disentangled representation learning approach posits that such an agent would benefit from separating out (disentangling) the underlying structure of…

Machine Learning · Computer Science 2018-12-07 Irina Higgins , David Amos , David Pfau , Sebastien Racaniere , Loic Matthey , Danilo Rezende , Alexander Lerchner

Agents that are aware of the separation between themselves and their environments can leverage this understanding to form effective representations of visual input. We propose an approach for learning such structured representations for RL…

Machine Learning · Computer Science 2023-09-06 Kevin Gmelin , Shikhar Bahl , Russell Mendonca , Deepak Pathak

We propose ID-to-3D, a method to generate identity- and text-guided 3D human heads with disentangled expressions, starting from even a single casually captured in-the-wild image of a subject. The foundation of our approach is anchored in…

Computer Vision and Pattern Recognition · Computer Science 2024-05-29 Francesca Babiloni , Alexandros Lattas , Jiankang Deng , Stefanos Zafeiriou

Facial expression and hand motions are necessary to express our emotions and interact with the world. Nevertheless, most of the 3D human avatars modeled from a casually captured video only support body motions without facial expressions and…

Computer Vision and Pattern Recognition · Computer Science 2024-08-01 Gyeongsik Moon , Takaaki Shiratori , Shunsuke Saito

The task of reconstructing detailed 3D human body models from images is interesting but challenging in computer vision due to the high freedom of human bodies. In order to tackle the problem, we propose a coarse-to-fine method to…

Computer Vision and Pattern Recognition · Computer Science 2020-12-14 Zhongguo Li , Magnus Oskarsson , Anders Heyden

Existing methods for human parsing into body parts and clothing often use fixed mask categories with broad labels that obscure fine-grained clothing types. Recent open-vocabulary segmentation approaches leverage pretrained text-to-image…

Computer Vision and Pattern Recognition · Computer Science 2025-12-18 Kiran Chhatre , Christopher Peters , Srikrishna Karanam

Developing meaningful and efficient representations that separate the fundamental structure of the data generation mechanism is crucial in representation learning. However, Disentangled Representation Learning has not fully shown its…

Computer Vision and Pattern Recognition · Computer Science 2025-06-27 Jacopo Dapueto , Nicoletta Noceti , Francesca Odone

We present Vid2Avatar, a method to learn human avatars from monocular in-the-wild videos. Reconstructing humans that move naturally from monocular in-the-wild videos is difficult. Solving it requires accurately separating humans from…

Computer Vision and Pattern Recognition · Computer Science 2023-02-23 Chen Guo , Tianjian Jiang , Xu Chen , Jie Song , Otmar Hilliges

We present a novel method for generating 3D garment deformations from given body poses, which is key to a wide range of applications, including virtual try-on and extended reality. To simplify the cloth dynamics, existing methods mostly…

Computer Vision and Pattern Recognition · Computer Science 2025-12-08 Rong Wang , Wei Mao , Changsheng Lu , Hongdong Li

Recent 3D Gaussian Splatting (3DGS) GANs for human heads synthesize and render photorealistic 3D models in real-time and offer a vast variety in identity and appearance. However, controlling specific semantic attributes such as hair color…

Computer Vision and Pattern Recognition · Computer Science 2026-05-26 Florian Barthel , Shalini De Mello , Koki Nagano , Wieland Morgenstern , Anna Hilsmann , Peter Eisert

Parametric body models offer expressive 3D representation of humans across a wide range of poses, shapes, and facial expressions, typically derived by learning a basis over registered 3D meshes. However, existing human mesh modeling…

Computer Vision and Pattern Recognition · Computer Science 2025-08-22 Jinhyung Park , Javier Romero , Shunsuke Saito , Fabian Prada , Takaaki Shiratori , Yichen Xu , Federica Bogo , Shoou-I Yu , Kris Kitani , Rawal Khirodkar

Telecommunication with photorealistic avatars in virtual or augmented reality is a promising path for achieving authentic face-to-face communication in 3D over remote physical distances. In this work, we present the Pixel Codec Avatars…

Computer Vision and Pattern Recognition · Computer Science 2021-04-13 Shugao Ma , Tomas Simon , Jason Saragih , Dawei Wang , Yuecheng Li , Fernando De La Torre , Yaser Sheikh

We show, for the first time, that neural networks trained only on synthetic data achieve state-of-the-art accuracy on the problem of 3D human pose and shape (HPS) estimation from real images. Previous synthetic datasets have been small,…

Computer Vision and Pattern Recognition · Computer Science 2023-06-30 Michael J. Black , Priyanka Patel , Joachim Tesch , Jinlong Yang
‹ Prev 1 8 9 10 Next ›