English
Related papers

Related papers: HumMorph: Generalized Dynamic Human Neural Fields …

200 papers

We present BlazePose GHUM Holistic, a lightweight neural network pipeline for 3D human body landmarks and pose estimation, specifically tailored to real-time on-device inference. BlazePose GHUM Holistic enables motion capture from a single…

Predicting human motion from historical pose sequence is crucial for a machine to succeed in intelligent interactions with humans. One aspect that has been obviated so far, is the fact that how we represent the skeletal pose has a critical…

Computer Vision and Pattern Recognition · Computer Science 2022-01-03 Zhenguang Liu , Shuang Wu , Shuyuan Jin , Shouling Ji , Qi Liu , Shijian Lu , Li Cheng

We propose a new self-supervised method for predicting 3D human body pose from a single image. The prediction network is trained from a dataset of unlabelled images depicting people in typical poses and a set of unpaired 2D poses. By…

Computer Vision and Pattern Recognition · Computer Science 2023-04-06 Jose Sosa , David Hogg

Recovering 3D human pose from 2D joints is still a challenging problem, especially without any 3D annotation, video information, or multi-view information. In this paper, we present an unsupervised GAN-based model consisting of multiple…

Computer Vision and Pattern Recognition · Computer Science 2022-04-14 Yicheng Deng , Cheng Sun , Jiahui Zhu , Yongqi Sun

The end-to-end Human Mesh Recovery (HMR) approach has been successfully used for 3D body reconstruction. However, most HMR-based frameworks reconstruct human body by directly learning mesh parameters from images or videos, while lacking…

Computer Vision and Pattern Recognition · Computer Science 2021-03-19 Tianyu Luan , Yali Wang , Junhao Zhang , Zhe Wang , Zhipeng Zhou , Yu Qiao

We introduce UPose3D, a novel approach for multi-view 3D human pose estimation, addressing challenges in accuracy and scalability. Our method advances existing pose estimation frameworks by improving robustness and flexibility without…

Computer Vision and Pattern Recognition · Computer Science 2024-07-11 Vandad Davoodnia , Saeed Ghorbani , Marc-André Carbonneau , Alexandre Messier , Ali Etemad

The neural rendering of humans is a topic of great research significance. However, previous works mostly focus on achieving photorealistic details, neglecting the exploration of human parsing. Additionally, classical semantic work are all…

Computer Vision and Pattern Recognition · Computer Science 2023-08-22 Jie Zhang , Pengcheng Shi , Zaiwang Gu , Yiyang Zhou , Zhi Wang

Image-based volumetric humans using pixel-aligned features promise generalization to unseen poses and identities. Prior work leverages global spatial encodings and multi-view geometric consistency to reduce spatial ambiguity. However,…

Computer Vision and Pattern Recognition · Computer Science 2022-07-22 Marko Mihajlovic , Aayush Bansal , Michael Zollhoefer , Siyu Tang , Shunsuke Saito

Dynamic human rendering from video sequences has achieved remarkable progress by formulating the rendering as a mapping from static poses to human images. However, existing methods focus on the human appearance reconstruction of every…

Computer Vision and Pattern Recognition · Computer Science 2024-04-03 Tao Hu , Fangzhou Hong , Ziwei Liu

Human motion capture either requires multi-camera systems or is unreliable when using single-view input due to depth ambiguities. Meanwhile, mirrors are readily available in urban environments and form an affordable alternative by recording…

Computer Vision and Pattern Recognition · Computer Science 2024-05-17 Daniel Ajisafe , James Tang , Shih-Yang Su , Bastian Wandt , Helge Rhodin

Recent advances in 3D human shape reconstruction from single images have shown impressive results, leveraging on deep networks that model the so-called implicit function to learn the occupancy status of arbitrarily dense 3D points in space.…

Computer Vision and Pattern Recognition · Computer Science 2022-05-10 Nicolas Ugrinovic , Albert Pumarola , Alberto Sanfeliu , Francesc Moreno-Noguer

Human actions are comprised of a sequence of poses. This makes videos of humans a rich and dense source of human poses. We propose an unsupervised method to learn pose features from videos that exploits a signal which is complementary to…

Computer Vision and Pattern Recognition · Computer Science 2016-09-20 Senthil Purushwalkam , Abhinav Gupta

In the rapidly evolving field of computer vision, the task of accurately estimating the poses of multiple individuals from various viewpoints presents a formidable challenge, especially if the estimations should be reliable as well. This…

Computer Vision and Pattern Recognition · Computer Science 2024-12-23 Daniel Bermuth , Alexander Poeppel , Wolfgang Reif

Reconstructing dynamic humans interacting with real-world environments from monocular videos is an important and challenging task. Despite considerable progress in 4D neural rendering, existing approaches either model dynamic scenes…

Computer Vision and Pattern Recognition · Computer Science 2025-11-14 Wenqing Wang , Haosen Yang , Josef Kittler , Xiatian Zhu

Generating realistic human motion is essential for many computer vision and graphics applications. The wide variety of human body shapes and sizes greatly impacts how people move. However, most existing motion models ignore these…

Computer Vision and Pattern Recognition · Computer Science 2025-04-04 Shashank Tripathi , Omid Taheri , Christoph Lassner , Michael J. Black , Daniel Holden , Carsten Stoll

Human motion prediction is crucial for human-centric multimedia understanding and interacting. Current methods typically rely on ground truth human poses as observed input, which is not practical for real-world scenarios where only raw…

Computer Vision and Pattern Recognition · Computer Science 2024-10-02 Xiao Han , Yiming Ren , Yichen Yao , Yujing Sun , Yuexin Ma

This paper presents a novel method for generating diverse 3D human poses in scenes with semantic control. Existing methods heavily rely on the human-scene interaction dataset, resulting in a limited diversity of the generated human poses.…

Computer Vision and Pattern Recognition · Computer Science 2024-06-11 Bowen Dang , Xi Zhao

We tackle the problem of highly-accurate, holistic performance capture for the face, body and hands simultaneously. Motion-capture technologies used in film and game production typically focus only on face, body or hand capture…

Creating a photorealistic scene and human reconstruction from a single monocular in-the-wild video figures prominently in the perception of a human-centric 3D world. Recent neural rendering advances have enabled holistic human-scene…

Computer Vision and Pattern Recognition · Computer Science 2025-04-21 Zetong Zhang , Manuel Kaufmann , Lixin Xue , Jie Song , Martin R. Oswald

We present the first method to capture the 3D total motion of a target person from a monocular view input. Given an image or a monocular video, our method reconstructs the motion from body, face, and fingers represented by a 3D deformable…

Computer Vision and Pattern Recognition · Computer Science 2018-12-05 Donglai Xiang , Hanbyul Joo , Yaser Sheikh