English
Related papers

Related papers: FactorizedHMR: A Hybrid Framework for Video Human …

200 papers

Although existing video-based 3D human mesh recovery methods have made significant progress, simultaneously estimating human pose and shape from low-resolution image features limits their performance. These image features lack sufficient…

Computer Vision and Pattern Recognition · Computer Science 2024-10-22 Tao Tang , Hong Liu , Yingxuan You , Ti Wang , Wenhao Li

Recovering whole-body mesh by inferring the abstract pose and shape parameters from visual content can obtain 3D bodies with realistic structures. However, the inferring process is highly non-linear and suffers from image-mesh misalignment,…

Computer Vision and Pattern Recognition · Computer Science 2023-04-13 Jiefeng Li , Siyuan Bian , Chao Xu , Zhicun Chen , Lixin Yang , Cewu Lu

Recovering the 3D geometry of a scene from a sparse set of uncalibrated images is a long-standing problem in computer vision. While recent learning-based approaches such as DUSt3R and MASt3R have demonstrated impressive results by directly…

Computer Vision and Pattern Recognition · Computer Science 2025-08-25 Sara Rojas , Matthieu Armando , Bernard Ghamen , Philippe Weinzaepfel , Vincent Leroy , Gregory Rogez

Multi-frame human pose estimation in complicated situations is challenging. Although state-of-the-art human joints detectors have demonstrated remarkable results for static images, their performances come short when we apply these models to…

Computer Vision and Pattern Recognition · Computer Science 2021-03-22 Zhenguang Liu , Haoming Chen , Runyang Feng , Shuang Wu , Shouling Ji , Bailin Yang , Xun Wang

The ability to sense, localize, and estimate the 3D position and orientation of the human body is critical in virtual reality (VR) and extended reality (XR) applications. This becomes more important and challenging with the deployment of…

Human-Computer Interaction · Computer Science 2024-05-14 Nguyen Quang Hieu , Dinh Thai Hoang , Diep N. Nguyen , Mohammad Abu Alsheikh

We present UniSH, a unified, feed-forward framework for joint metric-scale 3D scene and human reconstruction. A key challenge in this domain is the scarcity of large-scale, annotated real-world data, forcing a reliance on synthetic…

Computer Vision and Pattern Recognition · Computer Science 2026-01-06 Mengfei Li , Peng Li , Zheng Zhang , Jiahao Lu , Chengfeng Zhao , Wei Xue , Qifeng Liu , Sida Peng , Wenxiao Zhang , Wenhan Luo , Yuan Liu , Yike Guo

Recognizing human activities in videos is challenging due to the spatio-temporal complexity and context-dependence of human interactions. Prior studies often rely on single input modalities, such as RGB or skeletal data, limiting their…

Computer Vision and Pattern Recognition · Computer Science 2024-09-05 Tuyen Tran , Thao Minh Le , Hung Tran , Truyen Tran

Retrieving videos of a particular person with face image as a query via hashing technique has many important applications. While face images are typically represented as vectors in Euclidean space, characterizing face videos with some…

Computer Vision and Pattern Recognition · Computer Science 2019-11-05 Shishi Qiao , Ruiping Wang , Shiguang Shan , Xilin Chen

Human reconstruction and synthesis from monocular RGB videos is a challenging problem due to clothing, occlusion, texture discontinuities and sharpness, and framespecific pose changes. Many methods employ deferred rendering, NeRFs and…

Computer Vision and Pattern Recognition · Computer Science 2023-03-16 Rohit Jena , Pratik Chaudhari , James Gee , Ganesh Iyer , Siddharth Choudhary , Brandon M. Smith

Articulated 3D object generation is fundamental for creating realistic, functional, and interactable virtual assets which are not simply static. We introduce MeshArt, a hierarchical transformer-based approach to generate articulated 3D…

Computer Vision and Pattern Recognition · Computer Science 2025-06-10 Daoyi Gao , Yawar Siddiqui , Lei Li , Angela Dai

We present a method for recovering the shape and radiance of a scene consisting of multiple people given solely a few images. Multi-human scenes are complex due to additional occlusion and clutter. For single-human settings, existing…

Computer Vision and Pattern Recognition · Computer Science 2025-02-12 Qian li , Victoria Fernàndez Abrevaya , Franck Multon , Adnane Boukhayma

Generalizable rendering of an animatable human avatar from sparse inputs relies on data priors and inductive biases extracted from training on large data to avoid scene-specific optimization and to enable fast reconstruction. This raises…

Computer Vision and Pattern Recognition · Computer Science 2025-02-14 Jing Wen , Alexander G. Schwing , Shenlong Wang

Recent advancements in 3D human pose estimation from single-camera images and videos have relied on parametric models, like SMPL. However, these models oversimplify anatomical structures, limiting their accuracy in capturing true joint…

Computer Vision and Pattern Recognition · Computer Science 2025-01-15 Farnoosh Koleini , Muhammad Usama Saleem , Pu Wang , Hongfei Xue , Ahmed Helmy , Abbey Fenwick

Recent years have witnessed tremendous progress in the 3D reconstruction of dynamic humans from a monocular video with the advent of neural rendering techniques. This task has a wide range of applications, including the creation of virtual…

Computer Vision and Pattern Recognition · Computer Science 2024-09-24 Kanghao Chen , Zeyu Wang , Lin Wang

We propose an end-to-end unified 3D mesh recovery of humans and quadruped animals trained in a weakly-supervised way. Unlike recent work focusing on a single target class only, we aim to recover 3D mesh of broader classes with a single…

Computer Vision and Pattern Recognition · Computer Science 2021-11-05 Kim Youwang , Kim Ji-Yeon , Kyungdon Joo , Tae-Hyun Oh

We introduce a novel, data-driven approach for reconstructing temporally coherent 3D motion from unstructured and potentially partial observations of non-rigidly deforming shapes. Our goal is to achieve high-fidelity motion reconstructions…

Computer Vision and Pattern Recognition · Computer Science 2025-09-09 Aymen Merrouche , Stefanie Wuhrer , Edmond Boyer

3D Human Mesh Reconstruction (HMR) from 2D RGB images faces challenges in environments with poor lighting, privacy concerns, or occlusions. These weaknesses of RGB imaging can be complemented by acoustic signals, which are widely available,…

Computer Vision and Pattern Recognition · Computer Science 2024-12-17 Xiaoxuan Liang , Wuyang Zhang , Hong Zhou , Zhaolong Wei , Sicheng Zhu , Yansong Li , Rui Yin , Jiantao Yuan , Jeremy Gummeson

High Dynamic Range (HDR) imaging aims to reproduce the wide range of brightness levels present in natural scenes, which the human visual system can perceive but conventional digital cameras often fail to capture due to their limited dynamic…

Image and Video Processing · Electrical Eng. & Systems 2025-10-28 Kumbha Nagaswetha

Much progress has been made in the supervised learning of 3D reconstruction of rigid objects from multi-view images or a video. However, it is more challenging to reconstruct severely deformed objects from a single-view RGB image in an…

Computer Vision and Pattern Recognition · Computer Science 2022-01-25 Jie Mei , Jingxi Yu , Suzanne Romain , Craig Rose , Kelsey Magrane , Graeme LeeSon , Jenq-Neng Hwang

The rapid development of multi-view 3D human pose estimation (HPE) is attributed to the maturation of monocular 2D HPE and the geometry of 3D reconstruction. However, 2D detection outliers in occluded views due to neglect of view…

Computer Vision and Pattern Recognition · Computer Science 2023-02-24 Xiaoyue Wan , Zhuo Chen , Xu Zhao