English
Related papers

Related papers: XFormer: Fast and Accurate Monocular 3D Body Captu…

200 papers

In this paper, we aim to create generalizable and controllable neural signed distance fields (SDFs) that represent clothed humans from monocular depth observations. Recent advances in deep learning, especially neural implicit…

Computer Vision and Pattern Recognition · Computer Science 2022-01-21 Shaofei Wang , Marko Mihajlovic , Qianli Ma , Andreas Geiger , Siyu Tang

Recent advances in full-head reconstruction have been obtained by optimizing a neural field through differentiable surface or volume rendering to represent a single scene. While these techniques achieve an unprecedented accuracy, they take…

Computer Vision and Pattern Recognition · Computer Science 2024-04-08 Antonio Canela , Pol Caselles , Ibrar Malik , Eduard Ramon , Jaime García , Jordi Sánchez-Riera , Gil Triginer , Francesc Moreno-Noguer

Remote physiological signal measurement based on facial videos, also known as remote photoplethysmography (rPPG), involves predicting changes in facial vascular blood flow from facial videos. While most deep learning-based methods have…

Computer Vision and Pattern Recognition · Computer Science 2025-01-08 Jiachen Li , Shisheng Guo , Longzhen Tang , Cuolong Cui , Lingjiang Kong , Xiaobo Yang

We introduce SkelFormer, a novel markerless motion capture pipeline for multi-view human pose and shape estimation. Our method first uses off-the-shelf 2D keypoint estimators, pre-trained on large-scale in-the-wild data, to obtain 3D joint…

Computer Vision and Pattern Recognition · Computer Science 2024-04-22 Vandad Davoodnia , Saeed Ghorbani , Alexandre Messier , Ali Etemad

We propose a CNN-based approach for 3D human body pose estimation from single RGB images that addresses the issue of limited generalizability of models trained solely on the starkly limited publicly available 3D pose data. Using only the…

Computer Vision and Pattern Recognition · Computer Science 2017-10-05 Dushyant Mehta , Helge Rhodin , Dan Casas , Pascal Fua , Oleksandr Sotnychenko , Weipeng Xu , Christian Theobalt

Monocular dynamic video reconstruction faces significant challenges in dynamic human scenes due to geometric inconsistencies and resolution degradation issues. Existing methods lack 3D human structural understanding, producing geometrically…

Computer Vision and Pattern Recognition · Computer Science 2025-12-10 Weitao Xiong , Zhiyuan Yuan , Jiahao Lu , Chengfeng Zhao , Peng Li , Yuan Liu

We present a simple lightweight markerless facial performance capture framework using just a monocular video input that combines Active Appearance Models for feature tracking and prior constraints on 3D shapes into an integrated objective…

Computer Vision and Pattern Recognition · Computer Science 2019-01-17 Shridhar Ravikumar

The appearance of a human in clothing is driven not only by the pose but also by its temporal context, i.e., motion. However, such context has been largely neglected by existing monocular human modeling methods whose neural networks often…

Computer Vision and Pattern Recognition · Computer Science 2023-12-29 Hansol Lee , Junuk Cha , Yunhoe Ku , Jae Shin Yoon , Seungryul Baek

We present an end-to-end neural network-based model for inferring an approximate 3D mesh representation of a human face from single camera input for AR applications. The relatively dense mesh model of 468 vertices is well-suited for…

Computer Vision and Pattern Recognition · Computer Science 2019-07-17 Yury Kartynnik , Artsiom Ablavatski , Ivan Grishchenko , Matthias Grundmann

Immersive telepresence aims to transform human interaction in AR/VR applications by enabling lifelike full-body holographic representations for enhanced remote collaboration. However, existing systems rely on hardware-intensive multi-camera…

Computer Vision and Pattern Recognition · Computer Science 2026-01-13 Fangyu Lin , Yingdong Hu , Zhening Liu , Yufan Zhuang , Zehong Lin , Jun Zhang

3D human body reconstruction has been a challenge in the field of computer vision. Previous methods are often time-consuming and difficult to capture the detailed appearance of the human body. In this paper, we propose a new method called…

Computer Vision and Pattern Recognition · Computer Science 2024-03-19 Mingjin Chen , Junhao Chen , Xiaojun Ye , Huan-ang Gao , Xiaoxue Chen , Zhaoxin Fan , Hao Zhao

Rendering moving human bodies at free viewpoints only from a monocular video is quite a challenging problem. The information is too sparse to model complicated human body structures and motions from both view and pose dimensions. Neural…

Computer Vision and Pattern Recognition · Computer Science 2023-03-28 Taoran Yi , Jiemin Fang , Xinggang Wang , Wenyu Liu

We propose an approach for optimizing high-quality clothed human body shapes in minutes, using multi-view posed images. While traditional neural rendering methods struggle to disentangle geometry and appearance using only rendering loss,…

Computer Vision and Pattern Recognition · Computer Science 2023-10-31 Lixiang Lin , Songyou Peng , Qijun Gan , Jianke Zhu

We propose a real time deep learning framework for video-based facial expression capture. Our process uses a high-end facial capture pipeline based on FACEGOOD to capture facial expression. We train a convolutional neural network to produce…

Computer Vision and Pattern Recognition · Computer Science 2021-11-16 Hongwei Xu , Leijia Dai , Jianxing Fu , Xiangyuan Wang , Quanwei Wang

Recent advances in 3D foundation models have led to growing interest in reconstructing humans and their surrounding environments. However, most existing approaches focus on monocular inputs, and extending them to multi-view settings…

Computer Vision and Pattern Recognition · Computer Science 2026-03-19 Sangmin Kim , Minhyuk Hwang , Geonho Cha , Dongyoon Wee , Jaesik Park

In recent years, Neural Radiance Fields (NeRF) have achieved remarkable progress in dynamic human reconstruction and rendering. Part-based rendering paradigms, guided by human segmentation, allow for flexible parameter allocation based on…

Computer Vision and Pattern Recognition · Computer Science 2025-08-13 Yao Lu , Jiawei Li , Ming Jiang

Recently, deep learning-based tooth segmentation methods have been limited by the expensive and time-consuming processes of data collection and labeling. Achieving high-precision segmentation with limited datasets is critical. A viable…

Computer Vision and Pattern Recognition · Computer Science 2023-10-24 Yuan Li , Huan Liu , Yubo Tao , Xiangyang He , Haifeng Li , Xiaohu Guo , Hai Lin

In this paper, we propose a fully convolutional network for 3D human pose estimation from monocular images. We use limb orientations as a new way to represent 3D poses and bind the orientation together with the bounding box of each limb…

Computer Vision and Pattern Recognition · Computer Science 2018-12-06 Chenxu Luo , Xiao Chu , Alan Yuille

We present a new method, called MEsh TRansfOrmer (METRO), to reconstruct 3D human pose and mesh vertices from a single image. Our method uses a transformer encoder to jointly model vertex-vertex and vertex-joint interactions, and outputs 3D…

Computer Vision and Pattern Recognition · Computer Science 2021-06-16 Kevin Lin , Lijuan Wang , Zicheng Liu

Monocular scene reconstruction from posed images is challenging due to the complexity of a large environment. Recent volumetric methods learn to directly predict the TSDF volume and have demonstrated promising results in this task. However,…

Computer Vision and Pattern Recognition · Computer Science 2023-03-10 Weihao Yuan , Xiaodong Gu , Heng Li , Zilong Dong , Siyu Zhu