English
Related papers

Related papers: Decaf: Monocular Deformation Capture for Face and …

200 papers

We present a fully automatic approach to real-time 3D face reconstruction from monocular in-the-wild videos. With the use of a cascaded-regressor based face tracking and a 3D Morphable Face Model shape fitting, we obtain a semi-dense 3D…

Computer Vision and Pattern Recognition · Computer Science 2017-08-28 Patrik Huber , Philipp Kopp , Matthias Rätsch , William Christmas , Josef Kittler

As recent advances in Neural Radiance Fields (NeRF) have enabled high-fidelity 3D face reconstruction and novel view synthesis, its manipulation also became an essential task in 3D vision. However, existing manipulation methods require…

Computer Vision and Pattern Recognition · Computer Science 2023-08-21 Sungwon Hwang , Junha Hyung , Daejin Kim , Min-Jung Kim , Jaegul Choo

While recent work has shown progress on extracting clothed 3D human avatars from a single image, video, or a set of 3D scans, several limitations remain. Most methods use a holistic representation to jointly model the body and clothing,…

Computer Vision and Pattern Recognition · Computer Science 2022-10-06 Yao Feng , Jinlong Yang , Marc Pollefeys , Michael J. Black , Timo Bolkart

Photo-real digital human avatars are of enormous importance in graphics, as they enable immersive communication over the globe, improve gaming and entertainment experiences, and can be particularly beneficial for AR and VR settings.…

Computer Vision and Pattern Recognition · Computer Science 2023-07-20 Marc Habermann , Lingjie Liu , Weipeng Xu , Gerard Pons-Moll , Michael Zollhoefer , Christian Theobalt

Monocular SLAM in deformable scenes will open the way to multiple medical applications like computer-assisted navigation in endoscopy, automatic drug delivery or autonomous robotic surgery. In this paper we propose a novel method to…

Computer Vision and Pattern Recognition · Computer Science 2022-04-19 Juan J. Gomez Rodriguez , J. M. M Montiel , Juan D. Tardos

Building digital twins of articulated objects from monocular video presents an essential challenge in computer vision, which requires simultaneous reconstruction of object geometry, part segmentation, and articulation parameters from…

Computer Vision and Pattern Recognition · Computer Science 2025-09-23 Yu Liu , Baoxiong Jia , Ruijie Lu , Chuyue Gan , Huayu Chen , Junfeng Ni , Song-Chun Zhu , Siyuan Huang

Markerless motion capture and understanding of professional non-daily human movements is an important yet unsolved task, which suffers from complex motion patterns and severe self-occlusion, especially for the monocular setting. In this…

Computer Vision and Pattern Recognition · Computer Science 2021-07-19 Xin Chen , Anqi Pang , Wei Yang , Yuexin Ma , Lan Xu , Jingyi Yu

This paper presents a method to learn hand-object interaction prior for reconstructing a 3D hand-object scene from a single RGB image. The inference as well as training-data generation for 3D hand-object scene reconstruction is challenging…

Computer Vision and Pattern Recognition · Computer Science 2024-09-12 Hongsuk Choi , Nikhil Chavan-Dafle , Jiacheng Yuan , Volkan Isler , Hyunsoo Park

This work focuses on the 3D reconstruction of non-rigid objects based on monocular RGB video sequences. Concretely, we aim at building high-fidelity models for generic object categories and casually captured scenes. To this end, we do not…

Computer Vision and Pattern Recognition · Computer Science 2023-08-22 Yikai Wang , Yinpeng Dong , Fuchun Sun , Xiao Yang

In this paper, we present an approach for tracking people in monocular videos, by predicting their future 3D representations. To achieve this, we first lift people to 3D from a single frame in a robust way. This lifting includes information…

Computer Vision and Pattern Recognition · Computer Science 2021-12-09 Jathushan Rajasegaran , Georgios Pavlakos , Angjoo Kanazawa , Jitendra Malik

We build the first system to address the problem of reconstructing in-scene object manipulation from a monocular RGB video. It is challenging due to ill-posed scene reconstruction, ambiguous hand-object depth, and the need for physically…

Computer Vision and Pattern Recognition · Computer Science 2025-12-23 Dixuan Lin , Tianyou Wang , Zhuoyang Pan , Yufu Wang , Lingjie Liu , Kostas Daniilidis

Physical contact provides additional constraints for hand-object state reconstruction as well as a basis for further understanding of interaction affordances. Estimating these severely occluded regions from monocular images presents a…

Computer Vision and Pattern Recognition · Computer Science 2022-05-03 Zimeng Zhao , Binghui Zuo , Wei Xie , Yangang Wang

We propose an approach for 3D reconstruction and segmentation of a single object placed on a flat surface from an input video. Our approach is to perform dense depth map estimation for multiple views using a proposed objective function that…

Computer Vision and Pattern Recognition · Computer Science 2016-07-29 Tanmay Gupta , Daeyun Shin , Naren Sivagnanadasan , Derek Hoiem

Marker-less 3D human motion capture from a single colour camera has seen significant progress. However, it is a very challenging and severely ill-posed problem. In consequence, even the most accurate state-of-the-art approaches have…

Computer Vision and Pattern Recognition · Computer Science 2020-12-10 Soshi Shimada , Vladislav Golyanik , Weipeng Xu , Christian Theobalt

This paper presents a novel approach 4DRecons that takes a single camera RGB-D sequence of a dynamic subject as input and outputs a complete textured deforming 3D model over time. 4DRecons encodes the output as a 4D neural implicit surface…

Computer Vision and Pattern Recognition · Computer Science 2024-06-17 Xiaoyan Cong , Haitao Yang , Liyan Chen , Kaifeng Zhang , Li Yi , Chandrajit Bajaj , Qixing Huang

This paper describes how to obtain accurate 3D body models and texture of arbitrary people from a single, monocular video in which a person is moving. Based on a parametric body model, we present a robust processing pipeline achieving 3D…

Computer Vision and Pattern Recognition · Computer Science 2018-04-17 Thiemo Alldieck , Marcus Magnor , Weipeng Xu , Christian Theobalt , Gerard Pons-Moll

To address the ill-posed problem caused by partial observations in monocular human volumetric capture, we present AvatarCap, a novel framework that introduces animatable avatars into the capture pipeline for high-fidelity reconstruction in…

Computer Vision and Pattern Recognition · Computer Science 2022-07-13 Zhe Li , Zerong Zheng , Hongwen Zhang , Chaonan Ji , Yebin Liu

Human motion capture either requires multi-camera systems or is unreliable when using single-view input due to depth ambiguities. Meanwhile, mirrors are readily available in urban environments and form an affordable alternative by recording…

Computer Vision and Pattern Recognition · Computer Science 2024-05-17 Daniel Ajisafe , James Tang , Shih-Yang Su , Bastian Wandt , Helge Rhodin

We present a novel method to learn temporally consistent 3D reconstruction of clothed people from a monocular video. Recent methods for 3D human reconstruction from monocular video using volumetric, implicit or parametric human shape…

Computer Vision and Pattern Recognition · Computer Science 2021-04-20 Akin Caliskan , Armin Mustafa , Adrian Hilton

High-fidelity facial avatar reconstruction from a monocular video is a significant research problem in computer graphics and computer vision. Recently, Neural Radiance Field (NeRF) has shown impressive novel view rendering results and has…

Computer Vision and Pattern Recognition · Computer Science 2023-03-07 Yunpeng Bai , Yanbo Fan , Xuan Wang , Yong Zhang , Jingxiang Sun , Chun Yuan , Ying Shan