English
Related papers

Related papers: Decaf: Monocular Deformation Capture for Face and …

200 papers

Human-object interactions with articulated objects are common in everyday life. Despite much progress in single-view 3D reconstruction, it is still challenging to infer an articulated 3D object model from an RGB video showing a person…

Computer Vision and Pattern Recognition · Computer Science 2022-09-14 Sanjay Haresh , Xiaohao Sun , Hanxiao Jiang , Angel X. Chang , Manolis Savva

Since humans interact with diverse objects every day, the holistic 3D capture of these interactions is important to understand and model human behaviour. However, most existing methods for hand-object reconstruction from RGB either assume…

Computer Vision and Pattern Recognition · Computer Science 2023-12-01 Zicong Fan , Maria Parelli , Maria Eleni Kadoglou , Muhammed Kocabas , Xu Chen , Michael J. Black , Otmar Hilliges

Acquiring contact patterns between hands and nonrigid objects is a common concern in the vision and robotics community. However, existing learning-based methods focus more on contact with rigid ones from monocular images. When adopting them…

Computer Vision and Pattern Recognition · Computer Science 2023-08-31 Wei Xie , Zimeng Zhao , Shiying Li , Binghui Zuo , Yangang Wang

We present a near real-time method for 6-DoF tracking of an unknown object from a monocular RGBD video sequence, while simultaneously performing neural 3D reconstruction of the object. Our method works for arbitrary rigid objects, even when…

Computer Vision and Pattern Recognition · Computer Science 2023-03-27 Bowen Wen , Jonathan Tremblay , Valts Blukis , Stephen Tyree , Thomas Muller , Alex Evans , Dieter Fox , Jan Kautz , Stan Birchfield

We introduce a novel, data-driven approach for reconstructing temporally coherent 3D motion from unstructured and potentially partial observations of non-rigidly deforming shapes. Our goal is to achieve high-fidelity motion reconstructions…

Computer Vision and Pattern Recognition · Computer Science 2025-09-09 Aymen Merrouche , Stefanie Wuhrer , Edmond Boyer

In recent years, Neural Radiance Fields (NeRF) have achieved remarkable progress in dynamic human reconstruction and rendering. Part-based rendering paradigms, guided by human segmentation, allow for flexible parameter allocation based on…

Computer Vision and Pattern Recognition · Computer Science 2025-08-13 Yao Lu , Jiawei Li , Ming Jiang

The reconstruction of three-dimensional dynamic scenes is a well-established yet challenging task within the domain of computer vision. In this paper, we propose a novel approach that combines the domains of 3D geometry reconstruction and…

Computer Vision and Pattern Recognition · Computer Science 2025-09-11 David Stotko , Reinhard Klein

Modeling hand-object interactions is a fundamentally challenging task in 3D computer vision. Despite remarkable progress that has been achieved in this field, existing methods still fail to synthesize the hand-object interaction…

Computer Vision and Pattern Recognition · Computer Science 2024-02-12 Zhongqun Zhang , Jifei Song , Eduardo Pérez-Pellitero , Yiren Zhou , Hyung Jin Chang , Aleš Leonardis

We propose FaceVR, a novel image-based method that enables video teleconferencing in VR based on self-reenactment. State-of-the-art face tracking methods in the VR context are focused on the animation of rigged 3d avatars. While they…

Computer Vision and Pattern Recognition · Computer Science 2018-03-23 Justus Thies , Michael Zollhöfer , Marc Stamminger , Christian Theobalt , Matthias Nießner

We propose a method to build in real-time animated 3D head models using a consumer-grade RGB-D camera. Our proposed method is the first one to provide simultaneously comprehensive facial motion tracking and a detailed 3D model of the user's…

Computer Vision and Pattern Recognition · Computer Science 2020-04-23 Diego Thomas

Reconstructing interacting hands from monocular RGB data is a challenging task, as it involves many interfering factors, e.g. self- and mutual occlusion and similar textures. Previous works only leverage information from a single RGB image…

Computer Vision and Pattern Recognition · Computer Science 2024-01-08 Weichao Zhao , Hezhen Hu , Wengang Zhou , Li li , Houqiang Li

Neural rendering has demonstrated remarkable success in dynamic scene reconstruction. Thanks to the expressiveness of neural representations, prior works can accurately capture the motion and achieve high-fidelity reconstruction of the…

Computer Vision and Pattern Recognition · Computer Science 2024-04-05 Hengyi Wang , Jingwen Wang , Lourdes Agapito

Perceiving 3D objects from monocular inputs is crucial for robotic systems, given its economy compared to multi-sensor settings. It is notably difficult as a single image can not provide any clues for predicting absolute depth values.…

Computer Vision and Pattern Recognition · Computer Science 2023-03-02 Tai Wang , Jiangmiao Pang , Dahua Lin

Our work aims to reconstruct a 3D object that is held and rotated by a hand in front of a static RGB camera. Previous methods that use implicit neural representations to recover the geometry of a generic hand-held object from multi-view…

Computer Vision and Pattern Recognition · Computer Science 2023-12-29 Shijian Jiang , Qi Ye , Rengan Xie , Yuchi Huo , Xiang Li , Yang Zhou , Jiming Chen

Reconstructing hand-held objects from monocular RGB images is an appealing yet challenging task. In this task, contacts between hands and objects provide important cues for recovering the 3D geometry of the hand-held objects. Though recent…

Computer Vision and Pattern Recognition · Computer Science 2024-01-17 Junxing Hu , Hongwen Zhang , Zerui Chen , Mengcheng Li , Yunlong Wang , Yebin Liu , Zhenan Sun

The ubiquity of monocular videos capturing daily hand-object interactions presents a valuable resource for embodied intelligence. While 3D hand reconstruction from in-the-wild videos has seen significant progress, reconstructing the…

Computer Vision and Pattern Recognition · Computer Science 2026-02-09 Yuantao Chen , Jiahao Chang , Chongjie Ye , Chaoran Zhang , Zhaojie Fang , Chenghong Li , Xiaoguang Han

Objects manipulated by the hand (i.e., manipulanda) are particularly challenging to reconstruct from Internet videos. Not only does the hand occlude much of the object, but also the object is often only visible in a small number of image…

Computer Vision and Pattern Recognition · Computer Science 2026-01-01 Jane Wu , Georgios Pavlakos , Georgia Gkioxari , Jitendra Malik

We present Grasp in Gaussians (GraG), a fast and robust method for reconstructing dynamic 3D hand-object interactions from a single monocular video. Unlike recent approaches that optimize heavy neural representations, our method focuses on…

Computer Vision and Pattern Recognition · Computer Science 2026-04-15 Ayce Idil Aytekin , Xu Chen , Zhengyang Shen , Thabo Beeler , Helge Rhodin , Rishabh Dabral , Christian Theobalt

We present a method for reconstructing accurate and consistent 3D hands from a monocular video. We observe that detected 2D hand keypoints and the image texture provide important cues about the geometry and texture of the 3D hand, which can…

Computer Vision and Pattern Recognition · Computer Science 2023-03-21 Zhigang Tu , Zhisheng Huang , Yujin Chen , Di Kang , Linchao Bao , Bisheng Yang , Junsong Yuan

Semantic aware reconstruction is more advantageous than geometric-only reconstruction for future robotic and AR/VR applications because it represents not only where things are, but also what things are. Object-centric mapping is a task to…

Computer Vision and Pattern Recognition · Computer Science 2021-02-16 Kejie Li , Hamid Rezatofighi , Ian Reid