English
Related papers

Related papers: HandTailor: Towards High-Precision Monocular 3D Ha…

200 papers

We propose a novel 3D neural network architecture for 3D hand pose estimation from a single depth image. Different from previous works that mostly run on 2D depth image domain and require intermediate or post process to bring in the…

Computer Vision and Pattern Recognition · Computer Science 2017-04-10 Xiaoming Deng , Shuo Yang , Yinda Zhang , Ping Tan , Liang Chang , Hongan Wang

Virtual 3D try-on can provide an intuitive and realistic view for online shopping and has a huge potential commercial value. However, existing 3D virtual try-on methods mainly rely on annotated 3D human shapes and garment templates, which…

Computer Vision and Pattern Recognition · Computer Science 2021-08-12 Fuwei Zhao , Zhenyu Xie , Michael Kampffmeyer , Haoye Dong , Songfang Han , Tianxiang Zheng , Tao Zhang , Xiaodan Liang

Contemporary monocular 6D pose estimation methods can only cope with a handful of object instances. This naturally hampers possible applications as, for instance, robots seamlessly integrated in everyday processes necessarily require the…

Computer Vision and Pattern Recognition · Computer Science 2020-09-14 Fabian Manhardt , Gu Wang , Benjamin Busam , Manuel Nickel , Sven Meier , Luca Minciullo , Xiangyang Ji , Nassir Navab

Monocular 3D reconstruction of articulated object categories is challenging due to the lack of training data and the inherent ill-posedness of the problem. In this work we use video self-supervision, forcing the consistency of consecutive…

Computer Vision and Pattern Recognition · Computer Science 2021-04-28 Filippos Kokkinos , Iasonas Kokkinos

Reconstructing 3D hand meshes from monocular RGB images has attracted increasing amount of attention due to its enormous potential applications in the field of AR/VR. Most state-of-the-art methods attempt to tackle this task in an anonymous…

Computer Vision and Pattern Recognition · Computer Science 2022-09-23 Deying Kong , Linguang Zhang , Liangjian Chen , Haoyu Ma , Xiangyi Yan , Shanlin Sun , Xingwei Liu , Kun Han , Xiaohui Xie

Predicting camera-space hand meshes from single RGB images is crucial for enabling realistic hand interactions in 3D virtual and augmented worlds. Previous work typically divided the task into two stages: given a cropped image of the hand,…

Computer Vision and Pattern Recognition · Computer Science 2024-07-23 Eugene Valassakis , Guillermo Garcia-Hernando

We describe Human Mesh Recovery (HMR), an end-to-end framework for reconstructing a full 3D mesh of a human body from a single RGB image. In contrast to most current methods that compute 2D or 3D joint locations, we produce a richer and…

Computer Vision and Pattern Recognition · Computer Science 2018-06-26 Angjoo Kanazawa , Michael J. Black , David W. Jacobs , Jitendra Malik

Creating detailed 3D human avatars with fitted garments traditionally requires specialized expertise and labor-intensive workflows. While recent advances in generative AI have enabled text-to-3D human and clothing synthesis, existing…

Computer Vision and Pattern Recognition · Computer Science 2026-04-01 Zhiyao Sun , Yu-Hui Wen , Ho-Jui Fang , Sheng Ye , Matthieu Lin , Tian Lv , Yong-Jin Liu

Physical contact provides additional constraints for hand-object state reconstruction as well as a basis for further understanding of interaction affordances. Estimating these severely occluded regions from monocular images presents a…

Computer Vision and Pattern Recognition · Computer Science 2022-05-03 Zimeng Zhao , Binghui Zuo , Wei Xie , Yangang Wang

We introduce a novel method for human shape and pose recovery that can fully leverage multiple static views. We target fixed-multiview people monitoring, including elderly care and safety monitoring, in which calibrated cameras can be…

Computer Vision and Pattern Recognition · Computer Science 2025-03-26 Yuto Matsubara , Ko Nishino

We propose an online 3D Gaussian-based dense mapping framework for photorealistic details reconstruction from a monocular image stream. Our approach addresses two key challenges in monocular online reconstruction: distributing Gaussians…

Graphics · Computer Science 2025-05-15 Songyin Wu , Zhaoyang Lv , Yufeng Zhu , Duncan Frost , Zhengqin Li , Ling-Qi Yan , Carl Ren , Richard Newcombe , Zhao Dong

Manual assembly workers face increasing complexity in their work. Human-centered assistance systems could help, but object recognition as an enabling technology hinders sophisticated human-centered design of these systems. At the same time,…

Computer Vision and Pattern Recognition · Computer Science 2024-02-15 Christian Jauch , Timo Leitritz , Marco F. Huber

Existing methods for 3D tracking from monocular RGB videos predominantly consider articulated and rigid objects. Modelling dense non-rigid object deformations in this setting remained largely unaddressed so far, although such effects can…

Computer Vision and Pattern Recognition · Computer Science 2023-10-16 Soshi Shimada , Vladislav Golyanik , Patrick Pérez , Christian Theobalt

We reconstruct 3D deformable object through time, in the context of a live pottery making process where the crafter molds the object. Because the object suffers from heavy hand interaction, and is being deformed, classical techniques cannot…

Computer Vision and Pattern Recognition · Computer Science 2019-08-06 Raoul de Charette , Sotiris Manitsaris

In this paper, we propose to estimate 3D hand pose by recovering the 3D coordinates of joints in a group-wise manner, where less-related joints are automatically categorized into different groups and exhibit different features. This is…

Computer Vision and Pattern Recognition · Computer Science 2020-12-18 Moran Li , Yuan Gao , Nong Sang

We present HaPTIC, an approach that infers coherent 4D hand trajectories from monocular videos. Current video-based hand pose reconstruction methods primarily focus on improving frame-wise 3D pose using adjacent frames rather than studying…

Computer Vision and Pattern Recognition · Computer Science 2025-01-15 Yufei Ye , Yao Feng , Omid Taheri , Haiwen Feng , Shubham Tulsiani , Michael J. Black

Temporal 3D human pose estimation from monocular videos is a challenging task in human-centered computer vision due to the depth ambiguity of 2D-to-3D lifting. To improve accuracy and address occlusion issues, inertial sensor has been…

Computer Vision and Pattern Recognition · Computer Science 2024-04-30 Yiming Bao , Xu Zhao , Dahong Qian

Tracking and reconstructing the 3D pose and geometry of two hands in interaction is a challenging problem that has a high relevance for several human-computer interaction applications, including AR/VR, robotics, or sign language…

Computer Vision and Pattern Recognition · Computer Science 2021-06-23 Jiayi Wang , Franziska Mueller , Florian Bernard , Suzanne Sorli , Oleksandr Sotnychenko , Neng Qian , Miguel A. Otaduy , Dan Casas , Christian Theobalt

We build the first system to address the problem of reconstructing in-scene object manipulation from a monocular RGB video. It is challenging due to ill-posed scene reconstruction, ambiguous hand-object depth, and the need for physically…

Computer Vision and Pattern Recognition · Computer Science 2025-12-23 Dixuan Lin , Tianyou Wang , Zhuoyang Pan , Yufu Wang , Lingjie Liu , Kostas Daniilidis

Photorealistic human novel view synthesis from a single image is crucial for democratizing immersive 3D telepresence, eliminating the need for complex multi-camera setups. However, current rendering-centric methods prioritize visual…

Computer Vision and Pattern Recognition · Computer Science 2026-03-17 Fangyu Lin , Yingdong Hu , Lunjie Zhu , Zhening Liu , Yushi Huang , Zehong Lin , Jun Zhang