English
Related papers

Related papers: THOR-Net: End-to-end Graformer-based Realistic Two…

200 papers

In this research, we address the challenge faced by existing deep learning-based human mesh reconstruction methods in balancing accuracy and computational efficiency. These methods typically prioritize accuracy, resulting in large network…

Computer Vision and Pattern Recognition · Computer Science 2023-02-01 Ayman Ali , Ekkasit Pinyoanuntapong , Pu Wang , Mohsen Dorodchi

Accurately recovering the dense 3D mesh of both hands from monocular images poses considerable challenges due to occlusions and projection ambiguity. Most of the existing methods extract features from color images to estimate the…

Computer Vision and Pattern Recognition · Computer Science 2024-04-11 Jinwei Ren , Jianke Zhu

Monocular 3D clothed human reconstruction aims to create a complete 3D avatar from a single image. To tackle the human geometry lacking in one RGB image, current methods typically resort to a preceding model for an explicit geometric…

Computer Vision and Pattern Recognition · Computer Science 2025-06-17 Nanjie Yao , Gangjian Zhang , Wenhao Shen , Jian Shu , Hao Wang

Estimating the pose and shape of hands and objects under interaction finds numerous applications including augmented and virtual reality. Existing approaches for hand and object reconstruction require explicitly defined physical constraints…

Computer Vision and Pattern Recognition · Computer Science 2022-04-28 Tze Ho Elden Tse , Kwang In Kim , Ales Leonardis , Hyung Jin Chang

Significant advancements have been achieved in the realm of understanding poses and interactions of two hands manipulating an object. The emergence of augmented reality (AR) and virtual reality (VR) technologies has heightened the demand…

Computer Vision and Pattern Recognition · Computer Science 2025-02-28 Elkhan Ismayilzada , MD Khalequzzaman Chowdhury Sayem , Yihalem Yimolal Tiruneh , Mubarrat Tajoar Chowdhury , Muhammadjon Boboev , Seungryul Baek

In this work, we provide a solution for posturing the anthropomorphic Robonaut-2 hand and arm for grasping based on visual information. A mapping from visual features extracted from a convolutional neural network (CNN) to grasp points is…

Computer Vision and Pattern Recognition · Computer Science 2017-07-27 Li Yang Ku , Erik Learned-Miller , Rod Grupen

Graph convolutional networks (GCNs) have been widely used and achieved remarkable results in skeleton-based action recognition. We think the key to skeleton-based action recognition is a skeleton hanging in frames, so we focus on how the…

Computer Vision and Pattern Recognition · Computer Science 2023-12-07 Nguyen Huu Bao Long

Most of the existing deep learning-based methods for 3D hand and human pose estimation from a single depth map are based on a common framework that takes a 2D depth map and directly regresses the 3D coordinates of keypoints, such as hand or…

Computer Vision and Pattern Recognition · Computer Science 2018-08-17 Gyeongsik Moon , Ju Yong Chang , Kyoung Mu Lee

3D meshes are fundamental data representations for capturing complex geometric shapes in computer vision and graphics applications. While Convolutional Neural Networks (CNNs) have excelled in structured data like images, extending them to…

Graphics · Computer Science 2025-07-09 Saqib Nazir , Olivier Lézoray , Sébastien Bougleux

Reconstructing interacting hands from a single RGB image is a very challenging task. On the one hand, severe mutual occlusion and similar local appearance between two hands confuse the extraction of visual features, resulting in the…

Computer Vision and Pattern Recognition · Computer Science 2023-08-22 Pengfei Ren , Chao Wen , Xiaozheng Zheng , Zhou Xue , Haifeng Sun , Qi Qi , Jingyu Wang , Jianxin Liao

Objects manipulated by the hand (i.e., manipulanda) are particularly challenging to reconstruct from Internet videos. Not only does the hand occlude much of the object, but also the object is often only visible in a small number of image…

Computer Vision and Pattern Recognition · Computer Science 2026-01-01 Jane Wu , Georgios Pavlakos , Georgia Gkioxari , Jitendra Malik

This paper presents a method to learn hand-object interaction prior for reconstructing a 3D hand-object scene from a single RGB image. The inference as well as training-data generation for 3D hand-object scene reconstruction is challenging…

Computer Vision and Pattern Recognition · Computer Science 2024-09-12 Hongsuk Choi , Nikhil Chavan-Dafle , Jiacheng Yuan , Volkan Isler , Hyunsoo Park

Graph super-resolution, the task of inferring high-resolution (HR) graphs from low-resolution (LR) counterparts, is an underexplored yet crucial research direction that circumvents the need for costly data acquisition. This makes it…

Machine Learning · Computer Science 2025-11-13 Pragya Singh , Islem Rekik

3D hand shape and pose estimation from a single depth map is a new and challenging computer vision problem with many applications. Existing methods addressing it directly regress hand meshes via 2D convolutional neural networks, which leads…

Computer Vision and Pattern Recognition · Computer Science 2021-12-07 Jameel Malik , Soshi Shimada , Ahmed Elhayek , Sk Aziz Ali , Christian Theobalt , Vladislav Golyanik , Didier Stricker

Reconstructing a high-precision and high-fidelity 3D human hand from a color image plays a central role in replicating a realistic virtual hand in human-computer interaction and virtual reality applications. The results of current methods…

Computer Vision and Pattern Recognition · Computer Science 2021-07-30 Ping Chen , Yujin Chen , Dong Yang , Fangyin Wu , Qin Li , Qingpei Xia , Yong Tan

Nowadays, Transformers and Graph Convolutional Networks (GCNs) are the prevailing techniques for 3D human pose estimation. However, Transformer-based methods either ignore the spatial neighborhood relationships between the joints when used…

Computer Vision and Pattern Recognition · Computer Science 2025-05-05 Kamel Aouaidjia , Aofan Li , Wenhao Zhang , Chongsheng Zhang

A variety of modeling techniques have been developed in the past decade to reduce the computational expense and improve the accuracy of modeling. In this study, a new framework of modeling is suggested. Compared with other popular methods,…

Machine Learning · Computer Science 2018-09-06 Yu Li , Hu Wang , Kangjia Mo , Tao Zeng

We propose a novel attention-based 2D-to-3D pose estimation network for graph-structured data, named KOG-Transformer, and a 3D pose-to-shape estimation network for hand data, named GASE-Net. Previous 3D pose estimation methods have focused…

Computer Vision and Pattern Recognition · Computer Science 2022-09-27 Weixi Zhao , Weiqiang Wang

Wearable cameras are increasingly used as an observational and interventional tool for human behaviors by providing detailed visual data of hand-related activities. This data can be leveraged to facilitate memory recall for logging of…

Computer Vision and Pattern Recognition · Computer Science 2025-07-10 Soroush Shahi , Farzad Shahabi , Rama Nabulsi , Glenn Fernandes , Aggelos Katsaggelos , Nabil Alshurafa

Tracking and reconstructing the 3D pose and geometry of two hands in interaction is a challenging problem that has a high relevance for several human-computer interaction applications, including AR/VR, robotics, or sign language…

Computer Vision and Pattern Recognition · Computer Science 2021-06-23 Jiayi Wang , Franziska Mueller , Florian Bernard , Suzanne Sorli , Oleksandr Sotnychenko , Neng Qian , Miguel A. Otaduy , Dan Casas , Christian Theobalt