中文
相关论文

相关论文: THOR-Net: End-to-end Graformer-based Realistic Two…

200 篇论文

In this research, we address the challenge faced by existing deep learning-based human mesh reconstruction methods in balancing accuracy and computational efficiency. These methods typically prioritize accuracy, resulting in large network…

计算机视觉与模式识别 · 计算机科学 2023-02-01 Ayman Ali , Ekkasit Pinyoanuntapong , Pu Wang , Mohsen Dorodchi

Accurately recovering the dense 3D mesh of both hands from monocular images poses considerable challenges due to occlusions and projection ambiguity. Most of the existing methods extract features from color images to estimate the…

计算机视觉与模式识别 · 计算机科学 2024-04-11 Jinwei Ren , Jianke Zhu

Monocular 3D clothed human reconstruction aims to create a complete 3D avatar from a single image. To tackle the human geometry lacking in one RGB image, current methods typically resort to a preceding model for an explicit geometric…

计算机视觉与模式识别 · 计算机科学 2025-06-17 Nanjie Yao , Gangjian Zhang , Wenhao Shen , Jian Shu , Hao Wang

Estimating the pose and shape of hands and objects under interaction finds numerous applications including augmented and virtual reality. Existing approaches for hand and object reconstruction require explicitly defined physical constraints…

计算机视觉与模式识别 · 计算机科学 2022-04-28 Tze Ho Elden Tse , Kwang In Kim , Ales Leonardis , Hyung Jin Chang

Significant advancements have been achieved in the realm of understanding poses and interactions of two hands manipulating an object. The emergence of augmented reality (AR) and virtual reality (VR) technologies has heightened the demand…

In this work, we provide a solution for posturing the anthropomorphic Robonaut-2 hand and arm for grasping based on visual information. A mapping from visual features extracted from a convolutional neural network (CNN) to grasp points is…

计算机视觉与模式识别 · 计算机科学 2017-07-27 Li Yang Ku , Erik Learned-Miller , Rod Grupen

Graph convolutional networks (GCNs) have been widely used and achieved remarkable results in skeleton-based action recognition. We think the key to skeleton-based action recognition is a skeleton hanging in frames, so we focus on how the…

计算机视觉与模式识别 · 计算机科学 2023-12-07 Nguyen Huu Bao Long

Most of the existing deep learning-based methods for 3D hand and human pose estimation from a single depth map are based on a common framework that takes a 2D depth map and directly regresses the 3D coordinates of keypoints, such as hand or…

计算机视觉与模式识别 · 计算机科学 2018-08-17 Gyeongsik Moon , Ju Yong Chang , Kyoung Mu Lee

3D meshes are fundamental data representations for capturing complex geometric shapes in computer vision and graphics applications. While Convolutional Neural Networks (CNNs) have excelled in structured data like images, extending them to…

图形学 · 计算机科学 2025-07-09 Saqib Nazir , Olivier Lézoray , Sébastien Bougleux

Reconstructing interacting hands from a single RGB image is a very challenging task. On the one hand, severe mutual occlusion and similar local appearance between two hands confuse the extraction of visual features, resulting in the…

计算机视觉与模式识别 · 计算机科学 2023-08-22 Pengfei Ren , Chao Wen , Xiaozheng Zheng , Zhou Xue , Haifeng Sun , Qi Qi , Jingyu Wang , Jianxin Liao

Objects manipulated by the hand (i.e., manipulanda) are particularly challenging to reconstruct from Internet videos. Not only does the hand occlude much of the object, but also the object is often only visible in a small number of image…

计算机视觉与模式识别 · 计算机科学 2026-01-01 Jane Wu , Georgios Pavlakos , Georgia Gkioxari , Jitendra Malik

This paper presents a method to learn hand-object interaction prior for reconstructing a 3D hand-object scene from a single RGB image. The inference as well as training-data generation for 3D hand-object scene reconstruction is challenging…

计算机视觉与模式识别 · 计算机科学 2024-09-12 Hongsuk Choi , Nikhil Chavan-Dafle , Jiacheng Yuan , Volkan Isler , Hyunsoo Park

Graph super-resolution, the task of inferring high-resolution (HR) graphs from low-resolution (LR) counterparts, is an underexplored yet crucial research direction that circumvents the need for costly data acquisition. This makes it…

机器学习 · 计算机科学 2025-11-13 Pragya Singh , Islem Rekik

3D hand shape and pose estimation from a single depth map is a new and challenging computer vision problem with many applications. Existing methods addressing it directly regress hand meshes via 2D convolutional neural networks, which leads…

计算机视觉与模式识别 · 计算机科学 2021-12-07 Jameel Malik , Soshi Shimada , Ahmed Elhayek , Sk Aziz Ali , Christian Theobalt , Vladislav Golyanik , Didier Stricker

Reconstructing a high-precision and high-fidelity 3D human hand from a color image plays a central role in replicating a realistic virtual hand in human-computer interaction and virtual reality applications. The results of current methods…

计算机视觉与模式识别 · 计算机科学 2021-07-30 Ping Chen , Yujin Chen , Dong Yang , Fangyin Wu , Qin Li , Qingpei Xia , Yong Tan

Nowadays, Transformers and Graph Convolutional Networks (GCNs) are the prevailing techniques for 3D human pose estimation. However, Transformer-based methods either ignore the spatial neighborhood relationships between the joints when used…

计算机视觉与模式识别 · 计算机科学 2025-05-05 Kamel Aouaidjia , Aofan Li , Wenhao Zhang , Chongsheng Zhang

A variety of modeling techniques have been developed in the past decade to reduce the computational expense and improve the accuracy of modeling. In this study, a new framework of modeling is suggested. Compared with other popular methods,…

机器学习 · 计算机科学 2018-09-06 Yu Li , Hu Wang , Kangjia Mo , Tao Zeng

We propose a novel attention-based 2D-to-3D pose estimation network for graph-structured data, named KOG-Transformer, and a 3D pose-to-shape estimation network for hand data, named GASE-Net. Previous 3D pose estimation methods have focused…

计算机视觉与模式识别 · 计算机科学 2022-09-27 Weixi Zhao , Weiqiang Wang

Wearable cameras are increasingly used as an observational and interventional tool for human behaviors by providing detailed visual data of hand-related activities. This data can be leveraged to facilitate memory recall for logging of…

计算机视觉与模式识别 · 计算机科学 2025-07-10 Soroush Shahi , Farzad Shahabi , Rama Nabulsi , Glenn Fernandes , Aggelos Katsaggelos , Nabil Alshurafa

Tracking and reconstructing the 3D pose and geometry of two hands in interaction is a challenging problem that has a high relevance for several human-computer interaction applications, including AR/VR, robotics, or sign language…

计算机视觉与模式识别 · 计算机科学 2021-06-23 Jiayi Wang , Franziska Mueller , Florian Bernard , Suzanne Sorli , Oleksandr Sotnychenko , Neng Qian , Miguel A. Otaduy , Dan Casas , Christian Theobalt