中文
相关论文

相关论文: RenderIH: A Large-scale Synthetic Dataset for 3D I…

200 篇论文

We present a new method, called MEsh TRansfOrmer (METRO), to reconstruct 3D human pose and mesh vertices from a single image. Our method uses a transformer encoder to jointly model vertex-vertex and vertex-joint interactions, and outputs 3D…

计算机视觉与模式识别 · 计算机科学 2021-06-16 Kevin Lin , Lijuan Wang , Zicheng Liu

In natural conversation and interaction, our hands often overlap or are in contact with each other. Due to the homogeneous appearance of hands, this makes estimating the 3D pose of interacting hands from images difficult. In this paper we…

计算机视觉与模式识别 · 计算机科学 2021-11-30 Zicong Fan , Adrian Spurr , Muhammed Kocabas , Siyu Tang , Michael J. Black , Otmar Hilliges

Reconstructing 3D hand mesh robustly from a single image is very challenging, due to the lack of diversity in existing real-world datasets. While data synthesis helps relieve the issue, the syn-to-real gap still hinders its usage. In this…

计算机视觉与模式识别 · 计算机科学 2024-03-28 Hao Xu , Haipeng Li , Yinqiao Wang , Shuaicheng Liu , Chi-Wing Fu

3D hand pose estimation from single depth is a fundamental problem in computer vision, and has wide applications.However, the existing methods still can not achieve satisfactory hand pose estimation results due to view variation and…

计算机视觉与模式识别 · 计算机科学 2022-03-30 Jian Cheng , Yanguang Wan , Dexin Zuo , Cuixia Ma , Jian Gu , Ping Tan , Hongan Wang , Xiaoming Deng , Yinda Zhang

Estimating the 3D pose of a hand from a 2D image is a well-studied problem and a requirement for several real-life applications such as virtual reality, augmented reality, and hand gesture recognition. Currently, reasonable estimations can…

计算机视觉与模式识别 · 计算机科学 2022-05-10 Danilo Avola , Luigi Cinque , Alessio Fagioli , Gian Luca Foresti , Adriano Fragomeni , Daniele Pannone

We present a novel appearance-based approach for pose estimation of a human hand using the point clouds provided by the low-cost Microsoft Kinect sensor. Both the free-hand case, in which the hand is isolated from the surrounding…

计算机视觉与模式识别 · 计算机科学 2016-04-08 Pasquale Coscia , Francesco A. N. Palmieri , Francesco Castaldo , Alberto Cavallo

Capturing accurate 3D human pose in the wild would provide valuable data for training pose estimation and motion generation methods. While video-based estimation approaches have become increasingly accurate, they often fail in common…

计算机视觉与模式识别 · 计算机科学 2026-03-13 Maria-Paola Forte , Nikos Athanasiou , Giulia Ballardini , Jan Ulrich Bartels , Katherine J. Kuchenbecker , Michael J. Black

Most of the existing deep learning-based methods for 3D hand and human pose estimation from a single depth map are based on a common framework that takes a 2D depth map and directly regresses the 3D coordinates of keypoints, such as hand or…

计算机视觉与模式识别 · 计算机科学 2018-08-17 Gyeongsik Moon , Ju Yong Chang , Kyoung Mu Lee

We present a new approach for synthesizing novel views of people in new poses. Our novel differentiable renderer enables the synthesis of highly realistic images from any viewpoint. Rather than operating over mesh-based structures, our…

计算机视觉与模式识别 · 计算机科学 2022-02-22 Guillaume Rochette , Chris Russell , Richard Bowden

Recent generative models can synthesize high-quality images, but they often fail to generate humans interacting with objects using their hands. This arises mostly from the model's misunderstanding of such interactions and the hardships of…

计算机视觉与模式识别 · 计算机科学 2025-12-01 Patrick Kwon , Chen Chen , Hanbyul Joo

Hand pose estimation from monocular depth images is an important and challenging problem for human-computer interaction. Recently deep convolutional networks (ConvNet) with sophisticated design have been employed to address it, but the…

计算机视觉与模式识别 · 计算机科学 2019-03-04 Hengkai Guo , Guijin Wang , Xinghao Chen , Cairong Zhang , Fei Qiao , Huazhong Yang

Differentiable render is widely used in optimization-based 3D reconstruction which requires gradients from differentiable operations for gradient-based optimization. The existing differentiable renderers obtain the gradients of rendering…

计算机视觉与模式识别 · 计算机科学 2019-06-20 Zaiqiang Wu , Wei Jiang

Forecasting hand motion and pose from an egocentric perspective is essential for understanding human intention. However, existing methods focus solely on predicting positions without considering articulation, and only when the hands are…

计算机视觉与模式识别 · 计算机科学 2025-04-14 Masashi Hatano , Zhifan Zhu , Hideo Saito , Dima Damen

Estimating the 6D pose of objects from images is an important problem in various applications such as robot manipulation and virtual reality. While direct regression of images to object poses has limited accuracy, matching rendered images…

计算机视觉与模式识别 · 计算机科学 2019-10-03 Yi Li , Gu Wang , Xiangyang Ji , Yu Xiang , Dieter Fox

Hand pose estimation from the monocular 2D image is challenging due to the variation in lighting, appearance, and background. While some success has been achieved using deep neural networks, they typically require collecting a large dataset…

计算机视觉与模式识别 · 计算机科学 2019-04-17 Yikang Li , Chris Twigg , Yuting Ye , Lingling Tao , Xiaogang Wang

Following the success of deep convolutional networks, state-of-the-art methods for 3d human pose estimation have focused on deep end-to-end systems that predict 3d joint locations given raw image pixels. Despite their excellent performance,…

计算机视觉与模式识别 · 计算机科学 2017-08-08 Julieta Martinez , Rayat Hossain , Javier Romero , James J. Little

Automation in surgical robotics has the potential to improve patient safety and surgical efficiency, but it is difficult to achieve due to the need for robust perception algorithms. In particular, 6D pose estimation of surgical instruments…

机器人学 · 计算机科学 2025-01-24 Juan Antonio Barragan , Jintan Zhang , Haoying Zhou , Adnan Munawar , Peter Kazanzides

Generating natural hand-object interactions in 3D is challenging as the resulting hand and object motions are expected to be physically plausible and semantically meaningful. Furthermore, generalization to unseen objects is hindered by the…

计算机视觉与模式识别 · 计算机科学 2024-12-24 Sammy Christen , Shreyas Hampali , Fadime Sener , Edoardo Remelli , Tomas Hodan , Eric Sauser , Shugao Ma , Bugra Tekin

Multi-person pose estimation from a 2D image is an essential technique for human behavior understanding. In this paper, we propose a human pose refinement network that estimates a refined pose from a tuple of an input image and input pose.…

计算机视觉与模式识别 · 计算机科学 2019-03-12 Gyeongsik Moon , Ju Yong Chang , Kyoung Mu Lee

We present a technique for dynamically projecting 3D content onto human hands with short perceived motion-to-photon latency. Computing the pose and shape of human hands accurately and quickly is a challenging task due to their articulated…

图形学 · 计算机科学 2024-09-09 Yotam Erel , Or Kozlovsky-Mordenfeld , Daisuke Iwai , Kosuke Sato , Amit H. Bermano