English
Related papers

Related papers: THOR-Net: End-to-end Graformer-based Realistic Two…

200 papers

Existing deep learning-based human mesh reconstruction approaches have a tendency to build larger networks in order to achieve higher accuracy. Computational complexity and model size are often neglected, despite being key characteristics…

Computer Vision and Pattern Recognition · Computer Science 2022-07-19 Ce Zheng , Matias Mendieta , Pu Wang , Aidong Lu , Chen Chen

We present a method for reconstructing accurate and consistent 3D hands from a monocular video. We observe that detected 2D hand keypoints and the image texture provide important cues about the geometry and texture of the 3D hand, which can…

Computer Vision and Pattern Recognition · Computer Science 2023-03-21 Zhigang Tu , Zhisheng Huang , Yujin Chen , Di Kang , Linchao Bao , Bisheng Yang , Junsong Yuan

Being able to grasp objects is a fundamental component of most robotic manipulation systems. In this paper, we present a new approach to simultaneously reconstruct a mesh and a dense grasp quality map of an object from a depth image. At the…

Robotics · Computer Science 2022-12-21 Nikhil Chavan-Dafle , Sergiy Popovych , Shubham Agrawal , Daniel D. Lee , Volkan Isler

3D interacting hand pose estimation from a single RGB image is a challenging task, due to serious self-occlusion and inter-occlusion towards hands, confusing similar appearance patterns between 2 hands, ill-posed joint position mapping from…

Computer Vision and Pattern Recognition · Computer Science 2023-04-10 Changlong Jiang , Yang Xiao , Cunlin Wu , Mingyang Zhang , Jinghong Zheng , Zhiguo Cao , Joey Tianyi Zhou

We introduce a novel 3D hand pose estimator that can accurately recover the shape and pose of people's hands in a room from afar, typically from fixed cameras at room corners, in extremely low-resolution and frequently occluded views. Our…

Computer Vision and Pattern Recognition · Computer Science 2026-05-22 Shu Nakamura , Ryo Kawahara , Genki Kinoshita , Ryosuke Hirai , Yasutomo Kawanishi , Shohei Nobuhara , Ko Nishino

We introduce InverseFaceNet, a deep convolutional inverse rendering framework for faces that jointly estimates facial pose, shape, expression, reflectance and illumination from a single input image. By estimating all parameters from just a…

Computer Vision and Pattern Recognition · Computer Science 2018-05-17 Hyeongwoo Kim , Michael Zollhöfer , Ayush Tewari , Justus Thies , Christian Richardt , Christian Theobalt

Human Mesh Reconstruction (HMR) from monocular video plays an important role in human-robot interaction and collaboration. However, existing video-based human mesh reconstruction methods face a trade-off between accurate reconstruction and…

Computer Vision and Pattern Recognition · Computer Science 2024-12-03 Tao Tang , Hong Liu , Yingxuan You , Ti Wang , Wenhao Li

We study the problem of imitating object interactions from Internet videos. This requires understanding the hand-object interactions in 4D, spatially in 3D and over time, which is challenging due to mutual hand-object occlusions. In this…

Computer Vision and Pattern Recognition · Computer Science 2022-11-24 Austin Patel , Andrew Wang , Ilija Radosavovic , Jitendra Malik

Reconstructing 3D poses from 2D poses lacking depth information is particularly challenging due to the complexity and diversity of human motion. The key is to effectively model the spatial constraints between joints to leverage their…

Computer Vision and Pattern Recognition · Computer Science 2023-08-11 Hongbo Kang , Yong Wang , Mengyuan Liu , Doudou Wu , Peng Liu , Wenming Yang

Humans perceive the 3D world as a set of distinct objects that are characterized by various low-level (geometry, reflectance) and high-level (connectivity, adjacency, symmetry) properties. Recent methods based on convolutional neural…

Computer Vision and Pattern Recognition · Computer Science 2020-04-03 Despoina Paschalidou , Luc van Gool , Andreas Geiger

We introduce a simple and effective network architecture for monocular 3D hand pose estimation consisting of an image encoder followed by a mesh convolutional decoder that is trained through a direct 3D hand mesh reconstruction loss. We…

Computer Vision and Pattern Recognition · Computer Science 2020-04-07 Dominik Kulon , Riza Alp Güler , Iasonas Kokkinos , Michael Bronstein , Stefanos Zafeiriou

David Marr's seminal theory of vision proposes that the human visual system operates through a sequence of three stages, known as the 2D sketch, the 2.5D sketch, and the 3D model. In recent years, Deep Neural Networks (DNN) have been widely…

Computer Vision and Pattern Recognition · Computer Science 2024-11-26 Xiangyu Zhu , Chang Yu , Jiankuo Zhao , Zhaoxiang Zhang , Stan Z. Li , Zhen Lei

Estimating 6D poses and reconstructing 3D shapes of objects in open-world scenes from RGB-depth image pairs is challenging. Many existing methods rely on learning geometric features that correspond to specific templates while disregarding…

Computer Vision and Pattern Recognition · Computer Science 2023-08-07 Haowen Wang , Zhipeng Fan , Zhen Zhao , Zhengping Che , Zhiyuan Xu , Dong Liu , Feifei Feng , Yakun Huang , Xiuquan Qiao , Jian Tang

Despite the recent progress, 3D multi-person pose estimation from monocular videos is still challenging due to the commonly encountered problem of missing information caused by occlusion, partially out-of-frame target persons, and…

Computer Vision and Pattern Recognition · Computer Science 2021-04-08 Yu Cheng , Bo Wang , Bo Yang , Robby T. Tan

We present a novel method for the upright adjustment of 360 images. Our network consists of two modules, which are a convolutional neural network (CNN) and a graph convolutional network (GCN). The input 360 images is processed with the CNN…

Computer Vision and Pattern Recognition · Computer Science 2024-06-04 Raehyuk Jung , Sungmin Cho , Junseok Kwon

Advances in biosignal signal processing and machine learning, in particular Deep Neural Networks (DNNs), have paved the way for the development of innovative Human-Machine Interfaces for decoding the human intent and controlling artificial…

Machine Learning · Computer Science 2021-10-19 Elahe Rahimian , Soheil Zabihi , Amir Asif , Dario Farina , S. Farokh Atashzar , Arash Mohammadi

A lot of work has been done towards reconstructing the 3D facial structure from single images by capitalizing on the power of Deep Convolutional Neural Networks (DCNNs). In the recent works, the texture features either correspond to…

Computer Vision and Pattern Recognition · Computer Science 2022-03-28 Baris Gecer , Stylianos Ploumpis , Irene Kotsia , Stefanos Zafeiriou

Graph Neural Networks (GNNs) have gained momentum in graph representation learning and boosted the state of the art in a variety of areas, such as data mining (\emph{e.g.,} social network analysis and recommender systems), computer vision…

Computer Vision and Pattern Recognition · Computer Science 2024-08-15 Chaoqi Chen , Yushuang Wu , Qiyuan Dai , Hong-Yu Zhou , Mutian Xu , Sibei Yang , Xiaoguang Han , Yizhou Yu

This work provides an architecture that incorporates depth and tactile information to create rich and accurate 3D models useful for robotic manipulation tasks. This is accomplished through the use of a 3D convolutional neural network (CNN).…

Robotics · Computer Science 2023-02-13 David Watkins , Jacob Varley , Peter Allen

In recent years, 3D hand pose estimation methods have garnered significant attention due to their extensive applications in human-computer interaction, virtual reality, and robotics. In contrast, there has been a notable gap in hand…

Computer Vision and Pattern Recognition · Computer Science 2025-03-28 Rolandos Alexandros Potamias , Jinglei Zhang , Jiankang Deng , Stefanos Zafeiriou