中文
相关论文

相关论文: Rotation Equivariant 3D Hand Mesh Generation from …

200 篇论文

We propose the first approach to the problem of inferring the depth map of a human hand based on a single RGB image. We achieve this with a Convolutional Neural Network (CNN) that employs a stacked hourglass model as its main building…

计算机视觉与模式识别 · 计算机科学 2018-12-07 Vassilis C. Nicodemou , Iason Oikonomidis , Georgios Tzimiropoulos , Antonis Argyros

We present an approach that can reconstruct hands in 3D from monocular input. Our approach for Hand Mesh Recovery, HaMeR, follows a fully transformer-based architecture and can analyze hands with significantly increased accuracy and…

计算机视觉与模式识别 · 计算机科学 2023-12-11 Georgios Pavlakos , Dandan Shan , Ilija Radosavovic , Angjoo Kanazawa , David Fouhey , Jitendra Malik

Understanding three-dimensional (3D) geometries from two-dimensional (2D) images without any labeled information is promising for understanding the real world without incurring annotation cost. We herein propose a novel generative model,…

计算机视觉与模式识别 · 计算机科学 2020-05-26 Atsuhiro Noguchi , Tatsuya Harada

We consider the task of generating realistic 3D shapes, which is useful for a variety of applications such as automatic scene generation and physical simulation. Compared to other 3D representations like voxels and point clouds, meshes are…

图形学 · 计算机科学 2023-04-18 Zhen Liu , Yao Feng , Michael J. Black , Derek Nowrouzezahrai , Liam Paull , Weiyang Liu

Analyzing and understanding hand information from multimedia materials like images or videos is important for many real world applications and remains active in research community. There are various works focusing on recovering hand…

计算机视觉与模式识别 · 计算机科学 2021-07-29 Xiong Zhang , Hongsheng Huang , Jianchao Tan , Hongmin Xu , Cheng Yang , Guozhu Peng , Lei Wang , Ji Liu

Text-to-image generation models have achieved remarkable advancements in recent years, aiming to produce realistic images from textual descriptions. However, these models often struggle with generating anatomically accurate representations…

计算机视觉与模式识别 · 计算机科学 2024-12-24 Haozhuo Zhang , Bin Zhu , Yu Cao , Yanbin Hao

The recent advances in text and image synthesis show a great promise for the future of generative models in creative fields. However, a less explored area is the one of 3D model generation, with a lot of potential applications to game…

计算机视觉与模式识别 · 计算机科学 2023-12-14 Antoine Schnepf , Flavian Vasile , Ugo Tanielian

Estimating 3D from 2D is one of the central tasks in computer vision. In this work, we consider the monocular setting, i.e. single-view input, for 3D human pose estimation (HPE). Here, the task is to predict a 3D point set of human skeletal…

计算机视觉与模式识别 · 计算机科学 2026-01-21 Pavlo Melnyk , Cuong Le , Urs Waldmann , Per-Erik Forssén , Bastian Wandt

From early image processing to modern computational imaging, successful models and algorithms have relied on a fundamental property of natural signals: symmetry. Here symmetry refers to the invariance property of signal sets to…

信号处理 · 电气工程与系统科学 2022-09-07 Dongdong Chen , Mike Davies , Matthias J. Ehrhardt , Carola-Bibiane Schönlieb , Ferdia Sherry , Julián Tachella

Reconstructing a high-precision and high-fidelity 3D human hand from a color image plays a central role in replicating a realistic virtual hand in human-computer interaction and virtual reality applications. The results of current methods…

计算机视觉与模式识别 · 计算机科学 2021-07-30 Ping Chen , Yujin Chen , Dong Yang , Fangyin Wu , Qin Li , Qingpei Xia , Yong Tan

We present a method for reconstructing accurate and consistent 3D hands from a monocular video. We observe that detected 2D hand keypoints and the image texture provide important cues about the geometry and texture of the 3D hand, which can…

计算机视觉与模式识别 · 计算机科学 2023-03-21 Zhigang Tu , Zhisheng Huang , Yujin Chen , Di Kang , Linchao Bao , Bisheng Yang , Junsong Yuan

We present a method for recovering the dense 3D surface of the hand by regressing the vertex coordinates of a mesh model from a single depth map. To this end, we use a two-stage 2D fully convolutional network architecture. In the first…

计算机视觉与模式识别 · 计算机科学 2019-07-26 Chengde Wan , Thomas Probst , Luc Van Gool , Angela Yao

Recent advances in differentiable rendering have sparked an interest in learning generative models of textured 3D meshes from image collections. These models natively disentangle pose and appearance, enable downstream applications in…

计算机视觉与模式识别 · 计算机科学 2021-08-18 Dario Pavllo , Jonas Kohler , Thomas Hofmann , Aurelien Lucchi

Recent advances in deep learning and Transformers have driven major breakthroughs in robotics by employing techniques such as imitation learning, reinforcement learning, and LLM-based multimodal perception and decision-making. However,…

We propose a multimodal, physically grounded approach for metric-scale amodal object reconstruction and pose estimation under severe hand occlusion. Unlike prior occlusion-aware 3D generation methods that rely only on vision, we leverage…

计算机视觉与模式识别 · 计算机科学 2026-04-13 Gabriele Mario Caddeo , Pasquale Marra , Lorenzo Natale

Soft robotic hand shows considerable promise for various grasping applications. However, the sensing and reconstruction of the robot pose will cause limitation during the design and fabrication. In this work, we present a novel 3D pose…

机器人学 · 计算机科学 2023-08-08 Haihang Wang , He Xu , Yihan Meng

Reliable hand mesh reconstruction (HMR) from commonly-used color and depth sensors is challenging especially under scenarios with varied illuminations and fast motions. Event camera is a highly promising alternative for its high dynamic…

计算机视觉与模式识别 · 计算机科学 2024-03-13 Jianping Jiang , Xinyu Zhou , Bingxuan Wang , Xiaoming Deng , Chao Xu , Boxin Shi

Polygonal meshes have become the standard for discretely approximating 3D shapes, thanks to their efficiency and high flexibility in capturing non-uniform shapes. This non-uniformity, however, leads to irregularity in the mesh structure,…

计算机视觉与模式识别 · 计算机科学 2023-07-04 Giuseppe Vecchio , Luca Prezzavento , Carmelo Pino , Francesco Rundo , Simone Palazzo , Concetto Spampinato

Reconstructing high-fidelity 3D hands from egocentric monocular videos remains a challenge due to the limitations in capturing high-resolution geometry, hand-object interactions, and complex objects on hands. Additionally, existing methods…

计算机视觉与模式识别 · 计算机科学 2026-04-13 Haoyu Zhu , Yi Zhang , Lei Yao , Lap-pui Chau , Yi Wang

We demonstrate an object tracking method for 3D images with fixed computational cost and state-of-the-art performance. Previous methods predicted transformation parameters from convolutional layers. We instead propose an architecture that…

计算机视觉与模式识别 · 计算机科学 2021-09-28 Daniel Moyer , Esra Abaci Turk , P Ellen Grant , William M. Wells , Polina Golland