中文
相关论文

相关论文: Predicting 4D Hand Trajectory from Monocular Video…

200 篇论文

This paper presents an approach that reconstructs a hand-held object from a monocular video. In contrast to many recent methods that directly predict object geometry by a trained network, the proposed approach does not require any learned…

计算机视觉与模式识别 · 计算机科学 2022-12-01 Di Huang , Xiaopeng Ji , Xingyi He , Jiaming Sun , Tong He , Qing Shuai , Wanli Ouyang , Xiaowei Zhou

We present a novel method for monocular hand shape and pose estimation at unprecedented runtime performance of 100fps and at state-of-the-art accuracy. This is enabled by a new learning based architecture designed such that it can make use…

计算机视觉与模式识别 · 计算机科学 2022-03-14 Yuxiao Zhou , Marc Habermann , Weipeng Xu , Ikhsanul Habibie , Christian Theobalt , Feng Xu

We propose an approach to estimate arm and hand dynamics from monocular video by utilizing the relationship between arm and hand. Although monocular full human motion capture technologies have made great progress in recent years, recovering…

计算机视觉与模式识别 · 计算机科学 2022-03-31 Shuying Liu , Wenbin Wu , Jiaxian Wu , Yue Lin

Unlike traditional robotic hands, underactuated compliant hands are challenging to model due to inherent uncertainties. Consequently, pose estimation of a grasped object is usually performed based on visual perception. However, visual…

机器人学 · 计算机科学 2024-01-18 Osher Azulay , Inbar Ben-David , Avishai Sintov

We present a method for reconstructing accurate and consistent 3D hands from a monocular video. We observe that detected 2D hand keypoints and the image texture provide important cues about the geometry and texture of the 3D hand, which can…

计算机视觉与模式识别 · 计算机科学 2023-03-21 Zhigang Tu , Zhisheng Huang , Yujin Chen , Di Kang , Linchao Bao , Bisheng Yang , Junsong Yuan

Despite the advent in 3D hand pose estimation, current methods predominantly focus on single-image 3D hand reconstruction in the camera frame, overlooking the world-space motion of the hands. Such limitation prohibits their direct use in…

计算机视觉与模式识别 · 计算机科学 2025-01-07 Jinglei Zhang , Jiankang Deng , Chao Ma , Rolandos Alexandros Potamias

We focus on the task of estimating a physically plausible articulated human motion from monocular video. Existing approaches that do not consider physics often produce temporally inconsistent output with motion artifacts, while…

计算机视觉与模式识别 · 计算机科学 2022-05-26 Erik Gärtner , Mykhaylo Andriluka , Hongyi Xu , Cristian Sminchisescu

We propose Dyn-HaMR, to the best of our knowledge, the first approach to reconstruct 4D global hand motion from monocular videos recorded by dynamic cameras in the wild. Reconstructing accurate 3D hand meshes from monocular videos is a…

计算机视觉与模式识别 · 计算机科学 2025-06-03 Zhengdi Yu , Stefanos Zafeiriou , Tolga Birdal

Forecasting how human hands move in egocentric views is critical for applications like augmented reality and human-robot policy transfer. Recently, several hand trajectory prediction (HTP) methods have been developed to generate future…

计算机视觉与模式识别 · 计算机科学 2026-05-12 Junyi Ma , Wentao Bao , Jingyi Xu , Guanzhong Sun , Yu Zheng , Erhang Zhang , Xieyuanli Chen , Hesheng Wang

We introduce the task of Reconstructing Objects along Hand Interaction Timelines (ROHIT). We first define the Hand Interaction Timeline (HIT) from a rigid object's perspective. In a HIT, an object is first static relative to the scene, then…

计算机视觉与模式识别 · 计算机科学 2025-12-09 Zhifan Zhu , Siddhant Bansal , Shashank Tripathi , Dima Damen

Recovering temporally consistent 3D human body pose, shape and motion from a monocular video is a challenging task due to (self-)occlusions, poor lighting conditions, complex articulated body poses, depth ambiguity, and limited availability…

计算机视觉与模式识别 · 计算机科学 2023-11-21 Sushovan Chanda , Amogh Tiwari , Lokender Tiwari , Brojeshwar Bhowmick , Avinash Sharma , Hrishav Barua

We propose to forecast future hand-object interactions given an egocentric video. Instead of predicting action labels or pixels, we directly predict the hand motion trajectory and the future contact points on the next active object (i.e.,…

计算机视觉与模式识别 · 计算机科学 2022-04-05 Shaowei Liu , Subarna Tripathi , Somdeb Majumdar , Xiaolong Wang

3D hand pose estimation from monocular videos is a long-standing and challenging problem, which is now seeing a strong upturn. In this work, we address it for the first time using a single event camera, i.e., an asynchronous vision sensor…

计算机视觉与模式识别 · 计算机科学 2021-10-12 Viktor Rudnev , Vladislav Golyanik , Jiayi Wang , Hans-Peter Seidel , Franziska Mueller , Mohamed Elgharib , Christian Theobalt

Forecasting hand motion and pose from an egocentric perspective is essential for understanding human intention. However, existing methods focus solely on predicting positions without considering articulation, and only when the hands are…

计算机视觉与模式识别 · 计算机科学 2025-04-14 Masashi Hatano , Zhifan Zhu , Hideo Saito , Dima Damen

Accurate depth estimation from monocular videos remains challenging due to ambiguities inherent in single-view geometry, as crucial depth cues like stereopsis are absent. However, humans often perceive relative depth intuitively by…

计算机视觉与模式识别 · 计算机科学 2025-04-22 Seokju Cho , Jiahui Huang , Seungryong Kim , Joon-Young Lee

3D hand tracking from a monocular video is a very challenging problem due to hand interactions, occlusions, left-right hand ambiguity, and fast motion. Most existing methods rely on RGB inputs, which have severe limitations under low-light…

计算机视觉与模式识别 · 计算机科学 2023-12-22 Christen Millerdurai , Diogo Luvizon , Viktor Rudnev , André Jonas , Jiayi Wang , Christian Theobalt , Vladislav Golyanik

Existing deep models predict 2D and 3D kinematic poses from video that are approximately accurate, but contain visible errors that violate physical constraints, such as feet penetrating the ground and bodies leaning at extreme angles. In…

计算机视觉与模式识别 · 计算机科学 2020-07-27 Davis Rempe , Leonidas J. Guibas , Aaron Hertzmann , Bryan Russell , Ruben Villegas , Jimei Yang

We ask whether everyday open-world monocular videos can be turned into reusable 4D interaction primitives: articulated hand motion, object shape with 6D pose over time, and the when/where of contact. Such a capability would enable scalable…

计算机视觉与模式识别 · 计算机科学 2026-05-22 Hao Xu , Yilin Liu , Yinqiao Wang , Chi-Wing Fu , Niloy J. Mitra

Monocular dynamic reconstruction is a challenging and long-standing vision problem due to the highly ill-posed nature of the task. Existing approaches depend on templates, are effective only in quasi-static scenes, or fail to model 3D…

计算机视觉与模式识别 · 计算机科学 2025-10-17 Qianqian Wang , Vickie Ye , Hang Gao , Weijia Zeng , Jake Austin , Zhengqi Li , Angjoo Kanazawa

We present HARP (HAnd Reconstruction and Personalization), a personalized hand avatar creation approach that takes a short monocular RGB video of a human hand as input and reconstructs a faithful hand avatar exhibiting a high-fidelity…

计算机视觉与模式识别 · 计算机科学 2023-07-06 Korrawe Karunratanakul , Sergey Prokudin , Otmar Hilliges , Siyu Tang
‹ 上一页 1 2 3 10 下一页 ›