中文
相关论文

相关论文: TOCH: Spatio-Temporal Object-to-Hand Correspondenc…

200 篇论文

Our work aims to obtain 3D reconstruction of hands and manipulated objects from monocular videos. Reconstructing hand-object manipulations holds a great potential for robotics and learning from human demonstrations. The supervised learning…

计算机视觉与模式识别 · 计算机科学 2022-03-15 Yana Hasson , Gül Varol , Ivan Laptev , Cordelia Schmid

Appearance-based generic object recognition is a challenging problem because all possible appearances of objects cannot be registered, especially as new objects are produced every day. Function of objects, however, has a comparatively small…

计算机视觉与模式识别 · 计算机科学 2017-09-13 Tadashi Matsuo , Nobutaka Shimada

Human-object contact (HOT) is designed to accurately identify the areas where humans and objects come into contact. Current methods frequently fail to account for scenarios where objects are frequently blocking the view, resulting in…

计算机视觉与模式识别 · 计算机科学 2024-12-17 Yuxiao Wang , Wenpeng Neng , Zhenao Wei , Yu Lei , Weiying Xue , Nan Zhuang , Yanwu Xu , Xinyu Jiang , Qi Liu

Wearable cameras are increasingly used as an observational and interventional tool for human behaviors by providing detailed visual data of hand-related activities. This data can be leveraged to facilitate memory recall for logging of…

计算机视觉与模式识别 · 计算机科学 2025-07-10 Soroush Shahi , Farzad Shahabi , Rama Nabulsi , Glenn Fernandes , Aggelos Katsaggelos , Nabil Alshurafa

Recent progress in the text-driven 3D stylization of a single object has been considerably promoted by CLIP-based methods. However, the stylization of multi-object 3D scenes is still impeded in that the image-text pairs used for…

计算机视觉与模式识别 · 计算机科学 2023-12-08 Xuying Zhang , Bo-Wen Yin , Yuming Chen , Zheng Lin , Yunheng Li , Qibin Hou , Ming-Ming Cheng

This paper presents a novel manipulation strategy that uses keypoint correspondences extracted from visuo-tactile sensor images to facilitate precise object manipulation. Our approach uses the visuo-tactile feedback to guide the robot's…

机器人学 · 计算机科学 2024-05-24 Jeong-Jung Kim , Doo-Yeol Koh , Chang-Hyun Kim

We study the problem of precisely swapping objects in videos, with a focus on those interacted with by hands, given one user-provided reference object image. Despite the great advancements that diffusion models have made in video editing…

计算机视觉与模式识别 · 计算机科学 2024-11-12 Zihui Xue , Mi Luo , Changan Chen , Kristen Grauman

3D hand tracking from a monocular video is a very challenging problem due to hand interactions, occlusions, left-right hand ambiguity, and fast motion. Most existing methods rely on RGB inputs, which have severe limitations under low-light…

计算机视觉与模式识别 · 计算机科学 2023-12-22 Christen Millerdurai , Diogo Luvizon , Viktor Rudnev , André Jonas , Jiayi Wang , Christian Theobalt , Vladislav Golyanik

The human hand is our primary interface to the physical world, yet egocentric perception rarely knows when, where, or how forcefully it makes contact. Robust wearable tactile sensors are scarce, and no existing in-the-wild datasets align…

This paper presents a comprehensive study of virtual 3D object manipulation along 4DoF on real surfaces in mixed reality (MR), using hand-based and tangible interactions. A custom cylindrical tangible proxy leverages affordances of physical…

人机交互 · 计算机科学 2025-11-18 Carlos Mosquera , Neven Elsayed , Ernst Kruijff , Joseph Newman , Eduardo Veas

In this work, we tackle the challenging task of jointly tracking hand object pose and reconstructing their shapes from depth point cloud sequences in the wild, given the initial poses at frame 0. We for the first time propose a point cloud…

计算机视觉与模式识别 · 计算机科学 2022-09-27 Jiayi Chen , Mi Yan , Jiazhao Zhang , Yinzhen Xu , Xiaolong Li , Yijia Weng , Li Yi , Shuran Song , He Wang

Jointly estimating hand and object shape facilitates the grasping task in human-to-robot handovers. However, relying on hand-crafted prior knowledge about the geometric structure of the object fails when generalising to unseen objects, and…

机器人学 · 计算机科学 2025-05-13 Yik Lung Pang , Alessio Xompero , Changjae Oh , Andrea Cavallaro

This paper focuses on visual motion-based invariants that result in a representation of 3D points in which the stationary environment remains invariant, ensuring shape constancy. This is achieved even as the images undergo constant change…

计算机视觉与模式识别 · 计算机科学 2023-10-17 Juan D. Yepes , Daniel Raviv

3D hand-object pose estimation is the key to the success of many computer vision applications. The main focus of this task is to effectively model the interaction between the hand and an object. To this end, existing works either rely on…

计算机视觉与模式识别 · 计算机科学 2023-01-09 Rong Wang , Wei Mao , Hongdong Li

While 3D hand reconstruction from monocular images has made significant progress, generating accurate and temporally coherent motion estimates from videos remains challenging, particularly during hand-object interactions. In this paper, we…

计算机视觉与模式识别 · 计算机科学 2025-08-05 Yufei Zhang , Zijun Cui , Jeffrey O. Kephart , Qiang Ji

Hand segmentation for hand-object interaction is a necessary preprocessing step in many applications such as augmented reality, medical application, and human-robot interaction. However, typical methods are based on color information which…

计算机视觉与模式识别 · 计算机科学 2018-01-11 Byeongkeun Kang , Kar-Han Tan , Nan Jiang , Hung-Shuo Tai , Daniel Tretter , Truong Q. Nguyen

Accurate 3D reconstruction of the hand and object shape from a hand-object image is important for understanding human-object interaction as well as human daily activities. Different from bare hand pose estimation, hand-object interaction…

计算机视觉与模式识别 · 计算机科学 2021-04-21 Yujin Chen , Zhigang Tu , Di Kang , Ruizhi Chen , Linchao Bao , Zhengyou Zhang , Junsong Yuan

We study the problem of imitating object interactions from Internet videos. This requires understanding the hand-object interactions in 4D, spatially in 3D and over time, which is challenging due to mutual hand-object occlusions. In this…

计算机视觉与模式识别 · 计算机科学 2022-11-24 Austin Patel , Andrew Wang , Ilija Radosavovic , Jitendra Malik

This paper addresses the problem of generating 3D interactive human motion from text. Given a textual description depicting the actions of different body parts in contact with static objects, we synthesize sequences of 3D body poses that…

计算机视觉与模式识别 · 计算机科学 2024-09-17 Sihan Ma , Qiong Cao , Jing Zhang , Dacheng Tao

Manual annotations of temporal bounds for object interactions (i.e. start and end times) are typical training input to recognition, localization and detection algorithms. For three publicly available egocentric datasets, we uncover…

计算机视觉与模式识别 · 计算机科学 2017-07-27 Davide Moltisanti , Michael Wray , Walterio Mayol-Cuevas , Dima Damen