中文
相关论文

相关论文: GRIP: Generating Interaction Poses Using Spatial C…

200 篇论文

This paper presents a grounded language-image pre-training (GLIP) model for learning object-level, language-aware, and semantic-rich visual representations. GLIP unifies object detection and phrase grounding for pre-training. The…

Humans naturally perform bimanual skills to handle large and heavy objects. To enhance robots' object manipulation capabilities, generating effective bimanual grasp poses is essential. Nevertheless, bimanual grasp synthesis for dexterous…

机器人学 · 计算机科学 2024-11-26 Yanming Shao , Chenxi Xiao

Hand motion capture has been an active research topic in recent years, following the success of full-body pose tracking. Despite similarities, hand tracking proves to be more challenging, characterized by a higher dimensionality, severe…

计算机视觉与模式识别 · 计算机科学 2017-04-04 Dimitrios Tzionas , Abhilash Srikantha , Pablo Aponte , Juergen Gall

Real-time interactive grasp synthesis for dynamic objects remains challenging as existing methods fail to achieve low-latency inference while maintaining promptability. To bridge this gap, we propose SPGrasp (spatiotemporal prompt-driven…

机器人学 · 计算机科学 2025-09-03 Yunpeng Mei , Hongjie Cao , Yinqiu Xia , Wei Xiao , Zhaohan Feng , Gang Wang , Jie Chen

We introduce the task of Reconstructing Objects along Hand Interaction Timelines (ROHIT). We first define the Hand Interaction Timeline (HIT) from a rigid object's perspective. In a HIT, an object is first static relative to the scene, then…

计算机视觉与模式识别 · 计算机科学 2025-12-09 Zhifan Zhu , Siddhant Bansal , Shashank Tripathi , Dima Damen

We present a novel method for real-time pose and shape reconstruction of two strongly interacting hands. Our approach is the first two-hand tracking solution that combines an extensive list of favorable properties, namely it is marker-less,…

计算机视觉与模式识别 · 计算机科学 2021-06-16 Franziska Mueller , Micah Davis , Florian Bernard , Oleksandr Sotnychenko , Mickeal Verschoor , Miguel A. Otaduy , Dan Casas , Christian Theobalt

Modeling hand-object manipulations is essential for understanding how humans interact with their environment. While of practical importance, estimating the pose of hands and objects during interactions is challenging due to the large mutual…

计算机视觉与模式识别 · 计算机科学 2020-04-29 Yana Hasson , Bugra Tekin , Federica Bogo , Ivan Laptev , Marc Pollefeys , Cordelia Schmid

Recent advances in dexterous grasping synthesis have demonstrated significant progress in producing reasonable and plausible grasps for many task purposes. But it remains challenging to generalize to unseen object categories and diverse…

计算机视觉与模式识别 · 计算机科学 2025-07-22 Juntao Jian , Xiuping Liu , Zixuan Chen , Manyi Li , Jian Liu , Ruizhen Hu

This paper presents a novel method for generating diverse 3D human poses in scenes with semantic control. Existing methods heavily rely on the human-scene interaction dataset, resulting in a limited diversity of the generated human poses.…

计算机视觉与模式识别 · 计算机科学 2024-06-11 Bowen Dang , Xi Zhao

We present a method for teaching machines to understand and model the underlying spatial common sense of diverse human-object interactions in 3D in a self-supervised way. This is a challenging task, as there exist specific manifolds of the…

计算机视觉与模式识别 · 计算机科学 2023-09-06 Sookwan Han , Hanbyul Joo

Grasp synthesis is one of the challenging tasks for any robot object manipulation task. In this paper, we present a new deep learning-based grasp synthesis approach for 3D objects. In particular, we propose an end-to-end 3D Convolutional…

机器人学 · 计算机科学 2020-09-15 Yikun Li , Lambert Schomaker , S. Hamidreza Kasaei

Recent works have shown that Large Language Models (LLMs) can facilitate the grounding of instructions for robotic task planning. Despite this progress, most existing works have primarily focused on utilizing raw images to aid LLMs in…

机器人学 · 计算机科学 2024-03-12 Zhe Ni , Xiaoxin Deng , Cong Tai , Xinyue Zhu , Qinghongbing Xie , Weihang Huang , Xiang Wu , Long Zeng

Hand motion capture is a popular research field, recently gaining more attention due to the ubiquity of RGB-D sensors. However, even most recent approaches focus on the case of a single isolated hand. In this work, we focus on hands that…

计算机视觉与模式识别 · 计算机科学 2016-03-29 Dimitrios Tzionas , Luca Ballan , Abhilash Srikantha , Pablo Aponte , Marc Pollefeys , Juergen Gall

Hands are central to interacting with our surroundings and conveying gestures, making their inclusion essential for full-body motion synthesis. Despite this, existing human motion synthesis methods fall short: some ignore hand motions…

计算机视觉与模式识别 · 计算机科学 2026-01-08 Enes Duran , Nikos Athanasiou , Muhammed Kocabas , Michael J. Black , Omid Taheri

Grasping and manipulating a wide variety of objects is a fundamental skill that would determine the success and wide spread adaptation of robots in homes. Several end-effector designs for robust manipulation have been proposed but they…

Grasping deformable objects is not well researched due to the complexity in modelling and simulating the dynamic behavior of such objects. However, with the rapid development of physics-based simulators that support soft bodies, the…

机器人学 · 计算机科学 2021-07-20 Tran Nguyen Le , Jens Lundell , Fares J. Abu-Dakka , Ville Kyrki

Can we make virtual characters in a scene interact with their surrounding objects through simple instructions? Is it possible to synthesize such motion plausibly with a diverse set of objects and instructions? Inspired by these questions,…

计算机视觉与模式识别 · 计算机科学 2023-02-28 Anindita Ghosh , Rishabh Dabral , Vladislav Golyanik , Christian Theobalt , Philipp Slusallek

This work proposes a novel generative design tool for passive grippers -- robot end effectors that have no additional actuation and instead leverage the existing degrees of freedom in a robotic arm to perform grasping tasks. Passive…

图形学 · 计算机科学 2023-06-07 Milin Kodnongbua , Ian Good Yu Lou , Jeffrey Lipton , Adriana Schulz

We present GRIP, a graph neural network accelerator architecture designed for low-latency inference. AcceleratingGNNs is challenging because they combine two distinct types of computation: arithmetic-intensive vertex-centric operations and…

硬件体系结构 · 计算机科学 2020-07-31 Kevin Kiningham , Christopher Re , Philip Levis

Humans naturally build mental models of object interactions and dynamics, allowing them to imagine how their surroundings will change if they take a certain action. While generative models today have shown impressive results on…

计算机视觉与模式识别 · 计算机科学 2024-08-15 Sruthi Sudhakar , Ruoshi Liu , Basile Van Hoorick , Carl Vondrick , Richard Zemel