English
Related papers

Related papers: Affordance-Guided Diffusion Prior for 3D Hand Reco…

200 papers

Recent successes in image synthesis are powered by large-scale diffusion models. However, most methods are currently limited to either text- or image-conditioned generation for synthesizing an entire image, texture transfer or inserting…

Computer Vision and Pattern Recognition · Computer Science 2023-05-23 Yufei Ye , Xueting Li , Abhinav Gupta , Shalini De Mello , Stan Birchfield , Jiaming Song , Shubham Tulsiani , Sifei Liu

Diffusion Handles is a novel approach to enabling 3D object edits on diffusion images. We accomplish these edits using existing pre-trained diffusion models, and 2D image depth estimation, without any fine-tuning or 3D object retrieval. The…

Computer Vision and Pattern Recognition · Computer Science 2023-12-08 Karran Pandey , Paul Guerrero , Matheus Gadelha , Yannick Hold-Geoffroy , Karan Singh , Niloy Mitra

We propose a novel diffusion-based framework for reconstructing 3D geometry of hand-held objects from monocular RGB images by leveraging hand-object interaction as geometric guidance. Our method conditions a latent diffusion model on an…

Computer Vision and Pattern Recognition · Computer Science 2025-08-26 Ayce Idil Aytekin , Helge Rhodin , Rishabh Dabral , Christian Theobalt

Generating human grasping poses that accurately reflect both object geometry and user-specified interaction semantics is essential for natural hand-object interactions in AR/VR and embodied AI. However, existing semantic grasping approaches…

Robotics · Computer Science 2026-03-31 Xiaofei Wu , Yi Zhang , Yumeng Liu , Yuexin Ma , Yujiao Shi , Xuming He

The malformed hands in the AI-generated images seriously affect the authenticity of the images. To refine malformed hands, existing depth-based approaches use a hand depth estimator to guide the refinement of malformed hands. Due to the…

Computer Vision and Pattern Recognition · Computer Science 2025-06-18 Chen-Bin Feng , Kangdao Liu , Jian Sun , Jiping Jin , Yiguo Jiang , Chi-Man Vong

Controllable affordance Hand-Object Interaction (HOI) generation has become an increasingly important area of research in computer vision. In HOI generation, the hand grasp generation is a crucial step for effectively controlling the…

Computer Vision and Pattern Recognition · Computer Science 2025-01-28 Ishant , Rongliang Wu , Joo Hwee Lim

We tackle the task of reconstructing hand-object interactions from short video clips. Given an input video, our approach casts 3D inference as a per-video optimization and recovers a neural 3D representation of the object shape, as well as…

Computer Vision and Pattern Recognition · Computer Science 2023-09-12 Yufei Ye , Poorvi Hebbar , Abhinav Gupta , Shubham Tulsiani

While 3D hand reconstruction from monocular images has made significant progress, generating accurate and temporally coherent motion estimates from videos remains challenging, particularly during hand-object interactions. In this paper, we…

Computer Vision and Pattern Recognition · Computer Science 2025-08-05 Yufei Zhang , Zijun Cui , Jeffrey O. Kephart , Qiang Ji

Dexterous grasp synthesis must jointly satisfy functional intent and physical feasibility, yet existing pipelines often decouple semantic grounding from refinement, yielding unstable or non-functional contacts under object and pose…

Robotics · Computer Science 2026-03-13 Yifan Han , Yichuan Peng , Pengfei Yi , Junyan Li , Hanqing Wang , Gaojing Zhang , Qi Peng Liu , Wenzhao Lian

Extracting keypoint locations from input hand frames, known as 3D hand pose estimation, is a critical task in various human-computer interaction applications. Essentially, the 3D hand pose estimation can be regarded as a 3D point subset…

Computer Vision and Pattern Recognition · Computer Science 2024-04-05 Wencan Cheng , Hao Tang , Luc Van Gool , Jong Hwan Ko

Understanding how humans interact with the surrounding environment, and specifically reasoning about object interactions and affordances, is a critical challenge in computer vision, robotics, and AI. Current approaches often depend on…

Computer Vision and Pattern Recognition · Computer Science 2026-02-12 Harry Zhang , Luca Carlone

Visual affordance segmentation identifies the surfaces of an object an agent can interact with. Common challenges for the identification of affordances are the variety of the geometry and physical properties of these surfaces as well as…

Computer Vision and Pattern Recognition · Computer Science 2025-05-12 Tommaso Apicella , Alessio Xompero , Edoardo Ragusa , Riccardo Berta , Andrea Cavallaro , Paolo Gastaldo

Deterministic models for 3D hand pose reconstruction, whether single-staged or cascaded, struggle with pose ambiguities caused by self-occlusions and complex hand articulations. Existing cascaded approaches refine predictions in a…

Computer Vision and Pattern Recognition · Computer Science 2025-10-02 Taeyun Woo , Jinah Park , Tae-Kyun Kim

3D human pose estimation from 2D images is a challenging problem due to depth ambiguity and occlusion. Because of these challenges the task is underdetermined, where there exists multiple -- possibly infinite -- poses that are plausible…

Computer Vision and Pattern Recognition · Computer Science 2026-02-04 Francis Snelgar , Ming Xu , Stephen Gould , Liang Zheng , Akshay Asthana

3D object affordance grounding aims to predict the touchable regions on a 3d object, which is crucial for human-object interaction, human-robot interaction, embodied perception, and robot learning. Recent advances tackle this problem via…

Computer Vision and Pattern Recognition · Computer Science 2025-08-05 Hanqing Wang , Zhenhao Zhang , Kaiyang Ji , Mingyu Liu , Wenti Yin , Yuchao Chen , Zhirui Liu , Xiangyu Zeng , Tianxiang Gui , Hangxing Zhang

Understanding the inherent human knowledge in interacting with a given environment (e.g., affordance) is essential for improving AI to better assist humans. While existing approaches primarily focus on human-object contacts during…

Computer Vision and Pattern Recognition · Computer Science 2024-07-24 Hyeonwoo Kim , Sookwan Han , Patrick Kwon , Hanbyul Joo

This paper addresses the problem of affordance grounding from RGBD images of an object, which aims to localize surface regions corresponding to a text query that describes an action on the object. While existing methods predict affordance…

Computer Vision and Pattern Recognition · Computer Science 2026-04-14 Chunghyun Park , Seunghyeon Lee , Minsu Cho

Hand pose estimation from a single image has many applications. However, approaches to full 3D body pose estimation are typically trained on day-to-day activities or actions. As such, detailed hand-to-hand interactions are poorly…

Computer Vision and Pattern Recognition · Computer Science 2023-08-21 Maksym Ivashechkin , Oscar Mendez , Richard Bowden

Although diffusion models can generate high-quality human images, their applications are limited by the instability in generating hands with correct structures. In this paper, we introduce RHanDS, a conditional diffusion-based framework…

Computer Vision and Pattern Recognition · Computer Science 2025-04-15 Chengrui Wang , Pengfei Liu , Min Zhou , Ming Zeng , Xubin Li , Tiezheng Ge , Bo zheng

Recent years have seen significant progress in human image generation, particularly with the advancements in diffusion models. However, existing diffusion methods encounter challenges when producing consistent hand anatomy and the generated…

Computer Vision and Pattern Recognition · Computer Science 2024-05-01 Anton Pelykh , Ozge Mercanoglu Sincan , Richard Bowden
‹ Prev 1 2 3 10 Next ›