English
Related papers

Related papers: SHOWMe: Benchmarking Object-agnostic Hand-Object 3…

200 papers

This paper introduces a dataset for training and evaluating methods for 6D pose estimation of hand-held tools in task demonstrations captured by a standard RGB camera. Despite the significant progress of 6D pose estimation methods, their…

Estimating the 3D pose of hand and potential hand-held object from monocular images is a longstanding challenge. Yet, existing methods are specialized, focusing on either bare-hand or hand interacting with object. No method can flexibly…

Computer Vision and Pattern Recognition · Computer Science 2025-03-18 Yinqiao Wang , Hao Xu , Pheng-Ann Heng , Chi-Wing Fu

Accurate and robust 3D scene reconstruction from casual, in-the-wild videos can significantly simplify robot deployment to new environments. However, reliable camera pose estimation and scene reconstruction from such unconstrained videos…

Computer Vision and Pattern Recognition · Computer Science 2025-04-30 Shuo Sun , Torsten Sattler , Malcolm Mielle , Achim J. Lilienthal , Martin Magnusson

A lensless camera is an imaging system that uses a mask in place of a lens, making it thinner, lighter, and less expensive than a lensed camera. However, additional complex computation and time are required for image reconstruction. This…

Computer Vision and Pattern Recognition · Computer Science 2022-10-26 Yinger Zhang , Zhouyi Wu , Peiying Lin , Yang Pan , Yuting Wu , Liufang Zhang , Jiangtao Huangfu

Learning the prior knowledge of the 3D human-object spatial relation is crucial for reconstructing human-object interaction from images and understanding how humans interact with objects in 3D space. Previous works learn this prior from…

Computer Vision and Pattern Recognition · Computer Science 2024-08-01 Chaofan Huo , Ye Shi , Jingya Wang

Reconstructing detailed 3D scenes from single-view images remains a challenging task due to limitations in existing approaches, which primarily focus on geometric shape recovery, overlooking object appearances and fine shape details. To…

Computer Vision and Pattern Recognition · Computer Science 2023-11-02 Yixin Chen , Junfeng Ni , Nan Jiang , Yaowei Zhang , Yixin Zhu , Siyuan Huang

Monocular 3D object detection (M3OD) is intrinsically ill-posed, hence training a high-performance deep learning based M3OD model requires a humongous amount of labeled data with complicated visual variation from diverse scenes, variety of…

Computer Vision and Pattern Recognition · Computer Science 2026-03-10 Zhaonian Kuang , Rui Ding , Meng Yang , Xinhu Zheng , Gang Hua

This work addresses the problem of recovering complete, simulatable object geometry from reconstructed real-world scenes, enabling physics-based interaction with objects embedded in the scene. While modern multi-view reconstruction methods…

Computer Vision and Pattern Recognition · Computer Science 2026-05-29 Xin Dong , Weijian Deng , Lihan Zhang , Tianru Dai , Wenfeng Deng , Yansong Tang

3D human pose estimation captures the human joint points in three-dimensional space while keeping the depth information and physical structure. That is essential for applications that require precise pose information, such as human-computer…

Computer Vision and Pattern Recognition · Computer Science 2024-03-26 Jianbin Jiao , Xina Cheng , Weijie Chen , Xiaoting Yin , Hao Shi , Kailun Yang

Tactile sensing is one of the modalities humans rely on heavily to perceive the world. Working with vision, this modality refines local geometry structure, measures deformation at the contact area, and indicates the hand-object contact…

Computer Vision and Pattern Recognition · Computer Science 2023-03-28 Wenqiang Xu , Zhenjun Yu , Han Xue , Ruolin Ye , Siqiong Yao , Cewu Lu

Humans constantly interact with objects in daily life tasks. Capturing such processes and subsequently conducting visual inferences from a fixed viewpoint suffers from occlusions, shape and texture ambiguities, motions, etc. To mitigate the…

Computer Vision and Pattern Recognition · Computer Science 2022-12-16 Juze Zhang , Haimin Luo , Hongdi Yang , Xinru Xu , Qianyang Wu , Ye Shi , Jingyi Yu , Lan Xu , Jingya Wang

Contactless hand pose estimation requires sensors that provide precise spatial information and low computational complexity for real-time processing. Unlike vision-based systems, radar offers lighting independence and direct motion…

Signal Processing · Electrical Eng. & Systems 2024-06-21 Johanna Bräunig , Vanessa Wirth , Marc Stamminger , Ingrid Ullmann , Martin Vossiek

We introduce the task of Reconstructing Objects along Hand Interaction Timelines (ROHIT). We first define the Hand Interaction Timeline (HIT) from a rigid object's perspective. In a HIT, an object is first static relative to the scene, then…

Computer Vision and Pattern Recognition · Computer Science 2025-12-09 Zhifan Zhu , Siddhant Bansal , Shashank Tripathi , Dima Damen

Object handover is a common human collaboration behavior that attracts attention from researchers in Robotics and Cognitive Science. Though visual perception plays an important role in the object handover task, the whole handover process…

Computer Vision and Pattern Recognition · Computer Science 2021-10-28 Ruolin Ye , Wenqiang Xu , Zhendong Xue , Tutian Tang , Yanfeng Wang , Cewu Lu

Current parametric models have made notable progress in 3D hand pose and shape estimation. However, due to the fixed hand topology and complex hand poses, current models are hard to generate meshes that are aligned with the image well. To…

Computer Vision and Pattern Recognition · Computer Science 2023-12-27 Hanhui Li , Xiaojian Lin , Xuan Huang , Zejun Yang , Zhisheng Wang , Xiaodan Liang

Accurate modelling of object deformations is crucial for a wide range of robotic manipulation tasks, where interacting with soft or deformable objects is essential. Current methods struggle to generalise to unseen forces or adapt to new…

Robotics · Computer Science 2025-05-20 Sean M. V. Collins , Brendan Tidd , Mahsa Baktashmotlagh , Peyman Moghadam

Progress has been achieved recently in object detection given advancements in deep learning. Nevertheless, such tools typically require a large amount of training data and significant manual effort to label objects. This limits their…

Robotics · Computer Science 2017-08-04 Chaitanya Mitash , Kostas E. Bekris , Abdeslam Boularias

We present a new multi-stream 3D mesh reconstruction network (MSMR-Net) for hand pose estimation from a single RGB image. Our model consists of an image encoder followed by a mesh-convolution decoder composed of connected graph convolution…

Computer Vision and Pattern Recognition · Computer Science 2021-04-27 Uri Wollner , Guy Ben-Yosef

Nowadays, the need for large amounts of carefully and complexly annotated data for the training of computer vision modules continues to grow. Furthermore, although the research community presents state of the art solutions to many problems,…

Computer Vision and Pattern Recognition · Computer Science 2022-10-20 Prodromos Boutis , Zisis Batzos , Konstantinos Konstantoudakis , Anastasios Dimou , Petros Daras

This work presents a flexible system to reconstruct 3D models of objects captured with an RGB-D sensor. A major advantage of the method is that our reconstruction pipeline allows the user to acquire a full 3D model of the object. This is…

Computer Vision and Pattern Recognition · Computer Science 2015-05-22 Aitor Aldoma , Johann Prankl , Alexander Svejda , Markus Vincze