中文
相关论文

相关论文: Towards unconstrained joint hand-object reconstruc…

200 篇论文

The ubiquity of monocular videos capturing daily hand-object interactions presents a valuable resource for embodied intelligence. While 3D hand reconstruction from in-the-wild videos has seen significant progress, reconstructing the…

计算机视觉与模式识别 · 计算机科学 2026-02-09 Yuantao Chen , Jiahao Chang , Chongjie Ye , Chaoran Zhang , Zhaojie Fang , Chenghong Li , Xiaoguang Han

Learning the prior knowledge of the 3D human-object spatial relation is crucial for reconstructing human-object interaction from images and understanding how humans interact with objects in 3D space. Previous works learn this prior from…

计算机视觉与模式识别 · 计算机科学 2024-08-01 Chaofan Huo , Ye Shi , Jingya Wang

We present an approach that can reconstruct hands in 3D from monocular input. Our approach for Hand Mesh Recovery, HaMeR, follows a fully transformer-based architecture and can analyze hands with significantly increased accuracy and…

计算机视觉与模式识别 · 计算机科学 2023-12-11 Georgios Pavlakos , Dandan Shan , Ilija Radosavovic , Angjoo Kanazawa , David Fouhey , Jitendra Malik

The hardware challenges associated with light-field(LF) imaging has made it difficult for consumers to access its benefits like applications in post-capture focus and aperture control. Learning-based techniques which solve the ill-posed…

图像与视频处理 · 电气工程与系统科学 2022-07-22 Shrisudhan Govindarajan , Prasan Shedligeri , Sarah , Kaushik Mitra

Reconstructing realistic 3D human avatars from monocular videos is a challenging task due to the limited geometric information and complex non-rigid motion involved. We present MonoCloth, a new method for reconstructing and animating…

计算机视觉与模式识别 · 计算机科学 2025-11-18 Daisheng Jin , Ying He

Our goal is to learn a deep network that, given a small number of images of an object of a given category, reconstructs it in 3D. While several recent works have obtained analogous results using synthetic data or assuming the availability…

计算机视觉与模式识别 · 计算机科学 2021-03-31 Philipp Henzler , Jeremy Reizenstein , Patrick Labatut , Roman Shapovalov , Tobias Ritschel , Andrea Vedaldi , David Novotny

We present a method to learn single-view reconstruction of the 3D shape, pose, and texture of objects from categorized natural images in a self-supervised manner. Since this is a severely ill-posed problem, carefully designing a training…

计算机视觉与模式识别 · 计算机科学 2019-11-21 Hiroharu Kato , Tatsuya Harada

We present a learning-based model to infer the personalized 3D shape of people from a few frames (1-8) of a monocular video in which the person is moving, in less than 10 seconds with a reconstruction accuracy of 5mm. Our model learns to…

计算机视觉与模式识别 · 计算机科学 2019-04-09 Thiemo Alldieck , Marcus Magnor , Bharat Lal Bhatnagar , Christian Theobalt , Gerard Pons-Moll

This paper describes how to obtain accurate 3D body models and texture of arbitrary people from a single, monocular video in which a person is moving. Based on a parametric body model, we present a robust processing pipeline achieving 3D…

计算机视觉与模式识别 · 计算机科学 2018-04-17 Thiemo Alldieck , Marcus Magnor , Weipeng Xu , Christian Theobalt , Gerard Pons-Moll

We reconstruct 3D deformable object through time, in the context of a live pottery making process where the crafter molds the object. Because the object suffers from heavy hand interaction, and is being deformed, classical techniques cannot…

计算机视觉与模式识别 · 计算机科学 2019-08-06 Raoul de Charette , Sotiris Manitsaris

We tackle the task of reconstructing hand-object interactions from short video clips. Given an input video, our approach casts 3D inference as a per-video optimization and recovers a neural 3D representation of the object shape, as well as…

计算机视觉与模式识别 · 计算机科学 2023-09-12 Yufei Ye , Poorvi Hebbar , Abhinav Gupta , Shubham Tulsiani

Rendering articulated objects while controlling their poses is critical to applications such as virtual reality or animation for movies. Manipulating the pose of an object, however, requires the understanding of its underlying structure,…

计算机视觉与模式识别 · 计算机科学 2022-04-07 Atsuhiro Noguchi , Umar Iqbal , Jonathan Tremblay , Tatsuya Harada , Orazio Gallo

We propose CrossHuman, a novel method that learns cross-guidance from parametric human model and multi-frame RGB images to achieve high-quality 3D human reconstruction. To recover geometry details and texture even in invisible regions, we…

计算机视觉与模式识别 · 计算机科学 2022-07-21 Liliang Chen , Jiaqi Li , Han Huang , Yandong Guo

Holistic 3D human-scene reconstruction is a crucial and emerging research area in robot perception. A key challenge in holistic 3D human-scene reconstruction is to generate a physically plausible 3D scene from a single monocular RGB image.…

计算机视觉与模式识别 · 计算机科学 2023-07-28 Sandika Biswas , Kejie Li , Biplab Banerjee , Subhasis Chaudhuri , Hamid Rezatofighi

Understanding the 3D world is a fundamental problem in computer vision. However, learning a good representation of 3D objects is still an open problem due to the high dimensionality of the data and many factors of variation involved. In…

计算机视觉与模式识别 · 计算机科学 2017-08-15 Xinchen Yan , Jimei Yang , Ersin Yumer , Yijie Guo , Honglak Lee

Human-robot object handovers have been an actively studied area of robotics over the past decade; however, very few techniques and systems have addressed the challenge of handing over diverse objects with arbitrary appearance, size, shape,…

机器人学 · 计算机科学 2021-06-07 Wei Yang , Chris Paxton , Arsalan Mousavian , Yu-Wei Chao , Maya Cakmak , Dieter Fox

While 3D hand reconstruction from monocular images has made significant progress, generating accurate and temporally coherent motion estimates from videos remains challenging, particularly during hand-object interactions. In this paper, we…

计算机视觉与模式识别 · 计算机科学 2025-08-05 Yufei Zhang , Zijun Cui , Jeffrey O. Kephart , Qiang Ji

Reconstructing 3D models of dynamic, real-world objects with high-fidelity textures from monocular frame sequences has been a challenging problem in recent years. This difficulty stems from factors such as shadows, indirect illumination,…

计算机视觉与模式识别 · 计算机科学 2025-05-12 Alakh Aggarwal , Ningna Wang , Xiaohu Guo

Recent approaches to jointly reconstruct 3D humans and objects from a single RGB image represent 3D shapes with template-based or coarse models, which fail to capture details of loose clothing on human bodies. In this paper, we introduce a…

计算机视觉与模式识别 · 计算机科学 2025-03-11 Ayushi Dutta , Marco Pesavento , Marco Volino , Adrian Hilton , Armin Mustafa

We present an approach for safe and object-independent human-to-robot handovers using real time robotic vision and manipulation. We aim for general applicability with a generic object detector, a fast grasp selection algorithm and by using…