中文
相关论文

相关论文: Capturing Head Avatar with Hand Contacts from a Mo…

200 篇论文

Reconstructing textured 3D human models from a single image is fundamental for AR/VR and digital human applications. However, existing methods mostly focus on single individuals and thus fail in multi-human scenes, where naive composition…

计算机视觉与模式识别 · 计算机科学 2026-04-08 Gwanghyun Kim , Junghun James Kim , Suh Yoon Jeon , Jason Park , Se Young Chun

The recent state of the art on monocular 3D face reconstruction from image data has made some impressive advancements, thanks to the advent of Deep Learning. However, it has mostly focused on input coming from a single RGB image,…

Monocular 3D face reconstruction plays a crucial role in avatar generation, with significant demand in web-related applications such as generating virtual financial advisors in FinTech. Current reconstruction methods predominantly rely on…

计算机视觉与模式识别 · 计算机科学 2024-03-28 Haoxin Xu , Zezheng Zhao , Yuxin Cao , Chunyu Chen , Hao Ge , Ziyao Liu

We present a unified framework for understanding 3D hand and object interactions in raw image sequences from egocentric RGB cameras. Given a single RGB image, our model jointly estimates the 3D hand and object poses, models their…

计算机视觉与模式识别 · 计算机科学 2019-04-11 Bugra Tekin , Federica Bogo , Marc Pollefeys

Acquiring contact patterns between hands and nonrigid objects is a common concern in the vision and robotics community. However, existing learning-based methods focus more on contact with rigid ones from monocular images. When adopting them…

计算机视觉与模式识别 · 计算机科学 2023-08-31 Wei Xie , Zimeng Zhao , Shiying Li , Binghui Zuo , Yangang Wang

Most RGB-based hand-object reconstruction methods rely on object templates, while template-free methods typically assume full object visibility. This assumption often breaks in real-world settings, where fixed camera viewpoints and static…

计算机视觉与模式识别 · 计算机科学 2025-08-08 Shibo Wang , Haonan He , Maria Parelli , Christoph Gebhardt , Zicong Fan , Jie Song

This paper presents a method to learn hand-object interaction prior for reconstructing a 3D hand-object scene from a single RGB image. The inference as well as training-data generation for 3D hand-object scene reconstruction is challenging…

计算机视觉与模式识别 · 计算机科学 2024-09-12 Hongsuk Choi , Nikhil Chavan-Dafle , Jiacheng Yuan , Volkan Isler , Hyunsoo Park

We present a novel approach for hand-object action recognition that leverages 2D point tracks as an additional motion cue. While most existing methods rely on RGB appearance, human pose estimation, or their combination, our work…

计算机视觉与模式识别 · 计算机科学 2026-01-12 Dennis Holzmann , Sven Wachsmuth

3D facial avatar reconstruction has been a significant research topic in computer graphics and computer vision, where photo-realistic rendering and flexible controls over poses and expressions are necessary for many related applications.…

计算机视觉与模式识别 · 计算机科学 2023-07-10 Wangbo Yu , Yanbo Fan , Yong Zhang , Xuan Wang , Fei Yin , Yunpeng Bai , Yan-Pei Cao , Ying Shan , Yang Wu , Zhongqian Sun , Baoyuan Wu

Physically-based simulation is a powerful approach for 3D facial animation as the resulting deformations are governed by physical constraints, allowing to easily resolve self-collisions, respond to external forces and perform realistic…

计算机视觉与模式识别 · 计算机科学 2024-09-04 Lingchen Yang , Gaspard Zoss , Prashanth Chandran , Markus Gross , Barbara Solenthaler , Eftychios Sifakis , Derek Bradley

Human avatar has become a novel type of 3D asset with various applications. Ideally, a human avatar should be fully customizable to accommodate different settings and environments. In this work, we introduce NECA, an approach capable of…

计算机视觉与模式识别 · 计算机科学 2024-03-18 Junjin Xiao , Qing Zhang , Zhan Xu , Wei-Shi Zheng

In this paper, we propose ARCH (Animatable Reconstruction of Clothed Humans), a novel end-to-end framework for accurate reconstruction of animation-ready 3D clothed humans from a monocular image. Existing approaches to digitize 3D humans…

图形学 · 计算机科学 2020-04-14 Zeng Huang , Yuanlu Xu , Christoph Lassner , Hao Li , Tony Tung

Recent advances in full-head reconstruction have been obtained by optimizing a neural field through differentiable surface or volume rendering to represent a single scene. While these techniques achieve an unprecedented accuracy, they take…

计算机视觉与模式识别 · 计算机科学 2024-04-08 Antonio Canela , Pol Caselles , Ibrar Malik , Eduard Ramon , Jaime García , Jordi Sánchez-Riera , Gil Triginer , Francesc Moreno-Noguer

Marker-less monocular 3D human motion capture (MoCap) with scene interactions is a challenging research topic relevant for extended reality, robotics and virtual avatar generation. Due to the inherent depth ambiguity of monocular settings,…

计算机视觉与模式识别 · 计算机科学 2022-07-27 Soshi Shimada , Vladislav Golyanik , Zhi Li , Patrick Pérez , Weipeng Xu , Christian Theobalt

We present SimXR, a method for controlling a simulated avatar from information (headset pose and cameras) obtained from AR / VR headsets. Due to the challenging viewpoint of head-mounted cameras, the human body is often clipped out of view,…

计算机视觉与模式识别 · 计算机科学 2024-04-26 Zhengyi Luo , Jinkun Cao , Rawal Khirodkar , Alexander Winkler , Jing Huang , Kris Kitani , Weipeng Xu

This paper addresses the problem of 3D hand pose estimation from a monocular RGB image. While previous methods have shown great success, the structure of hands has not been fully exploited, which is critical in pose estimation. To this end,…

计算机视觉与模式识别 · 计算机科学 2021-06-04 Yiming He , Wei Hu

Due to the increasing use of virtual avatars, the animation of head-hand interactions has recently gained attention. To this end, we present a novel volumetric and physics-based interaction simulation. In contrast to previous work, our…

图形学 · 计算机科学 2024-10-18 Nicolas Wagner , Mario Botsch , Ulrich Schwanecke

Recovering 4D human-object interaction (HOI) from monocular video is a key step toward scalable 3D content creation, embodied AI, and simulation-based learning. Recent methods can reconstruct temporally coherent human and object…

计算机视觉与模式识别 · 计算机科学 2026-05-15 Yubo Zhao , Yujin Chai , Yunao Dong , Chengfeng Zhao , Zijiao Zeng , Yuan Liu , Chi-Keung Tang

We propose Dyn-HaMR, to the best of our knowledge, the first approach to reconstruct 4D global hand motion from monocular videos recorded by dynamic cameras in the wild. Reconstructing accurate 3D hand meshes from monocular videos is a…

计算机视觉与模式识别 · 计算机科学 2025-06-03 Zhengdi Yu , Stefanos Zafeiriou , Tolga Birdal

We introduce a novel framework for modeling high-fidelity, animatable 3D human avatars from motion-blurred monocular video inputs. Motion blur is prevalent in real-world dynamic video capture, especially due to human movements in 3D human…

计算机视觉与模式识别 · 计算机科学 2025-06-17 Xianrui Luo , Juewen Peng , Zhongang Cai , Lei Yang , Fan Yang , Zhiguo Cao , Guosheng Lin