中文
相关论文

相关论文: Multi-view Hand Reconstruction with a Point-Embedd…

200 篇论文

Human mesh recovery (HMR) provides rich human body information for various real-world applications. While image-based HMR methods have achieved impressive results, they often struggle to recover humans in dynamic scenarios, leading to…

计算机视觉与模式识别 · 计算机科学 2025-08-20 Ce Zheng , Xianpeng Liu , Qucheng Peng , Tianfu Wu , Pu Wang , Chen Chen

Photometric Stereo methods seek to reconstruct the 3d shape of an object from motionless images obtained with varying illumination. Most existing methods solve a restricted problem where the physical reflectance model, such as Lambertian…

计算机视觉与模式识别 · 计算机科学 2021-10-06 Ofer Bartal , Nati Ofir , Yaron Lipman , Ronen Basri

Human motion recovery for real-world interaction demands both precise action details and metric-scale trajectories. Recovering absolute human pose from monocular input presents a viable solution, but faces two main challenges: (1) models'…

计算机视觉与模式识别 · 计算机科学 2026-03-16 Zhumei Wang , Zechen Hu , Ruoxi Guo , Huaijin Pi , Ziyong Feng , Liang Zhang , Mingtao Pei , Siyuan Huang

Estimating 3D hand meshes from single RGB images is challenging, due to intrinsic 2D-3D mapping ambiguities and limited training data. We adopt a compact parametric 3D hand model that represents deformable and articulated hand meshes. To…

计算机视觉与模式识别 · 计算机科学 2019-04-10 Seungryul Baek , Kwang In Kim , Tae-Kyun Kim

Neural approaches have shown a significant progress on camera-based reconstruction. But they require either a fairly dense sampling of the viewing sphere, or pre-training on an existing dataset, thereby limiting their generalizability. In…

计算机视觉与模式识别 · 计算机科学 2024-04-02 Mohammed Brahimi , Bjoern Haefner , Zhenzhang Ye , Bastian Goldluecke , Daniel Cremers

In this paper, we present DIREG3D, a holistic framework for 3D Hand Tracking. The proposed framework is capable of utilizing camera intrinsic parameters, 3D geometry, intermediate 2D cues, and visual information to regress parameters for…

计算机视觉与模式识别 · 计算机科学 2022-02-01 Ashar Ali , Upal Mahbub , Gokce Dane , Gerhard Reitmayr

Image-based scene understanding allows Augmented Reality systems to provide contextual visual guidance in unprepared, real-world environments. While effective on video see-through (VST) head-mounted displays (HMDs), such methods suffer on…

人机交互 · 计算机科学 2025-09-16 Gerlinde Emsenhuber , Tobias Langlotz , Denis Kalkofen , Markus Tatzgern

We propose Dyn-HaMR, to the best of our knowledge, the first approach to reconstruct 4D global hand motion from monocular videos recorded by dynamic cameras in the wild. Reconstructing accurate 3D hand meshes from monocular videos is a…

计算机视觉与模式识别 · 计算机科学 2025-06-03 Zhengdi Yu , Stefanos Zafeiriou , Tolga Birdal

Recent advancements in view synthesis have significantly enhanced immersive experiences across various computer graphics and multimedia applications, including telepresence and entertainment. By enabling the generation of new perspectives…

计算机视觉与模式识别 · 计算机科学 2025-10-14 Manu Gond , Emin Zerman , Sebastian Knorr , Mårten Sjöström

We propose a novel transformer-based framework that reconstructs two high fidelity hands from multi-view RGB images. Unlike existing hand pose estimation methods, where one typically trains a deep network to regress hand model parameters…

This work presents an innovative method for point set self-embedding, that encodes the structural information of a dense point set into its sparser version in a visual but imperceptible form. The self-embedded point set can function as the…

计算机视觉与模式识别 · 计算机科学 2022-03-01 Ruihui Li , Xianzhi Li , Tien-Tsin Wong , Chi-Wing Fu

We present the first approach to volumetric performance capture and novel-view rendering at real-time speed from monocular video, eliminating the need for expensive multi-view systems or cumbersome pre-acquisition of a personalized template…

计算机视觉与模式识别 · 计算机科学 2020-07-29 Ruilong Li , Yuliang Xiu , Shunsuke Saito , Zeng Huang , Kyle Olszewski , Hao Li

Emotion detection presents challenges to intelligent human-robot interaction (HRI). Foundational deep learning techniques used in emotion detection are limited by information-constrained datasets or models that lack the necessary complexity…

计算机视觉与模式识别 · 计算机科学 2023-12-19 David C. Jeong , Tianma Shen , Hongji Liu , Raghav Kapoor , Casey Nguyen , Song Liu , Christopher A. Kitts

We introduce a novel 3D hand pose estimator that can accurately recover the shape and pose of people's hands in a room from afar, typically from fixed cameras at room corners, in extremely low-resolution and frequently occluded views. Our…

计算机视觉与模式识别 · 计算机科学 2026-05-22 Shu Nakamura , Ryo Kawahara , Genki Kinoshita , Ryosuke Hirai , Yasutomo Kawanishi , Shohei Nobuhara , Ko Nishino

Multimodal foundation models that can holistically process text alongside images, video, audio, and other sensory modalities are increasingly used in a variety of real-world applications. However, it is challenging to characterize and study…

In this work, we introduce the Virtual In-Hand Eye Transformer (VIHE), a novel method designed to enhance 3D manipulation capabilities through action-aware view rendering. VIHE autoregressively refines actions in multiple stages by…

机器人学 · 计算机科学 2024-03-20 Weiyao Wang , Yutian Lei , Shiyu Jin , Gregory D. Hager , Liangjun Zhang

We aim to simultaneously estimate the 3D articulated pose and high fidelity volumetric occupancy of human performance, from multiple viewpoint video (MVV) with as few as two views. We use a multi-channel symmetric 3D convolutional…

计算机视觉与模式识别 · 计算机科学 2020-09-08 Andrew Gilbert , Matthew Trumble , Adrian Hilton , John Collomosse

Automatically assessing handwritten mathematical solutions is an important problem in educational technology with practical applications, but it remains a significant challenge due to the diverse formats, unstructured layouts, and symbolic…

计算与语言 · 计算机科学 2025-10-28 Thu Phuong Nguyen , Duc M. Nguyen , Hyotaek Jeon , Hyunwook Lee , Hyunmin Song , Sungahn Ko , Taehwan Kim

Existing approaches of hand reconstruction predominantly adhere to a multi-stage framework, encompassing detection, left-right classification, and pose estimation. This paradigm induces redundant computation and cumulative errors. In this…

计算机视觉与模式识别 · 计算机科学 2025-03-20 Xingyu Chen , Zhuheng Song , Xiaoke Jiang , Yaoqing Hu , Junzhi Yu , Lei Zhang

Accurately estimating hand pose and hand-object contact events is essential for robot data-collection, immersive virtual environments, and biomechanical analysis, yet remains challenging due to visual occlusion, subtle contact cues,…

人机交互 · 计算机科学 2025-08-25 Yuemin Mao , Uksang Yoo , Yunchao Yao , Shahram Najam Syed , Luca Bondi , Jonathan Francis , Jean Oh , Jeffrey Ichnowski