中文
相关论文

相关论文: SiMA-Hand: Boosting 3D Hand-Mesh Reconstruction by…

200 篇论文

Despite recent advancements in the Large Reconstruction Model (LRM) demonstrating impressive results, when extending its input from single image to multiple images, it exhibits inefficiencies, subpar geometric and texture quality, as well…

计算机视觉与模式识别 · 计算机科学 2024-12-03 Mengfei Li , Xiaoxiao Long , Yixun Liang , Weiyu Li , Yuan Liu , Peng Li , Wenhan Luo , Wenping Wang , Yike Guo

Human hands play a central role in interacting with other people and objects. For realistic replication of such hand motions, high-fidelity hand meshes have to be reconstructed. In this study, we firstly propose DeepHandMesh, a…

计算机视觉与模式识别 · 计算机科学 2020-08-20 Gyeongsik Moon , Takaaki Shiratori , Kyoung Mu Lee

3D reconstruction is a core task in many applications such as robot navigation or sites inspections. Finding the best poses to capture part of the scene is one of the most challenging topic that goes under the name of Next Best View.…

计算机视觉与模式识别 · 计算机科学 2018-05-17 Luca Morreale , Andrea Romanoni , Matteo Matteucci

3D face reconstruction technology aims to generate a face stereo model naturally and realistically. Previous deep face reconstruction approaches are typically designed to generate convincing textures and cannot generalize well to multiple…

计算机视觉与模式识别 · 计算机科学 2024-12-31 Dapeng Zhao , Yue Qi

We tackle the problem of forecasting bimanual 3D hand motion & articulation from a single image in everyday settings. To address the lack of 3D hand annotations in diverse settings, we design an annotation pipeline consisting of a diffusion…

计算机视觉与模式识别 · 计算机科学 2025-10-08 Aditya Prakash , David Forsyth , Saurabh Gupta

Reliable hand mesh reconstruction (HMR) from commonly-used color and depth sensors is challenging especially under scenarios with varied illuminations and fast motions. Event camera is a highly promising alternative for its high dynamic…

计算机视觉与模式识别 · 计算机科学 2024-03-13 Jianping Jiang , Xinyu Zhou , Bingxuan Wang , Xiaoming Deng , Chao Xu , Boxin Shi

Estimating the 3D hand articulation from a single color image is an important problem with applications in Augmented Reality (AR), Virtual Reality (VR), Human-Computer Interaction (HCI), and robotics. Apart from the absence of depth…

计算机视觉与模式识别 · 计算机科学 2025-07-18 Christos Pantazopoulos , Spyridon Thermos , Gerasimos Potamianos

Vision Transformers have made remarkable progress in recent years, achieving state-of-the-art performance in most vision tasks. A key component of this success is due to the introduction of the Multi-Head Self-Attention (MHSA) module, which…

计算机视觉与模式识别 · 计算机科学 2025-02-04 Tianxiao Zhang , Bo Luo , Guanghui Wang

In this paper, we present a new method for multi-view geometric reconstruction. In recent years, large vision models have rapidly developed, performing excellently across various tasks and demonstrating remarkable generalization…

计算机视觉与模式识别 · 计算机科学 2025-03-19 Haoyu Guo , He Zhu , Sida Peng , Haotong Lin , Yunzhi Yan , Tao Xie , Wenguan Wang , Xiaowei Zhou , Hujun Bao

While models derived from Vision Transformers (ViTs) have been phonemically surging, pre-trained models cannot seamlessly adapt to arbitrary resolution images without altering the architecture and configuration, such as sampling the…

计算机视觉与模式识别 · 计算机科学 2024-05-28 Song Zhang , Qingzhong Wang , Jiang Bian , Haoyi Xiong

In this paper, we introduce a method for reconstructing 3D humans from a single image using a biomechanically accurate skeleton model. To achieve this, we train a transformer that takes an image as input and estimates the parameters of the…

计算机视觉与模式识别 · 计算机科学 2025-03-28 Yan Xia , Xiaowei Zhou , Etienne Vouga , Qixing Huang , Georgios Pavlakos

Learning bimanual manipulation is challenging due to its high dimensionality and tight coordination required between two arms. Eye-in-hand imitation learning, which uses wrist-mounted cameras, simplifies perception by focusing on…

机器人学 · 计算机科学 2025-08-19 I-Chun Arthur Liu , Jason Chen , Gaurav Sukhatme , Daniel Seita

We propose a novel framework for 3D hand shape reconstruction and hand-object grasp optimization from a single RGB image. The representation of hand-object contact regions is critical for accurate reconstructions. Instead of approximating…

计算机视觉与模式识别 · 计算机科学 2022-11-28 Ziwei Yu , Linlin Yang , You Xie , Ping Chen , Angela Yao

Accurate in-hand pose estimation is crucial for robotic object manipulation, but visual occlusion remains a major challenge for vision-based approaches. This paper presents an approach to robotic in-hand object pose estimation, combining…

机器人学 · 计算机科学 2025-06-13 Felix Nonnengießer , Alap Kshirsagar , Boris Belousov , Jan Peters

We present a self-trainable method, Mask2Hand, which learns to solve the challenging task of predicting 3D hand pose and shape from a 2D binary mask of hand silhouette/shadow without additional manually-annotated data. Given the intrinsic…

计算机视觉与模式识别 · 计算机科学 2022-07-04 Li-Jen Chang , Yu-Cheng Liao , Chia-Hui Lin , Hwann-Tzong Chen

This paper proposes a do-it-all neural model of human hands, named LISA. The model can capture accurate hand shape and appearance, generalize to arbitrary hand subjects, provide dense surface correspondences, be reconstructed from images in…

计算机视觉与模式识别 · 计算机科学 2022-04-05 Enric Corona , Tomas Hodan , Minh Vo , Francesc Moreno-Noguer , Chris Sweeney , Richard Newcombe , Lingni Ma

Masked Image Modeling (MIM) achieves outstanding success in self-supervised representation learning. Unfortunately, MIM models typically have huge computational burden and slow learning process, which is an inevitable obstacle for their…

计算机视觉与模式识别 · 计算机科学 2023-03-10 Haoqing Wang , Yehui Tang , Yunhe Wang , Jianyuan Guo , Zhi-Hong Deng , Kai Han

Enable neural networks to capture 3D geometrical-aware features is essential in multi-view based vision tasks. Previous methods usually encode the 3D information of multi-view stereo into the 2D features. In contrast, we present a novel…

计算机视觉与模式识别 · 计算机科学 2023-05-25 Lixin Yang , Jian Xu , Licheng Zhong , Xinyu Zhan , Zhicheng Wang , Kejian Wu , Cewu Lu

Real-time occlusion handling is a major problem in outdoor mixed reality system because it requires great computational cost mainly due to the complexity of the scene. Using only segmentation, it is difficult to accurately render a virtual…

计算机视觉与模式识别 · 计算机科学 2017-08-01 Menandro Roxas , Tomoki Hori , Taiki Fukiage , Yasuhide Okamoto , Takeshi Oishi

Accurate hand and finger tracking from video has significant clinical applications for monitoring activities of daily living and measuring range of motion, yet monocular video approaches for obtaining hand biomechanics remain…

计算机视觉与模式识别 · 计算机科学 2026-05-12 R. James Cotton , Pouyan Firouzabadi , Wendy Murray