English
Related papers

Related papers: CasCalib: Cascaded Calibration for Motion Capture …

200 papers

Monocular 3D human performance capture is indispensable for many applications in computer graphics and vision for enabling immersive experiences. However, detailed capture of humans requires tracking of multiple aspects, including the…

Computer Vision and Pattern Recognition · Computer Science 2022-10-12 Yue Jiang , Marc Habermann , Vladislav Golyanik , Christian Theobalt

Several methods have been proposed to estimate 3D human pose from multi-view images, achieving satisfactory performance on public datasets collected under relatively simple conditions. However, there are limited approaches studying…

Computer Vision and Pattern Recognition · Computer Science 2026-05-08 Zhiyu Pan , Zhicheng Zhong , Wenxuan Guo , Yifan Chen , Jianjiang Feng , Jie Zhou

To tackle the challeging problem of multi-person 3D pose estimation from a single image, we propose a multi-view matching (MVM) method in this work. The MVM method generates reliable 3D human poses from a large-scale video dataset, called…

Computer Vision and Pattern Recognition · Computer Science 2020-04-14 Yeji Shen , C. -C. Jay Kuo

Perception is one of the key abilities of autonomous mobile robotic systems, which often relies on fusion of heterogeneous sensors. Although this heterogeneity presents a challenge for sensor calibration, it is also the main prospect for…

Robotics · Computer Science 2019-04-09 Juraj Peršić , Luka Petrović , Ivan Marković , Ivan Petrović

Existing human Motion Capture (MoCap) methods mostly focus on the visual similarity while neglecting the physical plausibility. As a result, downstream tasks such as driving virtual human in 3D scene or humanoid robots in real world suffer…

Computer Vision and Pattern Recognition · Computer Science 2026-05-27 Shenghao Ren , Yi Lu , Jiayi Huang , Jiayi Zhao , He Zhang , Tao Yu , Qiu Shen , Xun Cao

Human-centric visual understanding is an important desideratum for effective human-robot interaction. In order to navigate crowded public places, social robots must be able to interpret the activity of the surrounding humans. This paper…

Computer Vision and Pattern Recognition · Computer Science 2023-07-28 Shengnan Hu , Ce Zheng , Zixiang Zhou , Chen Chen , Gita Sukthankar

Motion capture has become increasingly important, not only in computer animation but also in emerging fields like the virtual reality, bioinformatics, and humanoid training. Capturing outdoor environments offers extended horizon scenes but…

Robotics · Computer Science 2024-12-31 Aditya Rauniyar , Micah Corah , Sebastian Scherer

Multi-frame human pose estimation in complicated situations is challenging. Although state-of-the-art human joints detectors have demonstrated remarkable results for static images, their performances come short when we apply these models to…

Computer Vision and Pattern Recognition · Computer Science 2021-03-22 Zhenguang Liu , Haoming Chen , Runyang Feng , Shuang Wu , Shouling Ji , Bailin Yang , Xun Wang

A Bayesian framework for 3D human pose estimation from monocular images based on sparse representation (SR) is introduced. Our probabilistic approach aims at simultaneously learning two overcomplete dictionaries (one for the visual input…

Computer Vision and Pattern Recognition · Computer Science 2014-12-02 Behnam Babagholami-Mohamadabadi , Amin Jourabloo , Ali Zarghami , Shohreh Kasaei

Accurate 3D human pose estimation is fundamental for applications such as augmented reality and human-robot interaction. State-of-the-art multi-view methods learn to fuse predictions across views by training on large annotated datasets,…

Computer Vision and Pattern Recognition · Computer Science 2025-12-03 Laura Bragagnolo , Leonardo Barcellona , Stefano Ghidoni

Open-world 3D generation has recently attracted considerable attention. While many single-image-to-3D methods have yielded visually appealing outcomes, they often lack sufficient controllability and tend to produce hallucinated regions that…

Computer Vision and Pattern Recognition · Computer Science 2024-08-20 Chao Xu , Ang Li , Linghao Chen , Yulin Liu , Ruoxi Shi , Hao Su , Minghua Liu

Human Pose estimation is a challenging problem, especially in the case of 3D pose estimation from 2D images due to many different factors like occlusion, depth ambiguities, intertwining of people, and in general crowds. 2D multi-person…

Computer Vision and Pattern Recognition · Computer Science 2019-04-26 Rohit Jena

Hand-eye calibration aims to estimate the transformation between a camera and a robot. Traditional methods rely on fiducial markers, which require considerable manual effort and precise setup. Recent advances in deep learning have…

Robotics · Computer Science 2025-12-01 Tutian Tang , Minghao Liu , Wenqiang Xu , Cewu Lu

In this paper, we address the problem of estimating a 3D human pose from a single image, which is important but difficult to solve due to many reasons, such as self-occlusions, wild appearance changes, and inherent ambiguities of 3D…

Computer Vision and Pattern Recognition · Computer Science 2019-10-08 Geonho Cha , Minsik Lee , Jungchan Cho , Songhwai Oh

Reconstructing 3D humans from images captured at multiple perspectives typically requires pre-calibration, like using checkerboards or MVS algorithms, which limits scalability and applicability in diverse real-world scenarios. In this work,…

Computer Vision and Pattern Recognition · Computer Science 2026-03-16 Xiaozhen Qiao , Wenjia Wang , Zhiyuan Zhao , Jiacheng Sun , Ping Luo , Hongyuan Zhang , Xuelong Li

Estimating 3D human poses from a monocular video is still a challenging task. Many existing methods' performance drops when the target person is occluded by other objects, or the motion is too fast/slow relative to the scale and speed of…

Computer Vision and Pattern Recognition · Computer Science 2020-10-20 Cheng Yu , Bo Wang , Bo Yang , Robby T. Tan

Many works in collaborative robotics and human-robot interaction focuses on identifying and predicting human behaviour while considering the information about the robot itself as given. This can be the case when sensors and the robot are…

Contrastive Language-Image Pre-training (CLIP) starts to emerge in many computer vision tasks and has achieved promising performance. However, it remains underexplored whether CLIP can be generalized to 3D hand pose estimation, as bridging…

Multimedia · Computer Science 2023-09-29 Shaoxiang Guo , Qing Cai , Lin Qi , Junyu Dong

This paper proposes a novel system to estimate and track the 3D poses of multiple persons in calibrated RGB-Depth camera networks. The multi-view 3D pose of each person is computed by a central node which receives the single-view outcomes…

Computer Vision and Pattern Recognition · Computer Science 2017-10-18 Marco Carraro , Matteo Munaro , Jeff Burke , Emanuele Menegatti

Markerless motion capture and understanding of professional non-daily human movements is an important yet unsolved task, which suffers from complex motion patterns and severe self-occlusion, especially for the monocular setting. In this…

Computer Vision and Pattern Recognition · Computer Science 2021-07-19 Xin Chen , Anqi Pang , Wei Yang , Yuexin Ma , Lan Xu , Jingyi Yu