中文
相关论文

相关论文: RemoCap: Disentangled Representation Learning for …

200 篇论文

Occlusions remain one of the key challenges in 3D body pose estimation from single-camera video sequences. Temporal consistency has been extensively used to mitigate their impact but the existing algorithms in the literature do not…

计算机视觉与模式识别 · 计算机科学 2024-02-20 Soumava Kumar Roy , Ilia Badanin , Sina Honari , Pascal Fua

Recovering high-quality 3D human motion in complex scenes from monocular videos is important for many applications, ranging from AR/VR to robotics. However, capturing realistic human-scene interactions, while dealing with occlusions and…

计算机视觉与模式识别 · 计算机科学 2021-08-25 Siwei Zhang , Yan Zhang , Federica Bogo , Marc Pollefeys , Siyu Tang

Motion capture now underpins content creation far beyond digital humans, yet most existing pipelines remain species- or template-specific. We formalize this gap as Category-Agnostic Motion Capture (CAMoCap): given a monocular video and an…

计算机视觉与模式识别 · 计算机科学 2026-05-01 Kehong Gong , Zhengyu Wen , Weixia He , Mingxi Xu , Qi Wang , Ning Zhang , Zhengyu Li , Dongze Lian , Wei Zhao , Xiaoyu He , Mingyuan Zhang

Parametric models of humans, faces, hands and animals have been widely used for a range of tasks such as image-based reconstruction, shape correspondence estimation, and animation. Their key strength is the ability to factor surface…

计算机视觉与模式识别 · 计算机科学 2020-07-23 Keyang Zhou , Bharat Lal Bhatnagar , Gerard Pons-Moll

Imitation learning from human hand motion data presents a promising avenue for imbuing robots with human-like dexterity in real-world manipulation tasks. Despite this potential, substantial challenges persist, particularly with the…

机器人学 · 计算机科学 2024-07-08 Chen Wang , Haochen Shi , Weizhuo Wang , Ruohan Zhang , Li Fei-Fei , C. Karen Liu

3D reconstruction of dynamic crowds in large scenes has become increasingly important for applications such as city surveillance and crowd analysis. However, current works attempt to reconstruct 3D crowds from a static image, causing a lack…

计算机视觉与模式识别 · 计算机科学 2025-08-19 Hao Wen , Hongbo Kang , Jian Ma , Jing Huang , Yuanwang Yang , Haozhe Lin , Yu-Kun Lai , Kun Li

Hand-object motion-capture (MoCap) repositories offer large-scale, contact-rich demonstrations and hold promise for scaling dexterous robotic manipulation. Yet demonstration inaccuracies and embodiment gaps between human and robot hands…

机器人学 · 计算机科学 2025-09-12 Sirui Xu , Yu-Wei Chao , Liuyu Bian , Arsalan Mousavian , Yu-Xiong Wang , Liang-Yan Gui , Wei Yang

We propose a new method to reconstruct the 3D human body from RGB-D images with occlusions. The foremost challenge is the incompleteness of the RGB-D data due to occlusions between the body and the environment, leading to implausible…

计算机视觉与模式识别 · 计算机科学 2023-10-17 Bowen Dang , Xi Zhao , Bowen Zhang , He Wang

Recently, deep learning-based 3D face reconstruction methods have demonstrated promising advancements in terms of quality and efficiency. Nevertheless, these techniques face challenges in effectively handling occluded scenes and fail to…

计算机视觉与模式识别 · 计算机科学 2025-03-18 Dapeng Zhao

In recent years, Neural Radiance Fields (NeRF) have achieved remarkable progress in dynamic human reconstruction and rendering. Part-based rendering paradigms, guided by human segmentation, allow for flexible parameter allocation based on…

计算机视觉与模式识别 · 计算机科学 2025-08-13 Yao Lu , Jiawei Li , Ming Jiang

Neural shape models can represent complex 3D shapes with a compact latent space. When applied to dynamically deforming shapes such as the human hands, however, they would need to preserve temporal coherence of the deformation as well as the…

计算机视觉与模式识别 · 计算机科学 2021-10-06 Binbin Xu , Lingni Ma , Yuting Ye , Tanner Schmidt , Christopher D. Twigg , Steven Lovegrove

Deep learning-based dense object detectors have achieved great success in the past few years and have been applied to numerous multimedia applications such as video understanding. However, the current training pipeline for dense detectors…

计算机视觉与模式识别 · 计算机科学 2021-07-28 Zehui Chen , Chenhongyi Yang , Qiaofei Li , Feng Zhao , Zheng-Jun Zha , Feng Wu

Advances in Deep Learning have recently made it possible to recover full 3D meshes of human poses from individual images. However, extension of this notion to videos for recovering temporally coherent poses still remains unexplored. A major…

计算机视觉与模式识别 · 计算机科学 2019-07-02 Jian Liu , Naveed Akhtar , Ajmal Mian

We propose a novel approach to jointly perform 3D shape retrieval and pose estimation from monocular images.In order to make the method robust to real-world image variations, e.g. complex textures and backgrounds, we learn an embedding…

计算机视觉与模式识别 · 计算机科学 2019-03-28 Kyaw Zaw Lin , Weipeng Xu , Qianru Sun , Christian Theobalt , Tat-Seng Chua

This paper addresses the challenging task of reconstructing the poses of multiple individuals engaged in close interactions, captured by multiple calibrated cameras. The difficulty arises from the noisy or false 2D keypoint detections due…

计算机视觉与模式识别 · 计算机科学 2024-01-30 Qing Shuai , Zhiyuan Yu , Zhize Zhou , Lixin Fan , Haijun Yang , Can Yang , Xiaowei Zhou

In multi-view human body capture systems, the recovered 3D geometry or even the acquired imagery data can be heavily corrupted due to occlusions, noise, limited field of- view, etc. Direct estimation of 3D pose, body shape or motion on…

计算机视觉与模式识别 · 计算机科学 2018-02-02 Zhong Li , Yu Ji , Wei Yang , Jinwei Ye , Jingyi Yu

Recent studies on motion estimation have advocated an optimized motion representation that is globally consistent across the entire video, preferably for every pixel. This is challenging as a uniform representation may not account for the…

计算机视觉与模式识别 · 计算机科学 2024-07-17 Rui Li , Dong Liu

Motion capture (mocap) and time-of-flight based sensing of human actions are becoming increasingly popular modalities to perform robust activity analysis. Applications range from action recognition to quantifying movement quality for health…

计算机视觉与模式识别 · 计算机科学 2020-12-04 Suhas Lohit , Rushil Anirudh , Pavan Turaga

Multi-modal MRIs are widely used in neuroimaging applications since different MR sequences provide complementary information about brain structures. Recent works have suggested that multi-modal deep learning analysis can benefit from…

计算机视觉与模式识别 · 计算机科学 2021-06-14 Jiahong Ouyang , Ehsan Adeli , Kilian M. Pohl , Qingyu Zhao , Greg Zaharchuk

Marker-based optical motion capture (MoCap), while long regarded as the gold standard for accuracy, faces practical challenges, such as time-consuming preparation and marker identification ambiguity, due to its reliance on dense marker…

计算机视觉与模式识别 · 计算机科学 2025-11-21 Hai Lan , Zongyan Li , Jianmin Hu , Jialing Yang , Houde Dai