中文
相关论文

相关论文: Towards Unstructured Unlabeled Optical Mocap: A Vi…

200 篇论文

We present EgoHDM, an online egocentric-inertial human motion capture (mocap), localization, and dense mapping system. Our system uses 6 inertial measurement units (IMUs) and a commodity head-mounted RGB camera. EgoHDM is the first human…

计算机视觉与模式识别 · 计算机科学 2024-09-06 Bonan Liu , Handi Yin , Manuel Kaufmann , Jinhao He , Sammy Christen , Jie Song , Pan Hui

Recent methods for arbitrary-skeleton motion capture from monocular video follow a factorized pipeline, where a Video-to-Pose network predicts joint positions and an analytical inverse-kinematics (IK) stage recovers joint rotations. While…

计算机视觉与模式识别 · 计算机科学 2026-05-15 Kehong Gong , Zhengyu Wen , Dao Thien Phong , Mingxi Xu , Weixia He , Qi Wang , Ning Zhang , Zhengyu Li , Guanli Hou , Dongze Lian , Xiaoyu He , Mingyuan Zhang , Hanwang Zhang

The goal of unpaired image captioning (UIC) is to describe images without using image-caption pairs in the training phase. Although challenging, we except the task can be accomplished by leveraging a training set of images aligned with…

计算机视觉与模式识别 · 计算机科学 2022-03-08 Peipei Zhu , Xiao Wang , Yong Luo , Zhenglong Sun , Wei-Shi Zheng , Yaowei Wang , Changwen Chen

Optical motion capture (mocap) systems are widely used for ground-truth capture in AR/VR, SLAM and robotics datasets. These datasets require extrinsic calibration to align mocap coordinates to external camera frames -- a step that is…

计算机视觉与模式识别 · 计算机科学 2026-04-27 Tianyi Liu , Christopher Twigg , Patrick Grady , Kevin Harris , Shangchen Han , Kun He

We tackle the problem of highly-accurate, holistic performance capture for the face, body and hands simultaneously. Motion-capture technologies used in film and game production typically focus only on face, body or hand capture…

Person images captured by surveillance cameras are often occluded by various obstacles, which lead to defective feature representation and harm person re-identification (Re-ID) performance. To tackle this challenge, we propose to…

计算机视觉与模式识别 · 计算机科学 2021-05-18 Shijie Yu , Dapeng Chen , Rui Zhao , Haobin Chen , Yu Qiao

Humans constantly interact with daily objects to accomplish tasks. To understand such interactions, computers need to reconstruct these from cameras observing whole-body interaction with scenes. This is challenging due to occlusion between…

计算机视觉与模式识别 · 计算机科学 2022-10-04 Yinghao Huang , Omid Tehari , Michael J. Black , Dimitrios Tzionas

Human performance capture is a highly important computer vision problem with many applications in movie production and virtual/augmented reality. Many previous performance capture approaches either required expensive multi-view setups or…

计算机视觉与模式识别 · 计算机科学 2020-03-19 Marc Habermann , Weipeng Xu , Michael Zollhoefer , Gerard Pons-Moll , Christian Theobalt

Motion capture now underpins content creation far beyond digital humans, yet most existing pipelines remain species- or template-specific. We formalize this gap as Category-Agnostic Motion Capture (CAMoCap): given a monocular video and an…

计算机视觉与模式识别 · 计算机科学 2026-05-01 Kehong Gong , Zhengyu Wen , Weixia He , Mingxi Xu , Qi Wang , Ning Zhang , Zhengyu Li , Dongze Lian , Wei Zhao , Xiaoyu He , Mingyuan Zhang

Privacy preservation is a prerequisite for using video data in Operating Room (OR) research. Effective anonymization relies on the exhaustive localization of every individual; even a single missed detection necessitates extensive manual…

计算机视觉与模式识别 · 计算机科学 2026-02-10 Keqi Chen , Vinkle Srivastav , Armine Vardazaryan , Cindy Rolland , Didier Mutter , Nicolas Padoy

Recently, occluded person re-identification(Re-ID) remains a challenging task that people are frequently obscured by other people or obstacles, especially in a crowd massing situation. In this paper, we propose a self-supervised deep…

计算机视觉与模式识别 · 计算机科学 2022-02-11 Mi Zhou , Hongye Liu , Zhekun Lv , Wei Hong , Xiai Chen

Optical motion capture systems have become a widely used technology in various fields, such as augmented reality, robotics, movie production, etc. Such systems use a large number of cameras to triangulate the position of optical markers.The…

机器学习 · 计算机科学 2018-09-26 Taras Kucherenko , Jonas Beskow , Hedvig Kjellström

Capturing challenging human motions is critical for numerous applications, but it suffers from complex motion patterns and severe self-occlusion under the monocular setting. In this paper, we propose ChallenCap -- a template-based approach…

计算机视觉与模式识别 · 计算机科学 2021-03-30 Yannan He , Anqi Pang , Xin Chen , Han Liang , Minye Wu , Yuexin Ma , Lan Xu

We introduce a new method that generates photo-realistic humans under novel views and poses given a monocular video as input. Despite the significant progress recently on this topic, with several methods exploring shared canonical neural…

计算机视觉与模式识别 · 计算机科学 2023-04-21 Tiantian Wang , Nikolaos Sarafianos , Ming-Hsuan Yang , Tony Tung

The proliferation of commercial egocentric devices offers a unique lens into human behavior, yet reconstructing full-body 3D motion remains difficult due to frequent self-occlusion and the 'out-of-sight' nature of the wearer's limbs. While…

计算机视觉与模式识别 · 计算机科学 2026-04-02 Kyungwon Cho , Hanbyul Joo

We present a novel locality-based learning method for cleaning and solving optical motion capture data. Given noisy marker data, we propose a new heterogeneous graph neural network which treats markers and joints as different types of…

Although significant progress has been achieved on monocular maker-less human motion capture in recent years, it is still hard for state-of-the-art methods to obtain satisfactory results in occlusion scenarios. There are two main reasons:…

计算机视觉与模式识别 · 计算机科学 2022-07-13 Buzhen Huang , Yuan Shu , Jingyi Ju , Yangang Wang

In recent years, Neural Radiance Fields (NeRF) have achieved remarkable progress in dynamic human reconstruction and rendering. Part-based rendering paradigms, guided by human segmentation, allow for flexible parameter allocation based on…

计算机视觉与模式识别 · 计算机科学 2025-08-13 Yao Lu , Jiawei Li , Ming Jiang

Most existing monocular 3D pose estimation approaches only focus on a single body part, neglecting the fact that the essential nuance of human motion is conveyed through a concert of subtle movements of face, hands, and body. In this paper,…

计算机视觉与模式识别 · 计算机科学 2021-08-24 Yu Rong , Takaaki Shiratori , Hanbyul Joo

Unsupervised Learning based monocular visual odometry (VO) has lately drawn significant attention for its potential in label-free leaning ability and robustness to camera parameters and environmental variations. However, partially due to…

计算机视觉与模式识别 · 计算机科学 2019-03-18 Yang Li , Yoshitaka Ushiku , Tatsuya Harada