中文
相关论文

相关论文: MoCapDeform: Monocular 3D Human Motion Capture in …

200 篇论文

Estimating human motion from video is an active research area due to its many potential applications. Most state-of-the-art methods predict human shape and posture estimates for individual images and do not leverage the temporal information…

计算机视觉与模式识别 · 计算机科学 2022-07-26 Dorian F. Henning , Tristan Laidlow , Stefan Leutenegger

Markerless human motion capture (mocap) from multiple RGB cameras is a widely studied problem. Existing methods either need calibrated cameras or calibrate them relative to a static camera, which acts as the reference frame for the mocap…

计算机视觉与模式识别 · 计算机科学 2023-04-04 Nitin Saini , Chun-hao P. Huang , Michael J. Black , Aamir Ahmad

This paper proposes a new approach for monocular dense 3D reconstruction of a complex dynamic scene from two perspective frames. By applying superpixel over-segmentation to the image, we model a generically dynamic (hence non-rigid) scene…

计算机视觉与模式识别 · 计算机科学 2017-12-21 Suryansh Kumar , Yuchao Dai , Hongdong Li

This paper addresses the challenge of novel view synthesis for a human performer from a very sparse set of camera views. Some recent works have shown that learning implicit neural representations of 3D scenes achieves remarkable view…

计算机视觉与模式识别 · 计算机科学 2021-03-30 Sida Peng , Yuanqing Zhang , Yinghao Xu , Qianqian Wang , Qing Shuai , Hujun Bao , Xiaowei Zhou

Robotic manipulation requires accurate perception of the environment, which poses a significant challenge due to its inherent complexity and constantly changing nature. In this context, RGB image and point-cloud observations are two…

机器人学 · 计算机科学 2024-09-10 Boshi An , Yiran Geng , Kai Chen , Xiaoqi Li , Qi Dou , Hao Dong

In this paper, we propose MoDGS, a new pipeline to render novel views of dy namic scenes from a casually captured monocular video. Previous monocular dynamic NeRF or Gaussian Splatting methods strongly rely on the rapid move ment of input…

计算机视觉与模式识别 · 计算机科学 2025-05-20 Qingming Liu , Yuan Liu , Jiepeng Wang , Xianqiang Lyv , Peng Wang , Wenping Wang , Junhui Hou

Training state-of-the-art models for human body pose and shape recovery from images or videos requires datasets with corresponding annotations that are really hard and expensive to obtain. Our goal in this paper is to study whether poses…

计算机视觉与模式识别 · 计算机科学 2021-10-19 Fabien Baradel , Thibault Groueix , Philippe Weinzaepfel , Romain Brégier , Yannis Kalantidis , Grégory Rogez

Quantifying human movement (kinematics) and musculoskeletal forces (kinetics) at scale, such as estimating quadriceps force during a sit-to-stand movement, could transform prediction, treatment, and monitoring of mobility-related…

计算机视觉与模式识别 · 计算机科学 2026-03-27 Selim Gilon , Emily Y. Miller , Scott D. Uhlrich

Photorealistic 3D head avatars are vital for telepresence, gaming, and VR. However, most methods focus solely on facial regions, ignoring natural hand-face interactions, such as a hand resting on the chin or fingers gently touching the…

计算机视觉与模式识别 · 计算机科学 2025-10-21 Haonan He , Yufeng Zheng , Jie Song

The challenge of dynamic view synthesis from dynamic monocular videos, i.e., synthesizing novel views for free viewpoints given a monocular video of a dynamic scene captured by a moving camera, mainly lies in accurately modeling the…

计算机视觉与模式识别 · 计算机科学 2024-08-22 Meng You , Junhui Hou

Visual understanding of the world goes beyond the semantics and flat structure of individual images. In this work, we aim to capture both the 3D structure and dynamics of real-world scenes from monocular real-world videos. Our Dynamic Scene…

计算机视觉与模式识别 · 计算机科学 2024-03-18 Maximilian Seitzer , Sjoerd van Steenkiste , Thomas Kipf , Klaus Greff , Mehdi S. M. Sajjadi

Tracking 3D human motion from egocentric multi-camera headset is challenged by severe egomotion, partial visibility or occlusions and lack of training data. Existing methods designed for monocular video often require static or slowly-moving…

计算机视觉与模式识别 · 计算机科学 2026-05-08 Nan Yang , Julian Straub , Fan Zhang , Richard Newcombe , Jakob Engel , Lingni Ma

Estimating human pose and shape from monocular images is a long-standing problem in computer vision. Since the release of statistical body models, 3D human mesh recovery has been drawing broader attention. With the same goal of obtaining…

计算机视觉与模式识别 · 计算机科学 2024-01-03 Yating Tian , Hongwen Zhang , Yebin Liu , Limin Wang

We explore the task of embodied view synthesis from monocular videos of deformable scenes. Given a minute-long RGBD video of people interacting with their pets, we render the scene from novel camera trajectories derived from the in-scene…

计算机视觉与模式识别 · 计算机科学 2023-10-03 Chonghyuk Song , Gengshan Yang , Kangle Deng , Jun-Yan Zhu , Deva Ramanan

Recent years have witnessed tremendous progress in the 3D reconstruction of dynamic humans from a monocular video with the advent of neural rendering techniques. This task has a wide range of applications, including the creation of virtual…

计算机视觉与模式识别 · 计算机科学 2024-09-24 Kanghao Chen , Zeyu Wang , Lin Wang

Humans naturally interact with both others and the surrounding multiple objects, engaging in various social activities. However, recent advances in modeling human-object interactions mostly focus on perceiving isolated individuals and…

计算机视觉与模式识别 · 计算机科学 2024-04-03 Juze Zhang , Jingyan Zhang , Zining Song , Zhanhe Shi , Chengfeng Zhao , Ye Shi , Jingyi Yu , Lan Xu , Jingya Wang

We present a generative method to estimate 3D human motion and body shape from monocular video. Under the assumption that starting from an initial pose optical flow constrains subsequent human motion, we exploit flow to find temporally…

计算机视觉与模式识别 · 计算机科学 2017-03-22 Thiemo Alldieck , Marc Kassubeck , Marcus Magnor

To address the ill-posed problem caused by partial observations in monocular human volumetric capture, we present AvatarCap, a novel framework that introduces animatable avatars into the capture pipeline for high-fidelity reconstruction in…

计算机视觉与模式识别 · 计算机科学 2022-07-13 Zhe Li , Zerong Zheng , Hongwen Zhang , Chaonan Ji , Yebin Liu

Perceiving humans in the context of Intelligent Transportation Systems (ITS) often relies on multiple cameras or expensive LiDAR sensors. In this work, we present a new cost-effective vision-based method that perceives humans' locations in…

计算机视觉与模式识别 · 计算机科学 2021-05-04 Lorenzo Bertoni , Sven Kreiss , Alexandre Alahi

We present a novel non-rigid reconstruction method using a moving RGB-D camera. Current approaches use only non-rigid part of the scene and completely ignore the rigid background. Non-rigid parts often lack sufficient geometric and…

计算机视觉与模式识别 · 计算机科学 2018-05-31 Shafeeq Elanattil , Peyman Moghadam , Sridha Sridharan , Clinton Fookes , Mark Cox