English
Related papers

Related papers: GenFusion: Feed-forward Human Performance Capture …

200 papers

We present a novel framework for reconstructing animatable human avatars from multiple images, termed CanonicalFusion. Our central concept involves integrating individual reconstruction results into the canonical space. To be specific, we…

Computer Vision and Pattern Recognition · Computer Science 2024-07-16 Jisu Shin , Junmyeong Lee , Seongmin Lee , Min-Gyu Park , Ju-Mi Kang , Ju Hong Yoon , Hae-Gon Jeon

High-quality 4D reconstruction of human performance with complex interactions to various objects is essential in real-world scenarios, which enables numerous immersive VR/AR applications. However, recent advances still fail to provide…

Computer Vision and Pattern Recognition · Computer Science 2021-05-03 Zhuo Su , Lan Xu , Dawei Zhong , Zhong Li , Fan Deng , Shuxue Quan , Lu Fang

We introduce a new method that generates photo-realistic humans under novel views and poses given a monocular video as input. Despite the significant progress recently on this topic, with several methods exploring shared canonical neural…

Computer Vision and Pattern Recognition · Computer Science 2023-04-21 Tiantian Wang , Nikolaos Sarafianos , Ming-Hsuan Yang , Tony Tung

In this paper, we aim at synthesizing a free-viewpoint video of an arbitrary human performance using sparse multi-view cameras. Recently, several works have addressed this problem by learning person-specific neural radiance fields (NeRF) to…

Computer Vision and Pattern Recognition · Computer Science 2021-09-16 Youngjoong Kwon , Dahun Kim , Duygu Ceylan , Henry Fuchs

3D human reconstruction and animation are long-standing topics in computer graphics and vision. However, existing methods typically rely on sophisticated dense-view capture and/or time-consuming per-subject optimization procedures. To…

Graphics · Computer Science 2025-06-04 Zhiyuan Yu , Zhe Li , Hujun Bao , Can Yang , Xiaowei Zhou

We present a generalizable feed-forward Gaussian splatting framework for human 3D reconstruction and real-time animation that operates directly on multi-view RGB images and their associated SMPL-X poses. Unlike prior methods that rely on…

Computer Vision and Pattern Recognition · Computer Science 2026-04-14 Devdoot Chatterjee , Zakaria Laskar , C. V. Jawahar

High-fidelity reconstruction of driving scenes is crucial for autonomous driving. While recent feedforward 3D Gaussian Splatting (3DGS) methods enable fast reconstruction, their per-pixel Gaussian prediction paradigm often suffers from…

Computer Vision and Pattern Recognition · Computer Science 2026-05-13 Cheng Chi , Xianqi Wang , Hongcheng Luo , Mingfei Tu , Gangwei Xu , Zehan Zhang , Bing Wang , Guang Chen , Hangjun Ye , Sida Peng , Xin Yang , Haiyang Sun

Real-time dense scene reconstruction during unstable camera motions is crucial for robotics, yet current RGB-D SLAM systems fail when cameras experience large viewpoint changes, fast motions, or sudden shaking. Classical optimization-based…

Robotics · Computer Science 2026-03-04 Siyan Dong , Zijun Wang , Lulu Cai , Yi Ma , Yanchao Yang

We present the first real-time human performance capture approach that reconstructs dense, space-time coherent deforming geometry of entire humans in general everyday clothing from just a single RGB video. We propose a novel two-stage…

Computer Vision and Pattern Recognition · Computer Science 2019-01-28 Marc Habermann , Weipeng Xu , Michael Zollhoefer , Gerard Pons-Moll , Christian Theobalt

We present the first marker-less approach for temporally coherent 3D performance capture of a human with general clothing from monocular video. Our approach reconstructs articulated human skeleton motion as well as medium-scale non-rigid…

Computer Vision and Pattern Recognition · Computer Science 2018-02-26 Weipeng Xu , Avishek Chatterjee , Michael Zollhöfer , Helge Rhodin , Dushyant Mehta , Hans-Peter Seidel , Christian Theobalt

Accurate reconstruction and tracking of dynamic human faces from image sequences is challenging because non-rigid deformations, expression changes, and viewpoint variations occur simultaneously, creating significant ambiguity in geometry…

Computer Vision and Pattern Recognition · Computer Science 2026-04-22 Umut Kocasari , Simon Giebenhain , Richard Shaw , Matthias Nießner

Human volumetric capture is a long-standing topic in computer vision and computer graphics. Although high-quality results can be achieved using sophisticated off-line systems, real-time human volumetric capture of complex scenarios,…

Computer Vision and Pattern Recognition · Computer Science 2021-05-07 Tao Yu , Zerong Zheng , Kaiwen Guo , Pengpeng Liu , Qionghai Dai , Yebin Liu

Human-centric volumetric videos offer immersive free-viewpoint experiences, yet existing methods focus either on replaying general dynamic scenes or animating human avatars, limiting their ability to re-perform general dynamic scenes. In…

Computer Vision and Pattern Recognition · Computer Science 2025-03-18 Yuheng Jiang , Zhehao Shen , Chengcheng Guo , Yu Hong , Zhuo Su , Yingliang Zhang , Marc Habermann , Lan Xu

Video-based human motion transfer creates video animations of humans following a source motion. Current methods show remarkable results for tightly-clad subjects. However, the lack of temporally consistent handling of plausible clothing…

Computer Vision and Pattern Recognition · Computer Science 2020-12-22 Moritz Kappel , Vladislav Golyanik , Mohamed Elgharib , Jann-Ole Henningson , Hans-Peter Seidel , Susana Castillo , Christian Theobalt , Marcus Magnor

Combining sparse IMUs and a monocular camera is a new promising setting to perform real-time human motion capture. This paper proposes a diffusion-based solution to learn human motion priors and fuse the two modalities of signals together…

Computer Vision and Pattern Recognition · Computer Science 2025-08-11 Shaohua Pan , Xinyu Yi , Yan Zhou , Weihua Jian , Yuan Zhang , Pengfei Wan , Feng Xu

We introduce a method to generate temporally coherent human animation from a single image, a video, or a random noise. This problem has been formulated as modeling of an auto-regressive generation, i.e., to regress past frames to decode…

Computer Vision and Pattern Recognition · Computer Science 2024-03-25 Tserendorj Adiya , Jae Shin Yoon , Jungeun Lee , Sanghun Kim , Hwasup Lim

There has been rapid progress recently on 3D human rendering, including novel view synthesis and pose animation, based on the advances of neural radiance fields (NeRF). However, most existing methods focus on person-specific training and…

Computer Vision and Pattern Recognition · Computer Science 2022-07-28 Xiangjun Gao , Jiaolong Yang , Jongyoo Kim , Sida Peng , Zicheng Liu , Xin Tong

Stochastic Human Motion Prediction (HMP) aims to predict multiple possible future human pose sequences from observed ones. Most prior works learn motion distributions through encoding-decoding in the latent space, which does not preserve…

Computer Vision and Pattern Recognition · Computer Science 2024-08-20 Jiarui Sun , Girish Chowdhary

Recent advancements in personalizing text-to-image (T2I) diffusion models have shown the capability to generate images based on personalized visual concepts using a limited number of user-provided examples. However, these models often…

Computer Vision and Pattern Recognition · Computer Science 2024-02-20 Yan Hong , Jianfu Zhang

Online novel view synthesis remains challenging, requiring robust scene reconstruction from sequential, often unposed, observations. We present ReCoSplat, an autoregressive feed-forward Gaussian Splatting model supporting posed or unposed…

Computer Vision and Pattern Recognition · Computer Science 2026-03-11 Freeman Cheng , Botao Ye , Xueting Li , Junqi You , Fangneng Zhan , Ming-Hsuan Yang
‹ Prev 1 2 3 10 Next ›