English
Related papers

Related papers: HumanRAM: Feed-forward Human Reconstruction and An…

200 papers

We introduce the Deformable Gaussian Splats Large Reconstruction Model (DGS-LRM), the first feed-forward method predicting deformable 3D Gaussian splats from a monocular posed video of any dynamic scene. Feed-forward scene reconstruction…

3D human body reconstruction has been a challenge in the field of computer vision. Previous methods are often time-consuming and difficult to capture the detailed appearance of the human body. In this paper, we propose a new method called…

Computer Vision and Pattern Recognition · Computer Science 2024-03-19 Mingjin Chen , Junhao Chen , Xiaojun Ye , Huan-ang Gao , Xiaoxue Chen , Zhaoxin Fan , Hao Zhao

Recent image-to-3D reconstruction models have greatly advanced geometry generation, but they still struggle to faithfully generate realistic appearance. To address this, we introduce ARM, a novel method that reconstructs high-quality 3D…

Computer Vision and Pattern Recognition · Computer Science 2024-11-19 Xiang Feng , Chang Yu , Zoubin Bi , Yintong Shang , Feng Gao , Hongzhi Wu , Kun Zhou , Chenfanfu Jiang , Yin Yang

Feed-forward 3D generative models like the Large Reconstruction Model (LRM) have demonstrated exceptional generation speed. However, the transformer-based methods do not leverage the geometric priors of the triplane component in their…

Computer Vision and Pattern Recognition · Computer Science 2024-03-11 Zhengyi Wang , Yikai Wang , Yifei Chen , Chendong Xiang , Shuo Chen , Dajiang Yu , Chongxuan Li , Hang Su , Jun Zhu

We propose RelitLRM, a Large Reconstruction Model (LRM) for generating high-quality Gaussian splatting representations of 3D objects under novel illuminations from sparse (4-8) posed images captured under unknown static lighting. Unlike…

Computer Vision and Pattern Recognition · Computer Science 2024-10-11 Tianyuan Zhang , Zhengfei Kuang , Haian Jin , Zexiang Xu , Sai Bi , Hao Tan , He Zhang , Yiwei Hu , Milos Hasan , William T. Freeman , Kai Zhang , Fujun Luan

The integration of neural rendering and the SLAM system recently showed promising results in joint localization and photorealistic view reconstruction. However, existing methods, fully relying on implicit representations, are so…

Computer Vision and Pattern Recognition · Computer Science 2024-04-09 Huajian Huang , Longwei Li , Hui Cheng , Sai-Kit Yeung

We present deep neural network methodology to reconstruct the 3d pose and shape of people, given an input RGB image. We rely on a recently introduced, expressivefull body statistical 3d human model, GHUM, trained end-to-end, and learn to…

Computer Vision and Pattern Recognition · Computer Science 2021-06-15 Andrei Zanfir , Eduard Gabriel Bazavan , Mihai Zanfir , William T. Freeman , Rahul Sukthankar , Cristian Sminchisescu

In this study, we introduce a methodology for human image animation by leveraging a 3D human parametric model within a latent diffusion framework to enhance shape alignment and motion guidance in curernt human generative techniques. The…

Computer Vision and Pattern Recognition · Computer Science 2024-06-04 Shenhao Zhu , Junming Leo Chen , Zuozhuo Dai , Qingkun Su , Yinghui Xu , Xun Cao , Yao Yao , Hao Zhu , Siyu Zhu

Estimation of 3D human pose from monocular image has gained considerable attention, as a key step to several human-centric applications. However, generalizability of human pose estimation models developed using supervision on large-scale…

Computer Vision and Pattern Recognition · Computer Science 2020-06-26 Jogendra Nath Kundu , Siddharth Seth , Rahul M , Mugalodi Rakesh , R. Venkatesh Babu , Anirban Chakraborty

In this paper, we aim to address the challenge of novel view rendering of human performers who wear clothes with complex texture patterns using a sparse set of camera views. Although some recent works have achieved remarkable rendering…

Computer Vision and Pattern Recognition · Computer Science 2023-10-24 Tiansong Zhou , Jing Huang , Tao Yu , Ruizhi Shao , Kun Li

In this paper, we propose ARCH (Animatable Reconstruction of Clothed Humans), a novel end-to-end framework for accurate reconstruction of animation-ready 3D clothed humans from a monocular image. Existing approaches to digitize 3D humans…

Graphics · Computer Science 2020-04-14 Zeng Huang , Yuanlu Xu , Christoph Lassner , Hao Li , Tony Tung

In this paper, we present WonderHuman to reconstruct dynamic human avatars from a monocular video for high-fidelity novel view synthesis. Previous dynamic human avatar reconstruction methods typically require the input video to have full…

Computer Vision and Pattern Recognition · Computer Science 2026-01-08 Zilong Wang , Zhiyang Dou , Yuan Liu , Cheng Lin , Xiao Dong , Yunhui Guo , Chenxu Zhang , Xin Li , Wenping Wang , Xiaohu Guo

Human mesh recovery (HMR) models 3D human body from monocular videos, with recent works extending it to world-coordinate human trajectory and motion reconstruction. However, most existing methods remain offline, relying on future frames or…

Computer Vision and Pattern Recognition · Computer Science 2026-03-19 Yiwen Zhao , Ce Zheng , Yufu Wang , Hsueh-Han Daniel Yang , Liting Wen , Laszlo A. Jeni

End-to-end human animation, such as audio-driven talking human generation, has undergone notable advancements in the recent few years. However, existing methods still struggle to scale up as large general video generation models, limiting…

Computer Vision and Pattern Recognition · Computer Science 2025-07-01 Gaojie Lin , Jianwen Jiang , Jiaqi Yang , Zerong Zheng , Chao Liang

In this paper, we propose PixelHuman, a novel human rendering model that generates animatable human scenes from a few images of a person with unseen identity, views, and poses. Previous work have demonstrated reasonable performance in novel…

Computer Vision and Pattern Recognition · Computer Science 2023-07-19 Gyumin Shim , Jaeseong Lee , Junha Hyung , Jaegul Choo

Video-based human motion transfer creates video animations of humans following a source motion. Current methods show remarkable results for tightly-clad subjects. However, the lack of temporally consistent handling of plausible clothing…

Computer Vision and Pattern Recognition · Computer Science 2020-12-22 Moritz Kappel , Vladislav Golyanik , Mohamed Elgharib , Jann-Ole Henningson , Hans-Peter Seidel , Susana Castillo , Christian Theobalt , Marcus Magnor

3D reconstruction and view synthesis are foundational problems in computer vision, graphics, and immersive technologies such as augmented reality (AR), virtual reality (VR), and digital twins. Traditional methods rely on computationally…

Animating virtual avatars with free-view control is crucial for various applications like virtual reality and digital entertainment. Previous studies have attempted to utilize the representation power of the neural radiance field (NeRF) to…

Computer Vision and Pattern Recognition · Computer Science 2023-04-05 Zhengming Yu , Wei Cheng , Xian Liu , Wayne Wu , Kwan-Yee Lin

In this paper, we propose a novel approach to reconstruct 3D human body shapes based on a sparse set of RGBD frames using a single RGBD camera. We specifically focus on the realistic settings where human subjects move freely during the…

Computer Vision and Pattern Recognition · Computer Science 2020-06-16 Xinxin Zuo , Sen Wang , Jiangbin Zheng , Weiwei Yu , Minglun Gong , Ruigang Yang , Li Cheng

Human motion recovery for real-world interaction demands both precise action details and metric-scale trajectories. Recovering absolute human pose from monocular input presents a viable solution, but faces two main challenges: (1) models'…

Computer Vision and Pattern Recognition · Computer Science 2026-03-16 Zhumei Wang , Zechen Hu , Ruoxi Guo , Huaijin Pi , Ziyong Feng , Liang Zhang , Mingtao Pei , Siyuan Huang