English
Related papers

Related papers: HumanRAM: Feed-forward Human Reconstruction and An…

200 papers

Recent progress in human shape learning, shows that neural implicit models are effective in generating 3D human surfaces from limited number of views, and even from a single RGB image. However, existing monocular approaches still struggle…

Computer Vision and Pattern Recognition · Computer Science 2024-03-20 Marco Pesavento , Yuanlu Xu , Nikolaos Sarafianos , Robert Maier , Ziyan Wang , Chun-Han Yao , Marco Volino , Edmond Boyer , Adrian Hilton , Tony Tung

We present a method for fast 3D reconstruction and real-time rendering of dynamic humans from monocular videos with accompanying parametric body fits. Our method can reconstruct a dynamic human in less than 3h using a single GPU, compared…

Computer Vision and Pattern Recognition · Computer Science 2023-03-22 Ignacio Rocco , Iurii Makarov , Filippos Kokkinos , David Novotny , Benjamin Graham , Natalia Neverova , Andrea Vedaldi

We propose TRAM, a two-stage method to reconstruct a human's global trajectory and motion from in-the-wild videos. TRAM robustifies SLAM to recover the camera motion in the presence of dynamic humans and uses the scene background to derive…

Computer Vision and Pattern Recognition · Computer Science 2024-09-04 Yufu Wang , Ziyun Wang , Lingjie Liu , Kostas Daniilidis

Human mesh recovery can be approached using either regression-based or optimization-based methods. Regression models achieve high pose accuracy but struggle with model-to-image alignment due to the lack of explicit 2D-3D correspondences. In…

Computer Vision and Pattern Recognition · Computer Science 2025-02-07 Chongyang Xu , Buzhen Huang , Chengfang Zhang , Ziliang Feng , Yangang Wang

In recent years, Neural Radiance Fields (NeRF) have achieved remarkable progress in dynamic human reconstruction and rendering. Part-based rendering paradigms, guided by human segmentation, allow for flexible parameter allocation based on…

Computer Vision and Pattern Recognition · Computer Science 2025-08-13 Yao Lu , Jiawei Li , Ming Jiang

Human Mesh Recovery (HMR) is an important yet challenging problem with applications across various domains including motion capture, augmented reality, and biomechanics. Accurately predicting human pose parameters from a single image…

Computer Vision and Pattern Recognition · Computer Science 2024-11-19 Jaewoo Heo , George Hu , Zeyu Wang , Serena Yeung-Levy

We present FlexNeRF, a method for photorealistic freeviewpoint rendering of humans in motion from monocular videos. Our approach works well with sparse views, which is a challenging scenario when the subject is exhibiting fast/complex…

Computer Vision and Pattern Recognition · Computer Science 2023-03-28 Vinoj Jayasundara , Amit Agrawal , Nicolas Heron , Abhinav Shrivastava , Larry S. Davis

We present a feed-forward human performance capture method that renders novel views of a performer from a monocular RGB stream. A key challenge in this setting is the lack of sufficient observations, especially for unseen regions. Assuming…

Computer Vision and Pattern Recognition · Computer Science 2026-04-01 Youngjoong Kwon , Yao He , Heejung Choi , Chen Geng , Zhengmao Liu , Jiajun Wu , Ehsan Adeli

While recent advancements in animatable human rendering have achieved remarkable results, they require test-time optimization for each subject which can be a significant limitation for real-world applications. To address this, we tackle the…

Computer Vision and Pattern Recognition · Computer Science 2024-04-23 Mana Masuda , Jinhyung Park , Shun Iwase , Rawal Khirodkar , Kris Kitani

Estimating human pose and shape from monocular images is a long-standing problem in computer vision. Since the release of statistical body models, 3D human mesh recovery has been drawing broader attention. With the same goal of obtaining…

Computer Vision and Pattern Recognition · Computer Science 2024-01-03 Yating Tian , Hongwen Zhang , Yebin Liu , Limin Wang

Animating realistic character interactions with the surrounding environment is important for autonomous agents in gaming, AR/VR, and robotics. However, current methods for human motion reconstruction struggle with accurately placing humans…

Computer Vision and Pattern Recognition · Computer Science 2025-10-20 Joshua Li , Brendan Chharawala , Chang Shu , Xue Bin Peng , Pengcheng Xi

Human reconstruction and synthesis from monocular RGB videos is a challenging problem due to clothing, occlusion, texture discontinuities and sharpness, and framespecific pose changes. Many methods employ deferred rendering, NeRFs and…

Computer Vision and Pattern Recognition · Computer Science 2023-03-16 Rohit Jena , Pratik Chaudhari , James Gee , Ganesh Iyer , Siddharth Choudhary , Brandon M. Smith

We present Large Inverse Rendering Model (LIRM), a transformer architecture that jointly reconstructs high-quality shape, materials, and radiance fields with view-dependent effects in less than a second. Our model builds upon the recent…

Computer Vision and Pattern Recognition · Computer Science 2025-04-29 Zhengqin Li , Dilin Wang , Ka Chen , Zhaoyang Lv , Thu Nguyen-Phuoc , Milim Lee , Jia-Bin Huang , Lei Xiao , Cheng Zhang , Yufeng Zhu , Carl S. Marshall , Yufeng Ren , Richard Newcombe , Zhao Dong

Animation of humanoid characters is essential in various graphics applications, but requires significant time and cost to create realistic animations. We propose an approach to synthesize 4D animated sequences of input static 3D humanoid…

Graphics · Computer Science 2025-03-21 Marc Benedí San Millán , Angela Dai , Matthias Nießner

We propose a Pose-Free Large Reconstruction Model (PF-LRM) for reconstructing a 3D object from a few unposed images even with little visual overlap, while simultaneously estimating the relative camera poses in ~1.3 seconds on a single A100…

Computer Vision and Pattern Recognition · Computer Science 2023-11-27 Peng Wang , Hao Tan , Sai Bi , Yinghao Xu , Fujun Luan , Kalyan Sunkavalli , Wenping Wang , Zexiang Xu , Kai Zhang

We present THUNDR, a transformer-based deep neural network methodology to reconstruct the 3d pose and shape of people, given monocular RGB images. Key to our methodology is an intermediate 3d marker representation, where we aim to combine…

Computer Vision and Pattern Recognition · Computer Science 2021-06-18 Mihai Zanfir , Andrei Zanfir , Eduard Gabriel Bazavan , William T. Freeman , Rahul Sukthankar , Cristian Sminchisescu

The field of image-to-video generation has made remarkable progress. However, challenges such as human limb twisting and facial distortion persist, especially when generating long videos or modeling intensive motions. Existing human image…

Computer Vision and Pattern Recognition · Computer Science 2026-05-12 Chang Liu , Mengting Chen , Yixuan Huang , Haoning Wu , Chen Ju , Shuai Xiao , Jinsong Lan , Yanfeng Wang

Recent techniques on implicit geometry representation learning and neural rendering have shown promising results for 3D clothed human reconstruction from sparse video inputs. However, it is still challenging to reconstruct detailed surface…

Computer Vision and Pattern Recognition · Computer Science 2024-04-23 Hao Wang , Qingshan Xu , Hongyuan Chen , Rui Ma

We propose RoHM, an approach for robust 3D human motion reconstruction from monocular RGB(-D) videos in the presence of noise and occlusions. Most previous approaches either train neural networks to directly regress motion in 3D or learn…

Computer Vision and Pattern Recognition · Computer Science 2024-04-16 Siwei Zhang , Bharat Lal Bhatnagar , Yuanlu Xu , Alexander Winkler , Petr Kadlecek , Siyu Tang , Federica Bogo

The default strategy for training single-view Large Reconstruction Models (LRMs) follows the fully supervised route using large-scale datasets of synthetic 3D assets or multi-view captures. Although these resources simplify the training…

Computer Vision and Pattern Recognition · Computer Science 2024-06-13 Hanwen Jiang , Qixing Huang , Georgios Pavlakos