English
Related papers

Related papers: Monocular Models are Strong Learners for Multi-Vie…

200 papers

The end-to-end Human Mesh Recovery (HMR) approach has been successfully used for 3D body reconstruction. However, most HMR-based frameworks reconstruct human body by directly learning mesh parameters from images or videos, while lacking…

Computer Vision and Pattern Recognition · Computer Science 2021-03-19 Tianyu Luan , Yali Wang , Junhao Zhang , Zhe Wang , Zhipeng Zhou , Yu Qiao

Monocular image-based 3D reconstruction of faces is a long-standing problem in computer vision. Since image data is a 2D projection of a 3D face, the resulting depth ambiguity makes the problem ill-posed. Most existing methods rely on…

Computer Vision and Pattern Recognition · Computer Science 2019-04-10 Ayush Tewari , Florian Bernard , Pablo Garrido , Gaurav Bharaj , Mohamed Elgharib , Hans-Peter Seidel , Patrick Pérez , Michael Zollhöfer , Christian Theobalt

In this paper, we present a novel framework designed to reconstruct long-sequence 3D human motion in the world coordinates from in-the-wild videos with multiple shot transitions. Such long-sequence in-the-wild motions are highly valuable to…

Computer Vision and Pattern Recognition · Computer Science 2025-03-11 Yuhong Zhang , Guanlin Wu , Ling-Hao Chen , Zhuokai Zhao , Jing Lin , Xiaoke Jiang , Jiamin Wu , Zhuoheng Li , Hao Frank Yang , Haoqian Wang , Lei Zhang

Single-image human mesh recovery provides a compact 3D, person-centric representation that supports analysis, animation, AR and VR, rehabilitation, and human-computer interaction. However, prevailing systems impose an intact-limb prior and…

Computer Vision and Pattern Recognition · Computer Science 2026-05-01 Jiaying Ying , Heming Du , Kaihao Zhang , Sean M. Tweedy , Xin Yu

Despite recent advancements in the Large Reconstruction Model (LRM) demonstrating impressive results, when extending its input from single image to multiple images, it exhibits inefficiencies, subpar geometric and texture quality, as well…

Computer Vision and Pattern Recognition · Computer Science 2024-12-03 Mengfei Li , Xiaoxiao Long , Yixun Liang , Weiyu Li , Yuan Liu , Peng Li , Wenhan Luo , Wenping Wang , Yike Guo

The default strategy for training single-view Large Reconstruction Models (LRMs) follows the fully supervised route using large-scale datasets of synthetic 3D assets or multi-view captures. Although these resources simplify the training…

Computer Vision and Pattern Recognition · Computer Science 2024-06-13 Hanwen Jiang , Qixing Huang , Georgios Pavlakos

Multi-view triangulation is the gold standard for 3D reconstruction from 2D correspondences given known calibration and sufficient views. However in practice, expensive multi-view setups -- involving tens sometimes hundreds of cameras --…

Computer Vision and Pattern Recognition · Computer Science 2022-10-06 Mosam Dabhi , Chaoyang Wang , Kunal Saluja , Laszlo Jeni , Ian Fasel , Simon Lucey

This paper presents a novel framework to recover detailed human body shapes from a single image. It is a challenging task due to factors such as variations in human shapes, body poses, and viewpoints. Prior methods typically attempt to…

Computer Vision and Pattern Recognition · Computer Science 2019-05-14 Hao Zhu , Xinxin Zuo , Sen Wang , Xun Cao , Ruigang Yang

Reconstructing physically plausible human motion from monocular videos remains a challenging problem in computer vision and graphics. Existing methods primarily focus on kinematics-based pose estimation, often leading to unrealistic results…

Computer Vision and Pattern Recognition · Computer Science 2025-10-06 Qiao Feng , Yiming Huang , Yufu Wang , Jiatao Gu , Lingjie Liu

Undoubtedly, high-fidelity 3D hair is crucial for achieving realism, artistic expression, and immersion in computer graphics. While existing 3D hair modeling methods have achieved impressive performance, the challenge of achieving…

Computer Vision and Pattern Recognition · Computer Science 2024-03-28 Keyu Wu , Lingchen Yang , Zhiyi Kuang , Yao Feng , Xutao Han , Yuefan Shen , Hongbo Fu , Kun Zhou , Youyi Zheng

Recent advances in 3D foundation models have led to growing interest in reconstructing humans and their surrounding environments. However, most existing approaches focus on monocular inputs, and extending them to multi-view settings…

Computer Vision and Pattern Recognition · Computer Science 2026-03-19 Sangmin Kim , Minhyuk Hwang , Geonho Cha , Dongyoon Wee , Jaesik Park

While recent advancements in animatable human rendering have achieved remarkable results, they require test-time optimization for each subject which can be a significant limitation for real-world applications. To address this, we tackle the…

Computer Vision and Pattern Recognition · Computer Science 2024-04-23 Mana Masuda , Jinhyung Park , Shun Iwase , Rawal Khirodkar , Kris Kitani

Remarkable strides have been made in reconstructing static scenes or human bodies from monocular videos. Yet, the two problems have largely been approached independently, without much synergy. Most visual SLAM methods can only reconstruct…

Computer Vision and Pattern Recognition · Computer Science 2024-05-24 Yizhou Zhao , Tuanfeng Y. Wang , Bhiksha Raj , Min Xu , Jimei Yang , Chun-Hao Paul Huang

Monocular video human mesh recovery faces fundamental challenges in maintaining metric consistency and temporal stability due to inherent depth ambiguities and scale uncertainties. While existing methods rely primarily on RGB features and…

Computer Vision and Pattern Recognition · Computer Science 2026-02-05 Jiaxin Cen , Xudong Mao , Guanghui Yue , Wei Zhou , Ruomei Wang , Fan Zhou , Baoquan Zhao

This paper reports on a novel template-free monocular non-rigid surface reconstruction approach. Existing techniques using motion and deformation cues rely on multiple prior assumptions, are often computationally expensive and do not…

Computer Vision and Pattern Recognition · Computer Science 2017-10-18 Mohammad Dawud Ansari , Vladislav Golyanik , Didier Stricker

We present the first marker-less approach for temporally coherent 3D performance capture of a human with general clothing from monocular video. Our approach reconstructs articulated human skeleton motion as well as medium-scale non-rigid…

Computer Vision and Pattern Recognition · Computer Science 2018-02-26 Weipeng Xu , Avishek Chatterjee , Michael Zollhöfer , Helge Rhodin , Dushyant Mehta , Hans-Peter Seidel , Christian Theobalt

Recovering 3D human pose from multi-view imagery typically relies on precise camera calibration, which is often unavailable in real-world scenarios, thereby severely limiting the applicability of existing methods. To overcome this…

Computer Vision and Pattern Recognition · Computer Science 2026-04-28 Xiaolin Qin , Qianlei Wang , Jiacen Liu , Chaoning Zhang , Fei Zhu , Zhang Yi

We consider the problem of estimating a parametric model of 3D human mesh from a single image. While there has been substantial recent progress in this area with direct regression of model parameters, these methods only implicitly exploit…

Computer Vision and Pattern Recognition · Computer Science 2020-07-15 Georgios Georgakis , Ren Li , Srikrishna Karanam , Terrence Chen , Jana Kosecka , Ziyan Wu

Human performance capture is a highly important computer vision problem with many applications in movie production and virtual/augmented reality. Many previous performance capture approaches either required expensive multi-view setups or…

Computer Vision and Pattern Recognition · Computer Science 2021-11-23 Marc Habermann , Weipeng Xu , Michael Zollhoefer , Gerard Pons-Moll , Christian Theobalt

Reconstructing textured 3D human models from a single image is fundamental for AR/VR and digital human applications. However, existing methods mostly focus on single individuals and thus fail in multi-human scenes, where naive composition…

Computer Vision and Pattern Recognition · Computer Science 2026-04-08 Gwanghyun Kim , Junghun James Kim , Suh Yoon Jeon , Jason Park , Se Young Chun