English
Related papers

Related papers: Reconstructing People, Places, and Cameras

200 papers

We present DuoMo, a generative method that recovers human motion in world-space coordinates from unconstrained videos with noisy or incomplete observations. Reconstructing such motion requires solving a fundamental trade-off: generalizing…

Computer Vision and Pattern Recognition · Computer Science 2026-03-04 Yufu Wang , Evonne Ng , Soyong Shin , Rawal Khirodkar , Yuan Dong , Zhaoen Su , Jinhyung Park , Kris Kitani , Alexander Richard , Fabian Prada , Michael Zollhofer

Structure from motion (SfM) enables us to reconstruct a scene via casual capture from cameras at different viewpoints, and novel view synthesis (NVS) allows us to render a captured scene from a new viewpoint. Both are hard with casual…

We present a method for generating a full 360{\deg} orbit video around a person from a single input image. Existing methods typically adapt image-based diffusion models for multi-view synthesis, but yield inconsistent results across views…

Computer Vision and Pattern Recognition · Computer Science 2026-03-02 Keito Suzuki , Kunyao Chen , Lei Wang , Bang Du , Runfa Blark Li , Peng Liu , Ning Bi , Truong Nguyen

In this paper, we propose an efficient human pose estimation network -- SFM (slender fusion model) by fusing multi-level features and adding lightweight attention blocks -- HSA (High-Level Spatial Attention). Many existing methods on…

Computer Vision and Pattern Recognition · Computer Science 2021-07-30 Zhiyuan Ren , Yaohai Zhou , Yizhe Chen , Ruisong Zhou , Yayu Gao

Most of the previous 3D human pose estimation work relied on the powerful memory capability of the network to obtain suitable 2D-3D mappings from the training data. Few works have studied the modeling of human posture deformation in motion.…

Computer Vision and Pattern Recognition · Computer Science 2023-08-22 Haorui Ji , Hui Deng , Yuchao Dai , Hongdong Li

We propose a scalable neural network framework to reconstruct the 3D mesh of a human body from multi-view images, in the subspace of the SMPL model. Use of multi-view images can significantly reduce the projection ambiguity of the problem,…

Computer Vision and Pattern Recognition · Computer Science 2019-08-27 Junbang Liang , Ming C. Lin

Reconstructing dynamic scenes with multiple interacting humans and objects from sparse-view inputs is a critical yet challenging task, essential for creating high-fidelity digital twins for robotics and VR/AR. This problem, which we term…

Computer Vision and Pattern Recognition · Computer Science 2026-04-09 Weiquan Wang , Jun Xiao , Feifei Shao , Yi Yang , Yueting Zhuang , Long Chen

Reconstructing accurate 3D human meshes in the world coordinate system from in-the-wild images remains challenging due to the lack of camera rotation information. While existing methods achieve promising results in the camera coordinate…

Computer Vision and Pattern Recognition · Computer Science 2025-12-18 Changhai Ma , Ziyu Wu , Yunkang Zhang , Qijun Ying , Boyan Liu , Xiaohui Cai

Human body restoration, as a specific application of image restoration, is widely applied in practice and plays a vital role across diverse fields. However, thorough research remains difficult, particularly due to the lack of benchmark…

Computer Vision and Pattern Recognition · Computer Science 2026-02-06 Jue Gong , Jingkai Wang , Zheng Chen , Xing Liu , Hong Gu , Yulun Zhang , Xiaokang Yang

We present TexMesh, a novel approach to reconstruct detailed human meshes with high-resolution full-body texture from RGB-D video. TexMesh enables high quality free-viewpoint rendering of humans. Given the RGB frames, the captured…

Computer Vision and Pattern Recognition · Computer Science 2020-09-22 Tiancheng Zhi , Christoph Lassner , Tony Tung , Carsten Stoll , Srinivasa G. Narasimhan , Minh Vo

Non-rigid structure-from-motion (NRSfM) has so far been mostly studied for recovering 3D structure of a single non-rigid/deforming object. To handle the real world challenging multiple deforming objects scenarios, existing methods either…

Computer Vision and Pattern Recognition · Computer Science 2017-05-16 Suryansh Kumar , Yuchao Dai , Hongdong Li

Segmenting humans in 3D indoor scenes has become increasingly important with the rise of human-centered robotics and AR/VR applications. To this end, we propose the task of joint 3D human semantic segmentation, instance segmentation and…

Computer Vision and Pattern Recognition · Computer Science 2023-08-21 Ayça Takmaz , Jonas Schult , Irem Kaftan , Mertcan Akçay , Bastian Leibe , Robert Sumner , Francis Engelmann , Siyu Tang

Due to visual ambiguities and inter-person occlusions, existing human pose estimation methods cannot recover plausible close interactions from in-the-wild videos. Even state-of-the-art large foundation models~(\eg, SAM) cannot accurately…

Computer Vision and Pattern Recognition · Computer Science 2025-07-04 Buzhen Huang , Chen Li , Chongyang Xu , Dongyue Lu , Jinnan Chen , Yangang Wang , Gim Hee Lee

Modeling and capturing the 3D spatial arrangement of the human and the object is the key to perceiving 3D human-object interaction from monocular images. In this work, we propose to use the Human-Object Offset between anchors which are…

Computer Vision and Pattern Recognition · Computer Science 2024-07-31 Chaofan Huo , Ye Shi , Yuexin Ma , Lan Xu , Jingyi Yu , Jingya Wang

This paper addresses the challenges of estimating a continuous-time human motion field from a stream of events. Existing Human Mesh Recovery (HMR) methods rely predominantly on frame-based approaches, which are prone to aliasing and…

Computer Vision and Pattern Recognition · Computer Science 2024-12-03 Ziyun Wang , Ruijun Zhang , Zi-Yan Liu , Yufu Wang , Kostas Daniilidis

Multi-view 3D reconstruction, namely, structure-from-motion followed by multi-view stereo, is a fundamental component of 3D computer vision. In general, multi-view 3D reconstruction suffers from an unknown scale ambiguity unless a reference…

Computer Vision and Pattern Recognition · Computer Science 2026-05-18 Lilika Makabe , Kohei Ashida , Hiroaki Santo , Fumio Okura , Yasuyuki Matsushita

The paper introduces an accurate solution to dense orthographic Non-Rigid Structure from Motion (NRSfM) in scenarios with severe occlusions or, likewise, inaccurate correspondences. We integrate a shape prior term into variational…

Computer Vision and Pattern Recognition · Computer Science 2017-12-21 Vladislav Golyanik , Torben Fetzer , Didier Stricker

In this paper, we present a complete refractive Structure-from-Motion (RSfM) framework for underwater 3D reconstruction using refractive camera setups (for both, flat- and dome-port underwater housings). Despite notable achievements in…

Computer Vision and Pattern Recognition · Computer Science 2025-01-16 Mengkun She , Felix Seegräber , David Nakath , Kevin Köser

Human pose estimation is one of the key problems in computer vision that has been studied in the recent years. The significance of human pose estimation is in the higher level tasks of understanding human actions applications such as…

Computer Vision and Pattern Recognition · Computer Science 2014-08-26 Oinam Binarani Devi , Nissi S. Paul , Y. Jayanta Singh

Despite significant progress in single image-based 3D human mesh recovery, accurately and smoothly recovering 3D human motion from a video remains challenging. Existing video-based methods generally recover human mesh by estimating the…

Computer Vision and Pattern Recognition · Computer Science 2023-08-22 Yingxuan You , Hong Liu , Ti Wang , Wenhao Li , Runwei Ding , Xia Li