English
Related papers

Related papers: RHINO: Reconstructing Human Interactions with Nove…

200 papers

We present a novel method to learn temporally consistent 3D reconstruction of clothed people from a monocular video. Recent methods for 3D human reconstruction from monocular video using volumetric, implicit or parametric human shape…

Computer Vision and Pattern Recognition · Computer Science 2021-04-20 Akin Caliskan , Armin Mustafa , Adrian Hilton

Generalized robots must learn from diverse, large-scale human-object interactions (HOI) to operate robustly in the real world. Monocular internet videos offer a nearly limitless and readily available source of data, capturing an…

Computer Vision and Pattern Recognition · Computer Science 2026-03-20 Boran Wen , Ye Lu , Sirui Wang , Keyan Wan , Jiahong Zhou , Junxuan Liang , Xinpeng Liu , Bang Xiao , Ruiyang Liu , Yong-Lu Li

Reconstructing dynamic humans together with static scenes from monocular videos remains difficult, especially under fast motion, where RGB frames suffer from motion blur. Event cameras exhibit distinct advantages, e.g., microsecond temporal…

Computer Vision and Pattern Recognition · Computer Science 2025-09-24 Xiaoting Yin , Hao Shi , Kailun Yang , Jiajun Zhai , Shangwei Guo , Lin Wang , Kaiwei Wang

We introduce HOSNeRF, a novel 360{\deg} free-viewpoint rendering method that reconstructs neural radiance fields for dynamic human-object-scene from a single monocular in-the-wild video. Our method enables pausing the video at any frame and…

Computer Vision and Pattern Recognition · Computer Science 2023-04-25 Jia-Wei Liu , Yan-Pei Cao , Tianyuan Yang , Eric Zhongcong Xu , Jussi Keppo , Ying Shan , Xiaohu Qie , Mike Zheng Shou

Reconstructing 3D clothed human involves creating a detailed geometry of individuals in clothing, with applications ranging from virtual try-on, movies, to games. To enable practical and widespread applications, recent advances propose to…

Computer Vision and Pattern Recognition · Computer Science 2024-04-22 Yifan Yang , Dong Liu , Shuhai Zhang , Zeshuai Deng , Zixiong Huang , Mingkui Tan

Inferring human-scene contact (HSC) is the first step toward understanding how humans interact with their surroundings. While detecting 2D human-object interaction (HOI) and reconstructing 3D human pose and shape (HPS) have enjoyed…

Computer Vision and Pattern Recognition · Computer Science 2022-06-22 Chun-Hao P. Huang , Hongwei Yi , Markus Höschle , Matvey Safroshkin , Tsvetelina Alexiadis , Senya Polikovsky , Daniel Scharstein , Michael J. Black

Animating realistic character interactions with the surrounding environment is important for autonomous agents in gaming, AR/VR, and robotics. However, current methods for human motion reconstruction struggle with accurately placing humans…

Computer Vision and Pattern Recognition · Computer Science 2025-10-20 Joshua Li , Brendan Chharawala , Chang Shu , Xue Bin Peng , Pengcheng Xi

Articulated 3D reconstruction has valuable applications in various domains, yet it remains costly and demands intensive work from domain experts. Recent advancements in template-free learning methods show promising results with monocular…

Computer Vision and Pattern Recognition · Computer Science 2023-12-11 Tao Tu , Ming-Feng Li , Chieh Hubert Lin , Yen-Chi Cheng , Min Sun , Ming-Hsuan Yang

We propose Dyn-HaMR, to the best of our knowledge, the first approach to reconstruct 4D global hand motion from monocular videos recorded by dynamic cameras in the wild. Reconstructing accurate 3D hand meshes from monocular videos is a…

Computer Vision and Pattern Recognition · Computer Science 2025-06-03 Zhengdi Yu , Stefanos Zafeiriou , Tolga Birdal

This paper introduces a novel pipeline to reconstruct the geometry of interacting multi-person in clothing on a globally coherent scene space from a single image. The main challenge arises from the occlusion: a part of a human body is not…

Computer Vision and Pattern Recognition · Computer Science 2024-04-03 Junuk Cha , Hansol Lee , Jaewon Kim , Nhat Nguyen Bao Truong , Jae Shin Yoon , Seungryul Baek

Neural rendering has demonstrated remarkable success in dynamic scene reconstruction. Thanks to the expressiveness of neural representations, prior works can accurately capture the motion and achieve high-fidelity reconstruction of the…

Computer Vision and Pattern Recognition · Computer Science 2024-04-05 Hengyi Wang , Jingwen Wang , Lourdes Agapito

Our work aims to reconstruct hand-object interactions from a single-view image, which is a fundamental but ill-posed task. Unlike methods that reconstruct from videos, multi-view images, or predefined 3D templates, single-view…

Computer Vision and Pattern Recognition · Computer Science 2025-12-22 Yumeng Liu , Xiaoxiao Long , Zemin Yang , Yuan Liu , Marc Habermann , Christian Theobalt , Yuexin Ma , Wenping Wang

Remarkable progress has been made in 3D reconstruction of rigid structures from a video or a collection of images. However, it is still challenging to reconstruct nonrigid structures from RGB inputs, due to its under-constrained nature.…

Computer Vision and Pattern Recognition · Computer Science 2021-05-10 Gengshan Yang , Deqing Sun , Varun Jampani , Daniel Vlasic , Forrester Cole , Huiwen Chang , Deva Ramanan , William T. Freeman , Ce Liu

We focus on the task of estimating a physically plausible articulated human motion from monocular video. Existing approaches that do not consider physics often produce temporally inconsistent output with motion artifacts, while…

Computer Vision and Pattern Recognition · Computer Science 2022-05-26 Erik Gärtner , Mykhaylo Andriluka , Hongyi Xu , Cristian Sminchisescu

We present a method for fast 3D reconstruction and real-time rendering of dynamic humans from monocular videos with accompanying parametric body fits. Our method can reconstruct a dynamic human in less than 3h using a single GPU, compared…

Computer Vision and Pattern Recognition · Computer Science 2023-03-22 Ignacio Rocco , Iurii Makarov , Filippos Kokkinos , David Novotny , Benjamin Graham , Natalia Neverova , Andrea Vedaldi

We introduce MotioNet, a deep neural network that directly reconstructs the motion of a 3D human skeleton from monocular video.While previous methods rely on either rigging or inverse kinematics (IK) to associate a consistent skeleton with…

Computer Vision and Pattern Recognition · Computer Science 2024-05-14 Mingyi Shi , Kfir Aberman , Andreas Aristidou , Taku Komura , Dani Lischinski , Daniel Cohen-Or , Baoquan Chen

Human relighting is a highly desirable yet challenging task. Existing works either require expensive one-light-at-a-time (OLAT) captured data using light stage or cannot freely change the viewpoints of the rendered body. In this work, we…

Computer Vision and Pattern Recognition · Computer Science 2022-09-21 Zhaoxi Chen , Ziwei Liu

Existing multi-person human reconstruction approaches mainly focus on recovering accurate poses or avoiding penetration, but overlook the modeling of close interactions. In this work, we tackle the task of reconstructing closely interactive…

Computer Vision and Pattern Recognition · Computer Science 2024-04-18 Buzhen Huang , Chen Li , Chongyang Xu , Liang Pan , Yangang Wang , Gim Hee Lee

In this paper, we present WonderHuman to reconstruct dynamic human avatars from a monocular video for high-fidelity novel view synthesis. Previous dynamic human avatar reconstruction methods typically require the input video to have full…

Computer Vision and Pattern Recognition · Computer Science 2026-01-08 Zilong Wang , Zhiyang Dou , Yuan Liu , Cheng Lin , Xiao Dong , Yunhui Guo , Chenxu Zhang , Xin Li , Wenping Wang , Xiaohu Guo

Advances in deep learning techniques have allowed recent work to reconstruct the shape of a single object given only one RBG image as input. Building on common encoder-decoder architectures for this task, we propose three extensions: (1)…

Computer Vision and Pattern Recognition · Computer Science 2020-08-06 Stefan Popov , Pablo Bauszat , Vittorio Ferrari
‹ Prev 1 4 5 6 7 8 10 Next ›