English
Related papers

Related papers: RHINO: Reconstructing Human Interactions with Nove…

200 papers

We present a method for predicting dense depth in scenarios where both a monocular camera and people in the scene are freely moving. Existing methods for recovering depth for dynamic, non-rigid objects from monocular video impose strong…

Computer Vision and Pattern Recognition · Computer Science 2019-04-26 Zhengqi Li , Tali Dekel , Forrester Cole , Richard Tucker , Noah Snavely , Ce Liu , William T. Freeman

We present Knowledge NeRF to synthesize novel views for dynamic scenes. Reconstructing dynamic 3D scenes from few sparse views and rendering them from arbitrary perspectives is a challenging problem with applications in various domains.…

Computer Vision and Pattern Recognition · Computer Science 2024-04-09 Wenxiao Cai , Xinyue Lei , Xinyu He , Junming Leo Chen , Yangang Wang

We present a unified framework for understanding 3D hand and object interactions in raw image sequences from egocentric RGB cameras. Given a single RGB image, our model jointly estimates the 3D hand and object poses, models their…

Computer Vision and Pattern Recognition · Computer Science 2019-04-11 Bugra Tekin , Federica Bogo , Marc Pollefeys

Human interaction recognition is a challenging problem in computer vision and has been researched over the years due to its important applications. With the development of deep models for the human pose estimation problem, this work aims to…

Computer Vision and Pattern Recognition · Computer Science 2016-12-14 Marcel Sheeny de Moraes , Sankha Mukherjee , Neil M Robertson

We introduce an active 3D reconstruction method which integrates visual perception, robot-object interaction, and 3D scanning to recover both the exterior and interior, i.e., unexposed, geometries of a target 3D object. Unlike other works…

Computer Vision and Pattern Recognition · Computer Science 2023-10-24 Zihao Yan , Fubao Su , Mingyang Wang , Ruizhen Hu , Hao Zhang , Hui Huang

Reconstructing 3D Radiance Field (RF) scenes through opaque obstacles is a long-standing goal, yet it is fundamentally constrained by a laborious data acquisition process requiring thousands of static measurements, which treats human motion…

Networking and Internet Architecture · Computer Science 2025-11-24 Yiheng Bian , Zechen Li , Lanqing Yang , Hao Pan , Yezhou Wang , Longyuan Ge , Jeffery Wu , Ruiheng Liu , Yongjian Fu , Yichao chen , Guangtao xue

Recognizing human actions in untrimmed videos is an important challenging task. An effective 3D motion representation and a powerful learning model are two key factors influencing recognition performance. In this paper we introduce a new…

Computer Vision and Pattern Recognition · Computer Science 2018-12-31 Huy-Hieu Pham , Louahdi Khoudour , Alain Crouzil , Pablo Zegers , Sergio A. Velastin

The ubiquity of monocular videos capturing daily hand-object interactions presents a valuable resource for embodied intelligence. While 3D hand reconstruction from in-the-wild videos has seen significant progress, reconstructing the…

Computer Vision and Pattern Recognition · Computer Science 2026-02-09 Yuantao Chen , Jiahao Chang , Chongjie Ye , Chaoran Zhang , Zhaojie Fang , Chenghong Li , Xiaoguang Han

Digital human motion synthesis is a vibrant research field with applications in movies, AR/VR, and video games. Whereas methods were proposed to generate natural and realistic human motions, most only focus on modeling humans and largely…

Computer Vision and Pattern Recognition · Computer Science 2023-11-07 Quanzhou Li , Jingbo Wang , Chen Change Loy , Bo Dai

Analyzing the interactions between humans and objects from a video includes identification of the relationships between humans and the objects present in the video. It can be thought of as a specialized version of Visual Relationship…

Computer Vision and Pattern Recognition · Computer Science 2020-12-18 Sai Praneeth Reddy Sunkesula , Rishabh Dabral , Ganesh Ramakrishnan

A long-standing challenge in scene analysis is the recovery of scene arrangements under moderate to heavy occlusion, directly from monocular video. While the problem remains a subject of active research, concurrent advances have been made…

Graphics · Computer Science 2019-07-19 Aron Monszpart , Paul Guerrero , Duygu Ceylan , Ersin Yumer , Niloy J. Mitra

We present Non-Rigid Neural Radiance Fields (NR-NeRF), a reconstruction and novel view synthesis approach for general non-rigid dynamic scenes. Our approach takes RGB images of a dynamic scene as input (e.g., from a monocular video…

Computer Vision and Pattern Recognition · Computer Science 2021-08-24 Edgar Tretschk , Ayush Tewari , Vladislav Golyanik , Michael Zollhöfer , Christoph Lassner , Christian Theobalt

Joint reconstruction of 3D human and object from a single image is an active research area, with pivotal applications in robotics and digital content creation. Despite recent advances, existing approaches suffer from two fundamental…

Computer Vision and Pattern Recognition · Computer Science 2026-02-24 Hyeongjin Nam , Daniel Sungho Jung , Kyoung Mu Lee

Recovering high-quality 3D human motion in complex scenes from monocular videos is important for many applications, ranging from AR/VR to robotics. However, capturing realistic human-scene interactions, while dealing with occlusions and…

Computer Vision and Pattern Recognition · Computer Science 2021-08-25 Siwei Zhang , Yan Zhang , Federica Bogo , Marc Pollefeys , Siyu Tang

Reconstructing metrically accurate humans and their surrounding scenes from a single image is crucial for virtual reality, robotics, and comprehensive 3D scene understanding. However, existing methods struggle with depth ambiguity,…

Computer Vision and Pattern Recognition · Computer Science 2025-10-14 Pradyumna Yalandur Muralidhar , Yuxuan Xue , Xianghui Xie , Margaret Kostyrko , Gerard Pons-Moll

Human-Object Interaction (HOI) video reenactment with realistic motion remains a frontier in expressive digital human creation. Existing approaches primarily handle simple image-plane motion (e.g., in-plane translations), struggling with…

Computer Vision and Pattern Recognition · Computer Science 2026-03-17 Jinguang Tong , Jinbo Wu , Kaisiyuan Wang , Zhelun Shen , Xuan Huang , Mochu Xiang , Xuesong Li , Yingying Li , Haocheng Feng , Chen Zhao , Hang Zhou , Wei He , Chuong Nguyen , Jingdong Wang , Hongdong Li

Human motion synthesis is an important problem with applications in graphics, gaming and simulation environments for robotics. Existing methods require accurate motion capture data for training, which is costly to obtain. Instead, we…

Computer Vision and Pattern Recognition · Computer Science 2022-08-15 Kevin Xie , Tingwu Wang , Umar Iqbal , Yunrong Guo , Sanja Fidler , Florian Shkurti

Reconstruction of the soft tissues in robotic surgery from endoscopic stereo videos is important for many applications such as intra-operative navigation and image-guided robotic surgery automation. Previous works on this task mainly rely…

Computer Vision and Pattern Recognition · Computer Science 2022-07-01 Yuehao Wang , Yonghao Long , Siu Hin Fan , Qi Dou

The reconstruction of multi-layer 3D garments typically requires expensive multi-view capture setups and specialized 3D editing efforts. To support the creation of life-like clothed human avatars, we introduce ReMu for reconstructing…

Graphics · Computer Science 2025-08-05 Onat Vuran , Hsuan-I Ho

Learning to understand dynamic 3D scenes from imagery is crucial for applications ranging from robotics to scene reconstruction. Yet, unlike other problems where large-scale supervised training has enabled rapid progress, directly…

Computer Vision and Pattern Recognition · Computer Science 2025-05-01 Linyi Jin , Richard Tucker , Zhengqi Li , David Fouhey , Noah Snavely , Aleksander Holynski