English
Related papers

Related papers: RHINO: Reconstructing Human Interactions with Nove…

200 papers

We introduce a novel framework for reconstructing dynamic human-object interactions from monocular video that overcomes challenges associated with occlusions and temporal inconsistencies. Traditional 3D reconstruction methods typically…

Computer Vision and Pattern Recognition · Computer Science 2025-09-16 Hyungjun Doh , Dong In Lee , Seunggeun Chi , Pin-Hao Huang , Kwonjoon Lee , Sangpil Kim , Karthik Ramani

Novel-View Human Action Synthesis aims to synthesize the movement of a body from a virtual viewpoint, given a video from a real viewpoint. We present a novel 3D reasoning to synthesize the target viewpoint. We first estimate the 3D mesh of…

Computer Vision and Pattern Recognition · Computer Science 2020-10-09 Mohamed Ilyes Lakhal , Davide Boscaini , Fabio Poiesi , Oswald Lanz , Andrea Cavallaro

This paper addresses the challenge of novel-view synthesis and motion reconstruction of dynamic scenes from monocular video, which is critical for many robotic applications. Although Neural Radiance Fields (NeRF) and 3D Gaussian Splatting…

Robotics · Computer Science 2025-08-12 Xuesong Li , Lars Petersson , Vivien Rolland

We introduce the task of Reconstructing Objects along Hand Interaction Timelines (ROHIT). We first define the Hand Interaction Timeline (HIT) from a rigid object's perspective. In a HIT, an object is first static relative to the scene, then…

Computer Vision and Pattern Recognition · Computer Science 2025-12-09 Zhifan Zhu , Siddhant Bansal , Shashank Tripathi , Dima Damen

4D modeling of human-object interactions is critical for numerous applications. However, efficient volumetric capture and rendering of complex interaction scenarios, especially from sparse inputs, remain challenging. In this paper, we…

Computer Vision and Pattern Recognition · Computer Science 2022-03-29 Yuheng Jiang , Suyi Jiang , Guoxing Sun , Zhuo Su , Kaiwen Guo , Minye Wu , Jingyi Yu , Lan Xu

Rendering realistic images from 3D reconstruction is an essential task of many Computer Vision and Robotics pipelines, notably for mixed-reality applications as well as training autonomous agents in simulated environments. However, the…

Computer Vision and Pattern Recognition · Computer Science 2024-07-19 Lukas Bösiger , Mihai Dusmanu , Marc Pollefeys , Zuria Bauer

Humans naturally interact with both others and the surrounding multiple objects, engaging in various social activities. However, recent advances in modeling human-object interactions mostly focus on perceiving isolated individuals and…

Computer Vision and Pattern Recognition · Computer Science 2024-04-03 Juze Zhang , Jingyan Zhang , Zining Song , Zhanhe Shi , Chengfeng Zhao , Ye Shi , Jingyi Yu , Lan Xu , Jingya Wang

Human Activity Recognition in RGB-D videos has been an active research topic during the last decade. However, no efforts have been found in the literature, for recognizing human activity in RGB-D videos where several performers are…

Computer Vision and Pattern Recognition · Computer Science 2018-07-10 Snehasis Mukherjee , Leburu Anvitha , T. Mohana Lahari

Recent advances have enabled 3d object reconstruction approaches using a single off-the-shelf RGB-D camera. Although these approaches are successful for a wide range of object classes, they rely on stable and distinctive geometric or…

Computer Vision and Pattern Recognition · Computer Science 2017-04-04 Dimitrios Tzionas , Juergen Gall

We present the first approach to volumetric performance capture and novel-view rendering at real-time speed from monocular video, eliminating the need for expensive multi-view systems or cumbersome pre-acquisition of a personalized template…

Computer Vision and Pattern Recognition · Computer Science 2020-07-29 Ruilong Li , Yuliang Xiu , Shunsuke Saito , Zeng Huang , Kyle Olszewski , Hao Li

Existing neural human rendering methods struggle with a single image input due to the lack of information in invisible areas and the depth ambiguity of pixels in visible areas. In this regard, we propose Monocular Neural Human Renderer…

Computer Vision and Pattern Recognition · Computer Science 2022-10-04 Hongsuk Choi , Gyeongsik Moon , Matthieu Armando , Vincent Leroy , Kyoung Mu Lee , Gregory Rogez

This paper introduces Motion-oriented Compositional Neural Radiance Fields (MoCo-NeRF), a framework designed to perform free-viewpoint rendering of monocular human videos via novel non-rigid motion modeling approach. In the context of…

Computer Vision and Pattern Recognition · Computer Science 2024-07-19 Jaehyeok Kim , Dongyoon Wee , Dan Xu

Reconstructing high-fidelity animatable 3D human avatars from monocular RGB videos remains challenging, particularly in unconstrained in-the-wild scenarios where camera parameters and human poses from off-the-shelf methods (e.g., COLMAP,…

Computer Vision and Pattern Recognition · Computer Science 2026-02-05 Zihan Lou , Jinlong Fan , Sihan Ma , Yuxiang Yang , Jing Zhang

Capturing and faithfully rendering photo-realistic humans from novel views is a fundamental problem for AR/VR applications. While prior work has shown impressive performance capture results in laboratory settings, it is non-trivial to…

Computer Vision and Pattern Recognition · Computer Science 2022-08-03 Phong Nguyen-Ha , Nikolaos Sarafianos , Christoph Lassner , Janne Heikkila , Tony Tung

It has long been challenging to recover the underlying dynamic 3D scene representations from a monocular RGB video. Existing works formulate this problem into finding a single most plausible solution by adding various constraints such as…

Computer Vision and Pattern Recognition · Computer Science 2024-07-09 Ziyang Song , Jinxi Li , Bo Yang

Recognizing human activities in videos is challenging due to the spatio-temporal complexity and context-dependence of human interactions. Prior studies often rely on single input modalities, such as RGB or skeletal data, limiting their…

Computer Vision and Pattern Recognition · Computer Science 2024-09-05 Tuyen Tran , Thao Minh Le , Hung Tran , Truyen Tran

3D reconstruction of deformable (or non-rigid) scenes from a set of monocular 2D image observations is a long-standing and actively researched area of computer vision and graphics. It is an ill-posed inverse problem, since -- without…

Computer Vision and Pattern Recognition · Computer Science 2023-03-28 Edith Tretschk , Navami Kairanda , Mallikarjun B R , Rishabh Dabral , Adam Kortylewski , Bernhard Egger , Marc Habermann , Pascal Fua , Christian Theobalt , Vladislav Golyanik

3D Reconstruction of moving articulated objects without additional information about object structure is a challenging problem. Current methods overcome such challenges by employing category-specific skeletal models. Consequently, they do…

Computer Vision and Pattern Recognition · Computer Science 2024-01-18 Hao Zhang , Fang Li , Samyak Rawlekar , Narendra Ahuja

Existing deep models predict 2D and 3D kinematic poses from video that are approximately accurate, but contain visible errors that violate physical constraints, such as feet penetrating the ground and bodies leaning at extreme angles. In…

Computer Vision and Pattern Recognition · Computer Science 2020-07-27 Davis Rempe , Leonidas J. Guibas , Aaron Hertzmann , Bryan Russell , Ruben Villegas , Jimei Yang

We present a novel framework to reconstruct human avatars from monocular videos. Recent approaches have struggled either to capture the fine-grained dynamic details from the input or to generate plausible details at novel viewpoints, which…

Computer Vision and Pattern Recognition · Computer Science 2025-09-03 Yushuo Chen , Ruizhi Shao , Youxin Pang , Hongwen Zhang , Xinyi Wu , Rihui Wu , Yebin Liu
‹ Prev 1 8 9 10 Next ›