English
Related papers

Related papers: Physics-guided Shape-from-Template: Monocular Vide…

200 papers

Recent works on implicit neural representations have shown promising results for multi-view surface reconstruction. However, most approaches are limited to relatively simple geometries and usually require clean object masks for…

Computer Vision and Pattern Recognition · Computer Science 2021-08-24 Jingyang Zhang , Yao Yao , Long Quan

Real-time free-view human rendering from sparse-view RGB inputs is a challenging task due to the sensor scarcity and the tight time budget. To ensure efficiency, recent methods leverage 2D CNNs operating in texture space to learn rendering…

Computer Vision and Pattern Recognition · Computer Science 2025-06-23 Guoxing Sun , Rishabh Dabral , Heming Zhu , Pascal Fua , Christian Theobalt , Marc Habermann

Applying single image Monocular Depth Estimation (MDE) models to video sequences introduces significant temporal instability and flickering artifacts. We propose a novel approach that adapts any state-of-the-art image-based (depth)…

Computer Vision and Pattern Recognition · Computer Science 2026-01-07 Ivan Sobko , Hayko Riemenschneider , Markus Gross , Christopher Schroers

In recent years, neural implicit surface reconstruction methods have become popular for multi-view 3D reconstruction. In contrast to traditional multi-view stereo methods, these approaches tend to produce smoother and more complete…

Computer Vision and Pattern Recognition · Computer Science 2022-10-13 Zehao Yu , Songyou Peng , Michael Niemeyer , Torsten Sattler , Andreas Geiger

Learning deformable 3D objects from 2D images is often an ill-posed problem. Existing methods rely on explicit supervision to establish multi-view correspondences, such as template shape models and keypoint annotations, which restricts…

Computer Vision and Pattern Recognition · Computer Science 2022-06-30 Shangzhe Wu , Tomas Jakab , Christian Rupprecht , Andrea Vedaldi

This paper proposes a new approach for monocular dense 3D reconstruction of a complex dynamic scene from two perspective frames. By applying superpixel over-segmentation to the image, we model a generically dynamic (hence non-rigid) scene…

Computer Vision and Pattern Recognition · Computer Science 2017-12-21 Suryansh Kumar , Yuchao Dai , Hongdong Li

Recovering the 3D structure of the scene from images yields useful information for tasks such as shape and scene recognition, object detection, or motion planning and object grasping in robotics. In this thesis, we introduce a general…

Computer Vision and Pattern Recognition · Computer Science 2010-07-20 Hoang Trinh

Recovering textured 3D models of non-rigid human body shapes is challenging due to self-occlusions caused by complex body poses and shapes, clothing obstructions, lack of surface texture, background clutter, sparse set of cameras with…

Computer Vision and Pattern Recognition · Computer Science 2018-09-19 Abbhinav Venkat , Sai Sagar Jinka , Avinash Sharma

Surgical simulation is essential for medical training, enabling practitioners to develop crucial skills in a risk-free environment while improving patient safety and surgical outcomes. However, conventional methods for building simulation…

Computer Vision and Pattern Recognition · Computer Science 2025-09-23 Zhenya Yang

This paper proposes a new method for Non-Rigid Structure-from-Motion (NRSfM) from a long monocular video sequence observing a non-rigid object performing recurrent and possibly repetitive dynamic action. Departing from the traditional idea…

Computer Vision and Pattern Recognition · Computer Science 2018-04-19 Xiu Li , Hongdong Li , Hanbyul Joo , Yebin Liu , Yaser Sheikh

We propose a novel method for 3D object reconstruction from a sparse set of views captured from a 360-degree calibrated camera rig. We represent the object surface through a hybrid model that uses both an MLP-based neural representation and…

Computer Vision and Pattern Recognition · Computer Science 2024-03-29 Llukman Cerkezi , Paolo Favaro

Our work aims to obtain 3D reconstruction of hands and manipulated objects from monocular videos. Reconstructing hand-object manipulations holds a great potential for robotics and learning from human demonstrations. The supervised learning…

Computer Vision and Pattern Recognition · Computer Science 2022-03-15 Yana Hasson , Gül Varol , Ivan Laptev , Cordelia Schmid

Reconstructing hand-held objects in 3D from monocular images remains a significant challenge in computer vision. Most existing approaches rely on implicit 3D representations, which produce overly smooth reconstructions and are…

Computer Vision and Pattern Recognition · Computer Science 2025-07-22 Zerui Chen , Rolandos Alexandros Potamias , Shizhe Chen , Cordelia Schmid

Recently, Gaussian Splatting has sparked a new trend in the field of computer vision. Apart from novel view synthesis, it has also been extended to the area of multi-view reconstruction. The latest methods facilitate complete, detailed…

Computer Vision and Pattern Recognition · Computer Science 2025-01-09 Han Huang , Yulun Wu , Chao Deng , Ge Gao , Ming Gu , Yu-Shen Liu

3D Gaussian Splatting (GS) enables highly photorealistic scene reconstruction from posed image sequences but struggles with viewpoint extrapolation due to its anisotropic nature, leading to overfitting and poor generalization, particularly…

Computer Vision and Pattern Recognition · Computer Science 2025-12-09 Shuohan Tao , Boyao Zhou , Hanzhang Tu , Yuwang Wang , Yebin Liu

We introduce D$^3$-Human, a method for reconstructing Dynamic Disentangled Digital Human geometry from monocular videos. Past monocular video human reconstruction primarily focuses on reconstructing undecoupled clothed human bodies or only…

Computer Vision and Pattern Recognition · Computer Science 2025-01-06 Honghu Chen , Bo Peng , Yunfan Tao , Juyong Zhang

Reconstructing high-quality 3D models from sparse 2D images has garnered significant attention in computer vision. Recently, 3D Gaussian Splatting (3DGS) has gained prominence due to its explicit representation with efficient training speed…

Computer Vision and Pattern Recognition · Computer Science 2024-12-31 Keng-Wei Chang , Zi-Ming Wang , Shang-Hong Lai

The recovery of 3D human mesh from monocular images has significantly been developed in recent years. However, existing models usually ignore spatial and temporal information, which might lead to mesh and image misalignment and temporal…

Computer Vision and Pattern Recognition · Computer Science 2024-01-04 Wei Yao , Hongwen Zhang , Yunlian Sun , Jinhui Tang

Dynamic scene reconstruction from multi-view videos remains a fundamental challenge in computer vision. While recent neural surface reconstruction methods have achieved remarkable results in static 3D reconstruction, extending these…

Graphics · Computer Science 2025-09-22 Mohamed Ebbed , Zorah Lähner

Multi-view surface reconstruction is an ill-posed, inverse problem in 3D vision research. It involves modeling the geometry and appearance with appropriate surface representations. Most of the existing methods rely either on explicit…

Computer Vision and Pattern Recognition · Computer Science 2024-01-09 Zhangjin Huang , Zhihao Liang , Haojie Zhang , Yangkai Lin , Kui Jia