English
Related papers

Related papers: MonoNeRF: Learning Generalizable NeRFs from Monocu…

200 papers

The neural radiance field (NeRF) for realistic novel view synthesis requires camera poses to be pre-acquired by a structure-from-motion (SfM) approach. This two-stage strategy is not convenient to use and degrades the performance because…

Computer Vision and Pattern Recognition · Computer Science 2022-10-04 Shu Chen , Yang Zhang , Yaxin Xu , Beiji Zou

We propose DFPNet -- an unsupervised, joint learning system for monocular Depth, Optical Flow and egomotion (Camera Pose) estimation from monocular image sequences. Due to the nature of 3D scene geometry these three components are coupled.…

Computer Vision and Pattern Recognition · Computer Science 2022-10-12 Dipan Mandal , Abhilash Jain

Neural Radiance Fields (NeRF) show impressive performance in photo-realistic free-view rendering of scenes. Recent improvements on the NeRF such as TensoRF and ZipNeRF employ explicit models for faster optimization and rendering, as…

Computer Vision and Pattern Recognition · Computer Science 2026-03-17 Nagabhushan Somraj , Sai Harsha Mupparaju , Adithyan Karanayil , Rajiv Soundararajan

We extend neural 3D representations to allow for intuitive and interpretable user control beyond novel view rendering (i.e. camera control). We allow the user to annotate which part of the scene one wishes to control with just a small…

Computer Vision and Pattern Recognition · Computer Science 2021-12-07 Kacper Kania , Kwang Moo Yi , Marek Kowalski , Tomasz Trzciński , Andrea Tagliasacchi

This work targets at using a general deep learning framework to synthesize free-viewpoint images of arbitrary human performers, only requiring a sparse number of camera views as inputs and skirting per-case fine-tuning. The large variation…

Computer Vision and Pattern Recognition · Computer Science 2022-04-26 Wei Cheng , Su Xu , Jingtan Piao , Chen Qian , Wayne Wu , Kwan-Yee Lin , Hongsheng Li

We present PartNerFace, a part-based neural radiance fields approach, for reconstructing animatable facial avatar from monocular RGB videos. Existing solutions either simply condition the implicit network with the morphable model parameters…

Computer Vision and Pattern Recognition · Computer Science 2026-04-16 Xianggang Yu , Lingteng Qiu , Xiaohang Ren , Guanying Chen , Shuguang Cui , Xiaoguang Han , Baoyuan Wang

A long-standing goal in scene understanding is to obtain interpretable and editable representations that can be directly constructed from a raw monocular RGB-D video, without requiring specialized hardware setup or priors. The problem is…

Computer Vision and Pattern Recognition · Computer Science 2023-06-22 Yu-Shiang Wong , Niloy J. Mitra

We present a simple yet powerful neural network that implicitly represents and renders 3D objects and scenes only from 2D observations. The network models 3D geometries as a general radiance field, which takes a set of 2D images with camera…

Computer Vision and Pattern Recognition · Computer Science 2021-08-12 Alex Trevithick , Bo Yang

We present FlexNeRF, a method for photorealistic freeviewpoint rendering of humans in motion from monocular videos. Our approach works well with sparse views, which is a challenging scenario when the subject is exhibiting fast/complex…

Computer Vision and Pattern Recognition · Computer Science 2023-03-28 Vinoj Jayasundara , Amit Agrawal , Nicolas Heron , Abhinav Shrivastava , Larry S. Davis

Neural radiance fields (NeRFs) show potential for transforming images captured worldwide into immersive 3D visual experiences. However, most of this captured visual data remains siloed in our camera rolls as these images contain personal…

Computer Vision and Pattern Recognition · Computer Science 2024-04-01 Zaid Tasneem , Akshat Dave , Abhishek Singh , Kushagra Tiwary , Praneeth Vepakomma , Ashok Veeraraghavan , Ramesh Raskar

Perceiving 3D information is of paramount importance in many applications of computer vision. Recent advances in monocular depth estimation have shown that gaining such knowledge from a single camera input is possible by training deep…

Computer Vision and Pattern Recognition · Computer Science 2021-10-28 Sai Shyam Chanduri , Zeeshan Khan Suri , Igor Vozniak , Christian Müller

We present an algorithm for reconstructing dense, geometrically consistent depth for all pixels in a monocular video. We leverage a conventional structure-from-motion reconstruction to establish geometric constraints on pixels in the video.…

Computer Vision and Pattern Recognition · Computer Science 2020-08-28 Xuan Luo , Jia-Bin Huang , Richard Szeliski , Kevin Matzen , Johannes Kopf

Monocular image-based 3D reconstruction of faces is a long-standing problem in computer vision. Since image data is a 2D projection of a 3D face, the resulting depth ambiguity makes the problem ill-posed. Most existing methods rely on…

Computer Vision and Pattern Recognition · Computer Science 2019-04-10 Ayush Tewari , Florian Bernard , Pablo Garrido , Gaurav Bharaj , Mohamed Elgharib , Hans-Peter Seidel , Patrick Pérez , Michael Zollhöfer , Christian Theobalt

Monocular depth reconstruction of complex and dynamic scenes is a highly challenging problem. While for rigid scenes learning-based methods have been offering promising results even in unsupervised cases, there exists little to no…

Computer Vision and Pattern Recognition · Computer Science 2021-10-29 Ayça Takmaz , Danda Pani Paudel , Thomas Probst , Ajad Chhatkuli , Martin R. Oswald , Luc Van Gool

We propose DistillNeRF, a self-supervised learning framework addressing the challenge of understanding 3D environments from limited 2D observations in outdoor autonomous driving scenes. Our method is a generalizable feedforward model that…

Computer Vision and Pattern Recognition · Computer Science 2024-11-01 Letian Wang , Seung Wook Kim , Jiawei Yang , Cunjun Yu , Boris Ivanovic , Steven L. Waslander , Yue Wang , Sanja Fidler , Marco Pavone , Peter Karkus

In this paper, we propose MonoRec, a semi-supervised monocular dense reconstruction architecture that predicts depth maps from a single moving camera in dynamic environments. MonoRec is based on a multi-view stereo setting which encodes the…

Computer Vision and Pattern Recognition · Computer Science 2022-09-22 Felix Wimbauer , Nan Yang , Lukas von Stumberg , Niclas Zeller , Daniel Cremers

We present an approach that learns to synthesize high-quality, novel views of 3D objects or scenes, while providing fine-grained and precise control over the 6-DOF viewpoint. The approach is self-supervised and only requires 2D images and…

Computer Vision and Pattern Recognition · Computer Science 2019-09-10 Xu Chen , Jie Song , Otmar Hilliges

Photo-realistic neural reconstruction and rendering of the human portrait are critical for numerous VR/AR applications. Still, existing solutions inherently rely on multi-view capture settings, and the one-shot solution to get rid of the…

Computer Vision and Pattern Recognition · Computer Science 2021-05-17 Ziyu Wang , Liao Wang , Fuqiang Zhao , Minye Wu , Lan Xu , Jingyi Yu

We present Monocular Neural Parametric Head Models (MonoNPHM) for dynamic 3D head reconstructions from monocular RGB videos. To this end, we propose a latent appearance space that parameterizes a texture field on top of a neural parametric…

Computer Vision and Pattern Recognition · Computer Science 2024-05-31 Simon Giebenhain , Tobias Kirschstein , Markos Georgopoulos , Martin Rünz , Lourdes Agapito , Matthias Nießner

We present GeoNeRF, a generalizable photorealistic novel view synthesis method based on neural radiance fields. Our approach consists of two main stages: a geometry reasoner and a renderer. To render a novel view, the geometry reasoner…

Computer Vision and Pattern Recognition · Computer Science 2022-03-22 Mohammad Mahdi Johari , Yann Lepoittevin , François Fleuret
‹ Prev 1 4 5 6 7 8 10 Next ›