English
Related papers

Related papers: Mixture of Volumetric Primitives for Efficient Neu…

200 papers

State-of-the-art 3D-aware generative models rely on coordinate-based MLPs to parameterize 3D radiance fields. While demonstrating impressive results, querying an MLP for every sample along each ray leads to slow rendering. Therefore,…

Computer Vision and Pattern Recognition · Computer Science 2022-11-11 Katja Schwarz , Axel Sauer , Michael Niemeyer , Yiyi Liao , Andreas Geiger

We present a method for recovering the shape and radiance of a scene consisting of multiple people given solely a few images. Multi-human scenes are complex due to additional occlusion and clutter. For single-human settings, existing…

Computer Vision and Pattern Recognition · Computer Science 2025-02-12 Qian li , Victoria Fernàndez Abrevaya , Franck Multon , Adnane Boukhayma

Neural fields, also known as coordinate-based or implicit neural representations, have shown a remarkable capability of representing, generating, and manipulating various forms of signals. For video representations, however, mapping…

Computer Vision and Pattern Recognition · Computer Science 2023-08-08 Joo Chan Lee , Daniel Rho , Jong Hwan Ko , Eunbyung Park

In this paper, we introduce the Volumetric Relightable Morphable Model (VRMM), a novel volumetric and parametric facial prior for 3D face modeling. While recent volumetric prior models offer improvements over traditional methods like 3D…

Computer Vision and Pattern Recognition · Computer Science 2024-05-09 Haotian Yang , Mingwu Zheng , Chongyang Ma , Yu-Kun Lai , Pengfei Wan , Haibin Huang

We present a novel way of approaching image-based 3D reconstruction based on radiance fields. The problem of volumetric reconstruction is formulated as a non-linear least-squares problem and solved explicitly without the use of neural…

Computer Vision and Pattern Recognition · Computer Science 2021-12-13 Sverker Rasmuson , Erik Sintorn , Ulf Assarsson

While rendering and animation of photorealistic 3D human body models have matured and reached an impressive quality over the past years, modeling the spatial audio associated with such full body models has been largely ignored so far. In…

Sound · Computer Science 2024-07-23 Chao Huang , Dejan Markovic , Chenliang Xu , Alexander Richard

We present a method for generating high-quality watertight manifold meshes from multi-view input images. Existing volumetric rendering methods are robust in optimization but tend to generate noisy meshes with poor topology. Differentiable…

Computer Vision and Pattern Recognition · Computer Science 2025-01-20 Xinyue Wei , Fanbo Xiang , Sai Bi , Anpei Chen , Kalyan Sunkavalli , Zexiang Xu , Hao Su

Multi-camera perception tasks have gained significant attention in the field of autonomous driving. However, existing frameworks based on Lift-Splat-Shoot (LSS) in the multi-camera setting cannot produce suitable dense 3D features due to…

Computer Vision and Pattern Recognition · Computer Science 2023-12-20 Junkai Xu , Liang Peng , Haoran Cheng , Linxuan Xia , Qi Zhou , Dan Deng , Wei Qian , Wenxiao Wang , Deng Cai

We present a central-peripheral vision-inspired framework (CVP), a simple yet effective multimodal model for spatial reasoning that draws inspiration from the two types of human visual fields -- central vision and peripheral vision.…

Computer Vision and Pattern Recognition · Computer Science 2025-12-10 Zeyuan Chen , Xiang Zhang , Haiyang Xu , Jianwen Xie , Zhuowen Tu

High-fidelity 3D scene reconstruction from monocular videos continues to be challenging, especially for complete and fine-grained geometry reconstruction. The previous 3D reconstruction approaches with neural implicit representations have…

Computer Vision and Pattern Recognition · Computer Science 2022-10-03 Zi-Xin Zou , Shi-Sheng Huang , Yan-Pei Cao , Tai-Jiang Mu , Ying Shan , Hongbo Fu

In this work, we enhance a professional end-to-end volumetric video production pipeline to achieve high-fidelity human body reconstruction using only passive cameras. While current volumetric video approaches estimate depth maps using…

Computer Vision and Pattern Recognition · Computer Science 2022-03-01 Decai Chen , Markus Worchel , Ingo Feldmann , Oliver Schreer , Peter Eisert

Texturing is a crucial step in the 3D asset production workflow, which enhances the visual appeal and diversity of 3D assets. Despite recent advancements in Text-to-Texture (T2T) generation, existing methods often yield subpar results,…

Computer Vision and Pattern Recognition · Computer Science 2024-11-05 Wei Cheng , Juncheng Mu , Xianfang Zeng , Xin Chen , Anqi Pang , Chi Zhang , Zhibin Wang , Bin Fu , Gang Yu , Ziwei Liu , Liang Pan

Multi-view videos are becoming widely used in different fields, but their high resolution and multi-camera shooting raise significant challenges for storage and transmission. In this paper, we propose MV-MGINR, a multi-grid implicit neural…

Image and Video Processing · Electrical Eng. & Systems 2025-09-23 Qingyue Ling , Zhengxue Cheng , Donghui Feng , Shen Wang , Chen Zhu , Guo Lu , Heming Sun , Jiro Katto , Li Song

Novel view synthesis (NVS) of multi-human scenes imposes challenges due to the complex inter-human occlusions. Layered representations handle the complexities by dividing the scene into multi-layered radiance fields, however, they are…

Computer Vision and Pattern Recognition · Computer Science 2023-09-22 Youssef Abdelkareem , Shady Shehata , Fakhri Karray

We present a framework, called MVG-NeRF, that combines classical Multi-View Geometry algorithms and Neural Radiance Fields (NeRF) for image-based 3D reconstruction. NeRF has revolutionized the field of implicit 3D representations, mainly…

Computer Vision and Pattern Recognition · Computer Science 2022-10-25 Marco Orsingher , Paolo Zani , Paolo Medici , Massimo Bertozzi

Reinforcement learning based post-training paradigms for Video Large Language Models (VideoLLMs) have achieved significant success by optimizing for visual-semantic tasks such as captioning or VideoQA. However, while these approaches…

Computer Vision and Pattern Recognition · Computer Science 2026-01-08 Xiaokun Sun , Zezhong Wu , Zewen Ding , Linli Xu

Volume data is found in many important scientific and engineering applications. Rendering this data for visualization at high quality and interactive rates for demanding applications such as virtual reality is still not easily achievable…

Graphics · Computer Science 2022-09-22 David Bauer , Qi Wu , Kwan-Liu Ma

Human reconstruction and synthesis from monocular RGB videos is a challenging problem due to clothing, occlusion, texture discontinuities and sharpness, and framespecific pose changes. Many methods employ deferred rendering, NeRFs and…

Computer Vision and Pattern Recognition · Computer Science 2023-03-16 Rohit Jena , Pratik Chaudhari , James Gee , Ganesh Iyer , Siddharth Choudhary , Brandon M. Smith

Embedding polygonal mesh assets within photorealistic Neural Radience Fields (NeRF) volumes, such that they can be rendered and their dynamics simulated in a physically consistent manner with the NeRF, is under-explored from the system…

Graphics · Computer Science 2023-09-12 Yi-Ling Qiao , Alexander Gao , Yiran Xu , Yue Feng , Jia-Bin Huang , Ming C. Lin

Neural Radiance Fields (NeRF) revolutionize the realm of visual media by providing photorealistic Free-Viewpoint Video (FVV) experiences, offering viewers unparalleled immersion and interactivity. However, the technology's significant…

Computer Vision and Pattern Recognition · Computer Science 2023-12-13 Minye Wu , Zehao Wang , Georgios Kouros , Tinne Tuytelaars