中文
相关论文

相关论文: iNVS: Repurposing Diffusion Inpainters for Novel V…

200 篇论文

In this study, we present a method for synthesizing novel views from a single 360-degree RGB-D image based on the neural radiance field (NeRF) . Prior studies relied on the neighborhood interpolation capability of multi-layer perceptrons to…

计算机视觉与模式识别 · 计算机科学 2023-09-08 Takayuki Hara , Tatsuya Harada

Diffusion models for single image novel view synthesis (NVS) can generate highly realistic and plausible images, but they are limited in the geometric consistency to the given relative poses. The generated images often show significant…

计算机视觉与模式识别 · 计算机科学 2025-04-14 Josef Bengtson , David Nilsson , Fredrik Kahl

We introduce Zero-1-to-3, a framework for changing the camera viewpoint of an object given just a single RGB image. To perform novel view synthesis in this under-constrained setting, we capitalize on the geometric priors that large-scale…

计算机视觉与模式识别 · 计算机科学 2023-03-21 Ruoshi Liu , Rundi Wu , Basile Van Hoorick , Pavel Tokmakov , Sergey Zakharov , Carl Vondrick

Given just a few glimpses of a scene, can you imagine the movie playing out as the camera glides through it? That's the lens we take on \emph{sparse-input novel view synthesis}, not only as filling spatial gaps between widely spaced views,…

计算机视觉与模式识别 · 计算机科学 2025-11-25 Yan Xu , Yixing Wang , Stella X. Yu

We present 3DiM, a diffusion model for 3D novel view synthesis, which is able to translate a single input view into consistent and sharp completions across many views. The core component of 3DiM is a pose-conditional image-to-image…

计算机视觉与模式识别 · 计算机科学 2022-10-11 Daniel Watson , William Chan , Ricardo Martin-Brualla , Jonathan Ho , Andrea Tagliasacchi , Mohammad Norouzi

Active object reconstruction is crucial for many robotic applications. A key aspect in these scenarios is generating object-specific view configurations to obtain informative measurements for reconstruction. One-shot view planning enables…

机器人学 · 计算机科学 2025-04-17 Sicong Pan , Liren Jin , Xuying Huang , Cyrill Stachniss , Marija Popović , Maren Bennewitz

Multi-View Stereo (MVS) is a core task in 3D computer vision. With the surge of novel deep learning methods, learned MVS has surpassed the accuracy of classical approaches, but still relies on building a memory intensive dense cost volume.…

计算机视觉与模式识别 · 计算机科学 2022-06-16 Radu Alexandru Rosu , Sven Behnke

In this paper, we propose RI3D, a novel 3DGS-based approach that harnesses the power of diffusion models to reconstruct high-quality novel views given a sparse set of input images. Our key contribution is separating the view synthesis…

计算机视觉与模式识别 · 计算机科学 2026-01-22 Avinash Paliwal , Xilong Zhou , Wei Ye , Jinhui Xiong , Rakesh Ranjan , Nima Khademi Kalantari

Novel view synthesis from a single image has recently achieved remarkable results, although the requirement of some form of 3D, pose, or multi-view supervision at training time limits the deployment in real scenarios. This work aims at…

计算机视觉与模式识别 · 计算机科学 2021-12-16 Pierluigi Zama Ramirez , Diego Martin Arroyo , Alessio Tonioni , Federico Tombari

We present Stable Video 3D (SV3D) -- a latent video diffusion model for high-resolution, image-to-multi-view generation of orbital videos around a 3D object. Recent work on 3D generation propose techniques to adapt 2D generative models for…

计算机视觉与模式识别 · 计算机科学 2024-03-19 Vikram Voleti , Chun-Han Yao , Mark Boss , Adam Letts , David Pankratz , Dmitry Tochilkin , Christian Laforte , Robin Rombach , Varun Jampani

We present a novel method for 3D surface reconstruction from multiple images where only a part of the object of interest is captured. Our approach builds on two recent developments: surface reconstruction using neural radiance fields for…

计算机视觉与模式识别 · 计算机科学 2023-12-11 Savva Ignatyev , Daniil Selikhanovych , Oleg Voynov , Yiqun Wang , Peter Wonka , Stamatios Lefkimmiatis , Evgeny Burnaev

3D Gaussian Splatting has recently emerged as a powerful tool for fast and accurate novel-view synthesis from a set of posed input images. However, like most novel-view synthesis approaches, it relies on accurate camera pose information,…

计算机视觉与模式识别 · 计算机科学 2024-10-14 Christian Schmidt , Jens Piekenbrinck , Bastian Leibe

This paper considers the problem of generative novel view synthesis (GNVS), generating novel, plausible views of a scene given a limited number of known views. Here, we propose a set-based generative model that can simultaneously generate…

计算机视觉与模式识别 · 计算机科学 2024-07-29 Jason J. Yu , Tristan Aumentado-Armstrong , Fereshteh Forghani , Konstantinos G. Derpanis , Marcus A. Brubaker

Scene-level novel view synthesis (NVS) is fundamental to many vision and graphics applications. Recently, pose-conditioned diffusion models have led to significant progress by extracting 3D information from 2D foundation models, but these…

计算机视觉与模式识别 · 计算机科学 2024-08-23 Joseph Tung , Gene Chou , Ruojin Cai , Guandao Yang , Kai Zhang , Gordon Wetzstein , Bharath Hariharan , Noah Snavely

Utilizing pre-trained 2D large-scale generative models, recent works are capable of generating high-quality novel views from a single in-the-wild image. However, due to the lack of information from multiple views, these works encounter…

计算机视觉与模式识别 · 计算机科学 2024-03-27 Yunhan Yang , Yukun Huang , Xiaoyang Wu , Yuan-Chen Guo , Song-Hai Zhang , Hengshuang Zhao , Tong He , Xihui Liu

Novel-view synthesis aims to generate novel views of a scene from multiple input images or videos, and recent advancements like 3D Gaussian splatting (3DGS) have achieved notable success in producing photorealistic renderings with efficient…

计算机视觉与模式识别 · 计算机科学 2024-10-22 Xi Liu , Chaoyi Zhou , Siyu Huang

Prior approaches injecting camera control into diffusion models have focused on specific subsets of 4D consistency tasks: novel view synthesis, text-to-video with camera control, image-to-video, amongst others. Therefore, these fragmented…

计算机视觉与模式识别 · 计算机科学 2026-01-26 Xiang Fan , Sharath Girish , Vivek Ramanujan , Chaoyang Wang , Ashkan Mirzaei , Petr Sushko , Aliaksandr Siarohin , Sergey Tulyakov , Ranjay Krishna

We present BetterScene, an approach to enhance novel view synthesis (NVS) quality for diverse real-world scenes using extremely sparse, unconstrained photos. BetterScene leverages the production-ready Stable Video Diffusion (SVD) model…

计算机视觉与模式识别 · 计算机科学 2026-02-27 Yuci Han , Charles Toth , John E. Anderson , William J. Shuart , Alper Yilmaz

Despite the emergence of successful NeRF inpainting methods built upon explicit RGB and depth 2D inpainting supervisions, these methods are inherently constrained by the capabilities of their underlying 2D inpainters. This is due to two key…

计算机视觉与模式识别 · 计算机科学 2024-05-07 Honghua Chen , Chen Change Loy , Xingang Pan

Photorealistic simulators are essential for the training and evaluation of vision-centric autonomous vehicles (AVs). At their core is Novel View Synthesis (NVS), a crucial capability that generates diverse unseen viewpoints to accommodate…

计算机视觉与模式识别 · 计算机科学 2025-03-14 Xiangyu Han , Zhen Jia , Boyi Li , Yan Wang , Boris Ivanovic , Yurong You , Lingjie Liu , Yue Wang , Marco Pavone , Chen Feng , Yiming Li