English
Related papers

Related papers: BulletGen: Improving 4D Reconstruction with Bullet…

200 papers

While dynamic novel view synthesis from 2D videos has seen progress, achieving efficient reconstruction and rendering of dynamic scenes remains a challenging task. In this paper, we introduce Disentangled 4D Gaussian Splatting…

Graphics · Computer Science 2025-10-31 Hao Feng , Hao Sun , Wei Xie , Zhi Zuo , Zhengzhe Liu

We present BetterScene, an approach to enhance novel view synthesis (NVS) quality for diverse real-world scenes using extremely sparse, unconstrained photos. BetterScene leverages the production-ready Stable Video Diffusion (SVD) model…

Computer Vision and Pattern Recognition · Computer Science 2026-02-27 Yuci Han , Charles Toth , John E. Anderson , William J. Shuart , Alper Yilmaz

We address the problem of dynamic scene reconstruction from sparse-view videos. Prior work often requires dense multi-view captures with hundreds of calibrated cameras (e.g. Panoptic Studio). Such multi-view setups are prohibitively…

Computer Vision and Pattern Recognition · Computer Science 2026-03-03 Zihan Wang , Jeff Tan , Tarasha Khurana , Neehar Peri , Deva Ramanan

Synthesizing consistent and photorealistic 3D scenes is an open problem in computer vision. Video diffusion models generate impressive videos but cannot directly synthesize 3D representations, i.e., lack 3D consistency in the generated…

Computer Vision and Pattern Recognition · Computer Science 2025-03-18 Katja Schwarz , Norman Mueller , Peter Kontschieder

Reconstructing dynamic scenes with large-scale and complex motions remains a significant challenge. Recent techniques like Neural Radiance Fields and 3D Gaussian Splatting (3DGS) have shown promise but still struggle with scenes involving…

Computer Vision and Pattern Recognition · Computer Science 2024-12-04 Qiankun Gao , Yanmin Wu , Chengxiang Wen , Jiarui Meng , Luyang Tang , Jie Chen , Ronggang Wang , Jian Zhang

We introduce a fully automatic pipeline for dynamic scene reconstruction from casually captured monocular RGB videos. Rather than designing a new scene representation, we enhance the priors that drive Dynamic Gaussian Splatting. Video…

Computer Vision and Pattern Recognition · Computer Science 2025-12-15 Meng-Li Shih , Ying-Huan Chen , Yu-Lun Liu , Brian Curless

Recent advancements in generative models have ignited substantial interest in dynamic 3D content creation (\ie, 4D generation). Existing approaches primarily rely on Score Distillation Sampling (SDS) to infer novel-view videos, typically…

Computer Vision and Pattern Recognition · Computer Science 2025-01-06 Hanxin Zhu , Tianyu He , Xiqian Yu , Junliang Guo , Zhibo Chen , Jiang Bian

The increasing demand for augmented and virtual reality applications has highlighted the importance of crafting immersive 3D scenes from a simple single-view image. However, due to the partial priors provided by single-view input, existing…

Computer Vision and Pattern Recognition · Computer Science 2025-04-01 Tianyi Gong , Boyan Li , Yifei Zhong , Fangxin Wang

We present a novel framework named NeuralRecon for real-time 3D scene reconstruction from a monocular video. Unlike previous methods that estimate single-view depth maps separately on each key-frame and fuse them later, we propose to…

Computer Vision and Pattern Recognition · Computer Science 2021-04-02 Jiaming Sun , Yiming Xie , Linghao Chen , Xiaowei Zhou , Hujun Bao

Instant reconstruction of dynamic 3D humans from uncalibrated sparse-view videos is critical for numerous downstream applications. Existing methods, however, are either limited by the slow reconstruction speeds or incapable of generating…

Computer Vision and Pattern Recognition · Computer Science 2025-09-30 Yingdong Hu , Yisheng He , Jinnan Chen , Weihao Yuan , Kejie Qiu , Zehong Lin , Siyu Zhu , Zilong Dong , Jun Zhang

Generative models have emerged as an essential building block for many image synthesis and editing tasks. Recent advances in this field have also enabled high-quality 3D or video content to be generated that exhibits either multi-view or…

Computer Vision and Pattern Recognition · Computer Science 2023-08-10 Sherwin Bahmani , Jeong Joon Park , Despoina Paschalidou , Hao Tang , Gordon Wetzstein , Leonidas Guibas , Luc Van Gool , Radu Timofte

Stereo video generation has been gaining increasing attention with recent advancements in video diffusion models. However, most existing methods focus on generating 3D stereoscopic videos from monocular 2D videos. These approaches typically…

Computer Vision and Pattern Recognition · Computer Science 2025-06-09 Xingchang Huang , Ashish Kumar Singh , Florian Dubost , Cristina Nader Vasconcelos , Sakar Khattar , Liang Shi , Christian Theobalt , Cengiz Oztireli , Gurprit Singh

Editing 4D scenes reconstructed from monocular videos based on text prompts is a valuable yet challenging task with broad applications in content creation and virtual environments. The key difficulty lies in achieving semantically precise…

Computer Vision and Pattern Recognition · Computer Science 2025-10-13 Jin-Chuan Shi , Chengye Su , Jiajun Wang , Ariel Shamir , Miao Wang

In this paper, we present Consistent4D, a novel approach for generating 4D dynamic objects from uncalibrated monocular videos. Uniquely, we cast the 360-degree dynamic object reconstruction as a 4D generation problem, eliminating the need…

Computer Vision and Pattern Recognition · Computer Science 2023-11-07 Yanqin Jiang , Li Zhang , Jin Gao , Weimin Hu , Yao Yao

The rapid advancement of diffusion models holds the promise of revolutionizing the application of VR and AR technologies, which typically require scene-level 4D assets for user experience. Nonetheless, existing diffusion models…

Computer Vision and Pattern Recognition · Computer Science 2025-05-14 Haiyang Zhou , Wangbo Yu , Jiawen Guan , Xinhua Cheng , Yonghong Tian , Li Yuan

We introduce Diff4Splat, a feed-forward method that synthesizes controllable and explicit 4D scenes from a single image. Our approach unifies the generative priors of video diffusion models with geometry and motion constraints learned from…

Computer Vision and Pattern Recognition · Computer Science 2026-04-08 Panwang Pan , Chenguo Lin , Jingjing Zhao , Chenxin Li , Yuchen Lin , Haopeng Li , Honglei Yan , Kairun Wen , Yunlong Lin , Yixuan Yuan , Yadong Mu

Reconstructing dynamic objects from monocular videos is a severely underconstrained and challenging problem, and recent work has approached it in various directions. However, owing to the ill-posed nature of this problem, there has been no…

Computer Vision and Pattern Recognition · Computer Science 2024-04-02 Devikalyan Das , Christopher Wewer , Raza Yunus , Eddy Ilg , Jan Eric Lenssen

This paper addresses the challenge of novel-view synthesis and motion reconstruction of dynamic scenes from monocular video, which is critical for many robotic applications. Although Neural Radiance Fields (NeRF) and 3D Gaussian Splatting…

Robotics · Computer Science 2025-08-12 Xuesong Li , Lars Petersson , Vivien Rolland

We present a latent diffusion model for fast feed-forward 3D scene generation. Given one or more images, our model Bolt3D directly samples a 3D scene representation in less than seven seconds on a single GPU. We achieve this by leveraging…

Reconstructing 4D dynamic scenes from casually captured monocular videos is valuable but highly challenging, as each timestamp is observed from a single viewpoint. We introduce Vivid4D, a novel approach that enhances 4D monocular video…

Computer Vision and Pattern Recognition · Computer Science 2025-04-21 Jiaxin Huang , Sheng Miao , BangBang Yang , Yuewen Ma , Yiyi Liao
‹ Prev 1 4 5 6 7 8 10 Next ›