English
Related papers

Related papers: Zero-1-to-G: Taming Pretrained 2D Diffusion Model …

200 papers

Despite recent successes in novel view synthesis using 3D Gaussian Splatting (3DGS), modeling scenes with sparse inputs remains a challenge. In this work, we address two critical yet overlooked issues in real-world sparse-input modeling:…

Computer Vision and Pattern Recognition · Computer Science 2025-03-10 Yingji Zhong , Zhihao Li , Dave Zhenyu Chen , Lanqing Hong , Dan Xu

Humans excel at forecasting the future dynamics of a scene given just a single image. Video generation models that can mimic this ability are an essential component for intelligent systems. Recent approaches have improved temporal coherence…

Computer Vision and Pattern Recognition · Computer Science 2026-05-18 Melonie de Almeida , Daniela Ivanova , Tong Shi , John H. Williamson , Paul Henderson

Despite recent advances in leveraging generative prior from pre-trained diffusion models for 3D scene reconstruction, existing methods still face two critical limitations. First, due to the lack of reliable geometric supervision, they…

Computer Vision and Pattern Recognition · Computer Science 2026-02-27 Junfeng Ni , Yixin Chen , Zhifei Yang , Yu Liu , Ruijie Lu , Song-Chun Zhu , Siyuan Huang

Novel view synthesis is a fundamental challenge in image-to-3D generation, requiring the generation of target view images from a set of conditioning images and their relative poses. While recent approaches like Zero-1-to-3 have demonstrated…

Computer Vision and Pattern Recognition · Computer Science 2024-11-26 Jack Yu , Xueying Jia , Charlie Sun , Prince Wang

Recently, 3D Gaussian Splatting (3DGS) has demonstrated remarkable success in 3D reconstruction and novel view synthesis. However, reconstructing 3D scenes from sparse viewpoints remains highly challenging due to insufficient visual…

Computer Vision and Pattern Recognition · Computer Science 2025-09-24 Zhaorui Wang , Yi Gu , Deming Zhou , Renjing Xu

Dimensionality reduction algorithms map high-dimensional data into visualizable 2D or 3D spaces, but traditionally rely on a discrete point-cloud paradigm. This discrete abstraction is susceptible to visual occlusion and artificial…

Graphics · Computer Science 2026-05-19 João Paulo Gois , Luis Gustavo Nonato

Recent text-guided generation of individual 3D object has achieved great success using diffusion priors. However, these methods are not suitable for object insertion and replacement tasks as they do not consider the background, leading to…

Computer Vision and Pattern Recognition · Computer Science 2025-08-25 Hanyuan Xiao , Yingshu Chen , Huajian Huang , Haolin Xiong , Jing Yang , Pratusha Prasad , Yajie Zhao

3D Gaussian Splatting (3DGS) has demonstrated impressive performance in synthesizing novel views after training on a given set of viewpoints. However, its rendering quality deteriorates when the synthesized view deviates significantly from…

Computer Vision and Pattern Recognition · Computer Science 2025-03-13 Jiatong Xia , Lingqiao Liu

We present DreamPolisher, a novel Gaussian Splatting based method with geometric guidance, tailored to learn cross-view consistency and intricate detail from textual descriptions. While recent progress on text-to-3D generation methods have…

Computer Vision and Pattern Recognition · Computer Science 2024-03-27 Yuanze Lin , Ronald Clark , Philip Torr

Distilling 3D representations from pretrained 2D diffusion models is essential for 3D creative applications across gaming, film, and interior design. Current SDS-based methods are hindered by inefficient information distillation from…

Computer Vision and Pattern Recognition · Computer Science 2025-03-13 Haoran Li , Yuli Tian , Yonghui Wang , Yong Liao , Lin Wang , Yuyang Wang , Peng Yuan Zhou

Achieving high-resolution novel view synthesis (HRNVS) from low-resolution input views is a challenging task due to the lack of high-resolution data. Previous methods optimize high-resolution Neural Radiance Field (NeRF) from low-resolution…

Computer Vision and Pattern Recognition · Computer Science 2024-06-17 Xiqian Yu , Hanxin Zhu , Tianyu He , Zhibo Chen

Feed-forward 3D Gaussian Splatting (3DGS) enables efficient one-pass scene reconstruction, providing 3D representations for novel view synthesis without per-scene optimization. However, existing methods typically predict pixel-aligned…

Computer Vision and Pattern Recognition · Computer Science 2025-12-23 Jongmin Park , Minh-Quan Viet Bui , Juan Luis Gonzalez Bello , Jaeho Moon , Jihyong Oh , Munchurl Kim

We present Turbo3D, an ultra-fast text-to-3D system capable of generating high-quality Gaussian splatting assets in under one second. Turbo3D employs a rapid 4-step, 4-view diffusion generator and an efficient feed-forward Gaussian…

Computer Vision and Pattern Recognition · Computer Science 2024-12-06 Hanzhe Hu , Tianwei Yin , Fujun Luan , Yiwei Hu , Hao Tan , Zexiang Xu , Sai Bi , Shubham Tulsiani , Kai Zhang

We present SparseGen, a novel framework for efficient image-to-3D generation, which exhibits low input-view bias while being significantly faster. Unlike traditional approaches that rely on dense volumetric grids, triplanes, or…

Computer Vision and Pattern Recognition · Computer Science 2026-04-16 Zhiyuan Xu , Jiuming Liu , Yuxin Chen , Masayoshi Tomizuka , Chenfeng Xu , Chensheng Peng

Reconstructing 3D scenes from sparse images remains a challenging task due to the difficulty of recovering accurate geometry and texture without optimization. Recent approaches leverage generalizable models to generate 3D scenes using 3D…

Computer Vision and Pattern Recognition · Computer Science 2026-02-04 Bing He , Jingnan Gao , Yunuo Chen , Ning Cao , Gang Chen , Zhengxue Cheng , Li Song , Wenjun Zhang

Automatic car damage detection has been a topic of significant interest for the auto insurance industry as it promises faster, accurate, and cost-effective damage assessments. However, few works have gone beyond 2D image analysis to…

Computer Vision and Pattern Recognition · Computer Science 2025-09-30 Dragoş-Andrei Chileban , Andrei-Ştefan Bulzan , Cosmin Cernǎzanu-Glǎvan

Pose-free feed-forward 3D Gaussian Splatting (3DGS) has opened a new frontier for rapid 3D modeling, enabling high-quality Gaussian representations to be generated from uncalibrated multi-view images in a single forward pass. The dominant…

Computer Vision and Pattern Recognition · Computer Science 2026-03-25 Hwasik Jeong , Seungryong Lee , Gyeongjin Kang , Seungkwon Yang , Xiangyu Sun , Seungtae Nam , Eunbyung Park

Recent advances in 3D Gaussian Splatting (3D-GS) have shown remarkable success in representing 3D scenes and generating high-quality, novel views in real-time. However, 3D-GS and its variants assume that input images are captured based on…

Computer Vision and Pattern Recognition · Computer Science 2025-03-14 Liao Shen , Tianqi Liu , Huiqiang Sun , Jiaqi Li , Zhiguo Cao , Wei Li , Chen Change Loy

2D Gaussian Splatting (2DGS) is an emerging explicit scene representation method with significant potential for image compression due to high fidelity and high compression ratios. However, existing low-light enhancement algorithms operate…

Computer Vision and Pattern Recognition · Computer Science 2026-01-23 Yuhan Chen , Wenxuan Yu , Guofa Li , Yijun Xu , Ying Fang , Yicui Shi , Long Cao , Wenbo Chu , Keqiang Li

We propose SelfSplat, a novel 3D Gaussian Splatting model designed to perform pose-free and 3D prior-free generalizable 3D reconstruction from unposed multi-view images. These settings are inherently ill-posed due to the lack of…

Computer Vision and Pattern Recognition · Computer Science 2025-04-08 Gyeongjin Kang , Jisang Yoo , Jihyeon Park , Seungtae Nam , Hyeonsoo Im , Sangheon Shin , Sangpil Kim , Eunbyung Park
‹ Prev 1 4 5 6 7 8 10 Next ›