中文
相关论文

相关论文: Zero-1-to-G: Taming Pretrained 2D Diffusion Model …

200 篇论文

Existing multi-view 3D object reconstruction methods heavily rely on sufficient overlap between input views, where occlusions and sparse coverage in practice frequently yield severe reconstruction incompleteness. Recent advancements in…

计算机视觉与模式识别 · 计算机科学 2025-10-28 Jiahao Chang , Chongjie Ye , Yushuang Wu , Yuantao Chen , Yidan Zhang , Zhongjin Luo , Chenghong Li , Yihao Zhi , Xiaoguang Han

In this paper, we explore the existing challenges in 3D artistic scene generation by introducing ART3D, a novel framework that combines diffusion models and 3D Gaussian splatting techniques. Our method effectively bridges the gap between…

计算机视觉与模式识别 · 计算机科学 2024-05-20 Pengzhi Li , Chengshuai Tang , Qinxuan Huang , Zhiheng Li

Recently, multi-view diffusion-based 3D generation methods have gained significant attention. However, these methods often suffer from shape and texture misalignment across generated multi-view images, leading to low-quality 3D generation…

计算机视觉与模式识别 · 计算机科学 2025-11-18 Zhuojiang Cai , Yiheng Zhang , Meitong Guo , Mingdao Wang , Yuwang Wang

We introduce DD3G, a formulation that Distills a multi-view Diffusion model (MV-DM) into a 3D Generator using gaussian splatting. DD3G compresses and integrates extensive visual and spatial geometric knowledge from the MV-DM by simulating…

计算机视觉与模式识别 · 计算机科学 2025-04-04 Hao Qin , Luyuan Chen , Ming Kong , Mengxu Lu , Qiang Zhu

We train a feed-forward text-to-3D diffusion generator for human characters using only single-view 2D data for supervision. Existing 3D generative models cannot yet match the fidelity of image or video generative models. State-of-the-art 3D…

计算机视觉与模式识别 · 计算机科学 2024-12-24 Souhaib Attaiki , Paul Guerrero , Duygu Ceylan , Niloy J. Mitra , Maks Ovsjanikov

Object-level 3D reconstruction play important roles across domains such as cultural heritage digitization, industrial manufacturing, and virtual reality. However, existing Gaussian Splatting-based approaches generally rely on full-scene…

计算机视觉与模式识别 · 计算机科学 2026-03-17 Shuai Guo , Ao Guo , Junchao Zhao , Qi Chen , Yuxiang Qi , Zechuan Li , Dong Chen , Tianjia Shao , Mingliang Xu

3D Gaussian Splatting (3DGS) has recently emerged as a powerful paradigm for photorealistic view synthesis, representing scenes with spatially distributed Gaussian primitives. While highly effective for rendering, achieving accurate and…

计算机视觉与模式识别 · 计算机科学 2025-09-23 Wenzhi Guo , Bing Wang

Diffusion-based 3D generation has made remarkable progress in recent years. However, existing 3D generative models often produce overly dense and unstructured meshes, which stand in stark contrast to the compact, structured, and…

计算机视觉与模式识别 · 计算机科学 2025-03-25 Yuan Li , Cheng Lin , Yuan Liu , Xiaoxiao Long , Chenxu Zhang , Ningna Wang , Xin Li , Wenping Wang , Xiaohu Guo

Feed-forward 3D Gaussian Splatting methods enable single-pass reconstruction and real-time rendering. However, they typically adopt rigid pixel-to-Gaussian or voxel-to-Gaussian pipelines that uniformly allocate Gaussians, leading to…

计算机视觉与模式识别 · 计算机科学 2026-03-26 Injae Kim , Chaehyeon Kim , Minseong Bae , Minseok Joo , Hyunwoo J. Kim

3D Gaussian Splatting (3DGS) has become a powerful representation for image-based object reconstruction, yet its performance drops sharply in sparse-view settings. Prior works address this limitation by employing diffusion models to repair…

计算机视觉与模式识别 · 计算机科学 2026-01-29 Hung Nguyen , Runfa Li , An Le , Truong Nguyen

Recent advances in diffusion-based generative models have shown incredible promise for zero shot image-to-image translation and editing. Most of these approaches work by combining or replacing network-specific features used in the…

计算机视觉与模式识别 · 计算机科学 2025-10-06 Zeqi Gu , Ethan Yang , Abe Davis

Score Distillation Sampling (SDS) leverages pretrained 2D diffusion models to advance text-to-3D generation but neglects multi-view correlations, being prone to geometric inconsistencies and multi-face artifacts in the generated 3D content.…

计算机视觉与模式识别 · 计算机科学 2025-12-23 Feng Yang , Wenliang Qian , Wangmeng Zuo , Hui Li

3D Gaussian splatting (3DGS) has demonstrated impressive performance in synthesizing high-fidelity novel views. Nonetheless, its effectiveness critically depends on the quality of the initialized point cloud. Specifically, achieving uniform…

计算机视觉与模式识别 · 计算机科学 2025-10-13 Yikang Zhang , Rui Fan

3D Gaussian Splatting (3DGS) has emerged as a powerful technique for novel view synthesis. However, existing methods struggle to adaptively optimize the distribution of Gaussian primitives based on scene characteristics, making it…

计算机视觉与模式识别 · 计算机科学 2025-06-23 Hongbi Zhou , Zhangkai Ni

Recent advancements in large generative models and real-time neural rendering using point-based techniques pave the way for a future of widespread visual data distribution through sharing synthesized 3D assets. However, while standardized…

计算机视觉与模式识别 · 计算机科学 2024-07-02 Chenxin Li , Hengyu Liu , Zhiwen Fan , Wuyang Li , Yifan Liu , Panwang Pan , Yixuan Yuan

3D Gaussian Splatting (3DGS) is a promising technique for 3D reconstruction, offering efficient training and rendering speeds, making it suitable for real-time applications.However, current methods require highly controlled environments (no…

计算机视觉与模式识别 · 计算机科学 2024-07-31 Sara Sabour , Lily Goli , George Kopanas , Mark Matthews , Dmitry Lagun , Leonidas Guibas , Alec Jacobson , David J. Fleet , Andrea Tagliasacchi

3D Gaussian Splatting (3DGS) has recently revolutionized radiance field reconstruction, achieving high quality novel view synthesis and fast rendering speed without baking. However, 3DGS fails to accurately represent surfaces due to the…

计算机视觉与模式识别 · 计算机科学 2025-02-25 Binbin Huang , Zehao Yu , Anpei Chen , Andreas Geiger , Shenghua Gao

The recent development of feedforward 3D Gaussian Splatting (3DGS) presents a new paradigm to reconstruct 3D scenes. Using neural networks trained on large-scale multi-view datasets, it can directly infer 3DGS representations from sparse…

计算机视觉与模式识别 · 计算机科学 2025-06-12 Zetian Song , Jiaye Fu , Jiaqi Zhang , Xiaohan Lu , Chuanmin Jia , Siwei Ma , Wen Gao

Recent GS-based rendering has made significant progress for LiDAR, surpassing Neural Radiance Fields (NeRF) in both quality and speed. However, these methods exhibit artifacts in extrapolated novel view synthesis due to the incomplete…

计算机视觉与模式识别 · 计算机科学 2025-11-18 Qifeng Chen , Jiarun Liu , Rengan Xie , Tao Tang , Sicong Du , Yiru Zhao , Yuchi Huo , Sheng Yang

3D asset generation is getting massive amounts of attention, inspired by the recent success of text-guided 2D content creation. Existing text-to-3D methods use pretrained text-to-image diffusion models in an optimization problem or…

计算机视觉与模式识别 · 计算机科学 2024-07-30 Lukas Höllein , Aljaž Božič , Norman Müller , David Novotny , Hung-Yu Tseng , Christian Richardt , Michael Zollhöfer , Matthias Nießner