English
Related papers

Related papers: Leveling3D: Leveling Up 3D Reconstruction with Fee…

200 papers

Generating pose-aligned 3D objects is challenging due to the spatial mismatches and transformation ambiguities inherent in decoupled canonical-then-rotate paradigms. To this end, we introduce Pose-Aware Diffusion (PAD), a novel end-to-end…

Computer Vision and Pattern Recognition · Computer Science 2026-05-04 Zihan Zhou , Luxi Chen , Jingzhi Zhou , Yuhao Wan , Min Zhao , Baoyu Fan , Chongxuan Li

We have introduced SegSplat, a novel framework designed to bridge the gap between rapid, feed-forward 3D reconstruction and rich, open-vocabulary semantic understanding. By constructing a compact semantic memory bank from multi-view 2D…

Computer Vision and Pattern Recognition · Computer Science 2025-11-25 Peter Siegel , Federico Tombari , Marc Pollefeys , Daniel Barath

Recent advances in 3D generation have improved the fidelity and geometric details of synthesized 3D assets. However, due to the inherent ambiguity of single-view observations and the lack of robust global structural priors caused by limited…

Computer Vision and Pattern Recognition · Computer Science 2026-03-25 Wenyue Chen , Wenjue Chen , Peng Li , Qinghe Wang , Xu Jia , Heliang Zheng , Rongfei Jia , Yuan Liu , Ronggang Wang

Reconstructing 3D objects from extremely sparse views is a long-standing and challenging problem. While recent techniques employ image diffusion models for generating plausible images at novel viewpoints or for distilling pre-trained…

Computer Vision and Pattern Recognition · Computer Science 2023-12-21 Zi-Xin Zou , Weihao Cheng , Yan-Pei Cao , Shi-Sheng Huang , Ying Shan , Song-Hai Zhang

We present PanoPlane, an approach for high-fidelity sparse-view indoor novel view synthesis that reconstructs closed room geometry via panoramic scene completion. Unlike perspective-based methods that generate training views from limited…

Computer Vision and Pattern Recognition · Computer Science 2026-05-15 Adil Qureshi , Dongki Jung , Jaehoon Choi , Dinesh Manocha

We present GaussFusion, a novel approach for improving 3D Gaussian splatting (3DGS) reconstructions in the wild through geometry-informed video generation. GaussFusion mitigates common 3DGS artifacts, including floaters, flickering, and…

Computer Vision and Pattern Recognition · Computer Science 2026-03-31 Liyuan Zhu , Manjunath Narayana , Michal Stary , Will Hutchcroft , Gordon Wetzstein , Iro Armeni

We investigate data augmentation for 3D object detection in autonomous driving. We utilize recent advancements in 3D reconstruction based on Gaussian Splatting for 3D object placement in driving scenes. Unlike existing diffusion-based…

Computer Vision and Pattern Recognition · Computer Science 2025-04-24 Farhad G. Zanjani , Davide Abati , Auke Wiggers , Dimitris Kalatzis , Jens Petersen , Hong Cai , Amirhossein Habibian

Reconstructing 3D scenes and synthesizing novel views has seen rapid progress in recent years. Neural Radiance Fields demonstrated that continuous volumetric radiance fields can achieve high-quality image synthesis, but their long training…

Computer Vision and Pattern Recognition · Computer Science 2025-09-30 Jan Held , Renaud Vandeghen , Sanghyun Son , Daniel Rebain , Matheus Gadelha , Yi Zhou , Ming C. Lin , Marc Van Droogenbroeck , Andrea Tagliasacchi

Accurate geometric surface reconstruction, providing essential environmental information for navigation and manipulation tasks, is critical for enabling robotic self-exploration and interaction. Recently, 3D Gaussian Splatting (3DGS) has…

Computer Vision and Pattern Recognition · Computer Science 2025-07-17 Tengfei Wang , Xin Wang , Yongmao Hou , Zhaoning Zhang , Yiwei Xu , Zongqian Zhan

Recent approaches integrating vision-language models (VLMs) as prompt encoders for generative model conditioning typically rely on expensive end-to-end training or map features to compressed representations, discarding the dense spatial…

Computer Vision and Pattern Recognition · Computer Science 2026-05-29 Polytimi Anna Gkotsi , Andrii Zadaianchuk , Mohammad Mahdi Derakhshani

3D Gaussian Splatting has achieved impressive performance in novel view synthesis with real-time rendering capabilities. However, reconstructing high-quality surfaces with fine details using 3D Gaussians remains a challenging task. In this…

Computer Vision and Pattern Recognition · Computer Science 2024-12-03 Jiepeng Wang , Yuan Liu , Peng Wang , Cheng Lin , Junhui Hou , Xin Li , Taku Komura , Wenping Wang

3D Gaussian Splatting (3DGS) and its subsequent works are restricted to specific hardware setups, either on only low-cost or on only high-end configurations. Approaches aimed at reducing 3DGS memory usage enable rendering on low-cost GPU…

Computer Vision and Pattern Recognition · Computer Science 2025-06-12 Yunji Seo , Young Sun Choi , Hyun Seung Son , Youngjung Uh

During the Gaussian Splatting optimization process, the scene's geometry can gradually deteriorate if its structure is not deliberately preserved, especially in non-textured regions such as walls, ceilings, and furniture surfaces. This…

Computer Vision and Pattern Recognition · Computer Science 2024-07-18 Yanyan Li , Chenyu Lyu , Yan Di , Guangyao Zhai , Gim Hee Lee , Federico Tombari

We present LangFlash, a feed-forward framework for 3D Language Gaussian Splatting that reconstructs 3D scenes parameterized by Gaussian primitives enriched with language-aligned semantic features from sparse unposed multi-view images.…

Computer Vision and Pattern Recognition · Computer Science 2026-05-25 Yilong Liu , Wanhua Li , Chen Zhu-Tian , Hanspeter Pfister

Reconstructing 3D scenes using 3D Gaussian Splatting (3DGS) from sparse views is an ill-posed problem due to insufficient information, often resulting in noticeable artifacts. While recent approaches have sought to leverage generative…

Computer Vision and Pattern Recognition · Computer Science 2025-08-14 Xingyilang Yin , Qi Zhang , Jiahao Chang , Ying Feng , Qingnan Fan , Xi Yang , Chi-Man Pun , Huaqi Zhang , Xiaodong Cun

Gaussian Splatting has emerged as a leading method for novel view synthesis, offering superior training efficiency and real-time inference compared to NeRF approaches, while still delivering high-quality reconstructions. Beyond view…

Computer Vision and Pattern Recognition · Computer Science 2025-11-25 Lorenzo Rutayisire , Nicola Capodieci , Fabio Pellacini

Intraoral 3D reconstruction is fundamental to digital orthodontics, yet conventional methods like intraoral scanning are inaccessible for remote tele-orthodontics, which typically relies on sparse smartphone imagery. While 3D Gaussian…

Computer Vision and Pattern Recognition · Computer Science 2025-11-19 Yiyi Miao , Taoyu Wu , Tong Chen , Ji Jiang , Zhe Tang , Zhengyong Jiang , Angelos Stefanidis , Limin Yu , Jionglong Su

The creation of high-quality 3D assets is paramount for applications in digital heritage preservation, entertainment, and robotics. Traditionally, this process necessitates skilled professionals and specialized software for the modeling,…

Computer Vision and Pattern Recognition · Computer Science 2024-07-30 Shen Chen , Jiale Zhou , Zhongyu Jiang , Tianfang Zhang , Zongkai Wu , Jenq-Neng Hwang , Lei Li

This paper presents a novel method for building scalable 3D generative models utilizing pre-trained video diffusion models. The primary obstacle in developing foundation 3D generative models is the limited availability of 3D data. Unlike…

Computer Vision and Pattern Recognition · Computer Science 2024-07-22 Junlin Han , Filippos Kokkinos , Philip Torr

Recent AI-based 3D content creation has largely evolved along two paths: feed-forward image-to-3D reconstruction approaches and 3D generative models trained with 2D or 3D supervision. In this work, we show that existing feed-forward…

Computer Vision and Pattern Recognition · Computer Science 2025-01-07 Suttisak Wizadwongsa , Jinfan Zhou , Edward Li , Jeong Joon Park