English
Related papers

Related papers: LATTICE: Democratize High-Fidelity 3D Generation a…

200 papers

Recent advances in sparse voxel representations have significantly improved the quality of 3D content generation, enabling high-resolution modeling with fine-grained geometry. However, existing frameworks suffer from severe computational…

Computer Vision and Pattern Recognition · Computer Science 2025-08-01 Yiwen Chen , Zhihao Li , Yikai Wang , Hu Zhang , Qin Li , Chi Zhang , Guosheng Lin

Recent advancements in 3D generative modeling have significantly improved the generation realism, yet the field is still hampered by existing representations, which struggle to capture assets with complex topologies and detailed appearance.…

Computer Vision and Pattern Recognition · Computer Science 2025-12-17 Jianfeng Xiang , Xiaoxue Chen , Sicheng Xu , Ruicheng Wang , Zelong Lv , Yu Deng , Hongyuan Zhu , Yue Dong , Hao Zhao , Nicholas Jing Yuan , Jiaolong Yang

We introduce a novel 3D generation method for versatile and high-quality 3D asset creation. The cornerstone is a unified Structured LATent (SLAT) representation which allows decoding to different output formats, such as Radiance Fields, 3D…

Computer Vision and Pattern Recognition · Computer Science 2025-06-02 Jianfeng Xiang , Zelong Lv , Sicheng Xu , Yu Deng , Ruicheng Wang , Bowen Zhang , Dong Chen , Xin Tong , Jiaolong Yang

Prevailing 3D texture generation methods, which often rely on multi-view fusion, are frequently hindered by inter-view inconsistencies and incomplete coverage of complex surfaces, limiting the fidelity and completeness of the generated…

Computer Vision and Pattern Recognition · Computer Science 2025-12-03 Yifei Zeng , Yajie Bao , Jiachen Qian , Shuang Wu , Youtian Lin , Hao Zhu , Buyu Li , Feihu Zhang , Xun Cao , Yao Yao

While generative artificial intelligence has advanced significantly across text, image, audio, and video domains, 3D generation remains comparatively underdeveloped due to fundamental challenges such as data scarcity, algorithmic…

Computer Vision and Pattern Recognition · Computer Science 2025-05-13 Weiyu Li , Xuanyang Zhang , Zheng Sun , Di Qi , Hao Li , Wei Cheng , Weiwei Cai , Shihao Wu , Jiarui Liu , Zihao Wang , Xiao Chen , Feipeng Tian , Jianxiong Pan , Zeming Li , Gang Yu , Xiangyu Zhang , Daxin Jiang , Ping Tan

In this paper, we introduce LATO, a novel topology-preserving latent representation that enables scalable, flow matching-based synthesis of explicit 3D meshes. LATO represents a mesh as a Vertex Displacement Field (VDF) anchored on surface,…

Computer Vision and Pattern Recognition · Computer Science 2026-03-09 Tianhao Zhao , Youjia Zhang , Hang Long , Jinshen Zhang , Wenbing Li , Yang Yang , Gongbo Zhang , Jozef Hladký , Matthias Nießner , Wei Yang

Despite the availability of large-scale 3D datasets and advancements in 3D generative models, the complexity and uneven quality of 3D geometry and texture data continue to hinder the performance of 3D generation techniques. In most existing…

Computer Vision and Pattern Recognition · Computer Science 2025-05-29 Xin Yang , Jiantao Lin , Yingjie Xu , Haodong Li , Yingcong Chen

We present XCube (abbreviated as $\mathcal{X}^3$), a novel generative model for high-resolution sparse 3D voxel grids with arbitrary attributes. Our model can generate millions of voxels with a finest effective resolution of up to $1024^3$…

Computer Vision and Pattern Recognition · Computer Science 2024-06-26 Xuanchi Ren , Jiahui Huang , Xiaohui Zeng , Ken Museth , Sanja Fidler , Francis Williams

Part-level 3D generation is essential for applications requiring decomposable and structured 3D synthesis. However, existing methods either rely on implicit part segmentation with limited granularity control or depend on strong external…

Computer Vision and Pattern Recognition · Computer Science 2026-03-30 Xufan He , Yushuang Wu , Xiaoyang Guo , Chongjie Ye , Jiaqing Zhou , Tianlei Hu , Xiaoguang Han , Dong Du

State-of-the-art 3D-aware generative models rely on coordinate-based MLPs to parameterize 3D radiance fields. While demonstrating impressive results, querying an MLP for every sample along each ray leads to slow rendering. Therefore,…

Computer Vision and Pattern Recognition · Computer Science 2022-11-11 Katja Schwarz , Axel Sauer , Michael Niemeyer , Yiyi Liao , Andreas Geiger

Recent text-to-3D generation approaches produce impressive 3D results but require time-consuming optimization that can take up to an hour per prompt. Amortized methods like ATT3D optimize multiple prompts simultaneously to improve…

Computer Vision and Pattern Recognition · Computer Science 2024-03-25 Kevin Xie , Jonathan Lorraine , Tianshi Cao , Jun Gao , James Lucas , Antonio Torralba , Sanja Fidler , Xiaohui Zeng

We propose ArtiLatent, a generative framework that synthesizes human-made 3D objects with fine-grained geometry, accurate articulation, and realistic appearance. Our approach jointly models part geometry and articulation dynamics by…

Computer Vision and Pattern Recognition · Computer Science 2025-10-27 Honghua Chen , Yushi Lan , Yongwei Chen , Xingang Pan

Autoregressive transformers have revolutionized high-fidelity image generation. One crucial ingredient lies in the tokenizer, which compresses high-resolution image patches into manageable discrete tokens with a scanning or hierarchical…

Computer Vision and Pattern Recognition · Computer Science 2024-12-04 Jinzhi Zhang , Feng Xiong , Mu Xu

This paper presents a novel latent 3D diffusion model for the generation of neural voxel fields, aiming to achieve accurate part-aware structures. Compared to existing methods, there are two key designs to ensure high-quality and accurate…

Computer Vision and Pattern Recognition · Computer Science 2025-05-05 Yuhang Huang , SHilong Zou , Xinwang Liu , Kai Xu

We present VoxScene, a novel anchor-conditioned voxel diffusion framework tailored for 3D scene synthesis. Current data-driven layout generation techniques typically rely on bounding proxies or implicit representations, which overlook…

We introduce DMTet, a deep 3D conditional generative model that can synthesize high-resolution 3D shapes using simple user guides such as coarse voxels. It marries the merits of implicit and explicit 3D representations by leveraging a novel…

Computer Vision and Pattern Recognition · Computer Science 2021-11-09 Tianchang Shen , Jun Gao , Kangxue Yin , Ming-Yu Liu , Sanja Fidler

3D generation has witnessed significant advancements, yet efficiently producing high-quality 3D assets from a single image remains challenging. In this paper, we present a triplane autoencoder, which encodes 3D models into a compact…

Computer Vision and Pattern Recognition · Computer Science 2024-03-21 Bowen Zhang , Tianyu Yang , Yu Li , Lei Zhang , Xi Zhao

High-fidelity 3D asset generation is crucial for various industries. While recent 3D pretrained models show strong capability in producing realistic content, most are built upon diffusion models and follow a two-stage pipeline that first…

Computer Vision and Pattern Recognition · Computer Science 2025-09-30 Guanjun Wu , Jiemin Fang , Chen Yang , Sikuang Li , Taoran Yi , Jia Lu , Zanwei Zhou , Jiazhong Cen , Lingxi Xie , Xiaopeng Zhang , Wei Wei , Wenyu Liu , Xinggang Wang , Qi Tian

High-resolution image synthesis remains a core challenge in generative modeling, particularly in balancing computational efficiency with the preservation of fine-grained visual detail. We present Latent Wavelet Diffusion (LWD), a…

Computer Vision and Pattern Recognition · Computer Science 2026-04-17 Luigi Sigillo , Shengfeng He , Danilo Comminiello

3D modeling is moving from virtual to physical. Existing 3D generation primarily emphasizes geometries and textures while neglecting physical-grounded modeling. Consequently, despite the rapid development of 3D generative models, the…

Computer Vision and Pattern Recognition · Computer Science 2025-12-01 Ziang Cao , Zhaoxi Chen , Liang Pan , Ziwei Liu
‹ Prev 1 2 3 10 Next ›