English
Related papers

Related papers: 3DTopia: Large Text-to-3D Generation Model with Hy…

200 papers

The field of neural rendering has witnessed significant progress with advancements in generative models and differentiable rendering techniques. Though 2D diffusion has achieved success, a unified 3D diffusion pipeline remains unsettled.…

Computer Vision and Pattern Recognition · Computer Science 2025-12-22 Yushi Lan , Fangzhou Hong , Shangchen Zhou , Shuai Yang , Xuyi Meng , Yongwei Chen , Zhaoyang Lyu , Bo Dai , Xingang Pan , Chen Change Loy

Text-to-3D is an emerging task that allows users to create 3D content with infinite possibilities. Existing works tackle the problem by optimizing a 3D representation with guidance from pre-trained diffusion models. An apparent drawback is…

Computer Vision and Pattern Recognition · Computer Science 2023-06-06 Yiji Cheng , Fei Yin , Xiaoke Huang , Xintong Yu , Jiaxiang Liu , Shikun Feng , Yujiu Yang , Yansong Tang

We introduce TurboPortrait3D: a method for low-latency novel-view synthesis of human portraits. Our approach builds on the observation that existing image-to-3D models for portrait generation, while capable of producing renderable 3D…

Computer Vision and Pattern Recognition · Computer Science 2025-10-29 Emily Kim , Julieta Martinez , Timur Bagautdinov , Jessica Hodgins

Text-to-Image (T2I) generation methods based on diffusion model have garnered significant attention in the last few years. Although these image synthesis methods produce visually appealing results, they frequently exhibit spelling errors…

Computer Vision and Pattern Recognition · Computer Science 2023-12-27 Yiming Zhao , Zhouhui Lian

In this paper, we introduce a novel 3D-aware image generation method that leverages 2D diffusion models. We formulate the 3D-aware image generation task as multiview 2D image set generation, and further to a sequential…

Computer Vision and Pattern Recognition · Computer Science 2023-04-03 Jianfeng Xiang , Jiaolong Yang , Binbin Huang , Xin Tong

The diffusion model performs remarkable in generating high-dimensional content but is computationally intensive, especially during training. We propose Progressive Growing of Diffusion Autoencoder (PaGoDA), a novel pipeline that reduces the…

Computer Vision and Pattern Recognition · Computer Science 2024-10-30 Dongjun Kim , Chieh-Hsin Lai , Wei-Hsiang Liao , Yuhta Takida , Naoki Murata , Toshimitsu Uesaka , Yuki Mitsufuji , Stefano Ermon

While high-quality texture maps are essential for realistic 3D asset rendering, few studies have explored learning directly in the texture space, especially on large-scale datasets. In this work, we depart from the conventional approach of…

Computer Vision and Pattern Recognition · Computer Science 2024-11-25 Xin Yu , Ze Yuan , Yuan-Chen Guo , Ying-Tian Liu , JianHui Liu , Yangguang Li , Yan-Pei Cao , Ding Liang , Xiaojuan Qi

Text-to-3D generation aims to create 3D assets from text-to-image diffusion models. However, existing methods face an inherent bottleneck in generation quality because the widely-used objectives such as Score Distillation Sampling (SDS)…

Computer Vision and Pattern Recognition · Computer Science 2024-06-24 Zixuan Chen , Ruijie Su , Jiahao Zhu , Lingxiao Yang , Jian-Huang Lai , Xiaohua Xie

We introduce DD3G, a formulation that Distills a multi-view Diffusion model (MV-DM) into a 3D Generator using gaussian splatting. DD3G compresses and integrates extensive visual and spatial geometric knowledge from the MV-DM by simulating…

Computer Vision and Pattern Recognition · Computer Science 2025-04-04 Hao Qin , Luyuan Chen , Ming Kong , Mengxu Lu , Qiang Zhu

In this paper, we present TEXTure, a novel method for text-guided generation, editing, and transfer of textures for 3D shapes. Leveraging a pretrained depth-to-image diffusion model, TEXTure applies an iterative scheme that paints a 3D…

Computer Vision and Pattern Recognition · Computer Science 2023-02-06 Elad Richardson , Gal Metzer , Yuval Alaluf , Raja Giryes , Daniel Cohen-Or

3D generation methods have shown visually compelling results powered by diffusion image priors. However, they often fail to produce realistic geometric details, resulting in overly smooth surfaces or geometric details inaccurately baked in…

Computer Vision and Pattern Recognition · Computer Science 2024-12-10 Ruihan Gao , Kangle Deng , Gengshan Yang , Wenzhen Yuan , Jun-Yan Zhu

The ability to generate 3D multiphase microstructures on-demand with targeted attributes can greatly accelerate the design of advanced materials. Here, we present a conditional latent diffusion model (LDM) framework that rapidly synthesizes…

While recent work on text-conditional 3D object generation has shown promising results, the state-of-the-art methods typically require multiple GPU-hours to produce a single sample. This is in stark contrast to state-of-the-art generative…

Computer Vision and Pattern Recognition · Computer Science 2022-12-20 Alex Nichol , Heewoo Jun , Prafulla Dhariwal , Pamela Mishkin , Mark Chen

The generation of medical images presents significant challenges due to their high-resolution and three-dimensional nature. Existing methods often yield suboptimal performance in generating high-quality 3D medical images, and there is…

Image and Video Processing · Electrical Eng. & Systems 2025-12-02 Haoshen Wang , Zhentao Liu , Kaicong Sun , Xiaodong Wang , Dinggang Shen , Zhiming Cui

Text-driven 3D scene generation is widely applicable to video gaming, film industry, and metaverse applications that have a large demand for 3D scenes. However, existing text-to-3D generation methods are limited to producing 3D objects with…

Computer Vision and Pattern Recognition · Computer Science 2024-02-01 Jingbo Zhang , Xiaoyu Li , Ziyu Wan , Can Wang , Jing Liao

We introduce Imagen 3, a latent diffusion model that generates high quality images from text prompts. We describe our quality and responsibility evaluations. Imagen 3 is preferred over other state-of-the-art (SOTA) models at the time of…

Computer Vision and Pattern Recognition · Computer Science 2024-12-24 Imagen-Team-Google , : , Jason Baldridge , Jakob Bauer , Mukul Bhutani , Nicole Brichtova , Andrew Bunner , Lluis Castrejon , Kelvin Chan , Yichang Chen , Sander Dieleman , Yuqing Du , Zach Eaton-Rosen , Hongliang Fei , Nando de Freitas , Yilin Gao , Evgeny Gladchenko , Sergio Gómez Colmenarejo , Mandy Guo , Alex Haig , Will Hawkins , Hexiang Hu , Huilian Huang , Tobenna Peter Igwe , Christos Kaplanis , Siavash Khodadadeh , Yelin Kim , Ksenia Konyushkova , Karol Langner , Eric Lau , Rory Lawton , Shixin Luo , Soňa Mokrá , Henna Nandwani , Yasumasa Onoe , Aäron van den Oord , Zarana Parekh , Jordi Pont-Tuset , Hang Qi , Rui Qian , Deepak Ramachandran , Poorva Rane , Abdullah Rashwan , Ali Razavi , Robert Riachi , Hansa Srinivasan , Srivatsan Srinivasan , Robin Strudel , Benigno Uria , Oliver Wang , Su Wang , Austin Waters , Chris Wolff , Auriel Wright , Zhisheng Xiao , Hao Xiong , Keyang Xu , Marc van Zee , Junlin Zhang , Katie Zhang , Wenlei Zhou , Konrad Zolna , Ola Aboubakar , Canfer Akbulut , Oscar Akerlund , Isabela Albuquerque , Nina Anderson , Marco Andreetto , Lora Aroyo , Ben Bariach , David Barker , Sherry Ben , Dana Berman , Courtney Biles , Irina Blok , Pankil Botadra , Jenny Brennan , Karla Brown , John Buckley , Rudy Bunel , Elie Bursztein , Christina Butterfield , Ben Caine , Viral Carpenter , Norman Casagrande , Ming-Wei Chang , Solomon Chang , Shamik Chaudhuri , Tony Chen , John Choi , Dmitry Churbanau , Nathan Clement , Matan Cohen , Forrester Cole , Mikhail Dektiarev , Vincent Du , Praneet Dutta , Tom Eccles , Ndidi Elue , Ashley Feden , Shlomi Fruchter , Frankie Garcia , Roopal Garg , Weina Ge , Ahmed Ghazy , Bryant Gipson , Andrew Goodman , Dawid Górny , Sven Gowal , Khyatti Gupta , Yoni Halpern , Yena Han , Susan Hao , Jamie Hayes , Jonathan Heek , Amir Hertz , Ed Hirst , Emiel Hoogeboom , Tingbo Hou , Heidi Howard , Mohamed Ibrahim , Dirichi Ike-Njoku , Joana Iljazi , Vlad Ionescu , William Isaac , Reena Jana , Gemma Jennings , Donovon Jenson , Xuhui Jia , Kerry Jones , Xiaoen Ju , Ivana Kajic , Christos Kaplanis , Burcu Karagol Ayan , Jacob Kelly , Suraj Kothawade , Christina Kouridi , Ira Ktena , Jolanda Kumakaw , Dana Kurniawan , Dmitry Lagun , Lily Lavitas , Jason Lee , Tao Li , Marco Liang , Maggie Li-Calis , Yuchi Liu , Javier Lopez Alberca , Matthieu Kim Lorrain , Peggy Lu , Kristian Lum , Yukun Ma , Chase Malik , John Mellor , Thomas Mensink , Inbar Mosseri , Tom Murray , Aida Nematzadeh , Paul Nicholas , Signe Nørly , João Gabriel Oliveira , Guillermo Ortiz-Jimenez , Michela Paganini , Tom Le Paine , Roni Paiss , Alicia Parrish , Anne Peckham , Vikas Peswani , Igor Petrovski , Tobias Pfaff , Alex Pirozhenko , Ryan Poplin , Utsav Prabhu , Yuan Qi , Matthew Rahtz , Cyrus Rashtchian , Charvi Rastogi , Amit Raul , Ali Razavi , Sylvestre-Alvise Rebuffi , Susanna Ricco , Felix Riedel , Dirk Robinson , Pankaj Rohatgi , Bill Rosgen , Sarah Rumbley , Moonkyung Ryu , Anthony Salgado , Tim Salimans , Sahil Singla , Florian Schroff , Candice Schumann , Tanmay Shah , Eleni Shaw , Gregory Shaw , Brendan Shillingford , Kaushik Shivakumar , Dennis Shtatnov , Zach Singer , Evgeny Sluzhaev , Valerii Sokolov , Thibault Sottiaux , Florian Stimberg , Brad Stone , David Stutz , Yu-Chuan Su , Eric Tabellion , Shuai Tang , David Tao , Kurt Thomas , Gregory Thornton , Andeep Toor , Cristian Udrescu , Aayush Upadhyay , Cristina Vasconcelos , Alex Vasiloff , Andrey Voynov , Amanda Walker , Luyu Wang , Miaosen Wang , Simon Wang , Stanley Wang , Qifei Wang , Yuxiao Wang , Ágoston Weisz , Olivia Wiles , Chenxia Wu , Xingyu Federico Xu , Andrew Xue , Jianbo Yang , Luo Yu , Mete Yurtoglu , Ali Zand , Han Zhang , Jiageng Zhang , Catherine Zhao , Adilet Zhaxybay , Miao Zhou , Shengqi Zhu , Zhenkai Zhu , Dawn Bloxwich , Mahyar Bordbar , Luis C. Cobo , Eli Collins , Shengyang Dai , Tulsee Doshi , Anca Dragan , Douglas Eck , Demis Hassabis , Sissie Hsiao , Tom Hume , Koray Kavukcuoglu , Helen King , Jack Krawczyk , Yeqing Li , Kathy Meier-Hellstern , Andras Orban , Yury Pinsky , Amar Subramanya , Oriol Vinyals , Ting Yu , Yori Zwols

Articulated object generation has seen increasing advancements, yet existing models often lack the ability to be conditioned on text prompts. To address the significant gap between textual descriptions and 3D articulated object…

Computer Vision and Pattern Recognition · Computer Science 2025-12-04 Hao Sun , Lei Fan , Donglin Di , Shaohui Liu

Diffusion and flow matching models have achieved remarkable success in text-to-image generation. However, these models typically rely on the predetermined denoising schedules for all prompts. The multi-step reverse diffusion process can be…

Computer Vision and Pattern Recognition · Computer Science 2025-03-06 Zilyu Ye , Zhiyang Chen , Tiancheng Li , Zemin Huang , Weijian Luo , Guo-Jun Qi

Traditional 3D content creation tools empower users to bring their imagination to life by giving them direct control over a scene's geometry, appearance, motion, and camera path. Creating computer-generated videos, however, is a tedious…

Computer Vision and Pattern Recognition · Computer Science 2023-12-05 Shengqu Cai , Duygu Ceylan , Matheus Gadelha , Chun-Hao Paul Huang , Tuanfeng Yang Wang , Gordon Wetzstein

We present LT3SD, a novel latent diffusion model for large-scale 3D scene generation. Recent advances in diffusion models have shown impressive results in 3D object generation, but are limited in spatial extent and quality when extended to…

Computer Vision and Pattern Recognition · Computer Science 2025-05-02 Quan Meng , Lei Li , Matthias Nießner , Angela Dai