中文
相关论文

相关论文: PoreDiT: A Scalable Generative Model for Large-Sca…

200 篇论文

Diffusion Transformers have recently shown remarkable effectiveness in generating high-quality 3D point clouds. However, training voxel-based diffusion models for high-resolution 3D voxels remains prohibitively expensive due to the cubic…

计算机视觉与模式识别 · 计算机科学 2023-12-13 Shentong Mo , Enze Xie , Yue Wu , Junsong Chen , Matthias Nießner , Zhenguo Li

Predicting molecular conformations from molecular graphs is a fundamental problem in cheminformatics and drug discovery. Recently, significant progress has been achieved with machine learning approaches, especially with deep generative…

机器学习 · 计算机科学 2022-03-16 Minkai Xu , Lantao Yu , Yang Song , Chence Shi , Stefano Ermon , Jian Tang

Porous media are ubiquitous in energy storage and conversion, catalysis, biomechanics, hydrogeology, as well as many other fields. These materials possess high surface-to-volume ratios and their complex channels can restrict and guide the…

流体动力学 · 物理学 2025-12-04 Olivier Guévremont , Lucka Barbeau , Vaiana Moreau , Federico Galli , Nick Virgilio , Bruno Blais

The increasing demand for high-quality 3D assets across various industries necessitates efficient and automated 3D content creation. Despite recent advancements in 3D generative models, existing methods still face challenges with…

计算机视觉与模式识别 · 计算机科学 2025-03-18 Zhaoxi Chen , Jiaxiang Tang , Yuhao Dong , Ziang Cao , Fangzhou Hong , Yushi Lan , Tengfei Wang , Haozhe Xie , Tong Wu , Shunsuke Saito , Liang Pan , Dahua Lin , Ziwei Liu

Pore-scale modeling and simulation of reactive flow in porous media has a range of diverse applications, and poses a number of research challenges. It is known that the morphology of a porous medium has significant influence on the local…

计算工程、金融与科学 · 计算机科学 2015-07-09 Oleg Iliev , Zahra Lakdawala , Katherine Leonard , Yavor Vutov

Pore-scale simulations accurately describe transport properties of fluids in the subsurface. These simulations enhance our understanding of applications such as assessing hydrogen storage efficiency and forecasting CO$_2$ sequestration…

Probabilistic super-resolution of high-dimensional spatial fields using diffusion models is often computationally prohibitive due to the cost of operating directly in pixel space. We propose PODiff, a structured conditional generative…

机器学习 · 计算机科学 2026-05-06 Onkar Jadhav , Tim French , Matthew Rayson , Nicole L. Jones

Microstructure reconstruction serves as a crucial foundation for establishing Process-Structure-Property (PSP) relationship in material design. Confronting the limitations of variational autoencoder and generative adversarial network within…

计算工程、金融与科学 · 计算机科学 2023-11-30 Xianrui Lyu , Xiaodan Ren

We introduce a novel 3D generative method, Generative 3D Reconstruction (G3DR) in ImageNet, capable of generating diverse and high-quality 3D objects from single images, addressing the limitations of existing methods. At the heart of our…

计算机视觉与模式识别 · 计算机科学 2024-04-04 Pradyumna Reddy , Ismail Elezi , Jiankang Deng

We introduce GeoDiT, a diffusion transformer designed for text-to-satellite image generation with point-based control. Existing controlled satellite image generative models often require pixel-level maps that are time-consuming to acquire,…

计算机视觉与模式识别 · 计算机科学 2026-03-03 Srikumar Sastry , Dan Cher , Brian Wei , Aayush Dhakal , Subash Khanal , Dev Gupta , Nathan Jacobs

In the visual generative area, discrete diffusion models are gaining traction for their efficiency and compatibility. However, pioneered attempts still fall behind their continuous counterparts, which we attribute to noise (absorbing state)…

计算机视觉与模式识别 · 计算机科学 2025-09-30 Tianren Ma , Xiaosong Zhang , Boyu Yang , Junlan Feng , Qixiang Ye

Generating ground-level views and coherent 3D site models from aerial-only imagery is challenging due to extreme viewpoint changes, missing intermediate observations, and large scale variations. Existing methods either refine renderings…

计算机视觉与模式识别 · 计算机科学 2026-04-14 Sirshapan Mitra , Yogesh S. Rawat

Field-of-view and resolution trade-offs in X-Ray micro-computed tomography (micro-CT) imaging limit the characterization, analysis and model development of multi-scale porous systems. To this end, we developed an applied methodology…

地球物理 · 物理学 2022-03-17 Samuel J. Jackson , Yufu Niu , Sojwal Manoorkar , Peyman Mostaghimi , Ryan T. Armstrong

This paper explores image modeling from the frequency space and introduces DCTdiff, an end-to-end diffusion generative paradigm that efficiently models images in the discrete cosine transform (DCT) space. We investigate the design space of…

计算机视觉与模式识别 · 计算机科学 2025-06-02 Mang Ning , Mingxiao Li , Jianlin Su , Haozhe Jia , Lanmiao Liu , Martin Beneš , Wenshuo Chen , Albert Ali Salah , Itir Onal Ertugrul

Digital modeling of the microstructure is important for studying the physical and transport properties of porous media. Multiscale modeling for porous media can accurately characterize macro-pores and micro-pores in a large-FoV (field of…

图像与视频处理 · 电气工程与系统科学 2023-05-03 Pengcheng Yan , Qizhi Teng , Xiaohai He , Zhenchuan Ma , Ningning Zhang

Recent breakthroughs in Diffusion Transformers (DiTs) have revolutionized the field of visual synthesis due to their superior scalability. To facilitate DiTs' capability of capturing meaningful internal representations, recent works such as…

计算机视觉与模式识别 · 计算机科学 2026-03-05 Mengping Yang , Zhiyu Tan , Binglei Li , Xiaomeng Yang , Hesen Chen , Hao Li

The Diffusion Transformer (DiT) architecture is the state-of-the-art paradigm for high-fidelity image generation, underpinning models like Stable Diffusion-3 and FLUX.1. However, deploying these models on resource-constrained mobile devices…

计算机视觉与模式识别 · 计算机科学 2026-05-18 Kunpeng Du , Haizhen Xie , Sen Lu , Lei Yu , Binglei Bao , Huaao Tang , Chuntao Liu , Hao Wu , Yang Zhao , Zhicai Huang , Heyuan Gao , Zhijun Tu , Jie Hu , Xinghao Chen

Speech super-resolution (SR) is the task that restores high-resolution speech from low-resolution input. Existing models employ simulated data and constrained experimental settings, which limit generalization to real-world SR. Predictive…

音频与语音处理 · 电气工程与系统科学 2024-01-26 Heming Wang , Eric W. Healy , DeLiang Wang

Reconstructing large-scale colored point clouds is an important task in robotics, supporting perception, navigation, and scene understanding. Despite advances in LiDAR inertial visual odometry (LIVO), its performance remains highly…

机器人学 · 计算机科学 2025-11-04 Lijie Wang , Lianjie Guo , Ziyi Xu , Qianhao Wang , Fei Gao , Xieyuanli Chen

Diffusion Models (DMs) have achieved State-Of-The-Art (SOTA) results in the Lidar point cloud generation task, benefiting from their stable training and iterative refinement during sampling. However, DMs often fail to realistically model…

计算机视觉与模式识别 · 计算机科学 2024-04-09 Hamed Haghighi , Amir Samadi , Mehrdad Dianati , Valentina Donzella , Kurt Debattista