中文
相关论文

相关论文: WDM: 3D Wavelet Diffusion Models for High-Resoluti…

200 篇论文

Diffusion models have emerged as the leading approach for image synthesis, demonstrating exceptional photorealism and diversity. However, training diffusion models at high resolutions remains computationally prohibitive, and existing…

计算机视觉与模式识别 · 计算机科学 2025-06-26 Tobias Vontobel , Seyedmorteza Sadat , Farnood Salehi , Romann M. Weber

This paper presents the second-placed solution for task 8 and the participation solution for task 7 of BraTS 2024. The adoption of automated brain analysis algorithms to support clinical practice is increasing. However, many of these…

计算机视觉与模式识别 · 计算机科学 2024-12-03 André Ferreira , Gijs Luijten , Behrus Puladi , Jens Kleesiek , Victor Alves , Jan Egger

Diffusion models are rising as a powerful solution for high-fidelity image generation, which exceeds GANs in quality in many circumstances. However, their slow training and inference speed is a huge bottleneck, blocking them from being used…

计算机视觉与模式识别 · 计算机科学 2023-03-24 Hao Phung , Quan Dao , Anh Tran

Diffusion models have demonstrated remarkable performance in image and video synthesis. However, scaling them to high-resolution inputs is challenging and requires restructuring the diffusion pipeline into multiple independent components,…

计算机视觉与模式识别 · 计算机科学 2024-06-13 Ivan Skorokhodov , Willi Menapace , Aliaksandr Siarohin , Sergey Tulyakov

We propose an effective denoising diffusion model for generating high-resolution images (e.g., 1024$\times$512), trained on small-size image patches (e.g., 64$\times$64). We name our algorithm Patch-DM, in which a new feature collage…

计算机视觉与模式识别 · 计算机科学 2023-08-03 Zheng Ding , Mengqi Zhang , Jiajun Wu , Zhuowen Tu

We present LatentCSI, a novel method for generating images of the physical environment from WiFi CSI measurements that leverages a pretrained latent diffusion model (LDM). Unlike prior approaches that rely on complex and computationally…

计算机视觉与模式识别 · 计算机科学 2025-09-08 Eshan Ramesh , Takayuki Nishio

Continuous Conditional Diffusion Model (CCDM) is a diffusion-based framework designed to generate high-quality images conditioned on continuous regression labels. Although CCDM has demonstrated clear advantages over prior approaches across…

计算机视觉与模式识别 · 计算机科学 2026-02-03 Xin Ding , Yun Chen , Sen Zhang , Kao Zhang , Nenglun Chen , Peibei Cao , Yongwei Wang , Fei Wu

Solving 3D medical inverse problems such as image restoration and reconstruction is crucial in modern medical field. However, the curse of dimensionality in 3D medical data leads mainstream volume-wise methods to suffer from high resource…

图像与视频处理 · 电气工程与系统科学 2024-05-27 Jia He , Bonan Li , Ge Yang , Ziwen Liu

Deep neural network (DNN)-based algorithms are emerging as an important tool for many physical and MAC layer functions in future wireless communication systems, including for large multi-antenna channels. However, training such models…

信息论 · 计算机科学 2025-10-17 Taekyun Lee , Juseong Park , Hyeji Kim , Jeffrey G. Andrews

Improving the quality of hyperspectral images (HSIs), such as through super-resolution, is a crucial research area. However, generative modeling for HSIs presents several challenges. Due to their high spectral dimensionality, HSIs are too…

计算机视觉与模式识别 · 计算机科学 2025-11-11 Sirui Wang , Jiang He , Natàlia Blasco Andreo , Xiao Xiang Zhu

Recently, computer-aided diagnosis systems have been developed to support diagnosis, but their performance depends heavily on the quality and quantity of training data. However, in clinical practice, it is difficult to collect the large…

图像与视频处理 · 电气工程与系统科学 2026-01-19 Kaito Urata , Maiko Nagao , Atsushi Teramoto , Kazuyoshi Imaizumi , Masashi Kondo , Hiroshi Fujita

We introduce Efficient Motion Diffusion Model (EMDM) for fast and high-quality human motion generation. Current state-of-the-art generative diffusion models have produced impressive results but struggle to achieve fast generation without…

计算机视觉与模式识别 · 计算机科学 2024-11-26 Wenyang Zhou , Zhiyang Dou , Zeyu Cao , Zhouyingcheng Liao , Jingbo Wang , Wenjia Wang , Yuan Liu , Taku Komura , Wenping Wang , Lingjie Liu

Diagnosing medical conditions from histopathology data requires a thorough analysis across the various resolutions of Whole Slide Images (WSI). However, existing generative methods fail to consistently represent the hierarchical structure…

图像与视频处理 · 电气工程与系统科学 2024-07-19 Sarah Cechnicka , James Ball , Matthew Baugh , Hadrien Reynaud , Naomi Simmonds , Andrew P. T. Smith , Catherine Horsfield , Candice Roufosse , Bernhard Kainz

Diffusion models have emerged as the new state-of-the-art generative model with high quality samples, with intriguing properties such as mode coverage and high flexibility. They have also been shown to be effective inverse problem solvers,…

计算机视觉与模式识别 · 计算机科学 2025-10-06 Hyungjin Chung , Dohoon Ryu , Michael T. McCann , Marc L. Klasky , Jong Chul Ye

We present a data-driven method for learning to generate animations of 3D garments using a 2D image diffusion model. In contrast to existing methods, typically based on fully connected networks, graph neural networks, or generative…

计算机视觉与模式识别 · 计算机科学 2025-03-25 Raquel Vidaurre , Elena Garces , Dan Casas

This work introduces a new latent diffusion model to generate high-quality 3D chest CT scans conditioned on 3D anatomical masks. The method synthesizes volumetric images of size 256x256x256 at 1 mm isotropic resolution using a single…

计算机视觉与模式识别 · 计算机科学 2025-10-22 Anna Oliveras , Roger Marí , Rafael Redondo , Oriol Guardià , Ana Tost , Bhalaji Nagarajan , Carolina Migliorelli , Vicent Ribas , Petia Radeva

The limited availability of 3D medical image datasets, due to privacy concerns and high collection or annotation costs, poses significant challenges in the field of medical imaging. While a promising alternative is the use of synthesized…

图像与视频处理 · 电气工程与系统科学 2024-05-27 Lingting Zhu , Noel Codella , Dongdong Chen , Zhenchao Jin , Lu Yuan , Lequan Yu

Cross-modality medical image synthesis is a critical topic and has the potential to facilitate numerous applications in the medical imaging field. Despite recent successes in deep-learning-based generative models, most current medical image…

图像与视频处理 · 电气工程与系统科学 2023-07-20 Lingting Zhu , Zeyue Xue , Zhenchao Jin , Xian Liu , Jingzhen He , Ziwei Liu , Lequan Yu

Brain MRI scans are often found in four modalities, consisting of T1-weighted with and without contrast enhancement (T1ce and T1w), T2-weighted imaging (T2w), and Flair. Leveraging complementary information from these different modalities…

图像与视频处理 · 电气工程与系统科学 2025-09-22 Bhavesh Sandbhor , Bheeshm Sharma , Balamurugan Palaniappan

We introduce the Fixed Point Diffusion Model (FPDM), a novel approach to image generation that integrates the concept of fixed point solving into the framework of diffusion-based generative modeling. Our approach embeds an implicit fixed…

计算机视觉与模式识别 · 计算机科学 2024-01-18 Xingjian Bai , Luke Melas-Kyriazi