English
Related papers

Related papers: 3D Wavelet-Based Structural Priors for Controlled …

200 papers

Convolutional Neural Networks (CNNs) are known for requiring extensive computational resources, and quantization is among the best and most common methods for compressing them. While aggressive quantization (i.e., less than 4-bits) performs…

Computer Vision and Pattern Recognition · Computer Science 2022-10-18 Shahaf E. Finder , Yair Zohav , Maor Ashkenazi , Eran Treister

Wavelets are waveform functions that describe transient and unstable variations, such as noises. In this work, we study the advantages of discrete and continuous wavelet transforms (DWT and CWT) of microlensing data to denoise them and…

Instrumentation and Methods for Astrophysics · Physics 2023-10-06 Sedighe Sajadian , Hossein Fatheddin

Myocardial perfusion imaging using SPECT is widely utilized to diagnose coronary artery diseases, but image quality can be negatively affected in low-dose and few-view acquisition settings. Although various deep learning methods have been…

Performance of deep learning models is strongly governed by architectural capacity, with width and depth as primary controls. However, in physical-science applications, models are often compared at a single fixed size or by separating…

Machine Learning · Computer Science 2026-05-07 Alexander I. Khrabry , Edward A. Startsev , Andrew T. Powis , Igor D. Kaganovich

Diffusion models have shown great promise in medical image denoising and reconstruction, but their application to Positron Emission Tomography (PET) imaging remains limited by tracer-specific contrast variability and high computational…

Image and Video Processing · Electrical Eng. & Systems 2026-02-03 Fumio Hashimoto , Kuang Gong

Objective. Limited access to breast cancer diagnosis globally leads to delayed treatment. Ultrasound, an effective yet underutilized method, requires specialized training for sonographers, which hinders its widespread use. Approach. Volume…

Image and Video Processing · Electrical Eng. & Systems 2023-11-21 Donya Khaledyan , Thomas J. Marini , Avice OConnell , Steven Meng , Jonah Kan , Galen Brennan , Yu Zhao , Timothy M. Baran , Kevin J. Parker

Although supervised convolutional neural networks (CNNs) often outperform conventional alternatives for denoising positron emission tomography (PET) images, they require many low- and high-quality reference PET image pairs. Herein, we…

Medical Physics · Physics 2021-09-29 Yuya Onishi , Fumio Hashimoto , Kibo Ote , Hiroyuki Ohba , Ryosuke Ota , Etsuji Yoshikawa , Yasuomi Ouchi

We present LooseControl to allow generalized depth conditioning for diffusion-based image generation. ControlNet, the SOTA for depth-conditioned image generation, produces remarkable results but relies on having access to detailed depth…

Computer Vision and Pattern Recognition · Computer Science 2023-12-07 Shariq Farooq Bhat , Niloy J. Mitra , Peter Wonka

Low-count positron emission tomography (LCPET) imaging can reduce patients' exposure to radiation but often suffers from increased image noise and reduced lesion detectability, necessitating effective denoising techniques. Diffusion models…

Image and Video Processing · Electrical Eng. & Systems 2025-03-24 Yinchi Zhou , Huidong Xie , Menghua Xia , Qiong Liu , Bo Zhou , Tianqi Chen , Jun Hou , Liang Guo , Xinyuan Zheng , Hanzhong Wang , Biao Li , Axel Rominger , Kuangyu Shi , Nicha C. Dvorneka , Chi Liu

Controlling the spatial and semantic structure of diffusion-generated images remains a challenge. Existing methods like ControlNet rely on handcrafted condition maps and retraining, limiting flexibility and generalization. Inversion-based…

Computer Vision and Pattern Recognition · Computer Science 2025-11-10 Jiang Lin , Xinyu Chen , Song Wu , Zhiqiu Zhang , Jizhi Zhang , Ye Wang , Qiang Tang , Qian Wang , Jian Yang , Zili Yi

Latent diffusion models for medical image super-resolution universally inherit variational autoencoders designed for natural photographs. We show that this default choice, not the diffusion architecture, is the dominant constraint on…

Spatial control methods using additional modules on pretrained diffusion models have gained attention for enabling conditional generation in natural images. These methods guide the generation process with new conditions while leveraging the…

Computer Vision and Pattern Recognition · Computer Science 2025-03-12 Suhyun Ahn , Wonjung Park , Jihoon Cho , Seunghyuck Park , Jinah Park

Due to the three-dimensional nature of CT- or MR-scans, generative modeling of medical images is a particularly challenging task. Existing approaches mostly apply patch-wise, slice-wise, or cascaded generation techniques to fit the…

Image and Video Processing · Electrical Eng. & Systems 2024-10-15 Paul Friedrich , Julia Wolleb , Florentin Bieder , Alicia Durrer , Philippe C. Cattin

We propose a dual-domain cascade of U-nets (i.e. a "W-net") operating in both the spatial frequency and image domains to enhance low-dose CT (LDCT) images without the need for proprietary x-ray projection data. The central slice theorem…

Image and Video Processing · Electrical Eng. & Systems 2020-05-28 Kevin J. Chung , Roberto Souza , Richard Frayne , Ting-Yim Lee

Accurate and automated lesion segmentation in Positron Emission Tomography / Computed Tomography (PET/CT) imaging is essential for cancer diagnosis and therapy planning. This paper presents a Swin Transformer UNet 3D (SwinUNet3D) framework…

Image and Video Processing · Electrical Eng. & Systems 2026-01-07 Shovini Guha , Dwaipayan Nandi

Low-dose computed tomography (LDCT) reduces patient radiation exposure but introduces substantial noise that degrades image quality and hinders diagnostic accuracy. Existing denoising approaches often require many diffusion steps, limiting…

Medical Physics · Physics 2025-10-29 Qiang Li , Mojtaba Safari , Shansong Wang , Huiqiao Xie , Jie Ding , Tonghe Wang , Xiaofeng Yang

Low-dose CT (LDCT) reduces radiation exposure but introduces protocol-dependent noise and artifacts that vary across institutions. While federated learning enables collaborative training without centralizing patient data, existing methods…

Image and Video Processing · Electrical Eng. & Systems 2026-03-17 Anas Zafar , Muhammad Waqas , Amgad Muneer , Rukhmini Bandyopadhyay , Jia Wu

In this paper, we present CCEdit, a versatile generative video editing framework based on diffusion models. Our approach employs a novel trident network structure that separates structure and appearance control, ensuring precise and…

Computer Vision and Pattern Recognition · Computer Science 2024-04-09 Ruoyu Feng , Wenming Weng , Yanhui Wang , Yuhui Yuan , Jianmin Bao , Chong Luo , Zhibo Chen , Baining Guo

Fully supervised change detection methods require difficult to procure pixel-level labels, while weakly supervised approaches can be trained with image-level labels. However, most of these approaches require a combination of changed and…

Computer Vision and Pattern Recognition · Computer Science 2020-11-10 Philipp Andermatt , Radu Timofte

Text-to-Image diffusion models have made tremendous progress over the past two years, enabling the generation of highly realistic images based on open-domain text descriptions. However, despite their success, text descriptions often…

Computer Vision and Pattern Recognition · Computer Science 2023-10-31 Shihao Zhao , Dongdong Chen , Yen-Chun Chen , Jianmin Bao , Shaozhe Hao , Lu Yuan , Kwan-Yee K. Wong