English
Related papers

Related papers: MCAD: Multi-modal Conditioned Adversarial Diffusio…

200 papers

In the diverse field of medical imaging, automatic segmentation has numerous applications and must handle a wide variety of input domains, such as different types of Computed Tomography (CT) scans and Magnetic Resonance (MR) images. This…

Image and Video Processing · Electrical Eng. & Systems 2024-11-26 Chengyin Li , Hui Zhu , Rafi Ibn Sultan , Hassan Bagher Ebadian , Prashant Khanduri , Chetty Indrin , Kundan Thind , Dongxiao Zhu

Recent advances in denoising diffusion probabilistic models have shown great success in image synthesis tasks. While there are already works exploring the potential of this powerful tool in image semantic segmentation, its application in…

Computer Vision and Pattern Recognition · Computer Science 2023-09-19 Xinrong Hu , Yu-Jen Chen , Tsung-Yi Ho , Yiyu Shi

Text-to-image generative models have achieved remarkable breakthroughs in recent years. However, their application in medical image generation still faces significant challenges, including small dataset sizes, and scarcity of medical…

Computer Vision and Pattern Recognition · Computer Science 2025-06-26 Changlu Guo , Anders Nymark Christensen , Morten Rieger Hannemose

X-ray computed tomography (CT) is widely used in clinical practice. The involved ionizing X-ray radiation, however, could increase cancer risk. Hence, the reduction of the radiation dose has been an important topic in recent years. Few-view…

Image and Video Processing · Electrical Eng. & Systems 2019-12-17 Huidong Xie , Hongming Shan , Ge Wang

Ultra-low-dose positron emission tomography (PET) reconstruction holds significant potential for reducing patient radiation exposure and shortening examination times. However, it may also lead to increased noise and reduced imaging detail,…

Computer Vision and Pattern Recognition · Computer Science 2025-09-03 Mengxiao Geng , Ran Hong , Bingxuan Li , Qiegen Liu

We present ENTED, a new framework for blind face restoration that aims to restore high-quality and realistic portrait images. Our method involves repairing a single degraded input image using a high-quality reference image. We utilize a…

Computer Vision and Pattern Recognition · Computer Science 2024-01-17 Yuen-Fui Lau , Tianjia Zhang , Zhefan Rao , Qifeng Chen

Discrete diffusion models have recently shown great promise for modeling complex discrete data, with masked diffusion models (MDMs) offering a compelling trade-off between quality and generation speed. MDMs denoise by progressively…

Machine Learning · Computer Science 2026-04-15 Tianyu Xie , Shuchen Xue , Zijin Feng , Tianyang Hu , Jiacheng Sun , Zhenguo Li , Cheng Zhang

The computational intensity of detector simulation and event reconstruction poses a significant difficulty for data analysis in collider experiments. This challenge inspires the continued development of machine learning techniques to serve…

High Energy Physics - Experiment · Physics 2024-11-22 Dmitrii Kobylianskii , Nathalie Soybelman , Nilotpal Kakati , Etienne Dreyer , Benjamin Nachman , Eilam Gross

Diffusion models have become the mainstream architecture for text-to-image generation, achieving remarkable progress in visual quality and prompt controllability. However, current inference pipelines generally lack interpretable semantic…

Computer Vision and Pattern Recognition · Computer Science 2025-05-27 Zheqi Lv , Junhao Chen , Qi Tian , Keting Yin , Shengyu Zhang , Fei Wu

In this paper, we formulate a potentially valuable panoramic depth completion (PDC) task as panoramic 3D cameras often produce 360{\deg} depth with missing data in complex scenes. Its goal is to recover dense panoramic depths from raw…

Computer Vision and Pattern Recognition · Computer Science 2022-07-13 Zhiqiang Yan , Xiang Li , Kun Wang , Zhenyu Zhang , Jun Li , Jian Yang

As a class of fruitful approaches, diffusion probabilistic models (DPMs) have shown excellent advantages in high-resolution image reconstruction. On the other hand, masked autoencoders (MAEs), as popular self-supervised vision learners,…

Computer Vision and Pattern Recognition · Computer Science 2023-12-14 Zhiyuan Ma , zhihuan yu , Jianjun Li , Bowen Zhou

Modeling metasurfaces with high accuracy and efficiency is challenging because they have features smaller than the wavelength but sizes much larger than the wavelength. Full wave simulation is accurate but very slow. Popular design…

Optics · Physics 2023-04-04 Zhicheng Wu , Xiaoyan Huang , Nanfang Yu , Zongfu Yu

Unsupervised anomaly segmentation aims to detect patterns that are distinct from any patterns processed during training, commonly called abnormal or out-of-distribution patterns, without providing any associated manual segmentations. Since…

Image and Video Processing · Electrical Eng. & Systems 2023-11-06 Ziyun Liang , Harry Anthony , Felix Wagner , Konstantinos Kamnitsas

Multimodal medical image fusion is a crucial task that combines complementary information from different imaging modalities into a unified representation, thereby enhancing diagnostic accuracy and treatment planning. While deep learning…

Image and Video Processing · Electrical Eng. & Systems 2024-11-19 Meng Zhou , Yuxuan Zhang , Xiaolan Xu , Jiayi Wang , Farzad Khalvati

Bitstream-corrupted video recovery aims to restore realistic content degraded during video storage or transmission. Existing methods typically assume that predefined masks of corrupted regions are available, but manually annotating these…

Computer Vision and Pattern Recognition · Computer Science 2026-04-16 Shuyun Wang , Hu Zhang , Xin Shen , Dadong Wang , Xin Yu

Video anomaly detection (VAD) is a vital yet complex open-set task in computer vision, commonly tackled through reconstruction-based methods. However, these methods struggle with two key limitations: (1) insufficient robustness in open-set…

Computer Vision and Pattern Recognition · Computer Science 2025-03-28 Xiaofeng Tan , Hongsong Wang , Xin Geng , Liang Wang

Continual Anomaly Detection (CAD) enables anomaly detection models in learning new classes while preserving knowledge of historical classes. CAD faces two key challenges: catastrophic forgetting and segmentation of small anomalous regions.…

Computer Vision and Pattern Recognition · Computer Science 2025-05-13 Lei Hu , Zhiyong Gan , Ling Deng , Jinglin Liang , Lingyu Liang , Shuangping Huang , Tianshui Chen

Low-count positron emission tomography (LCPET) imaging can reduce patients' exposure to radiation but often suffers from increased image noise and reduced lesion detectability, necessitating effective denoising techniques. Diffusion models…

Image and Video Processing · Electrical Eng. & Systems 2025-03-24 Yinchi Zhou , Huidong Xie , Menghua Xia , Qiong Liu , Bo Zhou , Tianqi Chen , Jun Hou , Liang Guo , Xinyuan Zheng , Hanzhong Wang , Biao Li , Axel Rominger , Kuangyu Shi , Nicha C. Dvorneka , Chi Liu

With the rapid advancement of deep learning in image generation, facial forgery techniques have achieved unprecedented realism, posing serious threats to cybersecurity and information authenticity. Most existing deepfake detection…

Computer Vision and Pattern Recognition · Computer Science 2026-04-17 Haotian Wu , Yue Cheng , Shan Bian

The analysis of multi-modality positron emission tomography and computed tomography (PET-CT) images for computer aided diagnosis applications requires combining the sensitivity of PET to detect abnormal regions with anatomical localization…

Computer Vision and Pattern Recognition · Computer Science 2019-10-29 Ashnil Kumar , Michael Fulham , Dagan Feng , Jinman Kim