中文
相关论文

相关论文: Zero-shot Low-Field MRI Enhancement via Diffusion-…

200 篇论文

Zero-shot, training-free, image-based text-to-video generation is an emerging area that aims to generate videos using existing image-based diffusion models. Current methods in this space require specific architectural changes to image…

计算机视觉与模式识别 · 计算机科学 2025-04-10 Diljeet Jagpal , Xi Chen , Vinay P. Namboodiri

Echo-planar imaging (EPI) remains the cornerstone of diffusion MRI, but it is prone to severe geometric distortions due to its rapid sampling scheme that renders the sequence highly sensitive to $B_{0}$ field inhomogeneities. While deep…

图像与视频处理 · 电气工程与系统科学 2026-03-30 Namgyu Han , Seong Dae Yun , Chaeeun Lim , Sunghyun Seok , Sunju Kim , Yoonhwan Kim , Yohan Jun , Tae Hyung Kim , Berkin Bilgic , Jaejin Cho

High Dynamic Range (HDR) imaging aims to generate an artifact-free HDR image with realistic details by fusing multi-exposure Low Dynamic Range (LDR) images. Caused by large motion and severe under-/over-exposure among input LDR images, HDR…

计算机视觉与模式识别 · 计算机科学 2024-08-30 Shuaikang Shang , Xuejing Kang , Anlong Ming

High dynamic range (HDR) video reconstruction aims to generate HDR videos from low dynamic range (LDR) frames captured with alternating exposures. Most existing works solely rely on the regression-based paradigm, leading to adverse effects…

计算机视觉与模式识别 · 计算机科学 2024-06-13 Yuanshen Guan , Ruikang Xu , Mingde Yao , Ruisheng Gao , Lizhi Wang , Zhiwei Xiong

Score-based diffusion models have significantly advanced generative deep learning for image processing. Measurement conditioned models have also been applied to inverse problems such as CT reconstruction. However, the conventional approach,…

医学物理 · 物理学 2025-02-24 Matthew Tivnan , Dufan Wu , Quanzheng Li

Text-guided diffusion models revolutionize audio generation by adapting source audio to specific text prompts. However, existing zero-shot audio editing methods such as DDIM inversion accumulate errors across diffusion steps, reducing the…

音频与语音处理 · 电气工程与系统科学 2025-11-06 Huadai Liu , Jialei Wang , Xiangtai Li , Wen Wang , Qian Chen , Rongjie Huang , Yang Liu , Jiayang Xu , Zhou Zhao

Accurately capturing the full-range response of structures is crucial in structural health monitoring (SHM) for ensuring safety and operational integrity. However, limited sensor deployment due to cost, accessibility, or scale often hinders…

计算工程、金融与科学 · 计算机科学 2025-09-25 Wingho Feng , Quanwang Li , Chen Wang , Jian-sheng Fan

Optical imaging systems are inherently imperfect due to diffraction limits, lens manufacturing tolerances, assembly misalignment, and other physical constraints. In addition, unavoidable camera shake and object motion further introduce…

计算机视觉与模式识别 · 计算机科学 2025-11-17 Yanlong Yang , Guanxiong Luo

Multi-source stationary computed tomography (CT) has recently attracted attention for its ability to achieve rapid image reconstruction, making it suitable for time-sensitive clinical and industrial applications. However, practical systems…

计算机视觉与模式识别 · 计算机科学 2025-11-19 Jiancheng Fang , Shaoyu Wang , Junlin Wang , Weiwen Wu , Yikun Zhang , Qiegen Liu

Light field (LF) images containing information for multiple views have numerous applications, which can be severely affected by low-light imaging. Recent learning-based methods for low-light enhancement have some disadvantages, such as a…

计算机视觉与模式识别 · 计算机科学 2023-08-16 Shansi Zhang , Nan Meng , Edmund Y. Lam

In this work, we introduce HeFT (Head-Frequency Tracker), a zero-shot point tracking framework that leverages the visual priors of pretrained video diffusion models. To better understand how they encode spatiotemporal information, we…

计算机视觉与模式识别 · 计算机科学 2026-03-24 Tianyu Yuan , Yuanbo Yang , Lin-Zhuo Chen , Yao Yao , Zhuzhong Qian

We propose Deep Distribution Transfer(DDT), a new transfer learning approach to address the problem of zero and few-shot transfer in the context of facial forgery detection. We examine how well a model (pre-)trained with one forgery…

计算机视觉与模式识别 · 计算机科学 2020-06-23 Shivangi Aneja , Matthias Nießner

The Diffusion Transformer (DiT) architecture is the state-of-the-art paradigm for high-fidelity image generation, underpinning models like Stable Diffusion-3 and FLUX.1. However, deploying these models on resource-constrained mobile devices…

计算机视觉与模式识别 · 计算机科学 2026-05-18 Kunpeng Du , Haizhen Xie , Sen Lu , Lei Yu , Binglei Bao , Huaao Tang , Chuntao Liu , Hao Wu , Yang Zhao , Zhicai Huang , Heyuan Gao , Zhijun Tu , Jie Hu , Xinghao Chen

Low-dose computed tomography (LDCT) reduces radiation exposure but also introduces substantial noise and structural degradation, making it difficult to suppress noise without erasing subtle anatomical details. In this paper, we present…

计算机视觉与模式识别 · 计算机科学 2026-03-24 Tangtangfang Fang , Yang Jiao , Xiangjian He , Jingxi Hu , Jiaqi Yang

There is an increasing need for label free methods that could reveal intracellular structures and dynamics. In this context, we develop a new optical tomography method working in transmission - Full-field optical transmission tomography…

Low-dose Computed Tomography (LDCT) reconstruction is an important task in medical image analysis. Recent years have seen many deep learning based methods, proved to be effective in this area. However, these methods mostly follow a…

图像与视频处理 · 电气工程与系统科学 2023-02-08 Runyi Li

Diffusion MRI is commonly performed using echo-planar imaging (EPI) due to its rapid acquisition time. However, the resolution of diffusion-weighted images is often limited by magnetic field inhomogeneity-related artifacts and blurring…

图像与视频处理 · 电气工程与系统科学 2023-09-26 Jaejin Cho , Yohan Jun , Xiaoqing Wang , Caique Kobayashi , Berkin Bilgic

This work presents TV-LoRA, a novel method for low-dose sparse-view CT reconstruction that combines a diffusion generative prior (NCSN++ with SDE modeling) and multi-regularization constraints, including anisotropic TV and nuclear norm…

计算机视觉与模式识别 · 计算机科学 2025-10-07 Zongyin Deng , Qing Zhou , Yuhao Fang , Zijian Wang , Yao Lu , Ye Zhang , Chun Li

Existing multi-modal image fusion methods fail to address the compound degradations presented in source images, resulting in fusion images plagued by noise, color bias, improper exposure, \textit{etc}. Additionally, these methods often…

计算机视觉与模式识别 · 计算机科学 2024-11-01 Hao Zhang , Lei Cao , Jiayi Ma

Diffusion transformers have recently delivered strong text-to-image generation around 1K resolution, but we show that extending them to native 4K across diverse aspect ratios exposes a tightly coupled failure mode spanning positional…

计算机视觉与模式识别 · 计算机科学 2025-11-25 Tian Ye , Song Fei , Lei Zhu