English
Related papers

Related papers: Stable Optimization for Large Vision Model Based D…

200 papers

The data-driven approach of supervised learning methods has limited applicability in solving dipole inversion in Quantitative Susceptibility Mapping (QSM) with varying scan parameters across different objects. To address this generalization…

Image and Video Processing · Electrical Eng. & Systems 2023-08-21 Zhuang Xiong , Yang Gao , Yin Liu , Amir Fazlollahi , Peter Nestor , Feng Liu , Hongfu Sun

Cone-beam computed tomography (CBCT) using only a few X-ray projection views enables faster scans with lower radiation dose, but the resulting severe under-sampling causes strong artifacts and poor spatial coverage. We address these…

Image and Video Processing · Electrical Eng. & Systems 2025-06-25 Minmin Yang , Huantao Ren , Senem Velipasalar

Although vision models such as Contrastive Language-Image Pre-Training (CLIP) show impressive generalization performance, their zero-shot robustness is still limited under Out-of-Distribution (OOD) scenarios without fine-tuning. Instead of…

Computer Vision and Pattern Recognition · Computer Science 2024-05-30 Zhuo Huang , Chang Liu , Yinpeng Dong , Hang Su , Shibao Zheng , Tongliang Liu

Spectral computed tomography (CT) has attracted much attention in radiation dose reduction, metal artifacts removal, tissue quantification and material discrimination. The x-ray energy spectrum is divided into several bins, each…

Image and Video Processing · Electrical Eng. & Systems 2021-08-26 Weiwen Wu , Dianlin Hu , Chuang Niu , Lieza Vanden Broeke , Anthony P. H. Butler , Peng Cao , James Atlas , Alexander Chernoglazov , Varut Vardhanabhuti , Ge Wang

Perspective distortion (PD) leads to substantial alterations in the shape, size, orientation, angles, and spatial relationships of visual elements in images. Accurately determining camera intrinsic and extrinsic parameters is challenging,…

Computer Vision and Pattern Recognition · Computer Science 2024-10-10 Meenakshi Subhash Chippa , Prakash Chandra Chhipa , Kanjar De , Marcus Liwicki , Rajkumar Saini

Dataset distillation compresses a large training set into a small synthetic set that preserves downstream training utility. While most existing methods target training networks from scratch, modern visual transfer learning often uses frozen…

Computer Vision and Pattern Recognition · Computer Science 2026-05-11 Bincheng Peng , Guang Li , Ping Liu , Takahiro Ogawa , Miki Haseyama

Diffusion models, as powerful generative models, have found a wide range of applications and shown great potential in solving image reconstruction problems. Some works attempted to solve MRI reconstruction with diffusion models, but these…

Image and Video Processing · Electrical Eng. & Systems 2025-06-09 Xingjian Tang , Jingwei Guan , Linge Li , Ran Shi , Youmei Zhang , Mengye Lyu , Li Yan

Deep learning-based methods have achieved significant successes on solving the blind super-resolution (BSR) problem. However, most of them request supervised pre-training on labelled datasets. This paper proposes an unsupervised kernel…

Image and Video Processing · Electrical Eng. & Systems 2024-04-29 Zhixiong Yang , Jingyuan Xia , Shengxi Li , Xinghua Huang , Shuanghui Zhang , Zhen Liu , Yaowen Fu , Yongxiang Liu

Medical image restoration tasks aim to recover high-quality images from degraded observations, exhibiting emergent desires in many clinical scenarios, such as low-dose CT image denoising, MRI super-resolution, and MRI artifact removal.…

Image and Video Processing · Electrical Eng. & Systems 2025-07-09 Pengcheng Zheng , Kecheng Chen , Jiaxin Huang , Bohao Chen , Ju Liu , Yazhou Ren , Xiaorong Pu

Vision models are often vulnerable to out-of-distribution (OOD) samples without adapting. While visual prompts offer a lightweight method of input-space adaptation for large-scale vision models, they rely on a high-dimensional additive…

Computer Vision and Pattern Recognition · Computer Science 2023-10-27 Yun-Yun Tsai , Chengzhi Mao , Junfeng Yang

Sparse-view computed tomography (CT) is known as a widely used approach to reduce radiation dose while accelerating imaging through lowered projection views and correlated calculations. However, its severe imaging noise and streaking…

Image and Video Processing · Electrical Eng. & Systems 2021-01-20 Yitong Liu , Ken Deng , Chang Sun , Hongwen Yang

Denoising diffusion models achieved impressive results on several image generation tasks often outperforming GAN based models. Recently, the generative capabilities of diffusion models have been employed for perceptual image compression,…

Image and Video Processing · Electrical Eng. & Systems 2025-05-20 Jonas Brenig , Radu Timofte

The main challenges limiting the adoption of deep learning-based solutions in medical workflows are the availability of annotated data and the lack of interpretability of such systems. Concept Bottleneck Models (CBMs) tackle the latter by…

Computer Vision and Pattern Recognition · Computer Science 2025-10-07 Cristiano Patrício , Isabel Rio-Torto , Jaime S. Cardoso , Luís F. Teixeira , João C. Neves

We present MInDI-3D (Medical Inversion by Direct Iteration in 3D), the first 3D conditional diffusion-based model for real-world sparse-view Cone Beam Computed Tomography (CBCT) artefact removal, aiming to reduce imaging radiation exposure.…

Computer Vision and Pattern Recognition · Computer Science 2025-10-10 Daniel Barco , Marc Stadelmann , Martin Oswald , Ivo Herzig , Lukas Lichtensteiger , Pascal Paysan , Igor Peterlik , Michal Walczak , Bjoern Menze , Frank-Peter Schilling

Multimodal Large Language Models (MLLMs) are undergoing rapid progress and represent the frontier of AI development. However, their training and inference efficiency have emerged as a core bottleneck in making MLLMs more accessible and…

Significant strides have been made using large vision-language models, like Stable Diffusion (SD), for a variety of downstream tasks, including image editing, image correspondence, and 3D shape generation. Inspired by these advancements, we…

Computer Vision and Pattern Recognition · Computer Science 2024-03-15 Aliasghar Khani , Saeid Asgari Taghanaki , Aditya Sanghi , Ali Mahdavi Amiri , Ghassan Hamarneh

We present an effective post-processing method to reduce the artifacts from sparsely reconstructed cone-beam CT (CBCT) images. The proposed method is based on the state-of-the-art, image-to-image generative models with a perceptual loss as…

Computer Vision and Pattern Recognition · Computer Science 2018-12-11 Haofu Liao , Zhimin Huo , William J. Sehnert , Shaohua Kevin Zhou , Jiebo Luo

Conventional Computed Tomography (CT) methods require large numbers of noise-free projections for accurate density reconstructions, limiting their applicability to the more complex class of Cone Beam Geometry CT (CBCT) reconstruction.…

Image and Video Processing · Electrical Eng. & Systems 2023-07-18 Samuele Papa , David M. Knigge , Riccardo Valperga , Nikita Moriakov , Miltos Kofinas , Jan-Jakob Sonke , Efstratios Gavves

Limited-view computed tomography (CT) presents significant potential for reducing radiation exposure and expediting the scanning process. While deep learning (DL) methods have exhibited promising results in mitigating streaking artifacts…

Medical Physics · Physics 2025-02-18 Changyu Chen , Li Zhang , Yuxiang Xing , Zhiqiang Chen

X-ray Computed Tomography (CT) is one of the most important diagnostic imaging techniques in clinical applications. Sparse-view CT imaging reduces the number of projection views to a lower radiation dose and alleviates the potential risk of…

Image and Video Processing · Electrical Eng. & Systems 2024-11-22 Xiaohong Fan , Ke Chen , Huaming Yi , Yin Yang , Jianping Zhang