English
Related papers

Related papers: Multiscale Structure-Guided Latent Diffusion for M…

200 papers

Most existing MRI reconstruction methods perform tar-geted reconstruction of the entire MR image without tak-ing specific tissue regions into consideration. This may fail to emphasize the reconstruction accuracy on im-portant tissues for…

Image and Video Processing · Electrical Eng. & Systems 2023-09-06 Yu Guan , Chuanming Yu , Shiyu Lu , Zhuoxu Cui , Dong Liang , Qiegen Liu

Magnetic resonance imaging (MRI), especially functional MRI (fMRI) and diffusion MRI (dMRI), is essential for studying neurodegenerative diseases. However, missing modalities pose a major barrier to their clinical use. Although GAN- and…

Computer Vision and Pattern Recognition · Computer Science 2025-11-10 Xiongri Shen , Jiaqi Wang , Yi Zhong , Zhenxi Song , Leilei Zhao , Yichen Wei , Lingyan Liang , Shuqiang Wang , Baiying Lei , Demao Deng , Zhiguo Zhang

Sign language translation (SLT) is challenging, as it involves converting sign language videos into natural language. Previous studies have prioritized accuracy over diversity. However, diversity is crucial for handling lexical and…

Computer Vision and Pattern Recognition · Computer Science 2024-11-27 JiHwan Moon , Jihoon Park , Jungeun Kim , Jongseong Bae , Hyeongwoo Jeon , Ha Young Kim

Neuroimaging data, particularly from techniques like MRI or PET, offer rich but complex information about brain structure and activity. To manage this complexity, latent representation models - such as Autoencoders, Generative Adversarial…

Computer Vision and Pattern Recognition · Computer Science 2024-12-31 C. Vázquez-García , F. J. Martínez-Murcia , F. Segovia Román , Juan M. Górriz

Virtual interventions enable the physics-based simulation of device deployment within coronary arteries. This framework allows for counterfactual reasoning by deploying the same device in different arterial anatomies. However, current…

Image and Video Processing · Electrical Eng. & Systems 2025-11-26 Karim Kadry , Shreya Gupta , Jonas Sogbadji , Michiel Schaap , Kersten Petersen , Takuya Mizukami , Carlos Collet , Farhad R. Nezami , Elazer R. Edelman

Multi-modal Magnetic Resonance Imaging (MRI) translation leverages information from source MRI sequences to generate target modalities, enabling comprehensive diagnosis while overcoming the limitations of acquiring all sequences. While…

Image and Video Processing · Electrical Eng. & Systems 2025-05-20 Jiyao Liu , Shangqi Gao , Yuxin Li , Lihao Liu , Xin Gao , Zhaohu Xing , Junzhi Ning , Yanzhou Su , Xiao-Yong Zhang , Junjun He , Ningsheng Xu , Xiahai Zhuang

Recent DiT-based text-to-image models increasingly adopt LLMs as text encoders, yet text conditioning remains largely static and often utilizes only a single LLM layer, despite pronounced semantic hierarchy across LLM layers and…

Computer Vision and Pattern Recognition · Computer Science 2026-02-04 Bozhou Li , Yushuo Guan , Haolin Li , Bohan Zeng , Yiyan Ji , Yue Ding , Pengfei Wan , Kun Gai , Yuanxing Zhang , Wentao Zhang

Chest X-Ray (CXR) imaging for pulmonary diagnosis raises significant challenges, primarily because bone structures can obscure critical details necessary for accurate diagnosis. Recent advances in deep learning, particularly with diffusion…

Image and Video Processing · Electrical Eng. & Systems 2025-08-06 Yifei Sun , Zhanghao Chen , Hao Zheng , Yuqing Lu , Lixin Duan , Fenglei Fan , Ahmed Elazab , Xiang Wan , Changmiao Wang , Ruiquan Ge

Purpose: To propose a domain-conditioned and temporal-guided diffusion modeling method, termed dynamic Diffusion Modeling (dDiMo), for accelerated dynamic MRI reconstruction, enabling diffusion process to characterize spatiotemporal…

Image and Video Processing · Electrical Eng. & Systems 2025-01-17 Liping Zhang , Iris Yuwen Zhou , Sydney B. Montesi , Li Feng , Fang Liu

Recovering degraded low-resolution text images is challenging, especially for Chinese text images with complex strokes and severe degradation in real-world scenarios. Ensuring both text fidelity and style realness is crucial for…

Computer Vision and Pattern Recognition · Computer Science 2024-03-05 Yuzhe Zhang , Jiawei Zhang , Hao Li , Zhouxia Wang , Luwei Hou , Dongqing Zou , Liheng Bian

Diffusion Language Models (DLMs) are rapidly emerging as a powerful and promising alternative to the dominant autoregressive (AR) paradigm. By generating tokens in parallel through an iterative denoising process, DLMs possess inherent…

Computation and Language · Computer Science 2025-12-08 Tianyi Li , Mingda Chen , Bowei Guo , Zhiqiang Shen

As the deep learning revolution marches on, masked modeling has emerged as a distinctive approach that involves predicting parts of the original data that are proportionally masked during training, and has demonstrated exceptional…

Image and Video Processing · Electrical Eng. & Systems 2025-06-25 Qinrong Cai , Yu Guan , Zhibo Chen , Dong Liang , Qiuyun Fan , Qiegen Liu

Visual information has been introduced for enhancing machine translation (MT), and its effectiveness heavily relies on the availability of large amounts of bilingual parallel sentence pairs with manual image annotations. In this paper, we…

Computation and Language · Computer Science 2025-01-07 Andong Chen , Yuchen Song , Kehai Chen , Muyun Yang , Tiejun Zhao , Min Zhang

Snapshot compressive spectral imaging reconstruction aims to reconstruct three-dimensional spatial-spectral images from a single-shot two-dimensional compressed measurement. Existing state-of-the-art methods are mostly based on deep…

Image and Video Processing · Electrical Eng. & Systems 2024-08-27 Zongliang Wu , Ruiying Lu , Ying Fu , Xin Yuan

Diffusion models have recently gained recognition for generating diverse and high-quality content, especially in image synthesis. These models excel not only in creating fixed-size images but also in producing panoramic images. However,…

Computer Vision and Pattern Recognition · Computer Science 2025-04-08 Xiaoyu Zhang , Teng Zhou , Xinlong Zhang , Jia Wei , Yongchuan Tang

Recently, the diffusion model has emerged as a superior generative model that can produce high quality and realistic images. However, for medical image translation, the existing diffusion models are deficient in accurately retaining…

Image and Video Processing · Electrical Eng. & Systems 2023-10-31 Yunxiang Li , Hua-Chieh Shao , Xiao Liang , Liyuan Chen , Ruiqi Li , Steve Jiang , Jing Wang , You Zhang

Referring image segmentation aims to predict the foreground mask of the object referred by a natural language sentence. Multimodal context of the sentence is crucial to distinguish the referent from the background. Existing methods either…

Computer Vision and Pattern Recognition · Computer Science 2020-10-06 Tianrui Hui , Si Liu , Shaofei Huang , Guanbin Li , Sansi Yu , Faxi Zhang , Jizhong Han

This paper does not describe a new method; instead, it provides a thorough exploration of an important yet understudied design space related to recent advances in text-to-image synthesis -- specifically, the deep fusion of large language…

Computer Vision and Pattern Recognition · Computer Science 2025-05-16 Bingda Tang , Boyang Zheng , Xichen Pan , Sayak Paul , Saining Xie

Service-level mobile traffic prediction for individual users is essential for network efficiency and quality of service enhancement. However, current prediction methods are limited in their adaptability across different urban environments…

Machine Learning · Computer Science 2025-07-25 Shiyuan Zhang , Tong Li , Zhu Xiao , Hongyang Du , Kaibin Huang

With the rapid advancement of remote sensing technology, super-resolution image reconstruction is of great research and practical significance. Existing deep learning methods have made progress but still face limitations in handling complex…

Computer Vision and Pattern Recognition · Computer Science 2025-07-24 Shijie Lyu
‹ Prev 1 4 5 6 7 8 10 Next ›