English
Related papers

Related papers: Cross-Modal Semantic-Enhanced Diffusion Framework …

200 papers

Deep convolutional semantic segmentation (DCSS) learning doesn't converge to an optimal local minimum with random parameters initializations; a pre-trained model on the same domain becomes necessary to achieve convergence.In this work, we…

Machine Learning · Computer Science 2017-10-24 Mrinal Haloi

Diabetic Retinopathy (DR) is a complication of long-standing, unchecked diabetes and one of the leading causes of blindness in the world. This paper focuses on improved and robust methods to extract some of the features of DR, viz. Blood…

Image and Video Processing · Electrical Eng. & Systems 2022-07-12 Soham Basu , Sayantan Mukherjee , Ankit Bhattacharya , Anindya Sen

Diabetic Retinopathy (DR), a prevalent and severe complication of diabetes, affects millions of individuals globally, underscoring the need for accurate and timely diagnosis. Recent advancements in imaging technologies, such as…

In this manuscript, we automate the procedure of grading of diabetic retinopathy and macular edema from fundus images using an ensemble of convolutional neural networks. The availability of limited amount of labeled data to perform…

Computer Vision and Pattern Recognition · Computer Science 2018-09-13 Avinash Kori , Sai Saketh Chennamsetty , Mohammed Safwan K. P. , Varghese Alex

Referring image segmentation (RIS) requires accurate segmentation of target regions in images according to language descriptions, which is a cross-modal task integrating vision and language. Existing RIS methods typically employ large-scale…

Computer Vision and Pattern Recognition · Computer Science 2026-04-21 Chen Yang

The accurate segmentation of brain tumors from multi-modal MRI is critical for clinical diagnosis and treatment planning. While integrating complementary information from various MRI sequences is a common practice, the frequent absence of…

Computer Vision and Pattern Recognition · Computer Science 2026-03-06 Dongqing Xie , Yonghuang Wu , Zisheng Ai , Jun Min , Zhencun Jiang , Shaojin Geng , Lei Wang

In recent years, the incidence of vision-threatening eye diseases has risen dramatically, necessitating scalable and accurate screening solutions. This paper presents a comprehensive study on deep learning architectures for the automated…

Computer Vision and Pattern Recognition · Computer Science 2025-12-12 Mohammad Sadegh Gholizadeh , Amir Arsalan Rezapour

Cross-domain recommendation (CDR) is crucial for improving recommendation accuracy and generalization, yet traditional methods are often hindered by the reliance on shared user/item IDs, which are unavailable in most real-world scenarios.…

Information Retrieval · Computer Science 2025-11-18 Peiyu Hu , Wayne Lu , Jia Wang

Diffusion model (DM) based Video Super-Resolution (VSR) approaches achieve impressive perceptual quality. However, they suffer from error accumulation, spatial artifacts, and a trade-off between perceptual quality and fidelity, primarily…

Computer Vision and Pattern Recognition · Computer Science 2026-03-30 Jingyi Xu , Meisong Zheng , Ying Chen , Minglang Qiao , Xin Deng , Mai Xu

Supervised deep learning for semantic segmentation has achieved excellent results in accurately identifying anatomical and pathological structures in medical images. However, it often requires large annotated training datasets, which limits…

Computer Vision and Pattern Recognition · Computer Science 2026-03-11 Luca Ciampi , Gabriele Lagani , Giuseppe Amato , Fabrizio Falchi

Diabetic Retinopathy (DR) is one of the leading causes of preventable blindness, yet rural regions often lack the specialists and infrastructure needed for early detection. Although cloud-based deep learning systems offer high accuracy,…

Computer Vision and Pattern Recognition · Computer Science 2026-05-15 Nishi Doshi , Shrey Shah

Timely and affordable computer-aided diagnosis of retinal diseases is pivotal in precluding blindness. Accurate retinal vessel segmentation plays an important role in disease progression and diagnosis of such vision-threatening diseases. To…

Image and Video Processing · Electrical Eng. & Systems 2023-04-26 Tariq M. Khan , Syed S. Naqvi , Antonio Robles-Kelly , Imran Razzak

Diffusion models, such as Stable Diffusion, have shown incredible performance on text-to-image generation. Since text-to-image generation often requires models to generate visual concepts with fine-grained details and attributes specified…

Computer Vision and Pattern Recognition · Computer Science 2024-04-26 Xuehai He , Weixi Feng , Tsu-Jui Fu , Varun Jampani , Arjun Akula , Pradyumna Narayana , Sugato Basu , William Yang Wang , Xin Eric Wang

Semi-supervised learning utilizes insights from unlabeled data to improve model generalization, thereby reducing reliance on large labeled datasets. Most existing studies focus on limited samples and fail to capture the overall data…

Computer Vision and Pattern Recognition · Computer Science 2025-03-13 Xiuzhen Guo , Lianyuan Yu , Ji Shi , Na Lei , Hongxiao Wang

Deep learning has achieved remarkable success in medical image analysis, however its adoption in clinical practice is limited by a lack of interpretability. These models often make correct predictions without explaining their reasoning.…

Image and Video Processing · Electrical Eng. & Systems 2025-10-03 Jhonatan Contreras , Thomas Bocklitz

Diffusion models have been used extensively for high quality image and video generation tasks. In this paper, we propose a novel conditional diffusion model with spatial attention and latent embedding (cDAL) for medical image segmentation.…

Image and Video Processing · Electrical Eng. & Systems 2025-02-21 Behzad Hejrati , Soumyanil Banerjee , Carri Glide-Hurst , Ming Dong

Convolutional Neural Networks (CNNs) have successfully been used to classify diabetic retinopathy (DR) fundus images in recent times. However, deeper representations in CNNs may capture higher-level semantics at the expense of spatial…

Computer Vision and Pattern Recognition · Computer Science 2021-02-03 Samuel Ofosu Mensah , Bubacarr Bah , Willie Brink

Visual attribution in medical imaging seeks to make evident the diagnostically-relevant components of a medical image, in contrast to the more common detection of diseased tissue deployed in standard machine vision pipelines (which are less…

Image and Video Processing · Electrical Eng. & Systems 2024-01-04 Ammar A. Siddiqui , Santosh Tirunagari , Tehseen Zia , David Windridge

Remote sensing images captured by different platforms exhibit significant disparities in spatial resolution. Large scale factor super-resolution (SR) algorithms are vital for maximizing the utilization of low-resolution (LR) satellite data…

Computer Vision and Pattern Recognition · Computer Science 2024-05-14 Ce Wang , Wanjie Sun

We propose GRAM-DIFF, a Gram-matrix-guided diffusion framework for semi-blind multiple input multiple output (MIMO) channel estimation. Recent diffusion-based estimators leverage learned generative priors to improve pilot-based channel…

Information Theory · Computer Science 2026-02-18 Xinyuan Wang , Krishna Narayanan
‹ Prev 1 8 9 10 Next ›