English
Related papers

Related papers: RetFiner: A Vision-Language Refinement Scheme for …

200 papers

Retinal lesions play a vital role in the accurate classification of retinal abnormalities. Many researchers have proposed deep lesion-aware screening systems that analyze and grade the progression of retinopathy. However, to the best of our…

Computer Vision and Pattern Recognition · Computer Science 2020-08-17 Taimur Hassan , Muhammad Usman Akram , Naoufel Werghi

Noise-robust automatic speech recognition (ASR) has been commonly addressed by applying speech enhancement (SE) at the waveform level before recognition. However, speech-level enhancement does not always translate into consistent…

Audio and Speech Processing · Electrical Eng. & Systems 2026-01-09 Da-Hee Yang , Joon-Hyuk Chang

Longitudinal imaging is capable of capturing the static ana\-to\-mi\-cal structures and the dynamic changes of the morphology resulting from aging or disease progression. Self-supervised learning allows to learn new representation from…

Image and Video Processing · Electrical Eng. & Systems 2019-10-25 Antoine Rivail , Ursula Schmidt-Erfurth , Wolf-Dieter Vogl , Sebastian M. Waldstein , Sophie Riedl , Christoph Grechenig , Zhichao Wu , Hrvoje Bogunović

Hallucinations in large language model (LLM) outputs severely limit their reliability in knowledge-intensive tasks such as question answering. To address this challenge, we introduce REFIND (Retrieval-augmented Factuality hallucINation…

Computation and Language · Computer Science 2025-04-09 DongGeon Lee , Hwanjo Yu

Layer segmentation is important to quantitative analysis of retinal optical coherence tomography (OCT). Recently, deep learning based methods have been developed to automate this task and yield remarkable performance. However, due to the…

Image and Video Processing · Electrical Eng. & Systems 2023-12-07 Hong Liu , Dong Wei , Donghuan Lu , Xiaoying Tang , Liansheng Wang , Yefeng Zheng

Multiple-surface segmentation in Optical Coherence Tomography (OCT) images is a challenge problem, further complicated by the frequent presence of weak image boundaries. Recently, many deep learning (DL) based methods have been developed…

Image and Video Processing · Electrical Eng. & Systems 2022-10-13 Hui Xie , Weiyu Xu , Xiaodong Wu

Retinal imaging is fast, non-invasive, and widely available, offering quantifiable structural and vascular signals for ophthalmic and systemic health assessment. This accessibility creates an opportunity to study how quantitative retinal…

Computer Vision and Pattern Recognition · Computer Science 2026-02-10 Zhonghua Wang , Lie Ju , Sijia Li , Wei Feng , Sijin Zhou , Ming Hu , Jianhao Xiong , Xiaoying Tang , Yifan Peng , Mingquan Lin , Yaodong Ding , Yong Zeng , Wenbin Wei , Li Dong , Zongyuan Ge

Weakly Supervised Semantic Segmentation (WSSS) relying only on image-level supervision is a promising approach to deal with the need for Segmentation networks, especially for generating a large number of pixel-wise masks in a given dataset.…

Computer Vision and Pattern Recognition · Computer Science 2023-09-21 Bharath Srinivas Prabakaran , Erik Ostrowski , Muhammad Shafique

Recent advancements in vision-language models (VLMs) have improved performance by increasing the number of visual tokens, which are often significantly longer than text tokens. However, we observe that most real-world scenarios do not…

Computer Vision and Pattern Recognition · Computer Science 2025-07-18 Senqiao Yang , Junyi Li , Xin Lai , Bei Yu , Hengshuang Zhao , Jiaya Jia

Despite the revolutionary impact of AI and the development of locally trained algorithms, achieving widespread generalized learning from multi-modal data in medical AI remains a significant challenge. This gap hinders the practical…

Computer Vision and Pattern Recognition · Computer Science 2024-01-24 Fatema-E Jannat , Sina Gholami , Minhaj Nur Alam , Hamed Tabkhi

Foundation models (FMs) have shown great promise in medical image analysis by improving generalization across diverse downstream tasks. In ophthalmology, several FMs have recently emerged, but there is still no clear answer to fundamental…

Optical coherence tomography (OCT) suffers from speckle noise, causing the deterioration of image quality, especially in high-resolution modalities like visible light OCT (vis-OCT). The potential of conventional supervised deep learning…

Image and Video Processing · Electrical Eng. & Systems 2024-05-16 Lingyun Wang , Jose A Sahel , Shaohua Pi

The rising prevalence of vision-threatening retinal diseases poses a significant burden on the global healthcare systems. Deep learning (DL) offers a promising solution for automatic disease screening but demands substantial data.…

Image and Video Processing · Electrical Eng. & Systems 2024-11-18 Ruoyu Chen , Weiyi Zhang , Bowen Liu , Xiaolan Chen , Pusheng Xu , Shunming Liu , Mingguang He , Danli Shi

This study introduces a novel framework for enhancing domain generalization in medical imaging, specifically focusing on utilizing unlabelled multi-view colour fundus photographs. Unlike traditional approaches that rely on single-view…

Computer Vision and Pattern Recognition · Computer Science 2024-03-26 Ze Chen , Gongyu Zhang , Jiayu Huo , Joan Nunez do Rio , Charalampos Komninos , Yang Liu , Rachel Sparks , Sebastien Ourselin , Christos Bergeles , Timothy Jackson

Recent advancements in Large Language Models (LLMs) have significantly improved their performance across various Natural Language Processing (NLP) tasks. However, LLMs still struggle with generating non-factual responses due to limitations…

Computation and Language · Computer Science 2024-09-10 Taeho Hwang , Soyeong Jeong , Sukmin Cho , SeungYoon Han , Jong C. Park

Fine-tuning Large Language Models (LLMs) on specific datasets is a common practice to improve performance on target tasks. However, this performance gain often leads to overfitting, where the model becomes too specialized in either the task…

Computation and Language · Computer Science 2024-09-10 Sonam Gupta , Yatin Nandwani , Asaf Yehudai , Mayank Mishra , Gaurav Pandey , Dinesh Raghu , Sachindra Joshi

Low-light image enhancement is challenging due to complex degradations, including amplified noise, artifacts, and color distortion. While Retinex-based deep learning methods have achieved promising results, they primarily rely on…

Computer Vision and Pattern Recognition · Computer Science 2026-05-14 Youssef Aboelwafa , Hicham G. Elmongui , Marwan Torki

In the last decades, scene text recognition has gained worldwide attention from both the academic community and actual users due to its importance in a wide range of applications. Despite achievements in optical character recognition, scene…

Computer Vision and Pattern Recognition · Computer Science 2022-01-04 Bao Hieu Tran , Thanh Le-Cong , Huu Manh Nguyen , Duc Anh Le , Thanh Hung Nguyen , Phi Le Nguyen

Optical Coherence Tomography (OCT) is a novel and effective screening tool for ophthalmic examination. Since collecting OCT images is relatively more expensive than fundus photographs, existing methods use multi-modal learning to complement…

Image and Video Processing · Electrical Eng. & Systems 2023-08-02 Lehan Wang , Weihang Dai , Mei Jin , Chubin Ou , Xiaomeng Li

Medical foundation models, pre-trained with large-scale clinical data, demonstrate strong performance in diverse clinically relevant applications. RETFound, trained on nearly one million retinal images, exemplifies this approach in…

‹ Prev 1 4 5 6 7 8 10 Next ›