中文
相关论文

相关论文: RetFiner: A Vision-Language Refinement Scheme for …

200 篇论文

We propose a new deep learning approach for automatic detection and segmentation of fluid within retinal OCT images. The proposed framework utilizes both ResNet and Encoder-Decoder neural network architectures. When training the network, we…

计算机视觉与模式识别 · 计算机科学 2017-08-21 Dustin Morley , Hassan Foroosh , Saad Shaikh , Ulas Bagci

Optical Coherence Tomography (OCT) is a non-invasive imaging modality essential for diagnosing various eye diseases. Despite its clinical significance, developing OCT-based diagnostic tools faces challenges, such as limited public datasets,…

计算机视觉与模式识别 · 计算机科学 2025-01-30 Mohammadreza Saraei , Igor Kozak , Eung-Joo Lee

In this paper, we present a new approach for uncertainty-aware retinal layer segmentation in Optical Coherence Tomography (OCT) scans using probabilistic signed distance functions (SDF). Traditional pixel-wise and regression-based methods…

图像与视频处理 · 电气工程与系统科学 2024-12-09 Mohammad Mohaiminul Islam , Coen de Vente , Bart Liefers , Caroline Klaver , Erik J Bekkers , Clara I. Sánchez

Gloss-free sign language translation (SLT) is hindered by two key challenges: **inadequate sign representation** that fails to capture nuanced visual cues, and **sentence-level semantic misalignment** in current LLM-based methods, which…

计算机视觉与模式识别 · 计算机科学 2025-12-09 Zhi Rao , Yucheng Zhou , Benjia Zhou , Yiqing Huang , Sergio Escalera , Jun Wan

Retinal diseases are a leading cause of vision impairment and blindness, with timely diagnosis being critical for effective treatment. Optical Coherence Tomography (OCT) has become a standard imaging modality for retinal disease diagnosis,…

图像与视频处理 · 电气工程与系统科学 2025-10-06 Jilan Cheng , Guoli Long , Zeyu Zhang , Zhenjia Qi , Hanyu Wang , Libin Lu , Shuihua Wang , Yudong Zhang , Jin Hong

Conventional Fourier-domain Optical Coherence Tomography (FD-OCT) systems depend on resampling into wavenumber (k) domain to extract the depth profile. This either necessitates additional hardware resources or amplifies the existing…

光学 · 物理学 2025-09-24 Maryam Viqar , Erdem Sahin , Elena Stoykova , Violeta Madjarova

Retinal optical coherence tomography (OCT) images provide crucial insights into the health of the posterior ocular segment. Therefore, the advancement of automated image analysis methods is imperative to equip clinicians and researchers…

图像与视频处理 · 电气工程与系统科学 2024-02-16 Jiahao Wang , Hong Peng , Shengchao Chen , Sufen Ren

Large vision foundation models have been widely adopted for retinal disease classification without systematic evidence justifying their parameter requirements. In the present work we address two critical questions: First, are large…

图像与视频处理 · 电气工程与系统科学 2025-12-01 David Isztl , Tahm Spitznagel , Gabor Mark Somfai , Rui Santos

With Transformers achieving outstanding performance on individual remote sensing (RS) tasks, we are now approaching the realization of a unified model that excels across multiple tasks through multi-task learning (MTL). Compared to…

计算机视觉与模式识别 · 计算机科学 2026-01-12 Qingyun Li , Shuran Ma , Junwei Luo , Yi Yu , Yue Zhou , Fengxiang Wang , Xudong Lu , Xiaoxing Wang , Xin He , Yushi Chen , Xue Yang

Inability to express the confidence level and detect unseen classes has limited the clinical implementation of artificial intelligence in the real-world. We developed a foundation model with uncertainty estimation (FMUE) to detect 11…

图像与视频处理 · 电气工程与系统科学 2024-06-26 Yuanyuan Peng , Aidi Lin , Meng Wang , Tian Lin , Ke Zou , Yinglin Cheng , Tingkun Shi , Xulong Liao , Lixia Feng , Zhen Liang , Xinjian Chen , Huazhu Fu , Haoyu Chen

Large language models (LLMs) have shown remarkable effectiveness across various domains, with data augmentation methods utilizing GPT for synthetic data generation becoming prevalent. However, the quality and utility of augmented data…

计算与语言 · 计算机科学 2025-01-27 Zeao Tu , Xiangdi Meng , Yu He , Zihan Yao , Tianyu Qi , Jun Liu , Ming Li

Purpose: We proposed a deep convolutional neural network (CNN), named Retinal Fluid Segmentation Network (ReF-Net) to segment volumetric retinal fluid on optical coherence tomography (OCT) volume. Methods: 3 x 3-mm OCT scans were acquired…

图像与视频处理 · 电气工程与系统科学 2020-10-28 Yukun Guo , Tristan T. Hormel , Honglian Xiong , Jie Wang , Thomas S. Hwang , Yali Jia

Self-Supervised Learning (SSL) is at the core of training modern large machine learning models, providing a scheme for learning powerful representations that can be used in a variety of downstream tasks. However, SSL strategies must be…

高能物理 - 唯象学 · 物理学 2025-02-26 Philip Harris , Michael Kagan , Jeffrey Krupa , Benedikt Maier , Nathaniel Woodward

In the world of medical diagnostics, the adoption of various deep learning techniques is quite common as well as effective, and its statement is equally true when it comes to implementing it into the retina Optical Coherence Tomography…

图像与视频处理 · 电气工程与系统科学 2022-03-03 Tasnim Sakib Apon , Mohammad Mahmudul Hasan , Abrar Islam , MD. Golam Rabiul Alam

Retinal foundation models aim to learn generalizable representations from diverse retinal images, facilitating label-efficient model adaptation across various ophthalmic tasks. Despite their success, current retinal foundation models are…

计算机视觉与模式识别 · 计算机科学 2024-08-13 Kai Yu , Yang Zhou , Yang Bai , Zhi Da Soh , Xinxing Xu , Rick Siow Mong Goh , Ching-Yu Cheng , Yong Liu

Semantic-rich features from Vision Foundation Models (VFMs) have been leveraged to enhance Latent Diffusion Models (LDMs). However, raw VFM features are typically high-dimensional and redundant, increasing the difficulty of learning and…

计算机视觉与模式识别 · 计算机科学 2026-05-15 Guanfang Dong , Luke Schultz , Negar Hassanpour , Chao Gao

Multimodal Large Language Models (MLLMs) have shown strong performance in document image tasks, especially Optical Character Recognition (OCR). However, they struggle with Document Image Machine Translation (DIMT), which requires handling…

计算与语言 · 计算机科学 2025-07-14 Yupu Liang , Yaping Zhang , Zhiyang Zhang , Zhiyuan Chen , Yang Zhao , Lu Xiang , Chengqing Zong , Yu Zhou

Foundation models (FMs), powered by self-supervised learning (SSL), have redefined the capabilities of artificial intelligence, demonstrating exceptional performance in domains like natural language processing and computer vision. These…

机器学习 · 计算机科学 2025-06-23 Hamdi Altaheri , Fakhri Karray , Md. Milon Islam , S M Taslim Uddin Raju , Amir-Hossein Karimi

Supervised fine-tuning (SFT) plays a crucial role in adapting large language models (LLMs) to specific domains or tasks. However, as demonstrated by empirical experiments, the collected data inevitably contains noise in practical…

计算与语言 · 计算机科学 2024-12-20 Junyu Luo , Xiao Luo , Kaize Ding , Jingyang Yuan , Zhiping Xiao , Ming Zhang

Retinal image matching plays a crucial role in monitoring disease progression and treatment response. However, datasets with matched keypoints between temporally separated pairs of images are not available in abundance to train…

计算机视觉与模式识别 · 计算机科学 2023-07-24 Sahar Almahfouz Nasser , Nihar Gupte , Amit Sethi