中文
相关论文

相关论文: Conquering the Retina: Bringing Visual in-Context …

200 篇论文

Conventional deep learning models deal with images one-by-one, requiring costly and time-consuming expert labeling in the field of medical imaging, and domain-specific restriction limits model generalizability. Visual in-context learning…

Retinal image analysis is crucial for diagnosing and treating eye diseases, yet generating accurate medical reports from images remains challenging due to variability in image quality and pathology, especially with limited labeled data.…

计算机视觉与模式识别 · 计算机科学 2024-12-24 Teja Krishna Cherukuri , Nagur Shareef Shaik , Jyostna Devi Bodapati , Dong Hye Ye

While traditional computer vision models have historically struggled to generalize to endoscopic domains, the emergence of foundation models has shown promising cross-domain performance. In this work, we present the first large-scale study…

Training deep learning models for corneal optical coherence tomography (OCT) imaging is limited by the availability of large, well-annotated datasets. We present a configurable Monte Carlo simulation framework that generates synthetic…

图像与视频处理 · 电气工程与系统科学 2026-02-04 Jinglun Yu , Yaning Wang , Rosalinda Xiong , Ziyi Huang , Kristina Irsch , Jin U. Kang

Purpose. Localizing structures and estimating the motion of a specific target region are common problems for navigation during surgical interventions. Optical coherence tomography (OCT) is an imaging modality with a high spatial and…

图像与视频处理 · 电气工程与系统科学 2020-04-22 Marcel Bengs , Nils Gessert , Matthias Schlüter , Alexander Schlaefer

In-context learning (ICL) enables Large Language Models (LLMs) to learn tasks from demonstration examples without parameter updates. Although it has been extensively studied in LLMs, its effectiveness in Vision-Language Models (VLMs)…

机器学习 · 计算机科学 2025-10-29 Gabriel O. dos Santos , Esther Colombini , Sandra Avila

Large vision-language models (VLMs) have become state-of-the-art for many computer vision tasks, with in-context learning (ICL) as a popular adaptation strategy for new ones. But can VLMs learn novel concepts purely from visual…

计算机视觉与模式识别 · 计算机科学 2024-09-26 Bowen Zhao , Leo Parker Dirac , Paulina Varshavskaya

Automatic and accurate segmentation for retinal and choroidal layers of Optical Coherence Tomography (OCT) is crucial for detection of various ocular diseases. However, because of the variations in different equipments, OCT data obtained…

计算机视觉与模式识别 · 计算机科学 2019-09-02 Jiexiang Wang , Cheng Bian , Meng Li , Xin Yang , Kai Ma , Wenao Ma , Jin Yuan , Xinghao Ding , Yefeng Zheng

Performing data augmentation for learning deep neural networks is well known to be important for training visual recognition systems. By artificially increasing the number of training examples, it helps reducing overfitting and improves…

计算机视觉与模式识别 · 计算机科学 2018-07-20 Nikita Dvornik , Julien Mairal , Cordelia Schmid

In recent years, deep learning models have revolutionized medical image interpretation, offering substantial improvements in diagnostic accuracy. However, these models often struggle with challenging images where critical features are…

计算机视觉与模式识别 · 计算机科学 2023-07-03 Pradeep Singh , Kishore Babu Nampalle , Uppala Vivek Narayan , Balasubramanian Raman

Visual In-Context Learning (VICL) uses input-output image pairs, referred to as in-context pairs (or examples), as prompts alongside query images to guide models in performing diverse vision tasks. However, VICL often suffers from…

计算机视觉与模式识别 · 计算机科学 2025-09-29 Jiahao Zhang , Bowen Wang , Hong Liu , Yuta Nakashima , Hajime Nagahara

Image to Image Translation (I2I) is a challenging computer vision problem used in numerous domains for multiple tasks. Recently, ophthalmology became one of the major fields where the application of I2I is increasing rapidly. One such…

图像与视频处理 · 电气工程与系统科学 2021-12-14 Hemanth Pasupuleti , G. N. Girish

Visual In-Context Learning (VICL) is a prevailing way to transfer visual foundation models to new tasks by leveraging contextual information contained in in-context examples to enhance learning and prediction of query sample. The…

计算机视觉与模式识别 · 计算机科学 2024-10-11 Chengming Xu , Chen Liu , Yikai Wang , Yuan Yao , Yanwei Fu

In this work, we explore the mechanism of in-context learning (ICL) on out-of-distribution (OOD) tasks that were not encountered during training. To achieve this, we conduct synthetic experiments where the objective is to learn OOD…

机器学习 · 计算机科学 2024-12-05 Qixun Wang , Yifei Wang , Yisen Wang , Xianghua Ying

30 million Optical Coherence Tomography (OCT) imaging tests are issued annually to diagnose various retinal diseases, but accurate diagnosis of OCT scans requires trained eye care professionals who are still prone to making errors. With…

图像与视频处理 · 电气工程与系统科学 2022-10-04 Evan Wen , Rebecca Sorenson , Max Ehrlich

Optical coherence tomography angiography (OCTA) is a new imaging modality to visualize retinal microvasculature and has been readily adopted in clinics. High-resolution OCT angiograms are important to qualitatively and quantitatively…

图像与视频处理 · 电气工程与系统科学 2023-05-11 Yuyan Ruan , Dawei Yang , Ziqi Tang , An Ran Ran , Carol Y. Cheung , Hao Chen

The rapid advancement of large language models (LLMs) has accelerated the emergence of in-context learning (ICL) as a cutting-edge approach in the natural language processing domain. Recently, ICL has been employed in visual understanding…

计算机视觉与模式识别 · 计算机科学 2024-03-19 Dianmo Sheng , Dongdong Chen , Zhentao Tan , Qiankun Liu , Qi Chu , Jianmin Bao , Tao Gong , Bin Liu , Shengwei Xu , Nenghai Yu

Early and accurate classification of retinal diseases is critical to counter vision loss and for guiding clinical management of retinal diseases. In this study, we proposed a deep learning method for retinal disease classification utilizing…

计算机视觉与模式识别 · 计算机科学 2026-02-24 Mohammad Tahmid Noor , Shayan Abrar , Jannatul Adan Mahi , Md Parvez Mia , Asaduzzaman Hridoy , Samanta Ghosh

Optical coherence tomography (OCT) is a non-invasive, micrometer-scale imaging modality that has become a clinical standard in ophthalmology. By raster-scanning the retina, sequential cross-sectional image slices are acquired to generate…

图像与视频处理 · 电气工程与系统科学 2024-02-20 Stefan Ploner , Jungeun Won , Julia Schottenhamml , Jessica Girgis , Kenneth Lam , Nadia Waheed , James Fujimoto , Andreas Maier

Optical coherence tomography (OCT) is a medical imaging modality that allows us to probe deeper substructures of skin. The state-of-the-art wound care prediction and monitoring methods are based on visual evaluation and focus on surface…

图像与视频处理 · 电气工程与系统科学 2023-06-05 Prashant Kumar , Swatantra Dhara , Ayan Gope , Jyotirmoy Chatterjee , Subhamoy Mandal