中文
相关论文

相关论文: COROLLA: An Efficient Multi-Modality Fusion Framew…

200 篇论文

Recent advances in image-text pretraining have significantly enhanced visual understanding by aligning visual and textual representations. Contrastive Language-Image Pretraining (CLIP) has played a pivotal role in multimodal learning.…

计算机视觉与模式识别 · 计算机科学 2026-02-20 Zihan Li , Yiqing Wang , Sina Farsiu , Paul Kinahan

Glaucoma is an eye disease that causes damage to the optic nerve, which can lead to visual loss and permanent blindness. Early glaucoma detection is therefore critical in order to avoid permanent blindness. The estimation of the cup-to-disc…

图像与视频处理 · 电气工程与系统科学 2023-11-27 Mehwish Mehmood , Khuram Naveed , Khursheed Aurangzeb , Haroon Ahmed Khan , Musaed Alhussein , Syed Saud Naqvi

Glaucoma is a progressive eye disease that leads to optic nerve damage, causing irreversible vision loss if left untreated. Optical coherence tomography (OCT) has become a crucial tool for glaucoma diagnosis, offering high-resolution 3D…

计算机视觉与模式识别 · 计算机科学 2025-10-02 Roshan Kenia , Anfei Li , Rishabh Srivastava , Kaveri A. Thakoor

Ophthalmic images and derivatives such as the retinal nerve fiber layer (RNFL) thickness map are crucial for detecting and monitoring ophthalmic diseases (e.g., glaucoma). For computer-aided diagnosis of eye diseases, the key technique is…

Background/Aims: Standard Automated Perimetry (SAP) is the gold standard to monitor visual field (VF) loss in glaucoma management, but is prone to intra-subject variability. We developed and validated a deep learning (DL) regression model…

图像与视频处理 · 电气工程与系统科学 2021-06-08 Ruben Hemelings , Bart Elen , João Barbosa Breda , Erwin Bellon , Matthew B Blaschko , Patrick De Boever , Ingeborg Stalmans

In high-stakes medical applications, consistent answering across diverse question phrasings is essential for reliable diagnosis. However, we reveal that current Medical Vision-Language Models (Med-VLMs) exhibit concerning fragility in…

计算与语言 · 计算机科学 2025-08-27 Songtao Jiang , Yuxi Chen , Sibo Song , Yan Zhang , Yeying Jin , Yang Feng , Jian Wu , Zuozhu Liu

With the rapid development of artificial intelligence (AI) in medical image processing, deep learning in color fundus photography (CFP) analysis is also evolving. Although there are some open-source, labeled datasets of CFPs in the…

Glaucomatous optic neuropathy (GON) is a prevalent ocular disease that can lead to irreversible vision loss if not detected early and treated. The traditional diagnostic approach for GON involves a set of ophthalmic examinations, which are…

Glaucoma remains among the leading causes of blindness despite many treatment options available today. Effective treatment requires early diagnosis, which is difficult to achieve with existing imaging technologies that detect the already…

Purpose: (1) To assess the performance of geometric deep learning (PointNet) in diagnosing glaucoma from a single optical coherence tomography (OCT) 3D scan of the optic nerve head (ONH); (2) To compare its performance to that obtained with…

图像与视频处理 · 电气工程与系统科学 2022-04-15 Alexandre H. Thiery , Fabian Braeu , Tin A. Tun , Tin Aung , Michael J. A. Girard

Accurate segmentation of the optic disc and cup is critical for the early diagnosis and management of ocular diseases such as glaucoma. However, segmentation models trained on one dataset often suffer significant performance degradation…

计算机视觉与模式识别 · 计算机科学 2025-09-15 Rini Smita Thakur , Rajeev Ranjan Dwivedi , Vinod K Kurmi

Rich temporal information and variations in viewpoints make video data an attractive choice for learning image representations using unsupervised contrastive learning (UCL) techniques. State-of-the-art (SOTA) contrastive learning techniques…

图像与视频处理 · 电气工程与系统科学 2022-07-28 Soumen Basu , Somanshu Singla , Mayank Gupta , Pratyaksha Rana , Pankaj Gupta , Chetan Arora

Fundus image segmentation on unseen domains is challenging, especially for the over-parameterized deep models trained on the small medical datasets. To address this challenge, we propose a method named Adaptive Feature-fusion Neural Network…

计算机视觉与模式识别 · 计算机科学 2024-04-03 Jiyuan Zhong , Hu Ke , Ming Yan

In this paper, we present a self-training-based framework for glaucoma grading using OCT B-scans under the presence of domain shift. Particularly, the proposed two-step learning methodology resorts to pseudo-labels generated during the…

计算机视觉与模式识别 · 计算机科学 2021-11-24 Gabriel García , Adrián Colomer , Rafael Verdú-Monedero , José Dolz , Valery Naranjo

Despite significant advancements in Vision-Language Models (VLMs), the performance of existing VLMs remains hindered by object hallucination, a critical challenge to achieving accurate visual understanding. To address this issue, we propose…

计算机视觉与模式识别 · 计算机科学 2026-03-31 Woohyeon Park , Woojin Kim , Jaeik Kim , Jaeyoung Do

Graph Neural Networks (GNNs) have become popular in Graph Representation Learning (GRL). One fundamental application is few-shot node classification. Most existing methods follow the meta learning paradigm, showing the ability of fast…

机器学习 · 计算机科学 2023-09-20 Hao Liu , Jiarui Feng , Lecheng Kong , Dacheng Tao , Yixin Chen , Muhan Zhang

Graph contrastive learning has emerged as a powerful technique for learning graph representations that are robust and discriminative. However, traditional approaches often neglect the critical role of subgraph structures, particularly the…

机器学习 · 计算机科学 2025-03-14 Tianhao Peng , Xuhong Li , Haitao Yuan , Yuchen Li , Haoyi Xiong

Learning medical visual representations directly from paired radiology reports has become an emerging topic in representation learning. However, existing medical image-text joint learning methods are limited by instance or local supervision…

计算机视觉与模式识别 · 计算机科学 2022-10-13 Fuying Wang , Yuyin Zhou , Shujun Wang , Varut Vardhanabhuti , Lequan Yu

In this paper, we propose a framework that incorporates experts diagnostics and insights into the analysis of Optical Coherence Tomography (OCT) using multi-modal learning. To demonstrate the effectiveness of this approach, we create a…

图像与视频处理 · 电气工程与系统科学 2022-03-22 Y. Logan , K. Kokilepersaud , G. Kwon , G. AlRegib , C. Wykoff , H. Yu

In this paper, we proposed Transferable Ranking Convolutional Neural Network (TRk-CNN) that can be effectively applied when the classes of images to be classified show a high correlation with each other. The multi-class classification…

计算机视觉与模式识别 · 计算机科学 2019-05-17 Tae Joon Jun , Youngsub Eom , Dohyeun Kim , Cherry Kim , Ji-Hye Park , Hoang Minh Nguyen , Daeyoung Kim