中文
相关论文

相关论文: Multi-Modal Multi-Instance Learning for Retinal Di…

200 篇论文

Multiple Instance Learning (MIL) has been widely applied to medical imaging diagnosis, where bag labels are known and instance labels inside bags are unknown. Traditional MIL assumes that instances in each bag are independent samples from a…

图像与视频处理 · 电气工程与系统科学 2023-07-19 Yunan Wu , Francisco M. Castro-Macías , Pablo Morales-Álvarez , Rafael Molina , Aggelos K. Katsaggelos

Multiple instance learning (MIL) is a supervised learning methodology that aims to allow models to learn instance class labels from bag class labels, where a bag is defined to contain multiple instances. MIL is gaining traction for learning…

计算机视觉与模式识别 · 计算机科学 2019-11-14 Samuel W. Remedios , Zihao Wu , Camilo Bermudez , Cailey I. Kerley , Snehashis Roy , Mayur B. Patel , John A. Butman , Bennett A. Landman , Dzung L. Pham

Medical foundation models (MFMs) aim to learn universal representations from multimodal medical images that can generalize effectively to diverse downstream clinical tasks. However, most existing MFMs suffer from information ambiguity that…

计算机视觉与模式识别 · 计算机科学 2026-05-19 Yihang Liu , Longzhen Yang , Jiaxiong Yang , Ying Wen , Lianghua He , Heng Tao Shen

WHO has declared that more than 2.2 billion people worldwide are suffering from visual disorders, such as media haze, glaucoma, and drusen. At least 1 billion of these cases could have been either prevented or successfully treated, yet they…

图像与视频处理 · 电气工程与系统科学 2024-05-30 Yavuz Selim Inan

While Large Language Models (LLMs) are emerging as a promising direction in computational pathology, the substantial computational cost of giga-pixel Whole Slide Images (WSIs) necessitates the use of Multi-Instance Learning (MIL) to enable…

计算机视觉与模式识别 · 计算机科学 2025-11-12 Zhenfeng Zhuang , Fangyu Zhou , Liansheng Wang

Current fundus image analysis models are predominantly built for specific tasks relying on individual datasets. The learning process is usually based on data-driven paradigm without prior knowledge, resulting in poor transferability and…

计算机视觉与模式识别 · 计算机科学 2025-08-27 Ruiqi Wu , Chenran Zhang , Jianle Zhang , Yi Zhou , Tao Zhou , Huazhu Fu

Classification and localization are two pillars of visual object detectors. However, in CNN-based detectors, these two modules are usually optimized under a fixed set of candidate (or anchor) bounding boxes. This configuration significantly…

计算机视觉与模式识别 · 计算机科学 2019-12-06 Wei Ke , Tianliang Zhang , Zeyi Huang , Qixiang Ye , Jianzhuang Liu , Dong Huang

Ultra-widefield (UWF) imaging is a promising modality that captures a larger retinal field of view compared to traditional fundus photography. Previous studies showed that deep learning (DL) models are effective for detecting retinal…

图像与视频处理 · 电气工程与系统科学 2022-03-14 Justin Engelmann , Alice D. McTrusty , Ian J. C. MacCormick , Emma Pead , Amos Storkey , Miguel O. Bernabeu

In recent years, the incidence of vision-threatening eye diseases has risen dramatically, necessitating scalable and accurate screening solutions. This paper presents a comprehensive study on deep learning architectures for the automated…

计算机视觉与模式识别 · 计算机科学 2025-12-12 Mohammad Sadegh Gholizadeh , Amir Arsalan Rezapour

Fundus imaging such as CFP, OCT and UWF is crucial for the early detection of retinal anomalies and diseases. Fundus image understanding, due to its knowledge-intensive nature, poses a challenging vision-language task. An emerging approach…

计算机视觉与模式识别 · 计算机科学 2026-04-10 Yuchuan Deng , Qijie Wei , Kaiheng Qian , Jiazhen Liu , Zijie Xin , Bangxiang Lan , Jingyu Liu , Jianfeng Dong , Xirong Li

Retinal imaging has emerged as a powerful, non-invasive modality for detecting and quantifying biomarkers of systemic diseases-ranging from diabetes and hypertension to Alzheimer's disease and cardiovascular disorders but current insights…

图像与视频处理 · 电气工程与系统科学 2025-05-28 Tariq M Khan , Toufique Ahmed Soomro , Imran Razzak

Convolutional Neural Network models have successfully detected retinal illness from optical coherence tomography (OCT) and fundus images. These CNN models frequently rely on vast amounts of labeled data for training, difficult to obtain,…

计算机视觉与模式识别 · 计算机科学 2022-01-28 Sourya Dipta Das , Saikat Dutta , Nisarg A. Shah , Dwarikanath Mahapatra , Zongyuan Ge

Multiple Instance Learning (MIL) is a cornerstone approach in computational pathology (CPath) for generating clinically meaningful slide-level embeddings from gigapixel tissue images. However, MIL often struggles with small, weakly…

计算机视觉与模式识别 · 计算机科学 2025-06-12 Daniel Shao , Richard J. Chen , Andrew H. Song , Joel Runevic , Ming Y. Lu , Tong Ding , Faisal Mahmood

This paper studies automated categorization of age-related macular degeneration (AMD) given a multi-modal input, which consists of a color fundus image and an optical coherence tomography (OCT) image from a specific eye. Previous work uses…

图像与视频处理 · 电气工程与系统科学 2019-07-30 Weisen Wang , Zhiyan Xu , Weihong Yu , Jianchun Zhao , Jingyuan Yang , Feng He , Zhikun Yang , Di Chen , Dayong Ding , Youxin Chen , Xirong Li

The retina provides a unique, noninvasive window into Alzheimer's disease (AD) and dementia, capturing early structural changes through morphometric features, while systemic and lifestyle risk factors reflect well-established contributors…

计算机视觉与模式识别 · 计算机科学 2026-04-22 Seowung Leem , Lin Gu , Chenyu You , Kuang Gong , Ruogu Fang

With the increasing demand for histopathological specimen examination and diagnostic reporting, Multiple Instance Learning (MIL) has received heightened research focus as a viable solution for AI-centric diagnostic aid. Recently, to improve…

计算机视觉与模式识别 · 计算机科学 2025-12-25 Sungrae Hong , Sol Lee , Jisu Shin , Jiwon Jeong , Mun Yong Yi

Multiple instance learning (MIL) significantly reduced annotation costs via bag-level weak labels for large-scale images, such as histopathological whole slide images (WSIs). However, its adaptability to continual tasks with minimal…

计算机视觉与模式识别 · 计算机科学 2025-07-09 Byung Hyun Lee , Wongi Jeong , Woojae Han , Kyoungbun Lee , Se Young Chun

Multimodal fusion learning has shown significant promise in classifying various diseases such as skin cancer and brain tumors. However, existing methods face three key limitations. First, they often lack generalizability to other diagnosis…

计算机视觉与模式识别 · 计算机科学 2024-12-03 Joy Dhar , Nayyar Zaidi , Maryam Haghighat , Puneet Goyal , Sudipta Roy , Azadeh Alavi , Vikas Kumar

Multimodal depression detection is an important research topic that aims to predict human mental states using multimodal data. Previous methods treat different modalities equally and fuse each modality by na\"ive mathematical operations…

计算与语言 · 计算机科学 2024-01-09 Yuntao Wei , Yuzhe Zhang , Shuyang Zhang , Hong Zhang

With the increasing amounts of high-dimensional heterogeneous data to be processed, multi-modality feature selection has become an important research direction in medical image analysis. Traditional methods usually depict the data structure…

计算机视觉与模式识别 · 计算机科学 2022-04-11 Yuang Shi , Chen Zu , Mei Hong , Luping Zhou , Lei Wang , Xi Wu , Jiliu Zhou , Daoqiang Zhang , Yan Wang