AI Workflow, External Validation, and Development in Eye Disease Diagnosis
Abstract
Timely disease diagnosis is challenging due to increasing disease burdens and limited clinician availability. AI shows promise in diagnosis accuracy but faces real-world application issues due to insufficient validation in clinical workflows and diverse populations. This study addresses gaps in medical AI downstream accountability through a case study on age-related macular degeneration (AMD) diagnosis and severity classification. We designed and implemented an AI-assisted diagnostic workflow for AMD, comparing diagnostic performance with and without AI assistance among 24 clinicians from 12 institutions with real patient data sampled from the Age-Related Eye Disease Study (AREDS). Additionally, we demonstrated continual enhancement of an existing AI model by incorporating approximately 40,000 additional medical images (named AREDS2 dataset). The improved model was then systematically evaluated using both AREDS and AREDS2 test sets, as well as an external test set from Singapore. AI assistance markedly enhanced diagnostic accuracy and classification for 23 out of 24 clinicians, with the average F1-score increasing by 20% from 37.71 (Manual) to 45.52 (Manual + AI) (P-value < 0.0001), achieving an improvement of over 50% in some cases. In terms of efficiency, AI assistance reduced diagnostic times for 17 out of the 19 clinicians tracked, with time savings of up to 40%. Furthermore, a model equipped with continual learning showed robust performance across three independent datasets, recording a 29% increase in accuracy, and elevating the F1-score from 42 to 54 in the Singapore population.
Cite
@article{arxiv.2409.15087,
title = {AI Workflow, External Validation, and Development in Eye Disease Diagnosis},
author = {Qingyu Chen and Tiarnan D L Keenan and Elvira Agron and Alexis Allot and Emily Guan and Bryant Duong and Amr Elsawy and Benjamin Hou and Cancan Xue and Sanjeeb Bhandari and Geoffrey Broadhead and Chantal Cousineau-Krieger and Ellen Davis and William G Gensheimer and David Grasic and Seema Gupta and Luis Haddock and Eleni Konstantinou and Tania Lamba and Michele Maiberger and Dimosthenis Mantopoulos and Mitul C Mehta and Ayman G Nahri and Mutaz AL-Nawaflh and Arnold Oshinsky and Brittany E Powell and Boonkit Purt and Soo Shin and Hillary Stiefel and Alisa T Thavikulwat and Keith James Wroblewski and Tham Yih Chung and Chui Ming Gemmy Cheung and Ching-Yu Cheng and Emily Y Chew and Michelle R. Hribar and Michael F. Chiang and Zhiyong Lu},
journal= {arXiv preprint arXiv:2409.15087},
year = {2025}
}
Comments
Published in JAMA Network Open, doi:10.1001/jamanetworkopen.2025.17204