English

Reproducing and Improving CheXNet: Deep Learning for Chest X-ray Disease Classification

Image and Video Processing 2026-02-25 v3 Computer Vision and Pattern Recognition Machine Learning

Abstract

Deep learning for radiologic image analysis is a rapidly growing field in biomedical research and is likely to become a standard practice in modern medicine. On the publicly available NIH ChestX-ray14 dataset, containing X-ray images that are classified by the presence or absence of 14 different diseases, we reproduced an algorithm known as CheXNet, as well as explored other algorithms that outperform CheXNet's baseline metrics. Model performance was primarily evaluated using the F1 score and AUC-ROC, both of which are critical metrics for imbalanced, multi-label classification tasks in medical imaging. The best model achieved an average AUC-ROC score of 0.85 and an average F1 score of 0.39 across all 14 disease classifications present in the dataset.

Keywords

Cite

@article{arxiv.2505.06646,
  title  = {Reproducing and Improving CheXNet: Deep Learning for Chest X-ray Disease Classification},
  author = {Daniel J. Strick and Carlos Garcia and Anthony Huang and Thomas Gardos},
  journal= {arXiv preprint arXiv:2505.06646},
  year   = {2026}
}

Comments

13 pages, 4 figures

R2 v1 2026-06-28T23:28:09.125Z