English

Rethinking Evaluation of Infrared Small Target Detection

Computer Vision and Pattern Recognition 2025-09-24 v2

Abstract

As an essential vision task, infrared small target detection (IRSTD) has seen significant advancements through deep learning. However, critical limitations in current evaluation protocols impede further progress. First, existing methods rely on fragmented pixel- and target-level specific metrics, which fails to provide a comprehensive view of model capabilities. Second, an excessive emphasis on overall performance scores obscures crucial error analysis, which is vital for identifying failure modes and improving real-world system performance. Third, the field predominantly adopts dataset-specific training-testing paradigms, hindering the understanding of model robustness and generalization across diverse infrared scenarios. This paper addresses these issues by introducing a hybrid-level metric incorporating pixel- and target-level performance, proposing a systematic error analysis method, and emphasizing the importance of cross-dataset evaluation. These aim to offer a more thorough and rational hierarchical analysis framework, ultimately fostering the development of more effective and robust IRSTD models. An open-source toolkit has be released to facilitate standardized benchmarking.

Keywords

Cite

@article{arxiv.2509.16888,
  title  = {Rethinking Evaluation of Infrared Small Target Detection},
  author = {Youwei Pang and Xiaoqi Zhao and Lihe Zhang and Huchuan Lu and Georges El Fakhri and Xiaofeng Liu and Shijian Lu},
  journal= {arXiv preprint arXiv:2509.16888},
  year   = {2025}
}

Comments

NeurIPS 2025; Evaluation Toolkit: https://github.com/lartpang/PyIRSTDMetrics; Correct a few typos

R2 v1 2026-07-01T05:47:54.555Z