English

Critical Challenges and Guidelines in Evaluating Synthetic Tabular Data: A Systematic Review

Machine Learning 2026-05-15 v3 Artificial Intelligence Computers and Society

Abstract

Generating synthetic tabular health data is challenging, and evaluating their quality is equally, if not more, complex. This systematic review highlights the critical importance of rigorous evaluation of synthetic health data to ensure reliability, clinical relevance, and appropriate use. From an initial identification of 2067 relevant papers published in the last ten years, 134 studies were selected for detailed analysis. Our review identifies key challenges, including lack of consensus on evaluation methods, inconsistent application of evaluation metrics, limited involvement of domain experts, inadequate reporting of dataset characteristics, and limited reproducibility of results. In response, we provide a structured consolidation of synthetic data generation and evaluation methods into taxonomies, alongside practical guidelines to support more robust and standardised evaluation practices. These findings aim to support the responsible development and use of synthetic health data, aligned with emerging expectations around transparency, reproducibility, and governance, ultimately enabling the community to fully harness its transformative potential and accelerate innovation.

Keywords

Cite

@article{arxiv.2504.18544,
  title  = {Critical Challenges and Guidelines in Evaluating Synthetic Tabular Data: A Systematic Review},
  author = {Nazia Nafis and Inaki Esnaola and Alvaro Martinez-Perez and Maria-Cruz Villa-Uriol and Venet Osmani},
  journal= {arXiv preprint arXiv:2504.18544},
  year   = {2026}
}

Comments

32 pages

R2 v1 2026-06-28T23:11:42.695Z