English

Towards objectively evaluating the quality of generated medical summaries

Computation and Language 2021-04-12 v1

Abstract

We propose a method for evaluating the quality of generated text by asking evaluators to count facts, and computing precision, recall, f-score, and accuracy from the raw counts. We believe this approach leads to a more objective and easier to reproduce evaluation. We apply this to the task of medical report summarisation, where measuring objective quality and accuracy is of paramount importance.

Keywords

Cite

@article{arxiv.2104.04412,
  title  = {Towards objectively evaluating the quality of generated medical summaries},
  author = {Francesco Moramarco and Damir Juric and Aleksandar Savkov and Ehud Reiter},
  journal= {arXiv preprint arXiv:2104.04412},
  year   = {2021}
}