English

Optimizing the Factual Correctness of a Summary: A Study of Summarizing Radiology Reports

Computation and Language 2020-04-29 v3

Abstract

Neural abstractive summarization models are able to generate summaries which have high overlap with human references. However, existing models are not optimized for factual correctness, a critical metric in real-world applications. In this work, we develop a general framework where we evaluate the factual correctness of a generated summary by fact-checking it automatically against its reference using an information extraction module. We further propose a training strategy which optimizes a neural summarization model with a factual correctness reward via reinforcement learning. We apply the proposed method to the summarization of radiology reports, where factual correctness is a key requirement. On two separate datasets collected from hospitals, we show via both automatic and human evaluation that the proposed approach substantially improves the factual correctness and overall quality of outputs over a competitive neural summarization system, producing radiology summaries that approach the quality of human-authored ones.

Keywords

Cite

@article{arxiv.1911.02541,
  title  = {Optimizing the Factual Correctness of a Summary: A Study of Summarizing Radiology Reports},
  author = {Yuhao Zhang and Derek Merck and Emily Bao Tsai and Christopher D. Manning and Curtis P. Langlotz},
  journal= {arXiv preprint arXiv:1911.02541},
  year   = {2020}
}

Comments

ACL2020. 13 pages with appendices