English

Development, Evaluation, and Deployment of a Multi-Agent System for Thoracic Tumor Board

Artificial Intelligence 2026-04-15 v1

Abstract

Tumor boards are multidisciplinary conferences dedicated to producing actionable patient care recommendations with live review of primary radiology and pathology data. Succinct patient case summaries are needed to drive efficient and accurate case discussions. We developed a manual AI-based workflow to generate patient summaries to display live at the Stanford Thoracic Tumor board. To improve on this manually intensive process, we developed several automated AI chart summarization methods and evaluated them against physician gold standard summaries and fact-based scoring rubrics. We report these comparative evaluations as well as our deployment of the final state automated AI chart summarization tool along with post-deployment monitoring. We also validate the use of an LLM as a judge evaluation strategy for fact-based scoring. This work is an example of integrating AI-based workflows into routine clinical practice.

Keywords

Cite

@article{arxiv.2604.12161,
  title  = {Development, Evaluation, and Deployment of a Multi-Agent System for Thoracic Tumor Board},
  author = {Tim Ellis-Caleo and Timothy Keyes and Nerissa Ambers and Faraah Bekheet and Wen-wai Yim and Nikesh Kotecha and Nigam H. Shah and Joel Neal},
  journal= {arXiv preprint arXiv:2604.12161},
  year   = {2026}
}

Comments

64 pages, 14 figures

R2 v1 2026-07-01T12:07:45.839Z