English

Evaluation of AI Chatbots for Patient-Specific EHR Questions

Computation and Language 2023-06-06 v1 Artificial Intelligence Information Retrieval

Abstract

This paper investigates the use of artificial intelligence chatbots for patient-specific question answering (QA) from clinical notes using several large language model (LLM) based systems: ChatGPT (versions 3.5 and 4), Google Bard, and Claude. We evaluate the accuracy, relevance, comprehensiveness, and coherence of the answers generated by each model using a 5-point Likert scale on a set of patient-specific questions.

Keywords

Cite

@article{arxiv.2306.02549,
  title  = {Evaluation of AI Chatbots for Patient-Specific EHR Questions},
  author = {Alaleh Hamidi and Kirk Roberts},
  journal= {arXiv preprint arXiv:2306.02549},
  year   = {2023}
}