English
Related papers

Related papers: EHRNoteQA: An LLM Benchmark for Real-World Clinica…

200 papers

Writing discharge summaries to transfer medical information is an important but time-consuming process that can be assisted by Large Language Models (LLMs). This prospective mixed methods pilot study evaluated an Electronic Health Record…

Background: Recent advancements in large language models (LLMs) offer potential benefits in healthcare, particularly in processing extensive patient records. However, existing benchmarks do not fully assess LLMs' capability in handling…

Generating discharge summaries is a crucial yet time-consuming task in clinical practice, essential for conveying pertinent patient information and facilitating continuity of care. Recent advancements in large language models (LLMs) have…

Computation and Language · Computer Science 2025-07-01 Yiming Li , Fang Li , Kirk Roberts , Licong Cui , Cui Tao , Hua Xu

Patients have distinct information needs about their hospitalization that can be addressed using clinical evidence from electronic health records (EHRs). While artificial intelligence (AI) systems show promise in meeting these needs, robust…

Computation and Language · Computer Science 2026-03-31 Sarvesh Soni , Dina Demner-Fushman

Structured Electronic Health Record (EHR) data stores patient information in relational tables and plays a central role in clinical decision-making. Recent advances have explored the use of large language models (LLMs) to process such data,…

Artificial Intelligence · Computer Science 2026-04-02 Xiao Yang , Xuejiao Zhao , Zhiqi Shen

Identifying medication discontinuations in electronic health records (EHRs) is vital for patient safety but is often hindered by information being buried in unstructured notes. This study aims to evaluate the capabilities of advanced…

Computation and Language · Computer Science 2025-11-10 Chong Shao , Douglas Snyder , Chiran Li , Bowen Gu , Kerry Ngan , Chun-Ting Yang , Jiageng Wu , Richard Wyss , Kueiyu Joshua Lin , Jie Yang

Electronic health records (EHRs) contain vast amounts of complex data, but harmonizing and processing this information remains a challenging and costly task requiring significant clinical expertise. While large language models (LLMs) have…

Computation and Language · Computer Science 2024-07-02 João Matos , Jack Gallifant , Jian Pei , A. Ian Wong

With the rapid advancement of Large Language Models (LLMs) and their outstanding performance in semantic and contextual comprehension, the potential of LLMs in specialized domains warrants exploration. This paper introduces the NoteAid EHR…

Computation and Language · Computer Science 2024-01-01 Xiaocheng Zhang , Zonghai Yao , Hong Yu

Clinical problem-solving requires processing of semantic medical knowledge such as illness scripts and numerical medical knowledge of diagnostic tests for evidence-based decision-making. As large language models (LLMs) show promising…

Electronic Health Records (EHRs) often lack explicit links between medications and diagnoses, making clinical decision-making and research more difficult. Even when links exist, diagnosis lists may be incomplete, especially during early…

Computation and Language · Computer Science 2025-03-31 Dina Albassam , Adam Cross , Chengxiang Zhai

Large language models (LLMs) hold great promise for medical applications and are evolving rapidly, with new models being released at an accelerated pace. However, benchmarking on large-scale real-world data such as electronic health records…

While increasing patients' access to medical documents improves medical care, this benefit is limited by varying health literacy levels and complex medical terminology. Large language models (LLMs) offer solutions by simplifying medical…

Computation and Language · Computer Science 2025-02-06 Amin Dada , Osman Alperen Koras , Marie Bauer , Amanda Butler , Kaleb E. Smith , Jens Kleesiek , Julian Friedrich

To improve the reliability of Large Language Models (LLMs) in clinical applications, retrieval-augmented generation (RAG) is extensively applied to provide factual medical knowledge. However, beyond general medical knowledge from open-ended…

Computation and Language · Computer Science 2025-05-29 Justice Ou , Tinglin Huang , Yilun Zhao , Ziyang Yu , Peiqing Lu , Rex Ying

Evaluating large language models (LLMs) in medicine is crucial because medical applications require high accuracy with little room for error. Current medical benchmarks have three main types: medical exam-based, comprehensive medical, and…

Patient summarization is essential for clinicians to provide coordinated care and practice effective communication. Automated summarization has the potential to save time, standardize notes, aid clinical decision making, and reduce medical…

Information Retrieval · Computer Science 2018-10-30 Emily Alsentzer , Anne Kim

Electronic Health Record (EHR) retrieval plays a pivotal role in various clinical tasks, but its development has been severely impeded by the lack of publicly available benchmarks. In this paper, we introduce a novel public EHR retrieval…

Information Retrieval · Computer Science 2025-04-09 Zhengyun Zhao , Hongyi Yuan , Jingjing Liu , Haichao Chen , Huaiyuan Ying , Songchi Zhou , Yue Zhong , Sheng Yu

Discharge communication is a critical yet underexplored component of patient care, where the goal shifts from diagnosis to education. While recent large language model (LLM) benchmarks emphasize in-visit diagnostic reasoning, they fail to…

Computation and Language · Computer Science 2025-09-22 Zonghai Yao , Michael Sun , Won Seok Jang , Sunjae Kwon , Soie Kwon , Hong Yu

Large language models (LLMs) are increasingly used to generate summaries from clinical notes. However, their ability to preserve essential diagnostic information remains underexplored, which could lead to serious risks for patient care.…

Computation and Language · Computer Science 2026-02-20 Heloisa Oss Boll , Antonio Oss Boll , Leticia Puttlitz Boll , Ameen Abu Hanna , Iacer Calixto

We introduce a novel question-answering (QA) dataset using echocardiogram reports sourced from the Medical Information Mart for Intensive Care database. This dataset is specifically designed to enhance QA systems in cardiology, consisting…

Artificial Intelligence · Computer Science 2025-03-07 Lama Moukheiber , Mira Moukheiber , Dana Moukheiiber , Jae-Woo Ju , Hyung-Chul Lee

Incomplete or inconsistent discharge documentation is a primary driver of care fragmentation and avoidable readmissions. Despite its critical role in patient safety, auditing discharge summaries relies heavily on manual review and is…

Artificial Intelligence · Computer Science 2026-04-08 Akshat Dasula , Prasanna Desikan , Jaideep Srivastava
‹ Prev 1 2 3 10 Next ›