English
Related papers

Related papers: Conversational Medical AI: Ready for Practice

200 papers

An effective healthcare agent must be able to recall and reason over a patient's longitudinal medical history. However, the absence of datasets with realistic long-term dialogue timelines limits systematic evaluation. Real clinical text is…

Computation and Language · Computer Science 2026-05-20 Hebin Hu , Renke Dai , Ah-Hwee Tan , Yilin Kang

Clinical communication skills are essential for preparing healthcare professionals to provide equitable care across cultures. However, traditional training with simulated patients can be resource intensive and difficult to scale, especially…

Human-Computer Interaction · Computer Science 2025-06-16 Sandro Radovanović , Shuangyu Li

Clinical decision-making in emergency medicine demands rapid, accurate diagnoses under uncertainty. Despite benchmark progress, evidence for LLMs as interactive aids in live physician workflows remains sparse. MedSyn lets physicians…

Large language models (LLMs) have shown remarkable progress in encoding clinical knowledge and responding to complex medical queries with appropriate clinical reasoning. However, their applicability in subspecialist or complex medical…

Background: Clinical guidelines are central to safe evidence-based medicine in modern healthcare, providing diagnostic criteria, treatment options and monitoring advice for a wide range of illnesses. LLM-empowered chatbots have shown great…

Computation and Language · Computer Science 2025-05-07 Julia Ive , Felix Jozsa , Nick Jackson , Paulina Bondaronek , Ciaran Scott Hill , Richard Dobson

Conversational AI chatbots have become increasingly common within the customer service industry. Despite improvements in their emotional development, they often lack the authenticity of real customer service interactions or the competence…

Human-Computer Interaction · Computer Science 2025-02-14 Antonin Brun , Ruying Liu , Aryan Shukla , Frances Watson , Jonathan Gratch

The promise of AI in medicine depends on learning from data that reflect what matters to patients and clinicians. Most existing models are trained on electronic health records (EHRs), which capture biological measures but rarely…

Language models excel at diagnostic assessments on curated medical case-studies and vignettes, performing on par with, or better than, clinical professionals. However, existing studies focus on complex scenarios with rich context making it…

Although artificial intelligence (AI) agents are increasingly proposed to support potentially longitudinal health tasks, such as symptom management, behavior change, and patient support, most current implementations fall short of…

Artificial Intelligence · Computer Science 2026-04-30 Georgianna Lin , Rencong Jiang , Noémie Elhadad , Xuhai "Orson" Xu

The rapid advancement of Large Language Models (LLMs), reasoning models, and agentic AI approaches coincides with a growing global mental health crisis, where increasing demand has not translated into adequate access to professional…

Human-Computer Interaction · Computer Science 2025-04-03 Kellie Yu Hui Sim , Kenny Tsu Wei Choo

Recent advances in large language models made it possible to achieve high conversational performance with substantially reduced computational demands, enabling practical on-site deployment in clinical environments. Such progress allows for…

Machine Learning · Computer Science 2025-11-25 Jan Benedikt Ruhland , Doguhan Bahcivan , Jan-Peter Sowa , Ali Canbay , Dominik Heider

Trustworthy clinical advice is crucial but burdensome when seeking health support from professionals. Inaccessibility and financial burdens present obstacles to obtaining professional clinical advice, even when healthcare is available.…

Human-Computer Interaction · Computer Science 2024-02-14 Delong Du , Richard Paluch , Gunnar Stevens , Claudia Müller

One challenge in technical interviews is the think-aloud process, where candidates verbalize their thought processes while solving coding tasks. Despite its importance, opportunities for structured practice remain limited. Conversational AI…

Human-Computer Interaction · Computer Science 2025-07-22 Taufiq Daryanto , Sophia Stil , Xiaohan Ding , Daniel Manesh , Sang Won Lee , Tim Lee , Stephanie Lunn , Sarah Rodriguez , Chris Brown , Eugenia Rho

Autonomous agents utilizing Large Language Models (LLMs) have demonstrated remarkable capabilities in isolated medical tasks like diagnosis and image analysis, but struggle with integrated clinical workflows that connect diagnostic…

Artificial Intelligence · Computer Science 2025-10-14 Hongjie Zheng , Zesheng Shi , Ping Yi

Biased AI-generated medical advice and misdiagnoses can jeopardize patient safety, making the integrity of AI in healthcare more critical than ever. As Large Language Models (LLMs) take on a growing role in medical decision-making,…

Computation and Language · Computer Science 2024-10-10 Pardis Sadat Zahraei , Zahra Shakeri

Large Language Models (LLMs) are increasingly utilized for mental health support; however, current safety benchmarks often fail to detect the complex, longitudinal risks inherent in therapeutic dialogue. We introduce an evaluation framework…

Computation and Language · Computer Science 2026-03-06 Ian Steenstra , Paola Pedrelli , Weiyan Shi , Stacy Marsella , Timothy W. Bickmore

Large Language Models (LLMs) have demonstrated remarkable proficiency in human interactions, yet their application within the medical field remains insufficiently explored. Previous works mainly focus on the performance of medical knowledge…

Computation and Language · Computer Science 2024-07-23 Yusheng Liao , Yutong Meng , Yuhao Wang , Hongcheng Liu , Yanfeng Wang , Yu Wang

Conversational AI systems have emerged as key enablers of human-like interactions across diverse sectors. Nevertheless, the balance between linguistic nuance and factual accuracy has proven elusive. In this paper, we first introduce…

Computation and Language · Computer Science 2024-06-18 Ahtsham Zafar , Venkatesh Balavadhani Parthasarathy , Chan Le Van , Saad Shahid , Aafaq Iqbal khan , Arsalan Shahid

There is a lack of benchmarks for evaluating large language models (LLMs) in long-form medical question answering (QA). Most existing medical QA evaluation benchmarks focus on automatic metrics and multiple-choice questions. While valuable,…

Computation and Language · Computer Science 2024-11-21 Pedram Hosseini , Jessica M. Sin , Bing Ren , Bryceton G. Thomas , Elnaz Nouri , Ali Farahanchi , Saeed Hassanpour

With advanced AI/ML, there has been growing research on explainable AI (XAI) and studies on how humans interact with AI and XAI for effective human-AI collaborative decision-making. However, we still have a lack of understanding of how AI…

Human-Computer Interaction · Computer Science 2024-05-28 Min Hun Lee , Silvana Xin Yi Choo , Shamala D/O Thilarajah
‹ Prev 1 4 5 6 7 8 10 Next ›