English

PLACID: Privacy-preserving Large language models for Acronym Clinical Inference and Disambiguation

Computation and Language 2026-03-26 v1 Artificial Intelligence

Abstract

Large Language Models (LLMs) offer transformative solutions across many domains, but healthcare integration is hindered by strict data privacy constraints. Clinical narratives are dense with ambiguous acronyms, misinterpretation these abbreviations can precipitate severe outcomes like life-threatening medication errors. While cloud-dependent LLMs excel at Acronym Disambiguation, transmitting Protected Health Information to external servers violates privacy frameworks. To bridge this gap, this study pioneers the evaluation of small-parameter models deployed entirely on-device to ensure privacy preservation. We introduce a privacy-preserving cascaded pipeline leveraging general-purpose local models to detect clinical acronyms, routing them to domain-specific biomedical models for context-relevant expansions. Results reveal that while general instruction-following models achieve high detection accuracy (~0.988), their expansion capabilities plummet (~0.655). Our cascaded approach utilizes domain-specific medical models to increase expansion accuracy to (~0.81). This novel work demonstrates that privacy-preserving, on-device (2B-10B) models deliver high-fidelity clinical acronym disambiguation support.

Keywords

Cite

@article{arxiv.2603.23678,
  title  = {PLACID: Privacy-preserving Large language models for Acronym Clinical Inference and Disambiguation},
  author = {Manjushree B. Aithal and Ph. D. and Alexander Kotz and James Mitchell and Ph. D},
  journal= {arXiv preprint arXiv:2603.23678},
  year   = {2026}
}

Comments

10 pages, 2 figures, Under review AMIA Symposium