English

A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks

Computation and Language 2025-02-11 v1 Artificial Intelligence

Abstract

Theory of Mind (ToM), the ability to attribute mental states to others and predict their behaviour, is fundamental to social intelligence. In this paper, we survey studies evaluating behavioural and representational ToM in Large Language Models (LLMs), identify important safety risks from advanced LLM ToM capabilities, and suggest several research directions for effective evaluation and mitigation of these risks.

Keywords

Cite

@article{arxiv.2502.06470,
  title  = {A Survey of Theory of Mind in Large Language Models: Evaluations, Representations, and Safety Risks},
  author = {Hieu Minh "Jord" Nguyen},
  journal= {arXiv preprint arXiv:2502.06470},
  year   = {2025}
}

Comments

Advancing Artificial Intelligence through Theory of Mind Workshop, AAAI 2025