Computation and Language · Computer Science
NegotiationToM: A Benchmark for Stress-testing Machine Theory of Mind on Negotiation Surrounding
Chunkit Chan, Cheng Jiayang, Yauwai Yim, Zheye Deng +6
2024-10-08
Computation and Language · Computer Science
PersuasiveToM: A Benchmark for Evaluating Machine Theory of Mind in Persuasive Dialogues
Fangxu Yu, Lai Jiang, Shenyi Huang, Zhen Wu +1
2025-05-27
Artificial Intelligence · Computer Science
OpenToM: A Comprehensive Benchmark for Evaluating Theory-of-Mind Reasoning Capabilities of Large Language Models
Hainiu Xu, Runcong Zhao, Lixing Zhu, Jinhua Du +1
2024-06-04
Computation and Language · Computer Science
Do Theory of Mind Benchmarks Need Explicit Human-like Reasoning in Language Models?
Yi-Long Lu, Chunhui Zhang, Jiajun Song, Lifeng Fan +1
2025-05-19
Human-Computer Interaction · Computer Science
Rethinking Theory of Mind Benchmarks for LLMs: Towards A User-Centered Perspective
Qiaosi Wang, Xuhui Zhou, Maarten Sap, Jodi Forlizzi +1
2025-04-16
Computation and Language · Computer Science
Understanding Social Reasoning in Language Models with Language Models
Kanishk Gandhi, Jan-Philipp Fränken, Tobias Gerstenberg, Noah D. Goodman
2023-12-06
Artificial Intelligence · Computer Science
CogToM: A Comprehensive Theory of Mind Benchmark inspired by Human Cognition for Large Language Models
Haibo Tong, Zeyang Yue, Feifei Zhao, Erliang Lin +5
2026-01-23
Computation and Language · Computer Science
Towards Dynamic Theory of Mind: Evaluating LLM Adaptation to Temporal Evolution of Human States
Yang Xiao, Jiashuo Wang, Qiancheng Xu, Changhe Song +4
2025-06-10
Computation and Language · Computer Science
ToMBench: Benchmarking Theory of Mind in Large Language Models
Zhuang Chen, Jincenzi Wu, Jinfeng Zhou, Bosi Wen +7
2024-12-10
Computation and Language · Computer Science
Theory of Mind in Large Language Models: Assessment and Enhancement
Ruirui Chen, Weifeng Jiang, Chengwei Qin, Cheston Tan
2025-08-26
Machine Learning · Computer Science
Explore Theory of Mind: Program-guided adversarial data generation for theory of mind reasoning
Melanie Sclar, Jane Yu, Maryam Fazel-Zarandi, Yulia Tsvetkov +3
2024-12-18
Computation and Language · Computer Science
DialToM: A Theory of Mind Benchmark for Forecasting State-Driven Dialogue Trajectories
Neemesh Yadav, Palakorn Achananuparp, Jing Jiang, Ee-Peng Lim
2026-05-29
Artificial Intelligence · Computer Science
OmniToM: Benchmarking Theory of Mind in LLMs via Explicit Belief Modeling
Adam Bawatneh, Sagar Sapkota, Amrit Singh Bedi, Santu Karmaker +1
2026-05-27
Computation and Language · Computer Science
XToM: Exploring the Multilingual Theory of Mind for Large Language Models
Chunkit Chan, Yauwai Yim, Hongchuan Zeng, Zhiying Zou +13
2025-06-04
Computation and Language · Computer Science
UniToMBench: Integrating Perspective-Taking to Improve Theory of Mind in LLMs
Prameshwar Thiyagarajan, Vaishnavi Parimi, Shamant Sai, Soumil Garg +4
2025-06-12
Computation and Language · Computer Science
Do Large Language Models Possess a Theory of Mind? A Comparative Evaluation Using the Strange Stories Paradigm
Anna Babarczy, Andras Lukacs, Peter Vedres, Zeteny Bujka
2026-03-20
Computation and Language · Computer Science
Beyond Words: Evaluating and Bridging Epistemic Divergence in User-Agent Interaction via Theory of Mind
Minyuan Ruan, Ziyue Wang, Kaiming Liu, Yunghwei Lai +2
2026-02-17
Computation and Language · Computer Science
Clever Hans or Neural Theory of Mind? Stress Testing Social Reasoning in Large Language Models
Natalie Shapira, Mosh Levy, Seyed Hossein Alavi, Xuhui Zhou +4
2023-05-25
Computation and Language · Computer Science
ToMATO: Verbalizing the Mental States of Role-Playing LLMs for Benchmarking Theory of Mind
Kazutoshi Shinoda, Nobukatsu Hojo, Kyosuke Nishida, Saki Mizuno +4
2025-01-16
Computation and Language · Computer Science
HI-TOM: A Benchmark for Evaluating Higher-Order Theory of Mind Reasoning in Large Language Models
Yinghui He, Yufan Wu, Yilin Jia, Rada Mihalcea +2
2023-10-26
Computation and Language · Computer Science
TactfulToM: Do LLMs Have the Theory of Mind Ability to Understand White Lies?
Yiwei Liu, Emma Jane Pretty, Jiahao Huang, Saku Sugawara
2025-09-26
Artificial Intelligence · Computer Science
Position: Theory of Mind Benchmarks are Broken for Large Language Models
Matthew Riemer, Zahra Ashktorab, Djallel Bouneffouf, Payel Das +3
2025-06-13
Artificial Intelligence · Computer Science
Think Twice: Perspective-Taking Improves Large Language Models' Theory-of-Mind Capabilities
Alex Wilf, Sihyun Shawn Lee, Paul Pu Liang, Louis-Philippe Morency
2023-11-20
Artificial Intelligence · Computer Science
EnactToM: An Evolving Benchmark for Functional Theory of Mind in Embodied Agents
Gurusha Juneja, Dylan Lu, Saaket Agashe, Parth Diwane +6
2026-05-19