English
Related papers

Related papers: Learning Theory of Mind via Dynamic Traits Attribu…

200 papers

With their recent development, large language models (LLMs) have been found to exhibit a certain level of Theory of Mind (ToM), a complex cognitive capacity that is related to our conscious mind and that allows us to infer another's beliefs…

Computation and Language · Computer Science 2023-09-06 Mohsen Jamali , Ziv M. Williams , Jing Cai

As the performance of larger, newer Large Language Models continues to improve for strategic Theory of Mind (ToM) tasks, the demand for these state-of-the-art models increases commensurately. However, their deployment is costly both in…

Computation and Language · Computer Science 2024-11-01 Nunzio Lore , Sepehr Ilami , Babak Heydari

We introduce StorySim, a programmable framework for synthetically generating stories to evaluate the theory of mind (ToM) and world modeling (WM) capabilities of large language models (LLMs). Unlike prior benchmarks that may suffer from…

Computation and Language · Computer Science 2026-04-28 Nathaniel Getachew , Abulhair Saparov

With the widespread application of Artificial Intelligence (AI) in human society, enabling AI to autonomously align with human values has become a pressing issue to ensure its sustainable development and benefit to humanity. One of the most…

Artificial Intelligence · Computer Science 2025-01-08 Haibo Tong , Enmeng Lu , Yinqian Sun , Zhengqiang Han , Chao Liu , Feifei Zhao , Yi Zeng

Theory of Mind is an essential ability of humans to infer the mental states of others. Here we provide a coherent summary of the potential, current progress, and problems of deep learning approaches to Theory of Mind. We highlight that many…

Machine Learning · Computer Science 2023-02-14 Jaan Aru , Aqeel Labash , Oriol Corcoll , Raul Vicente

Large language models (LLMs) are increasingly tested for a "Theory of Mind" (ToM) - the ability to attribute mental states to oneself and others. Yet most evaluations stop at explicit belief attribution in classical toy stories or stylized…

Computation and Language · Computer Science 2026-03-03 Yuling Gu , Oyvind Tafjord , Hyunwoo Kim , Jared Moore , Ronan Le Bras , Peter Clark , Yejin Choi

Accurate prediction of human behavior is essential for robust and safe human-AI collaboration. However, existing approaches for modeling people are often data-hungry and brittle because they either make unrealistic assumptions about…

Artificial Intelligence · Computer Science 2025-10-03 Kunal Jha , Aydan Yuenan Huang , Eric Ye , Natasha Jaques , Max Kleiman-Weiner

Recent advancements in large language models (LLMs) have demonstrated emergent capabilities in complex reasoning, largely spurred by rule-based Reinforcement Learning (RL) techniques applied during the post-training. This has raised the…

Machine Learning · Computer Science 2025-07-22 Sneheel Sarangi , Hanan Salam

Social learning is a powerful mechanism through which agents learn about the world from others. However, humans don't always choose to observe others, since social learning can carry time and cognitive resource costs. How do people balance…

Multiagent Systems · Computer Science 2025-07-15 Lance Ying , Ryan Truong , Joshua B. Tenenbaum , Samuel J. Gershman

Large Language Models (LLMs) have developed rapidly and are widely applied to both general-purpose and professional tasks to assist human users. However, they still struggle to comprehend and respond to the true user needs when intentions…

Computation and Language · Computer Science 2026-02-17 Minyuan Ruan , Ziyue Wang , Kaiming Liu , Yunghwei Lai , Peng Li , Yang Liu

Theory of Mind (ToM)-the ability to reason about the mental states of oneself and others-is a cornerstone of human social intelligence. As Large Language Models (LLMs) become ubiquitous in real-world applications, validating their capacity…

Computation and Language · Computer Science 2026-03-13 Ruirui Chen , Weifeng Jiang , Chengwei Qin , Cheston Tan

Do Large Language Models (LLMs) possess a Theory of Mind (ToM)? Research into this question has focused on evaluating LLMs against benchmarks and found success across a range of social tasks. However, these evaluations do not test for the…

Artificial Intelligence · Computer Science 2026-02-27 John Muchovej , Amanda Royka , Shane Lee , Julian Jara-Ettinger

Humans interacting with robots often form predictions of what the robot will do next. For instance, based on the recent behavior of an autonomous car, a nearby human driver might predict that the car is going to remain in the same lane. It…

Robotics · Computer Science 2025-03-04 Sagar Parekh , Lauren Bramblett , Nicola Bezzo , Dylan P. Losey

Evaluating the theory of mind (ToM) capabilities of language models (LMs) has recently received a great deal of attention. However, many existing benchmarks rely on synthetic data, which risks misaligning the resulting experiments with…

Computation and Language · Computer Science 2024-06-11 Adil Soubki , John Murzaku , Arash Yousefi Jordehi , Peter Zeng , Magdalena Markowska , Seyed Abolghasem Mirroshandel , Owen Rambow

In recent years, the role of artificially intelligent (AI) agents has evolved from being basic tools to socially intelligent agents working alongside humans towards common goals. In such scenarios, the ability to predict future behavior by…

Machine Learning · Computer Science 2022-11-17 Chinmai Basavaraj , Adarsh Pyarelal , Evan Carter

Autonomous transportation systems such as road vehicles or vessels require the consideration of the static and dynamic environment to dislocate without collision. Anticipating the behavior of an agent in a given situation is required to…

Machine Learning · Computer Science 2024-06-06 Kathrin Donandt , Dirk Söffker

Many social sciences such as psychology and economics try to learn the behaviour of complex agents such as humans, organisations and countries. The current statistical methods used for learning this behaviour try to infer generally valid…

Artificial Intelligence · Computer Science 2021-03-08 Benedikt T. Kleppmann

Modern Large Language Models (LLMs) exhibit impressive zero-shot and few-shot generalization capabilities across complex natural language tasks, enabling their widespread use as virtual assistants for diverse applications such as…

Computation and Language · Computer Science 2025-06-19 Arjun Vaithilingam Sudhakar

The ability to represent oneself and others as agents with knowledge, intentions, and belief states that guide their behavior - Theory of Mind - is a human universal that enables us to navigate - and manipulate - the social world. It is…

Machine Learning · Computer Science 2026-05-12 Christopher Ackerman

Large language models (LLMs) are transforming human-computer interaction and conceptions of artificial intelligence (AI) with their impressive capacities for conversing and reasoning in natural language. There is growing interest in whether…

Human-Computer Interaction · Computer Science 2024-05-15 Winnie Street