English
Related papers

Related papers: Brittle Minds, Fixable Activations: Understanding …

200 papers

Large language models (LLMs) are capable of generating plausible explanations of how they arrived at an answer to a question. However, these explanations can misrepresent the model's "reasoning" process, i.e., they can be unfaithful. This,…

Computation and Language · Computer Science 2025-05-21 Katie Matton , Robert Osazuwa Ness , John Guttag , Emre Kıcıman

Theory of Mind (ToM)$\unicode{x2014}$the ability to reason about the mental states of other people$\unicode{x2014}$is a key element of our social intelligence. Yet, despite their ever more impressive performance, large-scale neural language…

Computation and Language · Computer Science 2023-06-02 Melanie Sclar , Sachin Kumar , Peter West , Alane Suhr , Yejin Choi , Yulia Tsvetkov

Can large multimodal models have a human-like ability for emotional and social reasoning, and if so, how does it work? Recent research has discovered emergent theory-of-mind (ToM) reasoning capabilities in large language models (LLMs). LLMs…

Computer Vision and Pattern Recognition · Computer Science 2025-09-16 Zhawnen Chen , Tianchun Wang , Yizhou Wang , Michal Kosinski , Xiang Zhang , Yun Fu , Sheng Li

While Large Language Models (LLMs) have demonstrated impressive accomplishments in both reasoning and planning, their abilities in multi-agent collaborations remains largely unexplored. This study evaluates LLM-based agents in a multi-agent…

Computation and Language · Computer Science 2024-06-28 Huao Li , Yu Quan Chong , Simon Stepputtis , Joseph Campbell , Dana Hughes , Michael Lewis , Katia Sycara

Do Large Language Models (LLMs) possess a Theory of Mind (ToM)? Research into this question has focused on evaluating LLMs against benchmarks and found success across a range of social tasks. However, these evaluations do not test for the…

Artificial Intelligence · Computer Science 2026-02-27 John Muchovej , Amanda Royka , Shane Lee , Julian Jara-Ettinger

Theory of Mind (ToM)-the ability to reason about the mental states of oneself and others-is a cornerstone of human social intelligence. As Large Language Models (LLMs) become ubiquitous in real-world applications, validating their capacity…

Computation and Language · Computer Science 2026-03-13 Ruirui Chen , Weifeng Jiang , Chengwei Qin , Cheston Tan

The ability to understand and predict the mental states of oneself and others, known as the Theory of Mind (ToM), is crucial for effective social scenarios. Although recent studies have evaluated ToM in Large Language Models (LLMs),…

Computation and Language · Computer Science 2025-05-27 Fangxu Yu , Lai Jiang , Shenyi Huang , Zhen Wu , Xinyu Dai

Large language models (LLMs) can be controlled at inference time through prompts (in-context learning) and internal activations (activation steering). Different accounts have been proposed to explain these methods, yet their common goal of…

Machine Learning · Computer Science 2026-03-13 Eric Bigelow , Daniel Wurgaft , YingQiao Wang , Noah Goodman , Tomer Ullman , Hidenori Tanaka , Ekdeep Singh Lubana

Social reasoning necessitates the capacity of theory of mind (ToM), the ability to contextualise and attribute mental states to others without having access to their internal cognitive structure. Recent machine learning approaches to ToM…

Artificial Intelligence · Computer Science 2023-01-18 Dung Nguyen , Phuoc Nguyen , Hung Le , Kien Do , Svetha Venkatesh , Truyen Tran

Theory of Mind (ToM), the ability to attribute beliefs, intentions, or mental states to others, is a crucial feature of human social interaction. In complex environments, where the human sensory system reaches its limits, behaviour is…

Neural and Evolutionary Computing · Computer Science 2024-07-26 Francesca Bianco , Silvia Rigato , Maria Laura Filippetti , Dimitri Ognibene

Large language models (LLMs) exhibit strikingly conflicting behaviors: they can appear steadfastly overconfident in their initial answers whilst at the same time being prone to excessive doubt when challenged. To investigate this apparent…

Existing Theory of Mind (ToM) benchmarks diverge from real-world scenarios in three aspects: 1) they assess a limited range of mental states such as beliefs, 2) false beliefs are not comprehensively explored, and 3) the diverse personality…

Computation and Language · Computer Science 2025-01-16 Kazutoshi Shinoda , Nobukatsu Hojo , Kyosuke Nishida , Saki Mizuno , Keita Suzuki , Ryo Masumura , Hiroaki Sugiyama , Kuniko Saito

Theory of Mind (ToM), the capacity to comprehend the mental states of distinct individuals, is essential for numerous practical applications. With the development of large language models (LLMs), there is a heated debate about whether they…

Computation and Language · Computer Science 2024-10-29 Xiaomeng Ma , Lingyu Gao , Qihui Xu

Rapid advancements in large language models (LLMs) have sparked the question whether these models possess some form of consciousness. To tackle this challenge, Butlin et al. (2023) introduced a list of indicators for consciousness in…

Computation and Language · Computer Science 2026-02-03 Noam Steinmetz Yalon , Ariel Goldstein , Liad Mudrik , Mor Geva

There is a growing literature on reasoning by large language models (LLMs), but the discussion on the uncertainty in their responses is still lacking. Our aim is to assess the extent of confidence that LLMs have in their answers and how it…

Computation and Language · Computer Science 2024-12-23 Yudi Pawitan , Chris Holmes

User modeling has traditionally relied on inferring preferences, traits, or intents from observable behaviour. While effective in many adaptive systems, this paradigm treats behaviour as the primary object of modeling and leaves…

Human-Computer Interaction · Computer Science 2026-05-12 Cristina Gena

Theory of Mind (ToM), the ability to understand people's minds based on their behavior, is key to developing socially intelligent agents. Current approaches to ToM reasoning either rely on prompting Large Language Models (LLMs), which are…

Artificial Intelligence · Computer Science 2026-01-15 Zhining Zhang , Chuanyang Jin , Mung Yao Jia , Shunchi Zhang , Tianmin Shu

We consider the problem of aligning a large language model (LLM) to model the preferences of a human population. Modeling the beliefs, preferences, and behaviors of a specific population can be useful for a variety of different…

Computation and Language · Computer Science 2024-04-01 Keiichi Namikoshi , Alex Filipowicz , David A. Shamma , Rumen Iliev , Candice L. Hogan , Nikos Arechiga

As machine learning becomes more widespread and is used in more critical applications, it's important to provide explanations for these models, to prevent unintended behavior. Unfortunately, many current interpretability methods struggle…

Computation and Language · Computer Science 2024-11-28 Andreas Madsen

Large language models (LLMs) are increasingly employed in information-seeking and decision-making tasks. Despite their broad utility, LLMs tend to generate information that conflicts with real-world facts, and their persuasive style can…

Computation and Language · Computer Science 2024-09-19 Arslan Chaudhry , Sridhar Thiagarajan , Dilan Gorur