English
Related papers

Related papers: A Devastating Example for the Halfer Rule

200 papers

Whereas cognitive models of learning often assume direct experience with both the features of an event and with a true label or outcome, much of everyday learning arises from hearing the opinions of others, without direct access to either…

Artificial Intelligence · Computer Science 2025-12-05 Yun-Shiuan Chuang , Jerry Zhu , Timothy T. Rogers

Recently, LLM-based agents have become increasingly popular across many applications, including complex sequential decision-making problems. However, they inherit the tendency of LLMs to hallucinate, leading to incorrect decisions. In…

The classical conception of falsification presents scientific theories as entities that are decisively refuted when their predictions fail. This picture has long been challenged by both philosophical analysis and scientific practice, yet…

Other Statistics · Statistics 2025-12-09 Tommaso Costa

Large Language Models (LLMs) have shown impressive capabilities but still suffer from the issue of hallucinations. A significant type of this issue is the false premise hallucination, which we define as the phenomenon when LLMs generate…

Computation and Language · Computer Science 2024-03-01 Hongbang Yuan , Pengfei Cao , Zhuoran Jin , Yubo Chen , Daojian Zeng , Kang Liu , Jun Zhao

Large language models are increasingly deployed as autonomous agents in multi-agent settings where they communicate intentions and take consequential actions with limited human oversight. A critical safety question is whether agents that…

Computers and Society · Computer Science 2026-04-07 Jerick Shi , Terry Jingcheng Zhang , Zhijing Jin , Vincent Conitzer

For Large Language Models (LLMs) to be reliably deployed, models must effectively know when not to answer: abstain. Reasoning models, in particular, have gained attention for impressive performance on complex tasks. However, reasoning…

Artificial Intelligence · Computer Science 2026-04-03 Abinitha Gourabathina , Inkit Padhi , Manish Nagireddy , Subhajit Chaudhury , Prasanna Sattigeri

Reflexion-style agents rely on self-generated reflections as memory, implicitly assuming that agents can accurately diagnose their own failures.We show that this assumption can fail systematically: across ALFWorld and HumanEval, agents…

Machine Learning · Computer Science 2026-05-29 Prakhar Dixit , Sadia Kamal , Tim Oates

Humans do not just find mistakes after the fact -- we often catch them mid-stream because 'reflection' is tied to the goal and its constraints. Today's large language models produce reasoning tokens and 'reflective' text, but is it…

Artificial Intelligence · Computer Science 2025-10-24 Sion Weatherhead , Flora Salim , Aaron Belbasis

Stemming from de Finetti's work on finitely additive coherent probabilities, the paradigm of coherence has been applied to many uncertainty calculi in order to remove structural restrictions on the domain of the assessment. Three possible…

Probability · Mathematics 2021-06-30 Davide Petturiti , Barbara Vantaggi

We consider the problem of distributed inference where agents in a network observe a stream of private signals generated by an unknown state, and aim to uniquely identify this state from a finite set of hypotheses. We focus on scenarios…

Systems and Control · Electrical Eng. & Systems 2021-09-01 Aritra Mitra , John A. Richards , Saurabh Bagchi , Shreyas Sundaram

Like a criminal under investigation, Large Language Models (LLMs) might pretend to be aligned while evaluated and misbehave when they have a good opportunity. Can current interpretability methods catch these 'alignment fakers?' To answer…

Computation and Language · Computer Science 2024-05-14 Joshua Clymer , Caden Juang , Severin Field

There are things we know, things we know we don't know, and then there are things we don't know we don't know. In this paper we address the latter two issues in a Bayesian framework, introducing the notion of doubt to quantify the degree of…

Data Analysis, Statistics and Probability · Physics 2008-11-18 Glenn D Starkman , Roberto Trotta , Pascal M Vaudrevange

We show that continual pretraining on plausible misinformation can overwrite specific factual knowledge in large language models without degrading overall performance. Unlike prior poisoning work under static pretraining, we study repeated…

Machine Learning · Computer Science 2026-02-09 Svetlana Churina , Niranjan Chebrolu , Kokil Jaidka

Adequacy for estimation between an inferential method and a model can be de{\ldots}ned through two main requirements: {\ldots}rstly the inferential tool should de{\ldots}ne a well posed problem when applied to the model; secondly the…

Statistics Theory · Mathematics 2025-07-30 Michel Broniatowski , Justin Moutsouka

We develop a model of opinion dynamics where agents in a social network seek to learn a ground truth among a set of competing hypotheses. Agents in the network form private beliefs about such hypotheses by aggregating their neighbors'…

Physics and Society · Physics 2023-07-26 Diana Riazi , Giacomo Livan

Emergence of cooperation in self-centered individuals has been a major puzzle in the study of evolutionary ethics. Reciprocal altruism is one of explanations put forward and prisoner's dilemma has been a paradigm in this context. Emergence…

Statistical Mechanics · Physics 2009-10-07 M Ali Saif , P. M. Gade

Prior work on large language model (LLM) hallucinations has associated them with model uncertainty or inaccurate knowledge. In this work, we define and investigate a distinct type of hallucination, where a model can consistently answer a…

Computation and Language · Computer Science 2025-08-26 Adi Simhi , Itay Itzhak , Fazl Barez , Gabriel Stanovsky , Yonatan Belinkov

Large language models (LLMs) increasingly help people solve problems, from debugging code to repairing machinery. This process requires generating plausible hypotheses from partial descriptions, then updating them as more information…

Machine Learning · Computer Science 2026-05-08 Hua-Dong Xiong

Despite its centrality in the philosophy of cognitive science, there has been little prior philosophical work engaging with the notion of representation in contemporary NLP practice. This paper attempts to fill that lacuna: drawing on ideas…

Computation and Language · Computer Science 2023-11-21 Jacqueline Harding

Why do people who disagree about one subject tend to disagree about other subjects as well? In this paper, we introduce a model to explore this phenomenon of "epistemic factionization". Agents attempt to discover the truth about multiple…

Social and Information Networks · Computer Science 2020-05-04 James Owen Weatherall , Cailin O'Connor