中文
相关论文

相关论文: True Lies

200 篇论文

Proper scoring rules elicit truth-telling when making predictions, or otherwise revealing information. However, when multiple predictions are made of the same event, telling the truth is in general no longer optimal, as agents are motivated…

计算机科学与博弈论 · 计算机科学 2017-07-04 Amir Ban

We propose a multi-agent epistemic logic of asynchronous announcements, where truthful announcements are publicly sent but individually received by agents, and in the order in which they were sent. Additional to epistemic modalities the…

人工智能 · 计算机科学 2021-01-19 Philippe Balbiani , Hans van Ditmarsch , Saúl Fernández González

We analyze the informal notion of truth and conclude that it can be formalized in essentially two distinct ways: constructively, in terms of provability, or classically, as a hierarchy of concepts which satisfy Tarski's biconditional in…

历史与综述 · 数学 2011-12-30 Nik Weaver

There has been considerable recent interest in explainability in AI, especially with black-box machine learning models. As correctly observed by the planning community, when the application at hand is not a single-shot decision or…

人工智能 · 计算机科学 2025-02-14 Vaishak Belle

Large Language Models (LLMs) have impressive capabilities, but are prone to outputting falsehoods. Recent work has developed techniques for inferring whether a LLM is telling the truth by training probes on the LLM's internal activations.…

人工智能 · 计算机科学 2024-08-20 Samuel Marks , Max Tegmark

We propose a dynamic logic of lying, wherein a 'lie that phi' (where phi is a formula in the logic) is an action in the sense of dynamic modal logic, that is interpreted as a state transformer relative to the formula phi. The states that…

人工智能 · 计算机科学 2012-03-22 Hans van Ditmarsch

We formalise the notion of an anonymous public announcement in the tradition of public announcement logic. Such announcements can be seen as in-between a public announcement from ``the outside" (an announcement of $\phi$) and a public…

计算机科学中的逻辑 · 计算机科学 2025-04-22 Thomas Ågotnes , Rustam Galimullin , Ken Satoh , Satoshi Tojo

Large language models (LLMs) are increasingly deployed in high-stakes settings where good decisions require forming beliefs over the probability of unknown outcomes. However, it is unclear whether LLMs act as if they hold coherent beliefs…

人工智能 · 计算机科学 2026-05-12 Khurram Yamin , Jingjing Tang , Santiago Cortes-Gomez , Amit Sharma , Eric Horvitz , Bryan Wilder

Recent probing studies reveal that large language models exhibit linear subspaces that separate true from false statements, yet the mechanism behind their emergence is unclear. We introduce a transparent, one-layer transformer toy model…

计算与语言 · 计算机科学 2025-10-20 Shauli Ravfogel , Gilad Yehudai , Tal Linzen , Joan Bruna , Alberto Bietti

Every countable language which conforms to classical logic is shown to have an extension which conforms to classical logic, and has a definitional theory of truth. That extension has a semantical theory of truth, if every sentence of the…

逻辑 · 数学 2020-02-04 Seppo Heikkilä

We formulate the problem of fake news detection using distributed fact-checkers (agents) with unknown reliability. The stream of news/statements is modeled as an independent and identically distributed binary source (to represent true and…

最优化与控制 · 数学 2025-03-05 Ashwin Verma , Soheil Mohajer , Behrouz Touri

Our approach is basically a coherence approach, but we avoid the well-known pitfalls of coherence theories of truth. Consistency is replaced by reliability, which expresses support and attack, and, in principle, every theory (or agent,…

人工智能 · 计算机科学 2018-04-03 Karl Schlechta

Text-based misinformation permeates online discourses, yet evidence of people's ability to discern truth from such deceptive textual content is scarce. We analyze a novel TV game show data where conversations in a high-stake environment…

计算与语言 · 计算机科学 2024-04-09 Sanchaita Hazra , Bodhisattwa Prasad Majumder

While Large Language Models (LLMs) have shown exceptional performance in various tasks, one of their most prominent drawbacks is generating inaccurate or false information with a confident tone. In this paper, we provide evidence that the…

计算与语言 · 计算机科学 2023-10-18 Amos Azaria , Tom Mitchell

We propose new definitions of (causal) explanation, using structural equations to model counterfactuals. The definition is based on the notion of actual cause, as defined and motivated in a companion paper. Essentially, an explanation is a…

人工智能 · 计算机科学 2007-05-23 Joseph Y. Halpern , Judea Pearl

Ordinary and transfinite recursion and induction and ZF set theory are used to construct from a fully interpreted object language and from an extra formula a new language. It is fully interpreted under a suitably defined interpretation.…

逻辑 · 数学 2017-12-15 Seppo Heikkilä

Arbitrary public announcement logic (APAL) reasons about how the knowledge of a set of agents changes after true public announcements and after arbitrary announcements of true epistemic formulas. We consider a variant of arbitrary public…

计算机科学中的逻辑 · 计算机科学 2021-11-30 Hans van Ditmarsch , Tim French , James Hales

Large Language Models (LLMs) have revolutionised natural language processing, exhibiting impressive human-like capabilities. In particular, LLMs are capable of "lying", knowingly outputting false statements. Hence, it is of interest and…

计算与语言 · 计算机科学 2024-10-22 Lennart Bürger , Fred A. Hamprecht , Boaz Nadler

The paper proposes a new type of negation in multi-valued logics, providing a different way to answer the following question: what does it mean that some object language formula does not have a given truth-value. Along the way, the paper…

计算机科学中的逻辑 · 计算机科学 2022-04-01 Nissim Francez

Knowing the truth is rarely enough -- we also seek out reasons why the fact is true. While much is known about how we explain contingent truths, we understand less about how we explain facts, such as those in mathematics, that are true as a…

人工智能 · 计算机科学 2026-01-08 Gülce Kardeş , Simon DeDeo
‹ 上一页 1 2 3 10 下一页 ›