中文
相关论文

相关论文: Wanting to Be Understood Explains the Meta-Problem…

200 篇论文

Most research on the interpretability of machine learning systems focuses on the development of a more rigorous notion of interpretability. I suggest that a better understanding of the deficiencies of the intuitive notion of…

机器学习 · 统计学 2017-12-08 Fabian Offert

Establishing a communication system is hard because the intended meaning of a signal is unknown to its receiver when first produced, and the signaller also has no idea how that signal will be interpreted. Most theoretical accounts of the…

计算与语言 · 计算机科学 2025-11-13 Richard A. Blythe , Casimir Fisch

The increasing reliance on digital information necessitates advancements in conversational search systems, particularly in terms of information transparency. While prior research in conversational information-seeking has concentrated on…

信息检索 · 计算机科学 2024-05-07 Weronika Łajewska , Damiano Spina , Johanne Trippas , Krisztian Balog

Interpretability aims to explain the behavior of deep neural networks. Despite rapid growth, there is mounting concern that much of this work has not translated into practical impact, raising questions about its relevance and utility. This…

A model of knowledge representation is described in which propositional facts and the relationships among them can be supported by other facts. The set of knowledge which can be supported is called the set of cognitive units, each having…

人工智能 · 计算机科学 2013-04-12 A. Julian Craddock , Roger A. Browse

Supervised machine learning models boast remarkable predictive capabilities. But can you trust your model? Will it work in deployment? What else can it tell you about the world? We want models to be not only good, but interpretable. And yet…

机器学习 · 计算机科学 2017-03-07 Zachary C. Lipton

As machine learning is increasingly deployed in high-stakes contexts affecting people's livelihoods, there have been growing calls to open the black box and to make machine learning algorithms more explainable. Providing useful explanations…

计算机与社会 · 计算机科学 2020-07-13 Umang Bhatt , McKane Andrus , Adrian Weller , Alice Xiang

Attitudes about artificial intelligence and machine learning are recent victims of endemic misunderstanding; given our increasing reliance on these technologies, the need for widespread understanding and confidence in their use is…

图形学 · 计算机科学 2026-05-04 Bokang Wang , Yingxuan Liao , Leah Lee , Jack Wesson , Anlan Yang , Ruizi Wang , Yigang Wen

Purpose and meaning are necessary concepts for understanding mind and culture, but appear to be absent from the physical world and are not part of the explanatory framework of the natural sciences. Understanding how meaning (in the broad…

种群与进化 · 定量生物学 2015-02-04 J. H. van Hateren

We seek general principles of the structure of the cellular collective activity associated with conscious awareness. Can we obtain evidence for features of the optimal brain organization that allows for adequate processing of stimuli and…

神经元与认知 · 定量生物学 2017-12-27 D. M. Mateos , R. Wennberg , R. Guevara , J. L. Perez Velazquez

Safety and assurance cases risk becoming detached from the understanding needed for responsible engineering and governance decisions. More broadly, the production and evaluation of critical socio-technical systems increasingly face an…

软件工程 · 计算机科学 2026-04-08 Robin Bloomfield

Mechanistic interpretability (MI) aims to explain how neural networks work by uncovering their underlying mechanisms. As the field grows in influence, it is increasingly important to examine not just models themselves, but the assumptions,…

The search for a scientific theory of consciousness should result in theories that are falsifiable. However, here we show that falsification is especially problematic for theories of consciousness. We formally describe the standard…

神经元与认知 · 定量生物学 2021-04-29 Johannes Kleiner , Erik Hoel

Awareness and self-awareness are two different notions related to knowing the environment and itself. In a general context, the mechanism of self-awareness belongs to a class of co-called "self-issues" (self-* or self-star):…

机器人学 · 计算机科学 2011-11-23 Serge Kernbach

To build agents that can collaborate effectively with others, recent research has trained artificial agents to communicate with each other in Lewis-style referential games. However, this often leads to successful but uninterpretable…

计算与语言 · 计算机科学 2022-01-11 Jesse Mu , Noah Goodman

This paper argues that explainability is only one facet of a broader ideal that shapes our expectations towards artificial intelligence (AI). Fundamentally, the issue is to what extent AI exhibits systematicity--not merely in being…

人工智能 · 计算机科学 2025-07-31 Matthieu Queloz

In recent years, promising mathematical models have been suggested which aim to describe conscious experience and its relation to the physical domain. Whereas the axioms and metaphysical ideas of these theories have been carefully…

神经元与认知 · 定量生物学 2020-07-15 Johannes Kleiner

Explainability and comprehensibility of AI are important requirements for intelligent systems deployed in real-world domains. Users want and frequently need to understand how decisions impacting them are made. Similarly it is important to…

计算机与社会 · 计算机科学 2019-07-10 Roman V. Yampolskiy

Language understanding entails not just extracting the surface-level meaning of the linguistic input, but constructing rich mental models of the situation it describes. Here we propose that because processing within the brain's core…

计算与语言 · 计算机科学 2025-11-26 Colton Casto , Anna Ivanova , Evelina Fedorenko , Nancy Kanwisher

The web does not only enable new forms of science, it also creates new possibilities to study science and new digital scholarship. This paper brings together multiple perspectives: from individual researchers seeking the best options to…

数字图书馆 · 计算机科学 2013-04-23 Christophe Guéret , Tamy Chambers , Linda Reijnhoudt , Frank van der Most , Andrea Scharnhorst