中文
相关论文

相关论文: AI with Alien Content and Alien Metasemantics

200 篇论文

Despite the fact that beliefs are mental states that cannot be directly observed, humans talk about each others' beliefs on a regular basis, often using rich compositional language to describe what others think and know. What explains this…

人工智能 · 计算机科学 2024-07-10 Lance Ying , Tan Zhi-Xuan , Lionel Wong , Vikash Mansinghka , Joshua Tenenbaum

While XAI focuses on providing AI explanations to humans, can the reverse - humans explaining their judgments to AI - foster richer, synergistic human-AI systems? This paper explores various forms of human inputs to AI and examines how…

计算机与社会 · 计算机科学 2025-03-07 Alan Dix , Tommaso Turchi , Ben Wilson , Anna Monreale , Matt Roach

For artificial intelligence to be beneficial to humans the behaviour of AI agents needs to be aligned with what humans want. In this paper we discuss some behavioural issues for language agents, arising from accidental misspecification by…

人工智能 · 计算机科学 2021-03-30 Zachary Kenton , Tom Everitt , Laura Weidinger , Iason Gabriel , Vladimir Mikulik , Geoffrey Irving

Explainable AI (XAI) is often promoted with the idea of helping users understand how machine learning models function and produce predictions. Still, most of these benefits are reserved for those with specialized domain knowledge, such as…

人工智能 · 计算机科学 2023-04-26 Chinasa T. Okolo

We examine recent research that asks whether current AI systems may be developing a capacity for "scheming" (covertly and strategically pursuing misaligned goals). We compare current research practices in this field to those adopted in the…

As AI advances in text generation, human trust in AI generated content remains constrained by biases that go beyond concerns of accuracy. This study explores how bias shapes the perception of AI versus human generated content. Through three…

计算与语言 · 计算机科学 2025-08-07 Tiffany Zhu , Iain Weissburg , Kexun Zhang , William Yang Wang

This paper explores the potential of a multidisciplinary approach to testing and aligning artificial intelligence (AI), specifically focusing on large language models (LLMs). Due to the rapid development and wide application of LLMs,…

计算机与社会 · 计算机科学 2025-01-07 Ljubisa Bojic , Matteo Cinelli , Dubravko Culibrk , Boris Delibasic

Go has long been considered as a testbed for artificial intelligence. By introducing certain quantum features, such as superposition and collapse of wavefunction, we experimentally demonstrate a quantum version of Go by using correlated…

Recent developments in artificial intelligence (AI) have permeated through an array of different immersive environments, including virtual, augmented, and mixed realities. AI brings a wealth of potential that centers on its ability to…

人机交互 · 计算机科学 2024-05-10 Wangfan Li , Rohit Mallick , Carlos Toxtli-Hernandez , Christopher Flathmann , Nathan J. McNeese

This chapter outlines the relation between artificial intelligence (AI) / machine learning (ML) algorithms and digital games. This relation is two-fold: on one hand, AI/ML researchers can generate large, in-the-wild datasets of human…

人工智能 · 计算机科学 2021-05-10 Kostas Karpouzis , George Tsatiris

Artificial intelligence commonly refers to the science and engineering of artificial systems that can carry out tasks generally associated with requiring aspects of human intelligence, such as playing games, translating languages, and…

人工智能 · 计算机科学 2025-02-11 Andreas Krause , Jonas Hübotter

Explainable AI (XAI) is frequently positioned as a technical problem of revealing the inner workings of an AI model. This position is affected by unexamined onto-epistemological assumptions: meaning is treated as immanent to the model, the…

人工智能 · 计算机科学 2026-01-26 Fabio Morreale , Joan Serrà , Yuki Mitsufuji

Generative AI systems have entered everyday academic, professional, and personal life with remarkable speed, yet most users encounter them as mysterious artifacts rather than intelligible systems. This chapter discusses large language…

计算机与社会 · 计算机科学 2026-04-21 John T. Behrens

Have you ever read a blog or social media post and suspected that it was written--at least in part--by artificial intelligence (AI)? While transparently acknowledging contributors to writing is generally valued, why some writers choose to…

人机交互 · 计算机科学 2025-05-28 Jingchao Fang , Mina Lee

Introspection is a foundational cognitive ability, but its mechanism is not well understood. Recent work has shown that AI models can introspect. We study the mechanism of this introspection. We first extensively replicate Lindsey (2025)'s…

人工智能 · 计算机科学 2026-04-08 Harvey Lederman , Kyle Mahowald

This paper explores educational interactions involving humans and artificial intelligences not as sequences of prompts and responses, but as a social process of conversation and exploration. In this conception, learners continually converse…

计算机与社会 · 计算机科学 2023-06-21 Mike Sharples

The proliferation of Artificial Intelligence (AI) systems exhibiting complex and seemingly agentive behaviours necessitates a critical philosophical examination of their agency, autonomy, and moral status. In this paper we undertake a…

计算机与社会 · 计算机科学 2026-02-03 Paul Formosa , Inês Hipólito , Thomas Montefiore

Despite many recent advancements in language modeling, state-of-the-art language models lack grounding in the real world and struggle with tasks involving complex reasoning. Meanwhile, advances in the symbolic reasoning capabilities of AI…

计算与语言 · 计算机科学 2022-12-19 Andrew Lee , David Wu , Emily Dinan , Mike Lewis

Metareasoning, a branch of AI, focuses on reasoning about reasons. It has the potential to enhance robots' decision-making processes in unexpected situations. However, the concept has largely been confined to theoretical discussions and…

机器人学 · 计算机科学 2025-05-07 Adrian Lendinez , Renxi Qiu , Lanfranco Zanzi , Dayou Li

When working with generative artificial intelligence (AI), users may see productivity gains, but the AI-generated content may not match their preferences exactly. To study this effect, we introduce a Bayesian framework in which…

人工智能 · 计算机科学 2025-07-08 Francisco Castro , Jian Gao , Sébastien Martin