中文
相关论文

相关论文: Safety is Non-Compositional: A Formal Framework fo…

200 篇论文

We prove a theorem stating that any semantics can be encoded as a compositional semantics, which means that, essentially, the standard definition of compositionality is formally vacuous. We then show that when compositional semantics is…

cmp-lg · 计算机科学 2008-02-03 Wlodek Zadrozny

As AI systems become more advanced, companies and regulators will make difficult decisions about whether it is safe to train and deploy them. To prepare for these decisions, we investigate how developers could make a 'safety case,' which is…

计算机与社会 · 计算机科学 2024-03-20 Joshua Clymer , Nick Gabrieli , David Krueger , Thomas Larsen

Studies of discrete languages emerging when neural agents communicate to solve a joint task often look for evidence of compositional structure. This stems for the expectation that such a structure would allow languages to be acquired faster…

计算与语言 · 计算机科学 2020-04-28 Eugene Kharitonov , Marco Baroni

Autonomous AI systems generate responsibility gaps: consequential actions that cannot be satisfactorily attributed to developers, operators, or users under existing legal frameworks. The prevailing subject-object dichotomy fails to…

计算机与社会 · 计算机科学 2026-05-14 Karsten Brensing

In this paper, we focus on the synthesis of secure timed systems which are modelled as timed automata. The security property that the system must satisfy is a non-interference property. Intuitively, non-interference ensures the absence of…

计算机科学中的逻辑 · 计算机科学 2012-07-23 Gilles Benattar , Franck Cassez , Didier Lime , Olivier H. Roux

The evolution of generative AI systems exposes the challenges of traditional legal and ethical frameworks built around consent. This chapter examines how the conventional notion of consent, while fundamental to data protection and privacy…

计算机与社会 · 计算机科学 2025-07-03 Giada Pistilli , Bruna Trevelin

Self-organization is a process where a stable pattern is formed by the cooperative behavior between parts of an initially disordered system without external control or influence. It has been introduced to multi-agent systems as an internal…

人工智能 · 计算机科学 2021-05-27 Jieting Luo , Beishui Liao , John-Jules Meyer

We present a framework to formally describe probabilistic system behavior and symbolically reason about it. In particular we aim at reasoning about possible failures and fault tolerance. We regard systems which are composed of different…

软件工程 · 计算机科学 2015-03-20 Jan Olaf Blech

The integration of neural networks into safety-critical systems has shown great potential in recent years. However, the challenge of effectively verifying the safety of Neural Network Controlled Systems (NNCS) persists. This paper…

计算机科学中的逻辑 · 计算机科学 2024-03-28 Yuhao Zhou , Stavros Tripakis

The efficiency of an AI system is contingent upon its ability to align with the specified requirements of a given task. How-ever, the inherent complexity of tasks often introduces the potential for harmful implications or adverse actions.…

计算机与社会 · 计算机科学 2023-12-08 Kamalakar Karlapalem

This work explores a collaborative method for ensuring safety in multi-agent formation control problems. We formulate a control barrier function (CBF) based safety filter control law for a generic distributed formation controller and extend…

机器人学 · 计算机科学 2024-10-08 Brooks A. Butler , Chi Ho Leung , Philip E. Paré

When the meaning of a phrase cannot be inferred from the individual meanings of its words (e.g., hot dog), that phrase is said to be non-compositional. Automatic compositionality detection in multi-word phrases is critical in any…

计算与语言 · 计算机科学 2019-03-21 Dongsheng Wang , Quichi Li , Lucas Chaves Lima , Jakob grue Simonsen , Christina Lioma

In this brief paper we introduce a robust-safety notion for differential inclusions, and we propose a general framework to certify such a notion in terms of barrier functions. While existing literature studied only what we designate by…

最优化与控制 · 数学 2023-05-08 Mohamed Maghenem , Masoumeh Ghanbarpour , Adnane Saoud

AI safety has emerged as a critical priority as these systems are increasingly deployed in real-world applications. We propose the first domain-agnostic AI safety ensuring framework that achieves strong safety guarantees while preserving…

人工智能 · 计算机科学 2025-10-07 Beomjun Kim , Kangyeon Kim , Sunwoo Kim , Yeonsang Shin , Heejin Ahn

Recent findings in multi-agent deep learning systems point towards the emergence of compositional languages. These claims are often made without exact analysis or testing of the language. In this work, we analyze the emergent language…

机器学习 · 计算机科学 2020-01-24 Bence Keresztury , Elia Bruni

Despite many recent advances, reactive synthesis is still not really a practical technique. The grand challenge is to scale from small transition systems, where synthesis performs well, to complex multi-component designs. Compositional…

计算机科学中的逻辑 · 计算机科学 2020-10-09 Bernd Finkbeiner , Noemi Passing

Because human preferences are too complex to codify, AIs operate with misspecified objectives. Optimizing such objectives often produces undesirable outcomes; this phenomenon is known as reward hacking. Such outcomes are not necessarily…

人工智能 · 计算机科学 2026-04-27 Henrik Marklund , Alex Infanger , Benjamin Van Roy

When AI systems are granted the agency to take impactful actions in the real world, there is an inherent risk that these systems behave in ways that are harmful. Typically, humans specify constraints on the AI system to prevent harmful…

人机交互 · 计算机科学 2022-11-09 Travis Mandel , Jahnu Best , Randall H. Tanaka , Hiram Temple , Chansen Haili , Kayla Schlectinger , Roy Szeto

Ensuring that AI agents behave safely and beneficially when interacting with other parties has emerged as one of the central challenges of modern AI safety. While mechanism design, as the theory of designing rules to align individual and…

计算机科学与博弈论 · 计算机科学 2026-05-12 Xuanqiang Angelo Huang , Charlie Tharas , Samuele Marro , Van Q. Truong , Bernhard Schölkopf , Emanuele La Malfa , Zhijing Jin

This paper argues that AI alignment is not merely difficult, but is founded on a fundamental logical contradiction. We first establish The Enumeration Paradox: we use machine learning precisely because we cannot enumerate all necessary…

人工智能 · 计算机科学 2025-06-26 Jasper Yao