中文
相关论文

相关论文: Model Routing as a Trust Problem: Route Receipts f…

200 篇论文

Transparency and security are both central to Responsible AI, but they may conflict in adversarial settings. We investigate the strategic effect of transparency for agents through the lens of transferable adversarial example attacks. In…

机器学习 · 计算机科学 2025-11-18 Lucas Fenaux , Christopher Srinivasa , Florian Kerschbaum

The ability to steer AI behavior is crucial to preventing its long term dangerous and catastrophic potential. Representation Engineering (RepE) has emerged as a novel, powerful method to steer internal model behaviors, such as "honesty", at…

机器学习 · 计算机科学 2024-10-10 Akshat Kannan

Adaptive Computing is an application-agnostic outer loop framework to strategically deploy simulations and experiments to guide decision making for scale-up analysis. Resources are allocated over successive batches, which makes the…

Companies dealing with Artificial Intelligence (AI) models in Autonomous Systems (AS) face several problems, such as users' lack of trust in adverse or unknown conditions, gaps between software engineering and AI model development, and…

In the last years, AI systems, in particular neural networks, have seen a tremendous increase in performance, and they are now used in a broad range of applications. Unlike classical symbolic AI systems, neural networks are trained using…

计算机视觉与模式识别 · 计算机科学 2021-08-16 Christian Berghoff , Pavol Bielik , Matthias Neu , Petar Tsankov , Arndt von Twickel

Dominant approaches, e.g. the EU's "Trustworthy AI framework", treat trust as a property that can be designed for, evaluated, and governed according to normative and technical criteria. They do not address how trust is subjectively…

计算机与社会 · 计算机科学 2026-02-02 Lameck Mbangula Amugongo , Tutaleni Asino , Nicola J Bidwell

Artificial intelligence (AI) and Machine Learning (ML) have moved from research and pilot projects into everyday business operations, with generative AI accelerating adoption across processes, products, and services. This paper introduces…

计算机与社会 · 计算机科学 2026-02-17 Stephan Sandfuchs , Diako Farooghi , Janis Mohr , Sarah Grewe , Markus Lemmen , Jörg Frochte

While there is significant interest in using generative AI tools as general-purpose models for specific ML applications, discriminative models are much more widely deployed currently. One of the key shortcomings of these discriminative AI…

人工智能 · 计算机科学 2023-12-13 Son The Nguyen , Theja Tulabandhula , Mary Beth Watson-Manheim

An assurance case is a structured argument, typically produced by safety engineers, to communicate confidence that a critical or complex system, such as an aircraft, will be acceptably safe within its intended context. Assurance cases often…

计算机与社会 · 计算机科学 2023-06-07 Zoe Porter , Ibrahim Habli , John McDermid , Marten Kaas

One of the most relevant challenges regarding on-demand ridepooling relates to the spatial imbalances of the demand, which induce a mismatch between the position of the vehicles and the origins of the emerging requests. Most ridepooling…

系统与控制 · 电气工程与系统科学 2021-06-29 Andres Fielbaum , Maximilian Kronmuller , Javier Alonso-Mora

The aim of this paper is to demonstrate the feasibility of authenticated throughput-efficient routing in an unreliable and dynamically changing synchronous network in which the majority of malicious insiders try to destroy and alter…

密码学与安全 · 计算机科学 2009-01-04 Yair Amir , Paul Bunn , Rafail Ostrovksy

Place recognition, the ability to identify previously visited locations, is critical for both biological navigation and autonomous systems. This review synthesizes findings from robotic systems, animal studies, and human research to explore…

机器人学 · 计算机科学 2025-11-19 Michael Milford , Tobias Fischer

A key strategy for balancing performance and cost in modern machine learning systems is to dynamically route queries to either a low-cost model or a more expensive oracle (such as a large pretrained model or human expert), an approach known…

机器学习 · 计算机科学 2026-05-11 Charlotte Peale , Siddartha Devic , Parikshit Gopalan , Udi Wieder , Aravind Gollakota

In this paper, we identify and characterize the emerging area of representation engineering (RepE), an approach to enhancing the transparency of AI systems that draws on insights from cognitive neuroscience. RepE places population-level…

In recent years, Artificial intelligence products and services have been offered potential users as pilots. The acceptance intention towards artificial intelligence is greatly influenced by the experience with current AI products and…

计算机与社会 · 计算机科学 2023-06-27 Minsang Yi , Hanbyul Choi

We propose and analyze a recipient-anonymous stochastic routing model to study a fundamental trade-off between anonymity and routing delay. An agent wants to quickly reach a goal vertex in a network through a sequence of routing actions,…

计算机科学与博弈论 · 计算机科学 2021-01-01 Mine Su Erturk , Kuang Xu

The development of Artificial Intelligence (AI), including AI in Science (AIS), should be done following the principles of responsible AI. Progress in responsible AI is often quantified through evaluation metrics, yet there has been less…

计算机与社会 · 计算机科学 2025-10-31 Theresia Veronika Rampisela , Maria Maistro , Tuukka Ruotsalo , Christina Lioma

Artificial intelligence-driven adaptive learning systems are reshaping education through data-driven adaptation of learning experiences. Yet many of these systems lack transparency, offering limited insight into how decisions are made. Most…

人工智能 · 计算机科学 2025-08-04 Maryam Mosleh , Marie Devlin , Ellis Solaiman

Robot behavior is often validated through simulation-based testing, yet the replicability of such campaigns depends critically on transparent documentation of how tests are configured, executed, and post-processed. We argue that data…

机器人学 · 计算机科学 2026-05-29 Argentina Ortega , Samuel Wiest , Frederik Pasch , Nico Hochgeschwender

The development of privacy-enhancing technologies has made immense progress in reducing trade-offs between privacy and performance in data exchange and analysis. Similar tools for structured transparency could be useful for AI governance by…

人工智能 · 计算机科学 2023-03-22 Emma Bluemke , Tantum Collins , Ben Garfinkel , Andrew Trask