中文
相关论文

相关论文: Functional trustworthiness of AI systems by statis…

200 篇论文

Deployed AI systems often do not work. They can be constructed haphazardly, deployed indiscriminately, and promoted deceptively. However, despite this reality, scholars, the press, and policymakers pay too little attention to functionality.…

机器学习 · 计算机科学 2022-07-05 Inioluwa Deborah Raji , I. Elizabeth Kumar , Aaron Horowitz , Andrew D. Selbst

Trustworthy Artificial Intelligence (TAI) integrates ethics that align with human values, looking at their influence on AI behaviour and decision-making. Primarily dependent on self-assessment, TAI evaluation aims to ensure ethical…

计算机与社会 · 计算机科学 2024-09-13 Louise McCormack , Malika Bendechache

Operationalizing the EU AI Act requires clear technical documentation to ensure AI systems are transparent, traceable, and accountable. Existing documentation templates for AI systems do not fully cover the entire AI lifecycle while meeting…

机器学习 · 计算机科学 2025-08-13 Laura Lucaj , Alex Loosley , Hakan Jonsson , Urs Gasser , Patrick van der Smagt

The EU has become one of the vanguards in regulating the digital age. A particularly important regulation in the Artificial Intelligence (AI) domain is the EU AI Act, which entered into force in 2024. The AI Act specifies -- due to a…

计算机与社会 · 计算机科学 2025-06-05 Alina Wernick , Kristof Meding

Building trustworthy autonomous systems is challenging for many reasons beyond simply trying to engineer agents that 'always do the right thing.' There is a broader context that is often not considered within AI and HRI: that the problem of…

计算机与社会 · 计算机科学 2020-10-15 Alexis Morris , Hallie Siegel , Jonathan Kelly

Significant investment and development have gone into integrating Artificial Intelligence (AI) in medical and healthcare applications, leading to advanced control systems in medical technology. However, the opacity of AI systems raises…

人工智能 · 计算机科学 2024-10-21 Francesco Sovrano , Michael Lognoul , Giulia Vilone

Public attention towards explainability of artificial intelligence (AI) systems has been rising in recent years to offer methodologies for human oversight. This has translated into the proliferation of research outputs, such as from…

计算机与社会 · 计算机科学 2023-04-25 Luca Nannini , Agathe Balayn , Adam Leon Smith

The number and importance of AI-based systems in all domains is growing. With the pervasive use and the dependence on AI-based systems, the quality of these systems becomes essential for their practical usage. However, quality assurance for…

软件工程 · 计算机科学 2023-08-02 Michael Felderer , Rudolf Ramler

With growing concerns regarding bias and discrimination in predictive models, the AI community has increasingly focused on assessing AI system trustworthiness. Conventionally, trustworthy AI literature relies on the probabilistic framework…

机器学习 · 统计学 2024-01-05 Ritwik Vashistha , Arya Farahi

This paper presents a conceptual and operational framework for developing and operating safe and trustworthy AI agents based on a Three-Pillar Model grounded in transparency, accountability, and trustworthiness. Building on prior work in…

计算机与社会 · 计算机科学 2026-01-13 Edward C. Cheng , Jeshua Cheng , Alice Siu

Many decision-making processes have begun to incorporate an AI element, including prison sentence recommendations, college admissions, hiring, and mortgage approval. In all of these cases, AI models are being trained to help human decision…

计算机与社会 · 计算机科学 2019-12-06 Maryam Ashoori , Justin D. Weisz

AI agents -- systems that can independently take actions to pursue complex goals with only limited human oversight -- have entered the mainstream. These systems are now being widely used to produce software, conduct business activities, and…

计算机与社会 · 计算机科学 2026-03-25 Kathrin Gardhouse , Amin Oueslati , Noam Kolt

With ubiquitous exposure of AI systems today, we believe AI development requires crucial considerations to be deemed trustworthy. While the potential of AI systems is bountiful, though, is still unknown-as are their risks. In this work, we…

计算机与社会 · 计算机科学 2023-09-19 Jamell Dacon

As artificial intelligence (AI) and robotics increasingly permeate society, ensuring the ethical behavior of these systems has become paramount. This paper contends that transparency in AI decision-making processes is fundamental to…

计算机与社会 · 计算机科学 2025-08-11 Ahmad Farooq , Kamran Iqbal

What makes safety claims about general purpose AI systems such as large language models trustworthy? We show that rather than the capabilities of security tools such as alignment and red teaming procedures, it is security practices based on…

密码学与安全 · 计算机科学 2025-07-30 Petr Spelda , Vit Stritecky

The European Union Artificial Intelligence (EU AI) Act, which explicitly references fundamental rights and ethical principles, is a comprehensive regulatory framework for governing Artificial Intelligence (AI) systems. This study examines…

计算机与社会 · 计算机科学 2026-05-11 Mehmet Murat Albayrakoglu , Mehmet Nafiz Aydin

In this work-in-progress, we investigate the certification of AI systems, focusing on the practical application and limitations of existing certification catalogues in the light of the AI Act by attempting to certify a publicly available AI…

计算机与社会 · 计算机科学 2025-02-19 Gregor Autischer , Kerstin Waxnegger , Dominik Kowald

The Artificial intelligence in critical sectors-healthcare, finance, and public safety-has made system integrity paramount for maintaining societal trust. Current verification methods for AI systems lack comprehensive lifecycle assurance,…

密码学与安全 · 计算机科学 2024-11-04 Mahesh Vaijainthymala Krishnamoorthy

Safety has become the central value around which dominant AI governance efforts are being shaped. Recently, this culminated in the publication of the International AI Safety Report, written by 96 experts of which 30 nominated by the…

计算机与社会 · 计算机科学 2025-03-10 Roel Dobbe

Artificial intelligence (AI) governance is the body of standards and practices used to ensure that AI systems are deployed responsibly. Current AI governance approaches consist mainly of manual review and documentation processes. While such…

计算机与社会 · 计算机科学 2023-02-17 Sean McGregor , Jesse Hostetler