中文
相关论文

相关论文: Distrust in (X)AI -- Measurement Artifact or Disti…

200 篇论文

Knowing when a classifier's prediction can be trusted is useful in many applications and critical for safely using AI. While the bulk of the effort in machine learning research has been towards improving classifier performance,…

机器学习 · 统计学 2018-10-30 Heinrich Jiang , Been Kim , Melody Y. Guan , Maya Gupta

Synthetic images, audio, and video can now be generated and edited by Artificial Intelligence (AI). In particular, the malicious use of synthetic data has raised concerns about potential harms to cybersecurity, personal privacy, and public…

人机交互 · 计算机科学 2025-08-05 Yingfan Zhou , Ester Chen , Manasa Pisipati , Aiping Xiong , Sarah Rajtmajer

What makes safety claims about general purpose AI systems such as large language models trustworthy? We show that rather than the capabilities of security tools such as alignment and red teaming procedures, it is security practices based on…

密码学与安全 · 计算机科学 2025-07-30 Petr Spelda , Vit Stritecky

Generative AI (GenAI) has brought opportunities and challenges for higher education as it integrates into teaching and learning environments. As instructors navigate this new landscape, understanding their engagement with and attitudes…

人机交互 · 计算机科学 2025-02-11 Wenhan Lyu , Shuang Zhang , Tingting , Chung , Yifan Sun , Yixuan Zhang

Building trustworthy AI systems for mental health support is a shared priority across stakeholders from multiple disciplines. However, "trustworthy" remains loosely defined and inconsistently operationalized. AI research often focuses on…

EXplainable Artificial Intelligence (XAI) is a vibrant research topic in the artificial intelligence community, with growing interest across methods and domains. Much has been written about the subject, yet XAI still lacks shared…

人工智能 · 计算机科学 2023-06-16 Matteo Rizzo , Alberto Veneri , Andrea Albarelli , Claudio Lucchese , Marco Nobile , Cristina Conati

Artificial intelligence (AI) is expected to revolutionize the practice of medicine. Recent advancements in the field of deep learning have demonstrated success in a variety of clinical tasks: detecting diabetic retinopathy from images,…

机器学习 · 计算机科学 2025-04-04 Joshua Hatherley

This paper addresses the critical challenge of building consumer trust in AI-powered customer engagement by emphasising the necessity for transparency and accountability. Despite the potential of AI to revolutionise business operations and…

计算机与社会 · 计算机科学 2024-10-04 Tara DeZao

Explainable AI (XAI) builds trust in complex systems through model attribution methods that reveal the decision rationale. However, due to the absence of a unified optimal explanation, existing XAI methods lack a ground truth for objective…

机器学习 · 计算机科学 2025-08-06 Yuhan Guo , Lizhong Ding , Shihan Jia , Yanyu Ren , Pengqi Li , Jiarun Fu , Changsheng Li , Ye yuan , Guoren Wang

Building reliable deception detectors for AI systems -- methods that could predict when an AI system is being strategically deceptive without necessarily requiring behavioural evidence -- would be valuable in mitigating risks from advanced…

机器学习 · 计算机科学 2025-12-17 Lewis Smith , Bilal Chughtai , Neel Nanda

The rationale behind a deep learning model's output is often difficult to understand by humans. EXplainable AI (XAI) aims at solving this by developing methods that improve interpretability and explainability of machine learning models.…

人工智能 · 计算机科学 2023-08-08 Rafaël Brandt , Daan Raatjens , Georgi Gaydadjiev

The accelerating adoption of large language models, retrieval-augmented generation pipelines, and multi-agent AI workflows has created a structural governance crisis. Organizations cannot govern what they cannot see, and existing compliance…

It's widely expected that humanity will someday create AI systems vastly more intelligent than us, leading to the unsolved alignment problem of "how to control superintelligence." However, this commonly expressed problem is not only…

人工智能 · 计算机科学 2024-12-02 James M. Mazzu

Trust is one of the most important factors shaping whether and how people adopt and rely on artificial intelligence (AI). Yet most existing studies measure trust in terms of functionality, focusing on whether a system is reliable, accurate,…

人机交互 · 计算机科学 2025-10-14 Haocan Sun , Weizi Liu , Di Wu , Guoming Yu , Mike Yao

Artificial Intelligence (AI) technology epitomizes the complex challenges posed by human-made artifacts, particularly those widely integrated into society and exerting significant influence, highlighting potential benefits and their…

人工智能 · 计算机科学 2025-10-06 Michael Papademas , Xenia Ziouvelou , Antonis Troumpoukis , Vangelis Karkaletsis

Explainable artificial intelligence has received limited attention in construction despite its growing importance in various other industrial sectors. In this paper, we provide a narrative review of XAI to raise awareness about its…

人工智能 · 计算机科学 2025-09-09 Peter ED Love , Weili Fang , Jane Matthews , Stuart Porter , Hanbin Luo , Lieyun Ding

Displaying confidence scores in human-AI interaction has been shown to help build trust between humans and AI systems. However, most existing research uses only the confidence score as a form of communication. As confidence scores are just…

人工智能 · 计算机科学 2023-03-13 Thao Le , Tim Miller , Ronal Singh , Liz Sonenberg

Research into explainable artificial intelligence (XAI) for data analysis tasks suffer from a large number of contradictions and lack of concrete design recommendations stemming from gaps in understanding the tasks that require AI…

人工智能 · 计算机科学 2025-07-15 Hamzah Ziadeh , Hendrik Knoche

Why do explainable AI (XAI) explanations in radiology, despite their promise of transparency, still fail to gain human trust? Current XAI approaches provide justification for predictions, however, these do not meet practitioners' needs.…

人机交互 · 计算机科学 2023-04-10 Robert Kaufman , David Kirsh

Explainable Artificial Intelligence (XAI) plays a critical role in fostering user trust and understanding in AI-driven systems. However, the design of effective XAI interfaces presents significant challenges, particularly for UX…

人机交互 · 计算机科学 2025-06-23 Mohammad Naiseh , Huseyin Dogan , Stephen Giff , Nan Jiang