中文
相关论文

相关论文: RISED: A Pre-Deployment Safety Evaluation Framewor…

200 篇论文

Semi-supervised learning has become a dominant paradigm for reducing annotation costs. However, we argue that the current progress is clouded by a twofold overconfidence problem. Algorithmically, mainstream pseudo-labeling frameworks often…

计算机视觉与模式识别 · 计算机科学 2026-05-26 Jun Li , Ziwei Qin

Risk assessments for advanced AI systems require evaluating both the models themselves and their deployment contexts. We introduce the Societal Capacity Assessment Framework (SCAF), an indicators-based approach to measuring a society's…

计算机与社会 · 计算机科学 2025-09-30 Milan Gandhi , Peter Cihon , Owen Larter , Rebecca Anselmetti

Deep learning models achieve strong performance in chest radiograph (CXR) interpretation, yet fairness and reliability concerns persist. Models often show uneven accuracy across patient subgroups, leading to hidden failures not reflected in…

计算机视觉与模式识别 · 计算机科学 2025-10-03 Han-Jay Shu , Wei-Ning Chiu , Shun-Ting Chang , Meng-Ping Huang , Takeshi Tohyama , Ahram Han , Po-Chih Kuo

Scene understanding is a vital part of autonomous driving systems, which requires the use of deep learning models. Deep learning methods are intrinsically black box models, which lack transparency and safety in autonomous driving. To make…

计算机视觉与模式识别 · 计算机科学 2026-05-07 Maryam Sadat Hosseini Azad , Shahriar Baradaran Shokouhi

Accurate and interpretable survival analysis remains a core challenge in oncology. With growing multimodal data and the clinical need for transparent models to support validation and trust, this challenge increases in complexity. We propose…

人工智能 · 计算机科学 2025-09-29 Mafalda Malafaia , Peter A. N. Bosman , Coen Rasch , Tanja Alderliesten

Enterprise AI systems, built on large language models, retrieval pipelines and autonomous agents, introduce a class of risks that traditional software quality assurance was never designed to address. These systems are probabilistic,…

软件工程 · 计算机科学 2026-05-25 Chitra Badagi , Divye Singh , Animesh Sen , Adinath Shirsath

Machine learning models deployed in critical care settings exhibit demographic biases, particularly gender disparities, that undermine clinical trust and equitable treatment. This paper introduces FairMed-XGB, a novel framework that…

机器学习 · 计算机科学 2026-03-17 Mitul Goswami , Romit Chatterjee , Arif Ahmed Sekh

Machine learning models for chronic kidney disease (CKD) risk prediction often post strong discrimination scores on internal test sets. Calibration and uncertainty quantification get far less attention, leaving clinicians without reliable…

机器学习 · 计算机科学 2026-05-22 Michael O. Eniolade

Artificial intelligence (AI) holds great promise for supporting clinical trials, from patient recruitment and endpoint assessment to treatment response prediction. However, deploying AI without safeguards poses significant risks,…

机器学习 · 计算机科学 2025-10-09 Yao Chen , David Ohlssen , Aimee Readie , Gregory Ligozio , Ruvie Martin , Thibaud Coroller

BACKGROUND: General aviation fleet expansion demands intelligent health monitoring under computational constraints. Real-world aircraft health diagnosis requires balancing accuracy with computational constraints under extreme class…

机器学习 · 计算机科学 2026-04-23 Xinhang Chen , Zhihuan Wei , Yang Hu , Zhiguo Zeng , Kang Zeng , Wei Wang

Machine learning (ML) algorithms are increasingly deployed in high-stakes decision-making domains such as loan approvals, hiring, and recidivism predictions. While existing fairness metrics (e.g., statistical parity, equal opportunity)…

机器学习 · 计算机科学 2026-05-18 Gideon Popoola , John Sheppard

Fairness evaluation in face analysis systems (FAS) typically depends on automatic demographic attribute inference (DAI), which itself relies on predefined demographic segmentation. However, the validity of fairness auditing hinges on the…

计算机视觉与模式识别 · 计算机科学 2025-10-24 Alexandre Fournier-Montgieux , Hervé Le Borgne , Adrian Popescu , Bertrand Luvison

For safety, medical AI systems undergo thorough evaluations before deployment, validating their predictions against a ground truth which is assumed to be fixed and certain. However, this ground truth is often curated in the form of…

Explainable Artificial Intelligence (XAI) is essential for the transparency and clinical adoption of Clinical Decision Support Systems (CDSS). However, the real-world effectiveness of existing XAI methods remains limited and is…

机器学习 · 计算机科学 2026-01-26 Alessandro Gambetti , Qiwei Han , Hong Shen , Claudia Soares

Current large language models (LLMs) excel in verifiable domains where outputs can be checked before action but prove less reliable for high-stakes strategic decisions with uncertain outcomes. This gap, driven by mutually reinforcing…

人工智能 · 计算机科学 2025-11-12 Alejandro R. Jadad

AI-generated text detectors have recently gained adoption in educational and professional contexts. Prior research has uncovered isolated cases of bias, particularly against English Language Learners (ELLs) however, there is a lack of…

人工智能 · 计算机科学 2025-12-15 Priyam Basu , Yunfeng Zhang , Vipul Raheja

The European Union's Artificial Intelligence Act establishes comprehensive requirements for high-risk AI systems, yet the harmonized standards necessary for demonstrating compliance remain not fully developed. In this paper, we investigate…

计算机与社会 · 计算机科学 2026-01-14 Gregor Autischer , Kerstin Waxnegger , Dominik Kowald

Machine learning approaches often require training and evaluation datasets with a clear separation between positive and negative examples. This risks simplifying and even obscuring the inherent subjectivity present in many tasks. Preserving…

There are limitations of traditional methods and deep learning methods in terms of interpretability, generalization, and quantification of uncertainty in industrial fault diagnosis, and there are core problems of insufficient credibility in…

系统与控制 · 电气工程与系统科学 2025-10-07 Yue wu

Artificial intelligence (AI) is increasingly embedded in NHS workflows, but its probabilistic and adaptive behaviour conflicts with the deterministic assumptions underpinning existing clinical-safety standards. DCB0129 and DCB0160 provide…

计算机与社会 · 计算机科学 2025-11-19 Robert Gigiu