Computers and Society · Computer Science
How Should AI Safety Benchmarks Benchmark Safety?
Cheng Yu, Severin Engelmann, Ruoxuan Cao, Dalia Ali +1
2026-02-10
Machine Learning · Computer Science
Safetywashing: Do AI Safety Benchmarks Actually Measure Safety Progress?
Richard Ren, Steven Basart, Adam Khoja, Alice Gatti +8
2024-12-30
Computers and Society · Computer Science
Evaluating AI Providers' Frontier Safety Frameworks
Lily Stelling, Malcolm Murray, Bruno Galizzi, Max Schaffelder +2
2026-05-01
Artificial Intelligence · Computer Science
Can We Trust AI Benchmarks? An Interdisciplinary Review of Current Issues in AI Evaluation
Maria Eriksson, Erasmo Purificato, Arman Noroozian, Joao Vinagre +3
2025-05-27
Computers and Society · Computer Science
Catastrophic Liability: Managing Systemic Risks in Frontier AI Development
Aidan Kierans, Kaley Rittichier, Utku Sonsayar, Avijit Ghosh
2025-06-03
Artificial Intelligence · Computer Science
A Frontier AI Risk Management Framework: Bridging the Gap Between Current AI Practices and Established Risk Management
Simeon Campos, Henry Papadatos, Fabien Roger, Chloé Touzet +2
2025-02-20
Artificial Intelligence · Computer Science
ForesightSafety Bench: A Frontier Risk Evaluation and Governance Framework towards Safe AI
Haibo Tong, Feifei Zhao, Linghao Feng, Ruoyu Wu +17
2026-03-02
Computers and Society · Computer Science
Interoperability in AI Safety Governance: Ethics, Regulations, and Standards
Yik Chan Chin, David A. Raho, Hag-Min Kim, Chunli Bi +3
2026-01-13
Computers and Society · Computer Science
A Methodology for Quantitative AI Risk Modeling
Malcolm Murray, Steve Barrett, Henry Papadatos, Otter Quarks +4
2025-12-12
Computers and Society · Computer Science
Frontier AI Regulation: Managing Emerging Risks to Public Safety
Markus Anderljung, Joslyn Barnhart, Anton Korinek, Jade Leung +20
2023-11-09
Cryptography and Security · Computer Science
Never Compromise to Vulnerabilities: A Comprehensive Survey on AI Governance
Yuchu Jiang, Jian Zhao, Yuchen Yuan, Tianle Zhang +62
2025-08-19
Artificial Intelligence · Computer Science
Holistic Safety and Responsibility Evaluations of Advanced AI Models
Laura Weidinger, Joslyn Barnhart, Jenny Brennan, Christina Butterfield +15
2024-04-23
Human-Computer Interaction · Computer Science
Measuring What Matters: Connecting AI Ethics Evaluations to System Attributes, Hazards, and Harms
Shalaleh Rismani, Renee Shelby, Leah Davis, Negar Rostamzadeh +1
2025-10-14
Computers and Society · Computer Science
The Role of Risk Modeling in Advanced AI Risk Management
Chloé Touzet, Henry Papadatos, Malcolm Murray, Otter Quarks +5
2025-12-10