English
Related papers

Related papers: International AI Safety Report

200 papers

The International AI Safety Report 2026 synthesises the current scientific evidence on the capabilities, emerging risks, and safety of general-purpose AI systems. The report series was mandated by the nations attending the AI Safety Summit…

Computers and Society · Computer Science 2026-02-25 Yoshua Bengio , Stephen Clare , Carina Prunkl , Maksym Andriushchenko , Ben Bucknall , Malcolm Murray , Rishi Bommasani , Stephen Casper , Tom Davidson , Raymond Douglas , David Duvenaud , Philip Fox , Usman Gohar , Rose Hadshar , Anson Ho , Tiancheng Hu , Cameron Jones , Sayash Kapoor , Atoosa Kasirzadeh , Sam Manning , Nestor Maslej , Vasilios Mavroudis , Conor McGlynn , Richard Moulange , Jessica Newman , Kwan Yee Ng , Patricia Paskov , Shalaleh Rismani , Girish Sastry , Elizabeth Seger , Scott Singer , Charlotte Stix , Lucia Velasco , Nicole Wheeler , Daron Acemoglu , Vincent Conitzer , Thomas G. Dietterich , Fredrik Heintz , Geoffrey Hinton , Nick Jennings , Susan Leavy , Teresa Ludermir , Vidushi Marda , Helen Margetts , John McDermid , Jane Munga , Arvind Narayanan , Alondra Nelson , Clara Neppel , Sarvapali D. Ramchurn , Stuart Russell , Marietje Schaake , Bernhard Schölkopf , Alvaro Soto , Lee Tiedrich , Gaël Varoquaux , Andrew Yao , Ya-Qin Zhang , Leandro Angelo Aguirre , Olubunmi Ajala , Fahad Albalawi , Noora AlMalek , Christian Busch , Jonathan Collas , André Carlos Ponce de Leon Ferreira de Carvalho , Amandeep Gill , Ahmet Halit Hatip , Juha Heikkilä , Chris Johnson , Gill Jolly , Ziv Katzir , Mary N. Kerema , Hiroaki Kitano , Antonio Krüger , Kyoung Mu Lee , José Ramón López Portillo , Aoife McLysaght , Oleksii Molchanovskyi , Andrea Monti , Mona Nemer , Nuria Oliver , Raquel Pezoa , Audrey Plonk , Balaraman Ravindran , Hammam Riza , Crystal Rugege , Haroon Sheikh , Denise Wong , Yi Zeng , Liming Zhu , Daniel Privitera , Sören Mindermann

Safety has become the central value around which dominant AI governance efforts are being shaped. Recently, this culminated in the publication of the International AI Safety Report, written by 96 experts of which 30 nominated by the…

Computers and Society · Computer Science 2025-03-10 Roel Dobbe

Since the publication of the first International AI Safety Report, AI capabilities have continued to improve across key domains. New training techniques that teach AI systems to reason step-by-step and inference-time enhancements have…

Rapidly improving AI capabilities and autonomy hold significant promise of transformation, but are also driving vigorous debate on how to ensure that AI is safe, i.e., trustworthy, reliable, and secure. Building a trusted ecosystem is…

Our survey of 53 specialists across 105 AI reliability and security research areas identifies the most promising research prospects to guide strategic AI R&D investment. As companies are seeking to develop AI systems with broadly…

Computers and Society · Computer Science 2025-05-29 Joe O'Brien , Jeremy Dolan , Jay Kim , Jonah Dykhuizen , Jeba Sania , Sebastian Becker , Jam Kraprayoon , Cara Labrador

This paper maps Africa's distinctive AI risk profile, from deepfake fuelled electoral interference and data colonial dependency to compute scarcity, labour disruption and disproportionate exposure to climate driven environmental costs.…

AI safety benchmarks are pivotal for safety in advanced AI systems; however, they have significant technical, epistemic, and sociotechnical shortcomings. We present a review of 210 safety benchmarks that maps out common challenges in safety…

Computers and Society · Computer Science 2026-02-10 Cheng Yu , Severin Engelmann , Ruoxuan Cao , Dalia Ali , Orestis Papakyriakopoulos

Artificial Intelligence (AI) Safety Institutes and governments worldwide are deciding whether they evaluate advanced AI themselves, support a private evaluation ecosystem or do both. Evaluation regimes have been established in a wide range…

Computers and Society · Computer Science 2025-08-06 Merlin Stein , Milan Gandhi , Theresa Kriecherbauer , Amin Oueslati , Robert Trager

This second update to the 2025 International AI Safety Report assesses new developments in general-purpose AI risk management over the past year. It examines how researchers, public institutions, and AI developers are approaching risk…

The malicious use or malfunction of advanced general-purpose AI (GPAI) poses risks that, according to leading experts, could lead to the 'marginalisation or extinction of humanity.' To address these risks, there are an increasing number of…

Computers and Society · Computer Science 2025-03-26 Rebecca Scholefield , Samuel Martin , Otto Barten

Rapid advancements in artificial intelligence (AI) have sparked growing concerns among experts, policymakers, and world leaders regarding the potential for increasingly advanced AI systems to pose catastrophic risks. Although numerous risks…

Computers and Society · Computer Science 2023-10-11 Dan Hendrycks , Mantas Mazeika , Thomas Woodside

Recent discussions and research in AI safety have increasingly emphasized the deep connection between AI safety and existential risk from advanced AI systems, suggesting that work on AI safety necessarily entails serious consideration of…

Computers and Society · Computer Science 2025-02-17 Balint Gyevnar , Atoosa Kasirzadeh

In this work, we present and analyze reported failures of artificially intelligent systems and extrapolate our analysis to future AIs. We suggest that both the frequency and the seriousness of future AI failures will steadily increase. AI…

Artificial Intelligence · Computer Science 2016-10-26 Roman V. Yampolskiy , M. S. Spellchecker

Governments, industry, and other actors involved in governing AI technologies around the world agree that, while AI offers tremendous promise to benefit the world, appropriate guardrails are required to mitigate risks. Global institutions,…

Computers and Society · Computer Science 2024-09-18 A. Leone De Castris , C. Thomas

We introduce a conceptual framework and provide considerations for the institutional design of AI incident reporting systems, i.e., processes for collecting information about safety- and rights-related events caused by general-purpose AI.…

Computers and Society · Computer Science 2026-04-15 Kevin Wei , Lennart Heim

The EU Artificial Intelligence (AI) Act directs businesses to assess their AI systems to ensure they are developed in a way that is human-centered and trustworthy. The rapid adoption of AI in the industry has outpaced ethical evaluation…

Computers and Society · Computer Science 2025-09-30 Louise McCormack , Diletta Huyskes , Dave Lewis , Malika Bendechache

The development of artificial general intelligence (AGI) is likely to be one of humanity's most consequential technological advancements. Leading AI labs and scientists have called for the global prioritization of AI safety citing…

Computers and Society · Computer Science 2025-02-24 Severin Field

A number of leading AI companies, including OpenAI, Google DeepMind, and Anthropic, have the stated goal of building artificial general intelligence (AGI) - AI systems that achieve or exceed human performance across a wide range of…

Computers and Society · Computer Science 2023-05-15 Jonas Schuett , Noemi Dreksler , Markus Anderljung , David McCaffary , Lennart Heim , Emma Bluemke , Ben Garfinkel

The increasing use of AI technologies has led to increasing AI incidents, posing risks and causing harm to individuals, organizations, and society. This study recognizes and addresses the lack of standardized protocols for reliably and…

Computers and Society · Computer Science 2025-01-28 Avinash Agarwal , Manisha J Nene
‹ Prev 1 2 3 10 Next ›