English
Related papers

Related papers: Who Should Run Advanced AI Evaluations -- AISIs?

200 papers

Artificial intelligence now decides who receives a loan, who is flagged for criminal investigation, and whether an autonomous vehicle brakes in time. Governments have responded: the EU AI Act, the NIST Risk Management Framework, and the…

Artificial Intelligence · Computer Science 2026-04-24 Natan Levy , Gadi Perl

Recent advances in AI applications have raised growing concerns about the need for ethical guidelines and regulations to mitigate the risks posed by these technologies. In this paper, we present a mixed-methods survey study - combining…

Computers and Society · Computer Science 2025-12-16 Wilder Baldwin , Sepideh Ghanavati , Manuel Woersdoerfer

Rising concern for the societal implications of artificial intelligence systems has inspired a wave of academic and journalistic literature in which deployed systems are audited for harm by investigators from outside the organizations…

The most advanced future AI systems will first be deployed inside the frontier AI companies developing them. According to these companies and independent experts, AI systems may reach or even surpass human intelligence and capabilities by…

As artificial intelligence systems grow more capable and autonomous, frontier AI development poses potential systemic risks that could affect society at a massive scale. Current practices at many AI labs developing these systems lack…

Computers and Society · Computer Science 2025-06-03 Aidan Kierans , Kaley Rittichier , Utku Sonsayar , Avijit Ghosh

Industry actors in the United States have gained extensive influence in conversations about the regulation of general-purpose artificial intelligence (AI) systems. Although industry participation is an important part of the policy process,…

Computers and Society · Computer Science 2025-08-27 Kevin Wei , Carson Ezell , Nick Gabrieli , Chinmay Deshpande

Artificial Intelligence (AI) is one of the most discussed technologies today. There are many innovative applications such as the diagnosis and treatment of cancer, customer experience, new business, education, contagious diseases…

Computers and Society · Computer Science 2020-01-28 Richard Benjamins , Idoia Salazar

As AI systems advance and integrate into society, well-designed and transparent evaluations are becoming essential tools in AI governance, informing decisions by providing evidence about system capabilities and risks. Yet there remains a…

Artificial Intelligence (AI) solutions and technologies are being increasingly adopted in smart systems context, however, such technologies are continuously concerned with ethical uncertainties. Various guidelines, principles, and…

Computers and Society · Computer Science 2022-11-16 Arif Ali Khan , Muhammad Azeem Akbar , Mahdi Fahmideh , Peng Liang , Muhammad Waseem , Aakash Ahmad , Mahmood Niazi , Pekka Abrahamsson

The first International AI Safety Report comprehensively synthesizes the current evidence on the capabilities, risks, and safety of advanced AI systems. The report was mandated by the nations attending the AI Safety Summit in Bletchley, UK.…

Computers and Society · Computer Science 2025-01-30 Yoshua Bengio , Sören Mindermann , Daniel Privitera , Tamay Besiroglu , Rishi Bommasani , Stephen Casper , Yejin Choi , Philip Fox , Ben Garfinkel , Danielle Goldfarb , Hoda Heidari , Anson Ho , Sayash Kapoor , Leila Khalatbari , Shayne Longpre , Sam Manning , Vasilios Mavroudis , Mantas Mazeika , Julian Michael , Jessica Newman , Kwan Yee Ng , Chinasa T. Okolo , Deborah Raji , Girish Sastry , Elizabeth Seger , Theodora Skeadas , Tobin South , Emma Strubell , Florian Tramèr , Lucia Velasco , Nicole Wheeler , Daron Acemoglu , Olubayo Adekanmbi , David Dalrymple , Thomas G. Dietterich , Edward W. Felten , Pascale Fung , Pierre-Olivier Gourinchas , Fredrik Heintz , Geoffrey Hinton , Nick Jennings , Andreas Krause , Susan Leavy , Percy Liang , Teresa Ludermir , Vidushi Marda , Helen Margetts , John McDermid , Jane Munga , Arvind Narayanan , Alondra Nelson , Clara Neppel , Alice Oh , Gopal Ramchurn , Stuart Russell , Marietje Schaake , Bernhard Schölkopf , Dawn Song , Alvaro Soto , Lee Tiedrich , Gaël Varoquaux , Andrew Yao , Ya-Qin Zhang , Fahad Albalawi , Marwan Alserkal , Olubunmi Ajala , Guillaume Avrin , Christian Busch , André Carlos Ponce de Leon Ferreira de Carvalho , Bronwyn Fox , Amandeep Singh Gill , Ahmet Halit Hatip , Juha Heikkilä , Gill Jolly , Ziv Katzir , Hiroaki Kitano , Antonio Krüger , Chris Johnson , Saif M. Khan , Kyoung Mu Lee , Dominic Vincent Ligot , Oleksii Molchanovskyi , Andrea Monti , Nusu Mwamanzi , Mona Nemer , Nuria Oliver , José Ramón López Portillo , Balaraman Ravindran , Raquel Pezoa Rivera , Hammam Riza , Crystal Rugege , Ciarán Seoighe , Jerry Sheehan , Haroon Sheikh , Denise Wong , Yi Zeng

Generative AI is rapidly moving from research to deployment, elevating the need for responsible development, evaluation, and governance. We conduct a PRISMA guided review of 232 studies (November 2022 - December 2025), spanning large…

Prominent AI companies are producing 'safety frameworks' as a type of voluntary self-governance. These statements purport to establish risk thresholds and safety procedures for the development and deployment of highly capable AI.…

Computers and Society · Computer Science 2025-10-14 Sam Coggins , Alexander K. Saeri , Katherine A. Daniell , Lorenn P. Ruster , Jessie Liu , Jenny L. Davis

Public sector use of AI has been quietly on the rise for the past decade, but only recently have efforts to regulate it entered the cultural zeitgeist. While simple to articulate, promoting ethical and effective roll outs of AI systems in…

Computers and Society · Computer Science 2024-04-24 Tom Zick , Mason Kortz , David Eaves , Finale Doshi-Velez

This paper examines the responsible integration of artificial intelligence (AI) in human services organizations (HSOs), proposing a nuanced framework for evaluating AI applications across multiple dimensions of risk. The authors argue that…

Computers and Society · Computer Science 2025-01-22 Brian E. Perron , Lauri Goldkind , Zia Qi , Bryan G. Victor

Governance efforts for artificial intelligence (AI) are taking on increasingly more concrete forms, drawing on a variety of approaches and instruments from hard regulation to standardisation efforts, aimed at mitigating challenges from…

Computers and Society · Computer Science 2021-10-19 Charlotte Stix

The advent of advanced AI underscores the urgent need for comprehensive safety evaluations, necessitating collaboration across communities (i.e., AI, software engineering, and governance). However, divergent practices and terminologies…

Software Engineering · Computer Science 2024-05-17 Boming Xia , Qinghua Lu , Liming Zhu , Zhenchang Xing

In this paper, we develop the position that current frameworks for evaluating emotional intelligence (EI) in artificial intelligence (AI) systems need refinement because they do not adequately or comprehensively measure the various aspects…

Artificial Intelligence · Computer Science 2025-12-30 Max Parks , Kheli Atluru , Meera Vinod , Mike Kuniavsky , Jud Brewer , Sean White , Sarah Adler , Wendy Ju

Although artificial intelligence (AI) is solving real-world challenges and transforming industries, there are serious concerns about its ability to behave and make decisions in a responsible way. Many AI ethics principles and guidelines for…

Artificial Intelligence · Computer Science 2022-07-22 Qinghua Lu , Liming Zhu , Xiwei Xu , Jon Whittle , David Douglas , Conrad Sanderson

The rapid emergence of large language models (LLMs) has raised urgent questions across the modern workforce about this new technology's strengths, weaknesses, and capabilities. For privacy professionals, the question is whether these AI…

Computers and Society · Computer Science 2025-08-13 Zane Witherspoon , Thet Mon Aye , YingYing Hao

As artificial intelligence (AI) technologies continue to advance, effective risk assessment, regulation, and oversight are necessary to ensure that AI development and deployment align with ethical principles while preserving innovation and…

Computers and Society · Computer Science 2026-03-25 Georgios Pavlidis