English
Related papers

Related papers: Towards Frontier Safety Policies Plus

200 papers

Rapidly advancing artificial intelligence (AI) systems introduce novel, uncertain, and potentially catastrophic risks. Managing these risks requires a mature risk-management infrastructure whose cornerstone is rigorous risk modeling. We…

AI has become integral to safety-critical areas like autonomous driving systems (ADS) and robotics. The architecture of recent autonomous systems are trending toward end-to-end (E2E) monolithic architectures such as large language models…

Artificial Intelligence · Computer Science 2025-07-24 Mandar Pitale , Jelena Frtunikj , Abhinaw Priyadershi , Vasu Singh , Maria Spence

As Artificial Intelligence (AI) becomes increasingly integrated into our lives, the need for new norms is urgent. However, AI evolves at a much faster pace than the characteristic time of norm formation, posing an unprecedented challenge to…

Physics and Society · Physics 2024-07-01 Andrea Baronchelli

In financial applications, regulations or best practices often lead to specific requirements in machine learning relating to four key pillars: fairness, privacy, interpretability and greenhouse gas emissions. These all sit in the broader…

Machine Learning · Computer Science 2024-07-18 Roberto Pagliari , Peter Hill , Po-Yu Chen , Maciej Dabrowny , Tingsheng Tan , Francois Buet-Golfouse

Safety and responsibility evaluations of advanced AI models are a critical but developing field of research and practice. In the development of Google DeepMind's advanced AI models, we innovated on and applied a broad set of approaches to…

International agreements about AI development may be required to reduce catastrophic risks from advanced AI systems. However, agreements about such a high-stakes technology must be backed by verification mechanisms--processes or tools that…

Computers and Society · Computer Science 2025-06-23 Aaron Scher , Lisa Thiergart

As artificial intelligence (AI) reshapes industries and societies, ensuring its trustworthiness-through mitigating ethical risks like bias, opacity, and accountability deficits-remains a global challenge. International Organization for…

Computers and Society · Computer Science 2025-04-24 Sridharan Sankaran

Safety frameworks represent a significant development in AI governance: they are the first type of publicly shared catastrophic risk management framework developed by major AI companies and focus specifically on AI scaling decisions. I…

Computers and Society · Computer Science 2024-10-02 Atoosa Kasirzadeh

As artificial intelligence scales, the concepts of alignment, agency, and autonomy have become central to AI safety, governance, and control. However, even in human contexts, these terms lack universal definitions, varying across…

Computers and Society · Computer Science 2025-03-11 Krti Tallam

With the increasing integration of frontier large language models (LLMs) into society and the economy, decisions related to their training, deployment, and use have far-reaching implications. These decisions should not be left solely in the…

Recent advances in machine learning, particularly the emergence of foundation models, are leading to new opportunities to develop technology-based solutions to societal problems. However, the reasoning and inner workings of today's complex…

Computers and Society · Computer Science 2025-07-01 Rajeev Alur , Greg Durrett , Hadas Kress-Gazit , Corina Păsăreanu , René Vidal

As Artificial Intelligence (AI) continues to advance rapidly, Friendly AI (FAI) has been proposed to advocate for more equitable and fair development of AI. Despite its importance, there is a lack of comprehensive reviews examining FAI from…

Artificial Intelligence · Computer Science 2024-12-20 Qiyang Sun , Yupei Li , Emran Alturki , Sunil Munthumoduku Krishna Murthy , Björn W. Schuller

This community paper developed out of the NSF Workshop on the Future of Artificial Intelligence (AI) and the Mathematical and Physics Sciences (MPS), which was held in March 2025 with the goal of understanding how the MPS domains…

Artificial Intelligence · Computer Science 2026-03-17 Andrew Ferguson , Marisa LaFleur , Lars Ruthotto , Jesse Thaler , Yuan-Sen Ting , Pratyush Tiwary , Soledad Villar , E. Paulo Alves , Jeremy Avigad , Simon Billinge , Camille Bilodeau , Keith Brown , Emmanuel Candes , Arghya Chattopadhyay , Bingqing Cheng , Jonathan Clausen , Connor Coley , Andrew Connolly , Fred Daum , Sijia Dong , Chrisy Xiyu Du , Cora Dvorkin , Cristiano Fanelli , Eric B. Ford , Luis Manuel Frutos , Nicolás García Trillos , Cecilia Garraffo , Robert Ghrist , Rafael Gomez-Bombarelli , Gianluca Guadagni , Sreelekha Guggilam , Sergei Gukov , Juan B. Gutiérrez , Salman Habib , Johannes Hachmann , Boris Hanin , Philip Harris , Murray Holland , Elizabeth Holm , Hsin-Yuan Huang , Shih-Chieh Hsu , Nick Jackson , Olexandr Isayev , Heng Ji , Aggelos Katsaggelos , Jeremy Kepner , Yannis Kevrekidis , Michelle Kuchera , J. Nathan Kutz , Branislava Lalic , Ann Lee , Matt LeBlanc , Josiah Lim , Rebecca Lindsey , Yongmin Liu , Peter Y. Lu , Sudhir Malik , Vuk Mandic , Vidya Manian , Emeka P. Mazi , Pankaj Mehta , Peter Melchior , Brice Ménard , Jennifer Ngadiuba , Stella Offner , Elsa Olivetti , Shyue Ping Ong , Christopher Rackauckas , Philippe Rigollet , Chad Risko , Philip Romero , Grant Rotskoff , Brett Savoie , Uros Seljak , David Shih , Gary Shiu , Dima Shlyakhtenko , Eva Silverstein , Taylor Sparks , Thomas Strohmer , Christopher Stubbs , Stephen Thomas , Suriyanarayanan Vaikuntanathan , Rene Vidal , Francisco Villaescusa-Navarro , Gregory Voth , Benjamin Wandelt , Rachel Ward , Melanie Weber , Risa Wechsler , Stephen Whitelam , Olaf Wiest , Mike Williams , Zhuoran Yang , Yaroslava G. Yingling , Bin Yu , Shuwen Yue , Ann Zabludoff , Huimin Zhao , Tong Zhang

As frontier artificial intelligence (AI) models rapidly advance, benchmarks are integral to comparing different models and measuring their progress in different task-specific domains. However, there is a lack of guidance on when and how…

Computers and Society · Computer Science 2025-07-10 Ayrton San Joaquin , Rokas Gipiškis , Leon Staufer , Ariel Gil

Public attention towards explainability of artificial intelligence (AI) systems has been rising in recent years to offer methodologies for human oversight. This has translated into the proliferation of research outputs, such as from…

Computers and Society · Computer Science 2023-04-25 Luca Nannini , Agathe Balayn , Adam Leon Smith

Artificial intelligence (AI) is emerging as a foundational general-purpose technology, raising new dilemmas of sovereignty in an interconnected world. While governments seek greater control over it, the very foundations of AI--global data…

Computers and Society · Computer Science 2025-11-21 Shalabh Kumar Singh , Shubhashis Sengupta

Recent advancements in large language models (LLMs) and AI systems have led to a paradigm shift in the design and optimization of complex AI workflows. By integrating multiple components, compound AI systems have become increasingly adept…

Computation and Language · Computer Science 2025-10-08 Yu-Ang Lee , Guan-Ting Yi , Mei-Yi Liu , Jui-Chao Lu , Guan-Bo Yang , Yun-Nung Chen

The rise of Generative AI (GenAI) brings about transformative potential across sectors, but its dual-use nature also amplifies risks. Governments globally are grappling with the challenge of regulating GenAI, balancing innovation against…

We outline the principles of classical assurance for computer-based systems that pose significant risks. We then consider application of these principles to systems that employ Artificial Intelligence (AI) and Machine Learning (ML). A key…

Artificial Intelligence · Computer Science 2025-06-04 Robin Bloomfield , John Rushby

The widespread adoption of Artificial Intelligence (AI) technologies in the public and private sectors has resulted in them significantly impacting the lives of people in new and unexpected ways. In this context, it becomes important to…

Computers and Society · Computer Science 2024-07-19 Ambreesh Parthasarathy , Aditya Phalnikar , Ameen Jauhar , Dhruv Somayajula , Gokul S Krishnan , Balaraman Ravindran
‹ Prev 1 8 9 10 Next ›