English
Related papers

Related papers: Strategic commitments shape collective cybersecuri…

200 papers

Existing evaluations of AI misuse safeguards provide a patchwork of evidence that is often difficult to connect to real-world decisions. To bridge this gap, we describe an end-to-end argument (a "safety case") that misuse safeguards reduce…

Machine Learning · Computer Science 2025-05-26 Joshua Clymer , Jonah Weinbaum , Robert Kirk , Kimberly Mai , Selena Zhang , Xander Davies

Generative AI compresses within-task skill differences while shifting economic value toward concentrated complementary assets, creating an apparent paradox: the technology that equalizes individual performance may widen aggregate…

Machine Learning · Computer Science 2026-03-10 Xupeng Chen , Shuchen Meng

Adversarial artificial intelligence (AI) attacks pose a significant threat to autonomous transportation, such as maritime vessels, that rely on AI components. Malicious actors can exploit these systems to deceive and manipulate AI-driven…

Cryptography and Security · Computer Science 2025-05-29 Mathew J. Walter , Aaron Barrett , Kimberly Tam

This paper introduces a conversational interface system that enables participatory design of differentially private AI systems in public sector applications. Addressing the challenge of balancing mathematical privacy guarantees with…

Information Theory · Computer Science 2025-05-28 Wenjun Yang , Eyhab Al-Masri

AI alignment research aims to develop techniques to ensure that AI systems do not cause harm. However, every alignment technique has failure modes, which are conditions in which there is a non-negligible chance that the technique fails to…

Artificial Intelligence · Computer Science 2025-10-14 Leonard Dung , Florian Mai

The ability to exploit the opportunities offered by AI within UK Defence calls for an understanding of systemic issues required to achieve an effective operational capability. This paper provides the authors' views of issues which currently…

Artificial Intelligence · Computer Science 2018-10-01 Gavin Pearson , Phil Jolley , Geraint Evans

Cybersecurity planning supports the selection of and implementation of security controls in resource-constrained settings to manage risk. Doing so requires considering adaptive adversaries with different levels of strategic sophistication…

Optimization and Control · Mathematics 2023-02-07 Eric B. DuBois , Ashley Peper , Laura A. Albert

The positive impact of cooperative bots on cooperation within evolutionary game theory is well documented; however, existing studies have predominantly used discrete strategic frameworks, focusing on deterministic actions with a fixed…

Physics and Society · Physics 2024-06-24 Zehua Si , Zhixue He , Chen Shen , Jun Tanimoto

This paper examines the systemic risks posed by incremental advancements in artificial intelligence, developing the concept of `gradual disempowerment', in contrast to the abrupt takeover scenarios commonly discussed in AI safety. We…

Computers and Society · Computer Science 2025-01-30 Jan Kulveit , Raymond Douglas , Nora Ammann , Deger Turan , David Krueger , David Duvenaud

This work tackles a critical challenge in AI safety research under limited compute: given a fixed computation budget, how can one maximize the strength of iterative adversarial attacks? Coarsely reducing the number of attack iterations…

Machine Learning · Computer Science 2025-11-03 Zhichao Hou , Weizhi Gao , Xiaorui Liu

Over the past decade, the machine learning security community has developed a myriad of defenses for evasion attacks. An understudied question in that community is: for whom do these defenses defend? This work considers common approaches to…

Machine Learning · Computer Science 2023-08-24 Luke E. Richards , Edward Raff , Cynthia Matuszek

Undoubtedly, the evolution of Generative AI (GenAI) models has been the highlight of digital transformation in the year 2022. As the different GenAI models like ChatGPT and Google Bard continue to foster their complexity and capability,…

Cryptography and Security · Computer Science 2023-07-04 Maanak Gupta , CharanKumar Akiri , Kshitiz Aryal , Eli Parker , Lopamudra Praharaj

Uncertainty in artificial intelligence (AI) predictions poses urgent legal and ethical challenges for AI-assisted decision-making. We examine two algorithmic interventions that act as guardrails for human-AI collaboration: selective…

Computers and Society · Computer Science 2025-08-12 Holli Sargeant , Mackenzie Jorgensen , Arina Shah , Adrian Weller , Umang Bhatt

Successful deployment of artificial intelligence (AI) in various settings has led to numerous positive outcomes for individuals and society. However, AI systems have also been shown to harm parts of the population due to biased predictions.…

Computers and Society · Computer Science 2023-07-21 Ondrej Bohdal , Timothy Hospedales , Philip H. S. Torr , Fazl Barez

We aim to demonstrate the value of mathematical models for policy debates about technological progress in cybersecurity by considering phishing, vulnerability discovery, and the dynamics between patching and exploitation. We then adjust the…

Cryptography and Security · Computer Science 2022-07-29 Andrew J Lohn , Krystal Alex Jackson

A membership inference attack (MIA) against a machine-learning model enables an attacker to determine whether a given data record was part of the model's training data or not. In this paper, we provide an in-depth study of the phenomenon of…

Machine Learning · Computer Science 2021-09-20 Bogdan Kulynych , Mohammad Yaghini , Giovanni Cherubin , Michael Veale , Carmela Troncoso

Stealthy attacks are a major cyber-security threat. In practice, both attackers and defenders have resource constraints that could limit their capabilities. Hence, to develop robust defense strategies, a promising approach is to utilize…

Computer Science and Game Theory · Computer Science 2019-10-22 Ming Zhang , Zizhan Zheng , Ness B. Shroff

Artificial intelligence (AI) systems in high-stakes domains raise concerns about proxy discrimination, unfairness, and explainability. Existing audits often fail to reveal why unfairness arises, particularly when rooted in structural bias.…

Artificial Intelligence · Computer Science 2025-11-25 Belona Sonna , Alban Grastien

As the use of artificial intelligence (AI) in high-stakes decision-making increases, the ability to contest such decisions is being recognised in AI ethics guidelines as an important safeguard for individuals. Yet, there is little guidance…

Human-Computer Interaction · Computer Science 2021-02-23 Henrietta Lyons , Eduardo Velloso , Tim Miller

The rapid growth of distributed energy resources (DERs), such as renewable energy sources, generators, consumers, and prosumers in the smart grid infrastructure, poses significant cybersecurity and trust challenges to the grid controller.…

Cryptography and Security · Computer Science 2023-06-16 Md. Shirajum Munir , Sachin Shetty , Danda B. Rawat