English
Related papers

Related papers: Why Locking Up Cartel Members Does not Work

200 papers

The standard iterated prisoner's dilemma is an unrealistic model of social behaviour because it forces individuals to participate in the interaction. We analyse a model in which players have the option of ending their association. If the…

Optimization and Control · Mathematics 2007-05-23 L. A. Khodarinova , J. N. Webb

This paper considers a key agreement problem in which two parties aim to agree on a key by exchanging messages in the presence of adversarial tampering. The aim of the adversary is to disrupt the key agreement process, but there are no…

Information Theory · Computer Science 2009-01-30 Terence Chan , Ning Cai , Alex Grant

Can randomness be better than scheduled practices, for securing an event at a large venue such as a stadium or entertainment arena? Perhaps surprisingly, from several perspectives the answer is "yes." This note examines findings from an…

Computers and Society · Computer Science 2025-04-01 Paul B. Kantor , Fred S. Roberts

Two types of interventions are commonly implemented in networks: characteristic intervention, which influences individuals' intrinsic incentives, and structural intervention, which targets the social links among individuals. In this paper…

Theoretical Economics · Economics 2021-02-16 Yang Sun , Wei Zhao , Junjie Zhou

Ransomware is a growing threat to individuals and enterprises alike, constituting a major factor in cyber insurance and in the security planning of every organization. Although the game theoretic lens often frames the game as a competition…

Cryptography and Security · Computer Science 2022-06-28 Erick Galinkin

Recent years have seen an explosion of interest in autonomous cyber defence agents trained to defend computer networks using deep reinforcement learning. These agents are typically trained in cyber gym environments using dense, highly…

Machine Learning · Computer Science 2026-02-13 Elizabeth Bates , Chris Hicks , Vasilios Mavroudis

Previous studies show that information security breaches and privacy violations are important issues for organisations and people. It is acknowledged that decreasing the risk in this domain requires consideration of the technological…

Computers and Society · Computer Science 2019-03-29 Nader Sohrabi Safa , Carsten Maple , Steve Furnell , Muhammad Ajmal Azad , Charith Perera , Mohammad Dabbagh , Mehdi Sookhak

We study a security threat to batch reinforcement learning and control where the attacker aims to poison the learned policy. The victim is a reinforcement learner / controller which first estimates the dynamics and the rewards from a batch…

Machine Learning · Computer Science 2019-11-01 Yuzhe Ma , Xuezhou Zhang , Wen Sun , Xiaojin Zhu

Recently, bandit optimization has received significant attention in real-world safety-critical systems that involve repeated interactions with humans. While there exist various algorithms with performance guarantees in the literature,…

Machine Learning · Computer Science 2023-11-13 Amirhossein Afsharrad , Ahmadreza Moradipari , Sanjay Lall

We study the transition towards effective payoffs in the prisoner's dilemma game on scale-free networks by introducing a normalization parameter guiding the system from accumulated payoffs to payoffs normalized with the connectivity of each…

Biological Physics · Physics 2008-03-29 Attila Szolnoki , Matjaz Perc , Zsuzsa Danku

The increasing adoption of Reinforcement Learning in safety-critical systems domains such as autonomous vehicles, health, and aviation raises the need for ensuring their safety. Existing safety mechanisms such as adversarial training,…

Machine Learning · Computer Science 2021-11-11 Paulina Stevia Nouwou Mindom , Amin Nikanjam , Foutse Khomh , John Mullins

We generalize the setting of online clustering of bandits by allowing non-uniform distribution over user frequencies. A more efficient algorithm is proposed with simple set structures to represent clusters. We prove a regret bound for the…

Machine Learning · Computer Science 2019-07-03 Shuai Li , Wei Chen , Shuai Li , Kwong-Sak Leung

We consider the stochastic contextual bandit problem with additional regularization. The motivation comes from problems where the policy of the agent must be close to some baseline policy which is known to perform well on the task. To…

Machine Learning · Statistics 2019-06-06 Xavier Fontaine , Quentin Berthet , Vianney Perchet

Collective violence in direct confrontations between two opposing groups happens in short bursts wherein small subgroups briefly attack small numbers of opponents, while the others form a non-fighting audience. The mechanism is fighters'…

Physics and Society · Physics 2018-02-05 Jeroen Bruggeman

Room allocation is a challenging task in detention centers since lots of related people need to be held separately with limited rooms. It is extremely difficult and risky to allocate rooms manually, especially for organized crime groups…

Social and Information Networks · Computer Science 2021-07-19 Jingwei Wang , Chuan Liu , Yukai Zhao , Yunlong Ma , Min Liu , Weiming Shen

We consider the problem of reinforcement learning under safety requirements, in which an agent is trained to complete a given task, typically formalized as the maximization of a reward signal over time, while concurrently avoiding…

Machine Learning · Computer Science 2018-09-25 Tu-Hoa Pham , Giovanni De Magistris , Don Joven Agravante , Subhajit Chaudhury , Asim Munawar , Ryuki Tachibana

Motivated by tensions between data privacy for individual citizens, and societal priorities such as counterterrorism and the containment of infectious disease, we introduce a computational model that distinguishes between parties for whom…

Data Structures and Algorithms · Computer Science 2015-06-02 Michael Kearns , Aaron Roth , Zhiwei Steven Wu , Grigory Yaroslavtsev

In a backdoor attack, an attacker injects corrupted examples into the training set. The goal of the attacker is to cause the final trained model to predict the attacker's desired target label when a predefined trigger is added to test…

Machine Learning · Computer Science 2022-10-13 Jonathan Hayase , Sewoong Oh

We consider a new class of max flow network interdiction problems, where the defender is able to introduce new arcs to the network after the attacker has made their interdiction decisions. We prove properties of when this restructuring will…

Optimization and Control · Mathematics 2022-12-05 Daniel Kosmas , Thomas C. Sharkey , John E. Mitchell , Kayse Lee Maass , Lauren Martin

The deadly triad refers to the instability of a reinforcement learning algorithm when it employs off-policy learning, function approximation, and bootstrapping simultaneously. In this paper, we investigate the target network as a tool for…

Machine Learning · Computer Science 2023-10-02 Shangtong Zhang , Hengshuai Yao , Shimon Whiteson